Sobre esta vaga de Senior Site Reliability Operations Engineer - Finance na Truelogic
About Truelogic
At Truelogic we are a leading provider of nearshore staff augmentation services headquartered in New York. For over two decades, we’ve been delivering top-tier technology solutions to companies of all sizes, from innovative startups to industry leaders, helping them achieve their digital transformation goals.
Our team of 600+ highly skilled tech professionals, based in Latin America, drives digital disruption by partnering with U.S. companies on their most impactful projects. Whether collaborating with Fortune 500 giants or scaling startups, we deliver results that make a difference.
By applying for this position, you’re taking the first step in joining a dynamic team that values your expertise and aspirations. We aim to align your skills with opportunities that foster exceptional career growth and success while contributing to transformative projects that shape the future.
Our Client
A leading Financial Services Company
Job Summary
The Site Reliability Operations (SRO) team ensures 24/7 stability of the internal IT infrastructure and mission-critical backend systems. This role is not DevOps-focused, but is crucial in monitoring, coordinating, and restoring operations during incidents, particularly in a high-stakes, regulated environment.
The role balances incident command, technical troubleshooting, project leadership, and communication with multiple internal and external stakeholders.
Responsibilities
Oversee multi-platform IT infrastructure health using AWS CloudWatch, New Relic, Nagios, and SumoLogic. Continuously refine alert thresholds to minimize noise and enable proactive remediation.
Serve as an escalation point for complex technical issues. Perform troubleshooting across Linux/UNIX, Windows, virtual servers, and virtual desktop environments.
Coordinate, automate, and execute code deployments using Jenkins, GitLab, or similar CI/CD tools, driving toward non-disruptive releases and zero-downtime updates.
Collaborate closely with Application Developers, 3rd-party vendors, and internal Incident Management.
Drive medium- to large-scale infrastructure projects, including migrations, cloud upgrades, and performance tuning.
Maintain SOPs in team knowledge bases, and oversee enterprise backup operations (CommVault, Veeam, AWS Backup).
Qualifications & Requirements
Experience: 3–5+ years in an Operations Center (SRO/NOC) or cloud infrastructure environment with hands-on experience in full-stack application deployments.
Proficiency in both Windows and UNIX/Linux administration, troubleshooting (scripting, grepping logs, analyzing performance metrics), and virtual server/desktop management.
Practical experience with AWS cloud services (Storage, VMs, Networking) and enterprise monitoring tools (AWS CloudWatch, New Relic, Nagios, SumoLogic).
Scripting or programming capability in PowerShell, Python, or bash to automate repetitive tasks and optimize run-time operations.
Practical experience with CI/CD platforms (Jenkins, GitLab), ITSM ticket platforms (ServiceNow, Jira), and backup solutions (CommVault, Veeam, AWS Backup).
Good verbal and written communication skills with experience serving as a bridge between technical teams, executive stakeholders, and external vendors.
Pluses
Advanced AWS Certifications
Background in or exposure to AI/ML tools for infrastructure monitoring and predictive analytics.
Previous experience in ITIL-aligned environments or enterprise Change/Incident Management frameworks.
Bachelor’s Degree in Computer Science, Information Technology, or a related field.
What We Offer
100% Remote Work: Enjoy the freedom to work from the location that helps you thrive. All it takes is a laptop and a reliable internet connection.
Highly Competitive USD Pay: Earn an excellent, market-leading compensation in USD, that goes beyond typical market offerings.
Paid Time Off: We value your well-being. Our paid time off policies ensure you have the chance to unwind and recharge when needed.
Work with Autonomy: Enjoy the freedom to manage your time as long as the work gets done. Focus on results, not the clock.
Work with Top American Companies: Grow your expertise working on innovative, high-impact projects with Industry-Leading U.S. Companies.
Why You’ll Like Working Here
A Culture That Values You: We prioritize well-being and work-life balance, offering engagement activities and fostering dynamic teams to ensure you thrive both personally and professionally.
Diverse, Global Network: Connect with over 600 professionals in 25+ countries, expand your network, and collaborate with a multicultural team from Latin America.
Team Up with Skilled Professionals: Join forces with senior talent. All of our team members are seasoned experts, ensuring you're working with the best in your field.
Apply now!