Senior Site Reliability Operations Engineer - Finance

Posted 2 Days Ago
Be an Early Applicant
Hiring Remotely in Mexico City, Cuauhtémoc, Mexico City, MEX
In-Office or Remote
Senior level
Information Technology • Software
The Role
Leads incident response and service restoration for mission-critical infrastructure in a regulated financial environment. Responsibilities include incident command, root-cause analysis, executive reporting, observability improvements, Linux and Windows troubleshooting, CI/CD deployments, infrastructure migrations, change-risk management, documentation, vendor coordination, and on-call support.
Summary Generated by Built In
About Truelogic

At Truelogic we are a leading provider of nearshore staff augmentation services headquartered in New York. For over two decades, we’ve been delivering top-tier technology solutions to companies of all sizes, from innovative startups to industry leaders, helping them achieve their digital transformation goals.

Our team of 600+ highly skilled tech professionals, based in Latin America, drives digital disruption by partnering with U.S. companies on their most impactful projects. Whether collaborating with Fortune 500 giants or scaling startups, we deliver results that make a difference.

By applying for this position, you’re taking the first step in joining a dynamic team that values your expertise and aspirations. We aim to align your skills with opportunities that foster exceptional career growth and success while contributing to transformative projects that shape the future.

Our Client

A leading Financial Services Company


Job Summary

The Site Reliability Operations (SRO) team ensures 24/7 stability of the internal IT infrastructure and mission-critical backend systems. Some of the skills are monitoring, coordinating, and restoring operations during incidents, particularly in a high-stakes, regulated environment. This is an Operations swing shift from 2 PM to 10:30 PM PST timezone.

The role balances incident command, technical troubleshooting, project leadership, and communication with multiple internal and external stakeholders.

Responsibilities
  • Oversee multi-platform IT infrastructure health using AWS CloudWatch, New Relic, Nagios, and SumoLogic. Continuously refine alert thresholds to minimize noise and enable proactive remediation.

  • Serve as an escalation point for complex technical issues. Perform troubleshooting across Linux/UNIX, Windows, virtual servers, and virtual desktop environments.

  • Coordinate, automate, and execute code deployments using Jenkins, GitLab, or similar CI/CD tools, driving toward non-disruptive releases and zero-downtime updates.

  • Collaborate closely with Application Developers, 3rd-party vendors, and internal Incident Management.

  • Drive medium- to large-scale infrastructure projects, including migrations, cloud upgrades, and performance tuning.

  • Maintain SOPs in team knowledge bases, and oversee enterprise backup operations (CommVault, Veeam, AWS Backup).

 
Qualifications & Requirements
  • 5+ years in an Operations Center (SRO/NOC) or cloud infrastructure environment with hands-on experience in full-stack application deployments.

  • Proficiency in both Windows and UNIX/Linux administration, troubleshooting (scripting, grepping logs, analyzing performance metrics), and virtual server/desktop management.

  • Practical experience with AWS cloud services (Storage, VMs, Networking) and enterprise monitoring tools (AWS CloudWatch, New Relic, Nagios, SumoLogic).

  • Scripting or programming capability in PowerShell, Python, or bash to automate repetitive tasks and optimize run-time operations.

  • Practical experience with CI/CD platforms (Jenkins, GitLab), ITSM ticket platforms (ServiceNow, Jira), and backup solutions (CommVault, Veeam, AWS Backup).

  • Good verbal and written communication skills with experience serving as a bridge between technical teams, executive stakeholders, and external vendors.

What We Offer
  • 100% Remote Work: Enjoy the freedom to work from the location that helps you thrive. All it takes is a laptop and a reliable internet connection.

  • Highly Competitive USD Pay: Earn an excellent, market-leading compensation in USD, that goes beyond typical market offerings.

  • Paid Time Off: We value your well-being. Our paid time off policies ensure you have the chance to unwind and recharge when needed.

  • Work with Autonomy: Enjoy the freedom to manage your time as long as the work gets done. Focus on results, not the clock.

  • Work with Top American Companies: Grow your expertise working on innovative, high-impact projects with Industry-Leading U.S. Companies.

Why You’ll Like Working Here
  • A Culture That Values You: We prioritize well-being and work-life balance, offering engagement activities and fostering dynamic teams to ensure you thrive both personally and professionally.

  • Diverse, Global Network: Connect with over 600 professionals in 25+ countries, expand your network, and collaborate with a multicultural team from Latin America.

  • Team Up with Skilled Professionals: Join forces with senior talent. All of our team members are seasoned experts, ensuring you're working with the best in your field.

Apply now!

Skills Required

  • 5+ years of experience troubleshooting Windows and Linux environments
  • Strong knowledge of monitoring tools including AWS CloudWatch, New Relic, Nagios, and Sumo Logic
  • Practical experience with Jenkins, GitLab, or similar CI/CD tools
  • Practical experience with CommVault, AWS Backup, or similar backup tools
  • Strong scripting skills in PowerShell, Python, or equivalent
  • Outstanding communication skills, including executive reporting and communication under pressure
  • Experience working in high-paced environments and on-call support models
  • Autonomous and proactive attitude with the ability to manage complex tasks independently
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, NY
266 Employees
Year Founded: 2003

What We Do

Truelogic Software is a Nearshore tech firm specializing in Staff Augmentation Services and Innovation Projects. We build innovative digital products by extending the engineering teams of US companies with our elite group of 500 Latin American highly experienced tech talent. Our SERVICES Staff Augmentation Services Dedicated Agile Teams Innovation Projects Our EXPERTISE Mobile & Web Development Data Engineering DevOps QA Automation & Testing UX/UI Designing Project Management Email us [email protected]

Similar Jobs

Samsara Logo Samsara

Consultant

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
México
4000 Employees

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Temporary Associate

eCommerce • Fashion • Retail • Sales • Wearables • Design
Remote or Hybrid
14 Locations
16000 Employees
15-20 Hourly

GitLab Logo GitLab

Recruiter

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
7 Locations
2500 Employees

GitLab Logo GitLab

Recruiter

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
7 Locations
2500 Employees
86K-146K Annually

Similar Companies Hiring

Ford Energy Thumbnail
Automotive • Software • Energy • Utilities • Manufacturing • Renewable Energy
US
55 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
70 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account