Senior Site Reliability Operations Engineer - Finance

Posted Yesterday
Be an Early Applicant
Hiring Remotely in São Paulo, BRA
In-Office or Remote
Senior level
Information Technology • Software
The Role
Leads incident response and service restoration for mission-critical infrastructure in a regulated financial environment. Responsibilities include incident command, root-cause analysis, executive reporting, observability improvements, Linux and Windows troubleshooting, CI/CD deployments, infrastructure migrations, change management, documentation, vendor coordination, and participation in an overnight on-call rotation.
Summary Generated by Built In
About Truelogic

At Truelogic we are a leading provider of nearshore staff augmentation services headquartered in New York. For over two decades, we’ve been delivering top-tier technology solutions to companies of all sizes, from innovative startups to industry leaders, helping them achieve their digital transformation goals.

Our team of 600+ highly skilled tech professionals, based in Latin America, drives digital disruption by partnering with U.S. companies on their most impactful projects. Whether collaborating with Fortune 500 giants or scaling startups, we deliver results that make a difference.

By applying for this position, you’re taking the first step in joining a dynamic team that values your expertise and aspirations. We aim to align your skills with opportunities that foster exceptional career growth and success while contributing to transformative projects that shape the future.

Our Client

A leading Financial Services


Job Summary

The Site Reliability Operations (SRO) team ensures 24/7 stability of the internal IT infrastructure and mission-critical backend systems. This role is not DevOps-focused, but is crucial in monitoring, coordinating, and restoring operations during incidents, particularly in a high-stakes, regulated environment.

The role balances incident command, technical troubleshooting, project leadership, and communication with multiple internal and external stakeholders.

Responsibilities
  • Lead incident response as Incident Commander, coordinating teams, communications, and service restoration

  • Produce executive-level incident reports, run RCAs, and drive continuous improvement

  • Monitor and improve observability using tools like AWS CloudWatch and New Relic, reducing alert noise and gaps

  • Provide hands-on system support across Linux and Windows environments, including complex infrastructure issues

  • Manage and execute deployments via Jenkins, GitLab, or similar CI/CD platforms

  • Own infrastructure initiatives such as migrations, upgrades, and process improvements

  • Enforce change management and risk assessment for production changes

  • Maintain documentation and SOPs, acting as a key liaison between engineering teams and external vendors

  • On-call rotation: 1-week rotation, subject to critical incident call-ins between 6:00 PM and 6:00 AM PT.

Qualifications and Job Requirements
  • 5+ years of experience in Windows and Linux environments with proven troubleshooting capabilities.

  • Strong knowledge of monitoring tools like AWS CloudWatch, New Relic, Nagios, SumoLogic.

  • Practical experience with CI/CD tools (Jenkins, GitLab) and backup tools (CommVault, AWS Backup).

  • Strong scripting skills in PowerShell, Python, or equivalent.

  • Outstanding communication skills, especially under pressure, including executive reporting.

  • Experience in high-paced environments and with on-call support models.

  • Autonomous and proactive attitude; capable of managing complex tasks independently.

What We Offer
  • 100% Remote Work: Enjoy the freedom to work from the location that helps you thrive. All it takes is a laptop and a reliable internet connection.

  • Highly Competitive USD Pay: Earn an excellent, market-leading compensation in USD, that goes beyond typical market offerings.

  • Paid Time Off: We value your well-being. Our paid time off policies ensure you have the chance to unwind and recharge when needed.

  • Work with Autonomy: Enjoy the freedom to manage your time as long as the work gets done. Focus on results, not the clock.

  • Work with Top American Companies: Grow your expertise working on innovative, high-impact projects with Industry-Leading U.S. Companies.

Why You’ll Like Working Here
  • A Culture That Values You: We prioritize well-being and work-life balance, offering engagement activities and fostering dynamic teams to ensure you thrive both personally and professionally.

  • Diverse, Global Network: Connect with over 600 professionals in 25+ countries, expand your network, and collaborate with a multicultural team from Latin America.

  • Team Up with Skilled Professionals: Join forces with senior talent. All of our team members are seasoned experts, ensuring you're working with the best in your field.

Apply now!

Skills Required

  • 5+ years of experience working in Windows and Linux environments with proven troubleshooting capabilities
  • Strong knowledge of monitoring tools such as AWS CloudWatch, New Relic, Nagios, and Sumo Logic
  • Practical experience with CI/CD tools such as Jenkins and GitLab
  • Practical experience with backup tools such as CommVault and AWS Backup
  • Strong scripting skills in PowerShell, Python, or equivalent
  • Outstanding communication skills, particularly during high-pressure incidents and executive reporting
  • Experience working in fast-paced environments and with on-call support models
  • Autonomous and proactive approach with the ability to manage complex tasks independently
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, NY
266 Employees
Year Founded: 2003

What We Do

Truelogic Software is a Nearshore tech firm specializing in Staff Augmentation Services and Innovation Projects. We build innovative digital products by extending the engineering teams of US companies with our elite group of 500 Latin American highly experienced tech talent. Our SERVICES Staff Augmentation Services Dedicated Agile Teams Innovation Projects Our EXPERTISE Mobile & Web Development Data Engineering DevOps QA Automation & Testing UX/UI Designing Project Management Email us [email protected]

Similar Jobs

Remote
12 Locations
266 Employees

Wise Logo Wise

Customer Support Specialist

Fintech • Mobile • Payments • Software • Financial Services
Remote or Hybrid
São Paulo, BRA
9000 Employees
7K-7K Annually

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Temporary Associate

eCommerce • Fashion • Retail • Sales • Wearables • Design
Remote or Hybrid
14 Locations
16000 Employees
15-20 Hourly

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Temporary Associate

eCommerce • Fashion • Retail • Sales • Wearables • Design
Remote or Hybrid
14 Locations
16000 Employees
15-20 Hourly

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account