Senior Site Reliability Operations Engineer - Finance

Posted Yesterday
Be an Early Applicant
12 Locations
Remote
Senior level
Information Technology • Software
The Role
Oversee 24/7 reliability of internal IT infrastructure and mission-critical backend systems during a swing shift. Monitor platforms, troubleshoot Linux/UNIX and Windows environments, coordinate CI/CD deployments, lead incident response, manage backups, and support infrastructure migrations and cloud upgrades. The role also automates operational tasks, refines alerting, maintains procedures, and coordinates with developers, vendors, incident management teams, and executive stakeholders.
Summary Generated by Built In
About Truelogic

At Truelogic we are a leading provider of nearshore staff augmentation services headquartered in New York. For over two decades, we’ve been delivering top-tier technology solutions to companies of all sizes, from innovative startups to industry leaders, helping them achieve their digital transformation goals.

Our team of 600+ highly skilled tech professionals, based in Latin America, drives digital disruption by partnering with U.S. companies on their most impactful projects. Whether collaborating with Fortune 500 giants or scaling startups, we deliver results that make a difference.

By applying for this position, you’re taking the first step in joining a dynamic team that values your expertise and aspirations. We aim to align your skills with opportunities that foster exceptional career growth and success while contributing to transformative projects that shape the future.

Our Client

A leading Financial Services Company


Job Summary

The Site Reliability Operations (SRO) team ensures 24/7 stability of the internal IT infrastructure and mission-critical backend systems. Some of the skills are monitoring, coordinating, and restoring operations during incidents, particularly in a high-stakes, regulated environment. This is an Operations swing shift from 2 PM to 10:30 PM PST timezone.

The role balances incident command, technical troubleshooting, project leadership, and communication with multiple internal and external stakeholders.

Responsibilities
  • Oversee multi-platform IT infrastructure health using AWS CloudWatch, New Relic, Nagios, and SumoLogic. Continuously refine alert thresholds to minimize noise and enable proactive remediation.

  • Serve as an escalation point for complex technical issues. Perform troubleshooting across Linux/UNIX, Windows, virtual servers, and virtual desktop environments.

  • Coordinate, automate, and execute code deployments using Jenkins, GitLab, or similar CI/CD tools, driving toward non-disruptive releases and zero-downtime updates.

  • Collaborate closely with Application Developers, 3rd-party vendors, and internal Incident Management.

  • Drive medium- to large-scale infrastructure projects, including migrations, cloud upgrades, and performance tuning.

  • Maintain SOPs in team knowledge bases, and oversee enterprise backup operations (CommVault, Veeam, AWS Backup).

Qualifications & Requirements
  • 5+ years in an Operations Center (SRO/NOC) or cloud infrastructure environment with hands-on experience in full-stack application deployments.

  • Proficiency in both Windows and UNIX/Linux administration, troubleshooting (scripting, grepping logs, analyzing performance metrics), and virtual server/desktop management.

  • Practical experience with AWS cloud services (Storage, VMs, Networking) and enterprise monitoring tools (AWS CloudWatch, New Relic, Nagios, SumoLogic).

  • Scripting or programming capability in PowerShell, Python, or bash to automate repetitive tasks and optimize run-time operations.

  • Practical experience with CI/CD platforms (Jenkins, GitLab), ITSM ticket platforms (ServiceNow, Jira), and backup solutions (CommVault, Veeam, AWS Backup).

  • Good verbal and written communication skills with experience serving as a bridge between technical teams, executive stakeholders, and external vendors.

Pluses
  • Advanced AWS Certifications

  • Background in or exposure to AI/ML tools for infrastructure monitoring and predictive analytics.

  • Previous experience in ITIL-aligned environments or enterprise Change/Incident Management frameworks.

  • Bachelor’s Degree in Computer Science, Information Technology, or a related field.

What We Offer
  • 100% Remote Work: Enjoy the freedom to work from the location that helps you thrive. All it takes is a laptop and a reliable internet connection.

  • Highly Competitive USD Pay: Earn an excellent, market-leading compensation in USD, that goes beyond typical market offerings.

  • Paid Time Off: We value your well-being. Our paid time off policies ensure you have the chance to unwind and recharge when needed.

  • Work with Autonomy: Enjoy the freedom to manage your time as long as the work gets done. Focus on results, not the clock.

  • Work with Top American Companies: Grow your expertise working on innovative, high-impact projects with Industry-Leading U.S. Companies.

Why You’ll Like Working Here
  • A Culture That Values You: We prioritize well-being and work-life balance, offering engagement activities and fostering dynamic teams to ensure you thrive both personally and professionally.

  • Diverse, Global Network: Connect with over 600 professionals in 25+ countries, expand your network, and collaborate with a multicultural team from Latin America.

  • Team Up with Skilled Professionals: Join forces with senior talent. All of our team members are seasoned experts, ensuring you're working with the best in your field.

Apply now!

Skills Required

  • 5+ years of experience in an Operations Center, SRO/NOC, or cloud infrastructure environment
  • Hands-on experience with full-stack application deployments
  • Windows and UNIX/Linux administration and troubleshooting experience
  • Experience with scripting, log analysis, performance metrics, and virtual server/desktop management
  • Practical experience with AWS cloud services, including storage, virtual machines, and networking
  • Experience with AWS CloudWatch, New Relic, Nagios, and SumoLogic
  • Scripting or programming ability in PowerShell, Python, or Bash
  • Experience with Jenkins or GitLab CI/CD platforms
  • Experience with ServiceNow or Jira ITSM platforms
  • Experience with CommVault, Veeam, or AWS Backup
  • Strong verbal and written communication skills with technical, executive, and vendor stakeholders
  • Advanced AWS certifications
  • Experience with AI/ML tools for infrastructure monitoring or predictive analytics
  • Experience in ITIL-aligned environments or enterprise Change/Incident Management frameworks
  • Bachelor’s degree in Computer Science, Information Technology, or a related field
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, NY
266 Employees
Year Founded: 2003

What We Do

Truelogic Software is a Nearshore tech firm specializing in Staff Augmentation Services and Innovation Projects. We build innovative digital products by extending the engineering teams of US companies with our elite group of 500 Latin American highly experienced tech talent. Our SERVICES Staff Augmentation Services Dedicated Agile Teams Innovation Projects Our EXPERTISE Mobile & Web Development Data Engineering DevOps QA Automation & Testing UX/UI Designing Project Management Email us [email protected]

Similar Jobs

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Temporary Sales Support Associate

eCommerce • Fashion • Retail • Sales • Wearables • Design
Remote or Hybrid
14 Locations
16000 Employees
15-20 Hourly

AirDNA Logo AirDNA

Sales Executive

Software • Travel
Easy Apply
Remote or Hybrid
10 Locations
150 Employees

Hewlett Packard Enterprise Logo Hewlett Packard Enterprise

Architect

Artificial Intelligence • Cloud • Information Technology • Consulting
In-Office or Remote
3 Locations
85422 Employees
In-Office or Remote
12 Locations
125 Employees
75K-195K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account