Site Reliability Engineer

Posted Yesterday
Hiring Remotely in San Francisco, CA, USA
In-Office or Remote
120K-249K Annually
Mid level
Software
The Role
Improve platform reliability, observability, resilience, and operational readiness. Partner with development teams to establish reliability standards, implement observability practices, enable autonomous service deployment and support, mentor developers, and drive adoption of incident response, SLO, error budget, and service health practices across the organization.
Summary Generated by Built In

MaintainX is a leading mobile-first work execution platform for industrial and frontline teams. More than 13,000 customers, including Duracell, McDonald's, Shell, DHL and Volvo, use MaintainX to cut unplanned downtime and run better operations, across 13.9 million managed assets and 79.5 million completed work orders.

In August 2026 MaintainX became part of Autodesk, joining Autodesk Operations Solutions, the organization unifying Autodesk's operations platform alongside Tandem, FlexSim and Fusion Operations. Autodesk's strategy is to converge design, make and operate into one continuous lifecycle: design an asset, build it, run it, then feed what you learn running it back into the next design. Autodesk had design and make. Operate is the phase that tells you what actually happened, and it is ours.

We’re looking for a Site Reliability Engineer to help advance MaintainX’s reliability, observability, and developer autonomy as we scale our platform.

In this role, you’ll partner closely with product and platform development teams to improve the stability, resilience, and operational readiness of our services. You’ll work alongside teams to design for reliability from the start, establish clear ownership and standards, and build shared tooling that enables teams to operate their services with confidence.

You’ll also contribute to company-wide initiatives that define how MaintainX approaches reliability software development, including observability standards, incident response practices, and service health metrics, helping the organization adopt proven industry practices at scale.

This role is well-suited for an developer who enjoys working across teams, influencing technical direction through strong development practices, and turning reliability principles into practical, scalable systems.

What You'll Do:

  • Assess service maturity and provide insights to development teams

  • Partner with development teams to implement observability best practices

  • Enable development teams to become autonomous with their service deployment, support, and infrastructure

  • Mentor developers on reliability practices, focusing on making them self-sufficient

  • Act as the bridge, ear and eyes of the Platform Division teams to drive tooling and practice adoption across development teams

About You:

  • Deep understanding of observability practices in a distributed system environment and how it influences system design and team behaviour

  • Practical experience with SRE concepts (SLOs, error budgets, incident management)

  • 3–5+ years in software development, SRE, DevOps, or production development roles with experience operating production systems

  • Proficient in cloud-native platforms and infrastructure-as-code concepts and tools

  • Working knowledge of at least one programming language (TypeScript/Node.js is a plus)

  • Excellent communication and collaboration abilities across technical and non-technical teams

  • Ability to translate complex reliability concepts into actionable guidance

  • You enjoy enabling teams to succeed independently and measuring success by reduced dependency on you

About Us:

MaintainX is committed to creating a diverse environment. All qualified applicants will receive consideration for employment without regard to race, colour, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status.

 

Our mission is to keep the physical world running. Factories, fleets, hospitals and campuses stay up because the people who maintain them have tools worth using. That is what we build.

Compensation and benefits. Base pay is one part of the package. Depending on the role, compensation may also include commission, an annual bonus and equity. Benefits differ by country. For roles in the United States, Autodesk’s benefits are described at benefits.autodesk.com. For roles in Canada and other countries, the plan differs on health coverage, retirement and leave, and your recruiter will walk you through it.

Belonging. We take pride in a culture where everyone can thrive. More at autodesk.com/company/global-belonging. More on where this is going: Autodesk CEO Andrew Anagnost on building the future of connected operations, and AOS SVP Stephen Hooper on welcoming MaintainX to Autodesk.

Skills Required

  • Deep understanding of observability practices in distributed systems
  • Practical experience with SRE concepts, including SLOs, error budgets, and incident management
  • 3-5+ years of experience in software development, SRE, DevOps, or production development roles
  • Experience operating production systems
  • Proficiency in cloud-native platforms and infrastructure-as-code concepts and tools
  • Working knowledge of at least one programming language
  • Excellent communication and collaboration abilities across technical and non-technical teams
  • Ability to translate complex reliability concepts into actionable guidance
  • TypeScript or Node.js experience

MaintainX Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about MaintainX and has not been reviewed or approved by MaintainX.

  • Strong & Reliable Incentives Feedback suggests incentive structures with accelerators can significantly boost earnings for high performers, making roles feel rewarding when targets are met.
  • Healthcare Strength Company materials indicate comprehensive medical, dental, and vision coverage as core benefits, and feedback suggests employees value the breadth of this coverage.
  • Leave & Time Off Breadth Job materials describe flexible or 'take what you need' PTO policies, and feedback suggests this flexibility supports work-life balance.

MaintainX Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, California
300 Employees
Year Founded: 2018

What We Do

MaintainX is the leading maintenance and work execution software, designed specifically for industrial and frontline teams. We help companies streamline maintenance operations, improve asset management, and empower workers—all while delivering insights that can improve your bottom line. As a mobile-first platform, MaintainX delivers a modern, IoT-enabled solution for maintenance, reliability, and operations teams trusted by over 6,500 companies worldwide. If you’re looking for a CMMS solution that’s easy to use and implement, look no further. The MaintainX platform manages over 15 million work orders and 2.5 million assets, and is used by hundreds of thousands of workers globally. We help customers reduce unplanned downtime and increase asset availability, while meeting complex compliance needs and keeping workers safe. Ready to ditch the clipboard? Here's what we can help your team digitize: -Maintenance Work Orders -Preventive Maintenance -Safety Procedures -Safety and Environmental Audits -Multi-site Reporting -IoT & ERP Integrations -Auditing/Inspection Workflows -Training Checklists -Parts Order Management & Vendor Connections We’re proud to serve some of the world’s largest brands, including Duracell, AB InBev, Univar, Cintas, McDonalds, Titan America, and many more. To learn more, please visit www.maintainx.com

Similar Jobs

GitLab Logo GitLab

Site Reliability Engineer

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
United States
2500 Employees

CrowdStrike Logo CrowdStrike

Senior Engineer

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
USA
11000 Employees
140K-215K Annually
Remote
USA
68 Employees

Weekday, Inc. Logo Weekday, Inc.

Site Reliability Engineer

Artificial Intelligence • HR Tech • Professional Services • Software
Remote
United States
80-120 Hourly

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account