Principle SRE

Posted Yesterday
Be an Early Applicant
Hiring Remotely in Mississauga, ON, CAN
In-Office or Remote
175K-262K Annually
Expert/Leader
Information Technology • Consulting
The Role
Merge medical imaging solutions, offered by Merative, combine intelligent, scalable imaging workflow tools with deep and broad expertise to help healthcare organizations improve their confidence in patient outcomes and optimize care delivery. This principal-level individual contributor sets reliability, observability, and automation direction for business-critical Azure cloud platforms. The role establishes SLOs, leads incident management, designs resilience and disaster recovery, builds infrastructure automation, reduces operational toil, mentors engineers, and influences cross-functional teams without direct authority.
Summary Generated by Built In
Merge posting Verbiage (add as first line in the Job Summary) :
Merge medical imaging solutions, offered by Merative, combine intelligent, scalable imaging workflow tools with deep and broad expertise to help healthcare organizations improve their confidence in patient outcomes and optimize care delivery.
With minimal supervision, leverages deep technical expertise to set the reliability, observability, and operational automation direction for business-critical, PHI-bearing cloud platforms.
Establishes reliability standards, service level objectives, and automation practices across multiple teams, and leads their implementation through technical influence rather than direct reporting authority.

Responsibilities: 

People 

  • Provide technical guidance, mentorship, and leadership to engineers across development, QA, and operations teams. 

  • Interact regularly with lower and/or senior management on matters concerning multiple functional areas, departments, and/or customers. 

  • Act as a senior point of escalation for complex or high-severity production issues. 

  • Build reliability capability in others through design reviews, pairing, and blameless post-incident learning. 

  • Foster a culture and develop approaches that generate innovative ideas, products, and services. 

Reliability Engineering 

  • Define, publish, and govern service level objectives (SLOs) management framework, help to define service level indicators, and error budgets for the platform and its shared services. 

  • Set observability standards for metrics, logging, tracing, and alerting, and drive consistent adoption across teams. 

  • Own the technical side of the incident lifecycle: detection, response, escalation, and post-incident review; drive corrective actions to closure. 

  • Lead capacity planning, performance engineering, failure-domain isolation, and disaster recovery design. 

  • Drive resilience patterns into product architecture in partnership with development teams. 

  • Monitor and act on incoming issues from support, customers, or other stakeholders. 

Operational Automation Design and Implementation 

  • Design the automation strategy for platform operations and set the standards, patterns, and tooling other engineers build against. 

  • Design and implement automated remediation for recurring failure modes so routine faults are resolved without human intervention. 

  • Build and maintain infrastructure-as-code, environment provisioning, and deployment automation. 

  • Automate recurring operational work including patching, scaling, certificate rotation, backup and restore validation, and disaster recovery exercises. 

  • Build self-service tooling that lets development and support teams perform routine operational tasks safely and without escalation. 

  • Automate the collection of evidence and the verification of security and compliance controls that apply to PHI-bearing workloads. 

  • Identify, measure, and reduce operational toil; set and report against measurable toil-reduction targets. 

Process 

  • Provide guidance on company processes. 

  • Liaise with cross functional teams (development, product, program management, support, implementation, security, etc.) in delivering and supporting their projects. 

  • Support cross functional teams in resolving customer concerns. 

  • Participate in the creation and/or review and/or approval of architecture, design, and project documents. 

  • Plan, track, and deliver assigned reliability and automation initiatives, on-time and on-budget. 

  • Report initiative status and immediately escalate when work is varying from commitments. 

  • Effectively represent the platform’s reliability posture to the customer and to auditors if/as needed. 

  • Adhere to Merge methodology, quality system requirements, and good engineering practices. 

  • Provide input into applicable budgets, including cloud consumption and tooling spend. 

  • Pursue self-development as an employee to be better at their current role as well as to grow into their next role if/as appropriate. 

  • Adhere to all applicable legal requirements. 

Core Competencies

  • Organized: Demonstrates strong planning, coordination, and attention to detail while managing multiple priorities.

  • Analytical: Uses data-driven thinking and sound judgment to solve complex problems and make informed decisions.

  • Growth Mindset: Embraces feedback, continuous learning, and opportunities for personal and professional development.

  • Quick Learner: Rapidly acquires new technical knowledge, processes, and business context.

  • Systems Thinking: Understands and evaluates failures and performance across an entire distributed platform rather than focusing on isolated components.

  • Technical Acumen: Possesses deep expertise in cloud infrastructure and distributed systems, with the credibility to influence technical strategy across teams.

  • Automation-First Mindset: Continuously seeks opportunities to eliminate manual effort through scalable automation and process improvement.

  • Communication Skills: Effectively communicates complex ideas through clear, concise verbal and written communication.

  • Leadership & Influence: Builds alignment, drives outcomes, and influences stakeholders without relying on direct authority.

  • Priority Management: Effectively balances competing demands and focuses efforts on the highest-impact work.

  • Composure Under Pressure: Maintains sound judgment, clear decision-making, and calm leadership during production incidents and high-pressure situations.

  • Collaboration: Builds strong relationships and works effectively with diverse stakeholders across teams, functions, and external partners.

Technical Skills, (if applicable): 

  • Cloud infrastructure and architecture, with depth in Microsoft Azure 

  • Infrastructure as code and configuration management (e.g., Terraform, Bicep/ARM, Ansible) 

  • CI/CD pipeline design and release automation 

  • Containers and orchestration (Kubernetes/AKS); service mesh concepts 

  • Observability and telemetry tooling (e.g., Azure Monitor/KQL, Prometheus, Grafana, distributed tracing) 

  • Proficiency in at least one automation or systems language (e.g., Python, Go, Bash, Java) 

  • Linux administration, networking, and identity/authorization fundamentals 

  • Incident management and post-incident analysis practice 

  • Ability to understand software architecture and design patterns 

  • Technical Project Management 

  • Experience in an Agile Environment 

  • Microsoft Office 

Preferred Skills

  • Service mesh implementation (Istio) 

  • Event streaming and messaging platforms (Kafka) 

  • Relational and NoSQL data platform operations at scale 

  • Cybersecurity and security engineering 

  • Chaos engineering and resilience testing 

  • Data Engineering 

  • Machine Learning and Artificial Intelligence applied to operations 

Supervisory Skills, (if applicable): 

  • Individual contributor role; no direct reports. 

  • Possess some fundamental leadership skills including strategic thinking, team building, adaptability and conflict resolution. 

  • Ability to technically lead and influence experienced level professionals. 

Qualifications Required: 

a)  Education Requirements: 

  • Degree in Computer Science, Engineering or related field; or equivalent level of industry related experience. 

Preferred

  • Azure certification (e.g., Azure Solutions Architect Expert or DevOps Engineer Expert).

Experience Required

  • 10+ years’ experience in software engineering, systems engineering, or infrastructure operations with demonstrated progression into a principal or staff level technical role. 

  • Demonstrated experience informally leading teams, projects or people to successful outcomes. 

  • Demonstrated experience operating production SaaS at scale against formal availability commitments. 

  • Demonstrated experience designing and implementing operational automation that measurably reduced manual effort or recovery time. 

Preferred: 

  • Experience in a regulated environment (HIPAA/HITRUST, ISO 13485, IEC 62304). 

  • Medical imaging experience: DICOM, HL7 

  • Agile, Scrum 

Work Environment 

The work environment characteristics here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. 

  • Office environment: remote 

  • Travel: Minimal (~5%) 

  • Participation in an on-call escalation rotation, including occasional work outside standard business hours 

Compensation

The salary range provided in this job posting is intended to reflect the general market value for the position. The actual salary offered may vary based on factors such as the candidate’s experience, qualifications, skills, and the specific requirements of the role. This range may also be subject to change as market conditions evolve. We encourage open communication throughout the interview process to discuss compensation expectations. For base-salary + commission sales roles, the range represents On-Target Earnings.

Min – Max :

$174,688.00 - $262,032.00 (USD)

Benefits

The benefits described represent the current offerings at our organization, however, benefits are subject to change and may vary by location and employment status.  We strive to provide a comprehensive benefits package that supports our employees’ health, wellness, and financial goals.  Please note that benefits may be discussed in more detail during the hiring process.

  • Remote first / work from home culture

  • Flexible vacation to help you rest, recharge, and connect with loved ones

  • Paid leave benefits

  • Health, dental, and vision insurance

  • 401k retirement savings plan

  • Infertility benefits

  • Tuition reimbursement, life insurance, EAP – and more!



It is the policy of Merative to provide equal employment opportunity (EEO) to all persons regardless of age, color, national origin, citizenship status, physical or mental disability, race, religion, creed, gender, sex, sexual orientation, gender identity and/or expression, genetic information, marital status, status with regard to public assistance, veteran status, or any other characteristic protected by federal, state or local law. In addition, Merative will provide reasonable accommodations for qualified individuals with disabilities.

Merative participates in the federal E-Verify program to confirm the identity and employment authorization of all newly hired employees. For further information about the E-Verify program, please click here: http://www.uscis.gov/e-verify/employees

Skills Required

  • Degree in Computer Science, Engineering, or a related field, or equivalent industry experience
  • 10+ years of experience in software engineering, systems engineering, or infrastructure operations, with progression into a principal or staff-level technical role
  • Experience informally leading teams, projects, or people to successful outcomes
  • Experience operating production SaaS at scale against formal availability commitments
  • Experience designing and implementing operational automation that measurably reduced manual effort or recovery time
  • Cloud infrastructure and architecture expertise, with depth in Microsoft Azure
  • Infrastructure as code and configuration management experience using technologies such as Terraform, Bicep/ARM, or Ansible
  • CI/CD pipeline design and release automation experience
  • Experience with containers and orchestration, including Kubernetes or AKS
  • Observability and telemetry experience, including metrics, logging, tracing, and alerting
  • Proficiency in at least one automation or systems language, such as Python, Go, Bash, or Java
  • Linux administration, networking, and identity and authorization fundamentals
  • Incident management and post-incident analysis experience
  • Ability to understand software architecture and design patterns
  • Technical project management experience
  • Experience working in an Agile environment
  • Microsoft Office proficiency
  • Azure certification, such as Azure Solutions Architect Expert or DevOps Engineer Expert
  • Experience in a regulated environment, such as HIPAA/HITRUST, ISO 13485, or IEC 62304
  • Medical imaging experience with DICOM or HL7
  • Agile or Scrum experience
  • Service mesh implementation experience, such as Istio
  • Experience with event streaming and messaging platforms such as Kafka
  • Experience operating relational and NoSQL data platforms at scale
  • Cybersecurity and security engineering experience
  • Chaos engineering and resilience testing experience
  • Data engineering experience
  • Experience applying machine learning and artificial intelligence to operations

Merative Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Merative and has not been reviewed or approved by Merative.

  • Wellbeing & Lifestyle Benefits Remote and hybrid flexibility supports work–life balance. Flexible schedules and remote-first norms are positioned as a relative bright spot across many roles.
  • Leave & Time Off Breadth Flexible or unlimited PTO policies are available in many roles alongside paid sick time. Parental leave adds further breadth to time-off options.
  • Fair & Transparent Compensation Base pay is considered fair to good in select roles, with some positions characterized as having solid compensation. Competitive ranges are evident for certain senior product and architecture roles.

Merative Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Ann Arbor, MI
1,585 Employees
Year Founded: 2022

What We Do

Merative is a data, software and technology partner for the health and government social services industries, working with providers, health plans, employers, life sciences companies and governments. With trusted technology and human expertise, the company works with clients to drive real progress. Merative helps clients orient information and insights around the people they serve to improve decision-making and performance. Merative, formerly IBM Watson Health, became a new standalone company as part of Francisco Partners in 2022. Learn more at merative.com

Similar Jobs

GitLab Logo GitLab

Staff Commercial Pricing Strategist

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
2 Locations
2500 Employees
139K-235K Annually
Remote
2 Locations
529 Employees
145K-183K Annually

Block Logo Block

Senior GRC Engineer

Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
In-Office or Remote
8 Locations
12000 Employees
185K-327K Annually

Block Logo Block

Senior Director, Global Tax Planning & Controversy

Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
In-Office or Remote
8 Locations
12000 Employees
252K-377K Annually

Similar Companies Hiring

Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account