Senior Site Reliability Engineer (SRE)

Posted 18 Days Ago
Be an Early Applicant
2 Locations
In-Office
93K-204K Annually
Senior level
Fitness • Healthtech • Retail • Pharmaceutical
The Role
Ensure reliability, scalability, and performance of distributed retail and pharmacy systems. Implement observability, monitoring, SLOs, incident response, and reliability improvements. Lead cloud-native microservices and Kubernetes/OpenShift deployments, CI/CD automation, performance testing, and collaborate with engineering and operations. Participate in on-call rotation.
Summary Generated by Built In

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time.

Position Summary
The Senior Site Reliability Engineer is pivotal in ensuring the reliability, scalability, and performance of distributed store technologies that power thousands of locations nationwide — including pharmacy platforms, Point of Sale (POS) systems, handheld devices, store servers, dispensing locations, and connectivity infrastructure. You will drive improvements in observability, monitoring, and DevOps practices to enhance system resilience and accelerate lead time for core development initiatives. Participation in On call rotation is required.

The role is hybrid based out of our corporate offices in Woonsocket, RI or Scottsdale, AZ.

Responsibilities:

Observability & Monitoring

  • Architect and optimize observability solutions using tools such as Splunk, Dynatrace, Datadog, Prometheus, and Grafana to provide end-to-end visibility across retail and pharmacy platforms.

  • Develop proactive monitoring, alerting, and dashboarding strategies while managing SLOs, SLAs, and error budgets to ensure service reliability and performance.

Performance & Reliability Engineering

  • Partner with engineering, infrastructure, and operations teams to embed SRE best practices, improve application resiliency, and optimize performance across enterprise environments.

  • Lead incident response, root cause analysis, performance testing, and reliability reviews to drive continuous improvement and reduce operational risk.

Microservices & Deployments

  • Champion cloud-native architectures leveraging microservices, Kubernetes, and OpenShift to deliver scalable, highly available solutions across hybrid cloud environments.

  • Drive deployment excellence through CI/CD automation, streamlining release processes and improving deployment speed, stability, and quality.

Required Qualifications

  • 5+ years of experience in SRE, DevOps, or related technology roles with experience with observability and monitoring tools such as Splunk, Dynatrace, Datadog, Prometheus, Grafana, etc.

  • 3+ years of experience in delivering software in a large-scale environment with reliability and resilience concepts (multi-region, multi-cloud, containerization, etc.)

  • 2+ years of experience with AI/AIOps and Java/Python

  • 2+ years of experience with Cloud Technologies (AWS, Microsoft Azure, Google Cloud), Microservices concepts, and capabilities like Rancher, Docker, Kubernetes, and web API’s

  • 2+ years of experience with source control and continuous integration tools like GitHub, BitBucket, or Jenkins 

Preferred Qualifications

  • Experience in Incident Management, Change Management, Infrastructure Support, and Problem Management concepts and processes

  • Experience with retail SRE organizations, including experience with store systems; Point of Sale (POS), hand-helds, etc.

  • Experience supporting retail or pharmacy systems at scale

  • Expertise in cloud development and deployment technologies, including containerization and multi-cloud configurations

  • Demonstrated understanding of various API management and related platforms like Apigee, Vordel, Data power

Education

  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience (HS diploma + 4 years relevant experience)

Anticipated Weekly Hours

40

Time Type

Full time

Pay Range

The typical pay range for this role is:

$92,700.00 - $203,940.00

This pay range represents the base hourly rate or base annual full-time salary for all positions in the job grade within which this position falls.  The actual base salary offer will depend on a variety of factors including experience, education, geography and other relevant factors.  This position is eligible for a CVS Health bonus, commission or short-term incentive program in addition to the base pay range listed above. 
 

Our people fuel our future. Our teams reflect the customers, patients, members and communities we serve and we are committed to fostering a workplace where every colleague feels valued and that they belong.

Great benefits for great people

We take pride in offering a comprehensive and competitive mix of pay and benefits that reflects our commitment to our colleagues and their families.

This full‑time position is eligible for a comprehensive benefits package designed to support the physical, emotional, and financial well‑being of colleagues and their families. The benefits for this position include medical, dental, and vision coverage, paid time off, retirement savings options, wellness programs, and other resources, based on eligibility.


Additional details about available benefits are provided during the application process and on
Benefits Moments.

We anticipate the application window for this opening will close on: 08/06/2026

Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state and local laws.

Skills Required

  • 5+ years of experience in SRE, DevOps, or related technology roles with observability and monitoring tools such as Splunk, Dynatrace, Datadog, Prometheus, Grafana
  • 3+ years of experience delivering software in large-scale environments with reliability and resilience concepts (multi-region, multi-cloud, containerization)
  • 2+ years of experience with AI/AIOps and Java/Python
  • 2+ years of experience with Cloud Technologies (AWS, Microsoft Azure, Google Cloud), Microservices concepts, and tools like Rancher, Docker, Kubernetes, and web APIs
  • 2+ years of experience with source control and continuous integration tools like GitHub, BitBucket, or Jenkins
  • Participation in on-call rotation
  • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience
  • Experience in Incident Management, Change Management, Infrastructure Support, and Problem Management concepts and processes
  • Experience in retail SRE organizations and store systems (Point of Sale, handheld devices)
  • Experience supporting retail or pharmacy systems at scale
  • Expertise in cloud development and deployment technologies, including containerization and multi-cloud configurations
  • Understanding of API management and related platforms like Apigee, Vordel, DataPower
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Woonsocket, RI
119,959 Employees
Year Founded: 1963

What We Do

CVS Health is the leading health solutions company that delivers care in ways no one else can. We reach people in more ways and improve the health of communities across America through our local presence, digital channels and our nearly 300,000 dedicated colleagues – including more than 40,000 physicians, pharmacists, nurses and nurse practitioners. Wherever and whenever people need us, we help them with their health – whether that’s managing chronic diseases, staying compliant with their medications, or accessing affordable health and wellness services in the most convenient ways. We help people navigate the health care system – and their personal health care – by improving access, lowering costs and being a trusted partner for every meaningful moment of health. And we do it all with heart, each and every day.

Similar Jobs

Remote or Hybrid
United States
1750 Employees

Akamai Technologies Logo Akamai Technologies

Site Reliability Engineer

Cloud • Security • Software • Cybersecurity
In-Office or Remote
2 Locations
10285 Employees
146K-264K Annually

ServiceTitan Logo ServiceTitan

Senior Site Reliability Engineer

Artificial Intelligence • Cloud • Fintech • Machine Learning • Mobile • Software
Remote or Hybrid
US
2760 Employees
138K-221K Annually

Akamai Technologies Logo Akamai Technologies

Site Reliability Engineer

Cloud • Security • Software • Cybersecurity
In-Office or Remote
2 Locations
10285 Employees
146K-264K Annually

Similar Companies Hiring

Scotch Thumbnail
Artificial Intelligence • eCommerce • Fintech • Payments • Retail • Software • Analytics
US
35 Employees
OneImaging Thumbnail
Healthtech
Miami, FL
62 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account