Staff Site Reliability Engineer

Posted 2 Hours Ago
Be an Early Applicant
Kraków, Małopolskie, POL
In-Office
Senior level
Healthtech
The Role
Own reliability, availability, performance, and operational maturity for cloud-native platforms and data/AI services. Establish SLOs and error budgets, build observability systems, automate toil, lead incident response and postmortems, influence architecture, and improve distributed systems, infrastructure, and CI/CD pipelines. Set cross-team technical standards and mentor engineers across the organization.
Summary Generated by Built In

Bring more to life. 

  

At Danaher, our work saves lives. And each of us plays a part. Fueled by our culture of continuous improvement, we turn ideas into impact – innovating at the speed of life.  

  

Our 60,000+ associates work across the globe at more than 15 unique businesses within life sciences, diagnostics, and biotechnology.   

  

Are you ready to accelerate your potential and make a real difference? At Danaher, you can build an incredible career at a leading science and technology company, where we’re committed to hiring and developing from within. You’ll thrive in a culture of belonging where you and your unique viewpoint matter.  

  

Learn about the Danaher Business System which makes everything possible. 

 

The Staff Site Reliability Engineer is responsible for the availability, performance, and operational maturity of our cloud-native platform and the critical data and AI/ML services that run on it. This is a senior individual contributor role: you will set reliability direction across teams, raise the engineering bar through standards and mentorship, and do the hands-on work of making complex distributed systems predictable. 

This position reports to the Senior Director, Data and AI Platform and is part of the Chief Information Officer (CIO) Office and will be located onsite in Krakow, Poland.

This is a Danaher Corporate role, hosted by our Cytiva operating company in Kraków. 

In this role, you will have the opportunity to: 

  • Champion SRE practice at scale. Establish and monitor Service Level Objectives and error budgets for critical services, and use them to drive real decisions about reliability, availability, performance, and cost-efficiency — including when to slow down and when to ship. 

  • Own observability end to end. Design, implement, and maintain monitoring, logging, and distributed tracing for cloud-native applications and infrastructure (Kubernetes, microservices), and build the dashboards, alerts, and runbooks that give teams deep insight into system health rather than alert noise. 

  • Eliminate toil. Identify repetitive operational work and manual process across the production estate and automate it away, developing the tools, scripts, and pipeline improvements that make operations, deployment, and incident response faster and safer. 

  • Lead incident response and learning. Participate across the full incident lifecycle — detection, triage, mitigation, resolution — and lead thorough blameless postmortems that get to root cause and produce preventative measures that stick. 

  • Shape systems before they're built. Partner with development teams to influence the design of new services so operability, reliability, and cost-efficiency are engineered in from the start, and proactively surface performance bottlenecks and architectural weaknesses before they reach production. 

  • Set technical direction and grow the team. Drive cross-team architecture and reliability decisions, establish standards and documentation, mentor engineers across our operating companies, and help foster a culture of technical rigor, blameless learning, and collaboration. 

 

The essential requirements of the job include: 

  • 5+ years of hands-on experience in a Site Reliability Engineering, DevOps, or equivalent role focused on production system reliability and operations; CS/Engineering degree or equivalent practical experience. 

  • Strong understanding and practical application of SRE principles — SLOs, error budgets, toil reduction, and blameless culture — with proven experience participating in and improving incident management processes for business-critical systems. 

  • Expertise designing, implementing, and managing observability platforms for cloud-native environments (e.g., Prometheus, Grafana, Datadog, ELK stack, OpenTelemetry, Splunk), including the dashboards, alerting, and runbooks that make them actionable. 

  • Extensive hands-on experience with at least one major cloud platform (AWS, Azure, or GCP) across compute, networking, and database services; containerization and orchestration (Docker, Kubernetes); Infrastructure as Code (e.g., Terraform, OpenTofu, Pulumi); and proficiency in at least one language (Python, Go) for automation and tool development. 

  • Proven ability to set technical direction at platform scale — driving cross-team architecture and design decisions, establishing standards, and mentoring engineers — grounded in strong troubleshooting skills across complex distributed systems, including microservices, CI/CD pipelines, and large-scale data infrastructure. 

 

Travel, Motor Vehicle Record & Physical/Environment Requirements:  

  • Ability to travel – up to 10% 

  

It would be a plus if you also possess previous experience in: 

  • Life sciences, diagnostics, or biotechnology (e.g., partnering with R&D, quality, clinical, manufacturing, or commercial teams) 

  • Reliability and efficiency practices at scale — chaos engineering and resilience testing, capacity planning, or cloud cost optimization (FinOps). 

  • Working in a matrixed environment 

 

Danaher offers a broad array of comprehensive, competitive benefit programs that add value to our lives. Whether it’s a health care program or paid time off, our programs contribute to life beyond the job. Check out our benefits at Danaher Benefits Info. 

At Danaher, we believe in designing a better, more sustainable workforce. We recognize the benefits of flexible, remote working arrangements for eligible roles and are committed to providing enriching careers, no matter the work arrangement. This position is eligible for a remote work arrangement in which you can work remotely from your home. Additional information about this remote work arrangement will be provided by your interview team. Explore the flexibility and challenge that working for Danaher can provide.

#LI-KK1

Join our winning team today. Together, we’ll accelerate the real-life impact of tomorrow’s science and technology. We partner with customers across the globe to help them solve their most complex challenges, architecting solutions that bring the power of science to life.

For more information, visit www.danaher.com.

Skills Required

  • 5+ years of hands-on experience in Site Reliability Engineering, DevOps, or an equivalent production reliability and operations role
  • Computer Science or Engineering degree, or equivalent practical experience
  • Practical knowledge of SRE principles, including SLOs, error budgets, toil reduction, and blameless incident management
  • Experience designing and managing observability platforms for cloud-native environments, including dashboards, alerting, and runbooks
  • Hands-on experience with at least one major cloud platform: AWS, Azure, or GCP
  • Experience with compute, networking, database services, Docker, and Kubernetes
  • Experience with Infrastructure as Code such as Terraform, OpenTofu, or Pulumi
  • Proficiency in at least one programming language, including Python or Go, for automation and tool development
  • Ability to set technical direction, drive cross-team architecture decisions, establish standards, and mentor engineers
  • Strong troubleshooting skills across distributed systems, microservices, CI/CD pipelines, and large-scale data infrastructure
  • Experience in life sciences, diagnostics, or biotechnology
  • Experience with chaos engineering, resilience testing, capacity planning, or cloud cost optimization
  • Experience working in a matrixed environment
  • Ability to travel up to 10%

Danaher Corporation Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Danaher Corporation and has not been reviewed or approved by Danaher Corporation.

  • Healthcare Strength Healthcare coverage is described as comprehensive, including medical plan options alongside dental, vision, life, disability, and mental health support. Wellness initiatives and support programs such as an EAP and vaccination or fitness offerings add breadth beyond core insurance.
  • Retirement Support Retirement benefits include a 401(k) plan with employer matching and options such as pre-tax and Roth contributions. Broader financial rewards such as performance bonuses and access to equity or an employee stock purchase plan are also described.
  • Parental & Family Support Parental leave and family-building support are described as available, including maternity and paternity leave and fertility assistance. Childcare and eldercare support are also highlighted as part of the overall package.

Danaher Corporation Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Washington, DC
57,802 Employees
Year Founded: 1984

What We Do

Danaher is a global science and technology innovator committed to helping our customers solve complex challenges and improve quality of life around the world. A global network of more than 25 operating companies, we drive meaningful innovation in some of today’s most dynamic industries through our operating companies in four strategic platforms: Life Sciences, Diagnostics, Water Quality and Product Identification. The engine at the heart of our success is the Danaher Business System (DBS), a set of tools that enables continuous improvement around lean, growth and leadership. Through the ingenuity of our people, the power of DBS and the impact of our meaningful technologies, we help realize life’s potential in ourselves and for those we serve.

Similar Jobs

Alarm.com Logo Alarm.com

Staff Software Engineer

Internet of Things • Software
In-Office
Kraków, Małopolskie, POL
1100 Employees

Akamai Technologies Logo Akamai Technologies

Senior Site Reliability Engineer

Cloud • Security • Software • Cybersecurity
In-Office or Remote
2 Locations
10285 Employees

SailPoint Logo SailPoint

Consultant

Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Remote or Hybrid
2 Locations
2461 Employees

Ericsson Logo Ericsson

Architect

Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
In-Office
6 Locations
88000 Employees

Similar Companies Hiring

Sailor Health Thumbnail
Healthtech • Social Impact • Telehealth
New York City, NY
20 Employees
Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
OneImaging Thumbnail
Healthtech
Miami, FL
62 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account