Senior Site Reliability Engineer

Reposted 3 Days Ago
Be an Early Applicant
Brno, Brno-město, Jihomoravský kraj, CZE
In-Office
Senior level
Information Technology • Security • Software
The Role
Own reliability and uptime of production infrastructure and data platforms. Lead zero-downtime Kubernetes upgrades, manage ELK logging, ClickHouse, and Kafka at scale. Improve observability, runbooks, incident response, CI/CD and IaC automation, and collaborate with engineering, data, and product teams.
Summary Generated by Built In

At SolarWinds, we’re a people-first company. Our purpose is to enrich the lives of the people we serve—including our employees, customers, shareholders, partners, and communities. Join us in our mission to help customers accelerate business transformation with simple, powerful, and secure solutions.

The ideal candidate thrives in an innovative, fast-paced environment and is collaborative, accountable, ready, and empathetic. We’re looking for individuals who believe they can accomplish more as a team and create lasting growth for themselves and others. We hire based on attitude, competency, and commitment. Solarians are ready to advance our world-class solutions in a fast-paced environment and accept the challenge to lead with purpose. If you’re looking to build your career with an exceptional team, you’ve come to the right place. Join SolarWinds and grow with us!


We work in a hybrid mode 3+2, with a minimum of 3 days at the office (with mandatory Tuesdays and Wednesdays) and a maximum of 2 days at the home office.

The location of our office is Holandská 873/6, Brno – Štýřice, 639 00.

We employ only via an employment contract – full-time employment (HPP).


About the Role

We're looking for a Senior Site Reliability Engineer who takes extreme ownership of production systems and thrives in a collaborative, fast-paced environment. You bring deep hands-on experience across infrastructure, data platforms, and delivery pipelines — and you hold yourself accountable for reliability outcomes.

Responsibilities

  • Own reliability and uptime of production infrastructure including Kubernetes clusters and data platforms
  • Lead Kubernetes version upgrades with zero-downtime strategies across large-scale environments
  • Deep understanding of data pipelines and ensuring reliability, observability, and scalability end-to-end
  • Manage and scale ELK stack for centralized logging and observability across services
  • Operate and optimize ClickHouse clusters for high-throughput analytical workloads
  • Administer Kafka clusters — tuning, scaling, and ensuring fault-tolerant message delivery
  • Participate in and continuously improve on-call rotations, runbooks, and incident response processes
  • Drive automation across infrastructure provisioning, deployments, and operational toil
  • Collaborate closely with engineering, data, and product teams to embed reliability from day one

Required Qualifications

  • 8+ years of experience in SRE, DevOps, or infrastructure engineering
  • Hands-on experience with Kubernetes including version upgrade planning, node pool migrations, and zero-downtime rollouts
  • Strong experience with Kafka — operations, tuning, and scaling in production
  • Extensive hands-on experience on cloud platforms (AWS preferred)
  • Experience operating ClickHouse or similar columnar databases at scale
  • Solid background in building and maintaining data pipelines reliably in production
  • Infrastructure as Code with Terraform — modules, state management, and multi-environment setups
  • CI/CD expertise using Flux, Jenkins, or Spinnaker
  • Strong Git practices — branching strategies, GitOps workflows
  • Proficiency in at least one programming/scripting language (Python, Go, Bash)
  • Proven on-call experience with a track record of improving alert quality and reducing MTTR
  • Strong automation mindset — eliminate toil, build durable solutions

Preferred Qualifications

  • Extreme ownership — you don't wait to be asked, you drive problems to resolution
  • Collaborative team player who lifts those around them
  • Clear communicator across engineering and non-engineering stakeholders

Our benefits:

  • 25 days of vacation per year
  • 3 sick days per year
  • 10 study days per year
  • 2 volunteering days per year
  • 4 weeks’ holidays after 5-year tenure, Sabbatical Leave
  • Up to 48 300CZK personal education budget per year
  • Pension or life insurance matching donation up to 3% of the salary or 4000 CZK per month
  • Cash allowance for meals of 95 CZK per working day
  • Unlimited access to LinkedIn Learning
  • English/Czech classes
  • Multisport card
  • Solarian Referral Program
  • SolarWinds Appreciation Program
  • Giving – Donation Matching
  • Employee Assistance
  • Competitive Race Reimbursement
  • Breakfast on Wednesdays
  • Fresh fruits and snacks on Mondays

SolarWinds is an Equal Employment Opportunity Employer. SolarWinds will consider all qualified applicants for employment without regard to race, color, religion, sex, age, national origin, sexual orientation, gender identity, marital status, disability, veteran status or any other characteristic protected by law.

All applications are treated in accordance with the SolarWinds Privacy Notice: https://www.solarwinds.com/applicant-privacy-notice

Skills Required

  • 8+ years of experience in SRE, DevOps, or infrastructure engineering
  • Hands-on experience with Kubernetes including version upgrade planning, node pool migrations, and zero-downtime rollouts
  • Strong experience with Kafka operations, tuning, and scaling in production
  • Proficiency with ELK stack (Elasticsearch, Logstash, Kibana) for log management and observability
  • Experience operating ClickHouse or similar columnar databases at scale
  • Background in building and maintaining data pipelines reliably in production
  • Infrastructure as Code with Terraform including modules, state management, and multi-environment setups
  • CI/CD expertise using Flux, Jenkins, or Spinnaker
  • Strong Git practices including branching strategies and GitOps workflows
  • Proficiency in at least one programming/scripting language (Python, Go, Bash)
  • Proven on-call experience with a track record of improving alert quality and reducing MTTR
  • Strong automation mindset to eliminate toil and build durable solutions
  • Extreme ownership and drive problems to resolution
  • Collaborative team player who lifts those around them
  • Clear communicator across engineering and non-engineering stakeholders

SolarWinds Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about SolarWinds and has not been reviewed or approved by SolarWinds.

  • Healthcare Strength Health coverage is described as comprehensive, with some locations covering healthcare, dental, and vision at 100% and generally well-regarded plan quality. Feedback suggests wellness resources, counseling, and fitness reimbursements further strengthen the offering.
  • Leave & Time Off Breadth Paid time off, holidays, and a sabbatical after five years are part of the package. Feedback suggests global parental leave minimums add meaningful family support to the overall time-off mix.
  • Pay Growth & Progression Compensation is considered solid by many, with consistent raises noted alongside base pay. Feedback suggests bonuses and structured increases contribute to a sense of ongoing pay growth.

SolarWinds Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Austin, TX
2,299 Employees
Year Founded: 1999

What We Do

SolarWinds is a leading provider of powerful and affordable IT management software. Our products give organizations worldwide—regardless of type, size, or complexity—the power to monitor and manage their IT services, infrastructures, and applications; whether on-premises, in the cloud, or via hybrid models. We continuously engage with technology professionals—IT service and operations professionals, DevOps professionals, and managed services providers (MSPs)—to understand the challenges they face in maintaining high-performing and highly available IT infrastructures and applications. The insights we gain from them, in places like our THWACK® community, allow us to solve well-understood IT management challenges in the ways technology professionals want them solved. Our focus on the user and commitment to excellence in end-to-end hybrid IT management has established SolarWinds as a worldwide leader in solutions for network and IT service management, application performance, and managed services.

Similar Jobs

Zencoder Logo Zencoder

Senior Engineer

Artificial Intelligence • Information Technology • Software
In-Office or Remote
28 Locations
25 Employees

Nebius Logo Nebius

Senior Site Reliability Engineer

Artificial Intelligence • Information Technology • Consulting
In-Office or Remote
30 Locations
473 Employees

Nebius Logo Nebius

Senior Site Reliability Engineer

Artificial Intelligence • Information Technology • Consulting
In-Office or Remote
27 Locations
473 Employees
In-Office or Remote
28 Locations
164 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account