Senior Site Reliability Engineer

Posted 6 Days Ago
Be an Early Applicant
Raleigh, NC, USA
In-Office
119K-196K Annually
Senior level
Cloud • Information Technology • Internet of Things • Software • Consulting • Infrastructure as a Service (IaaS) • Automation
Creating better technology the open source way
The Role
Design, build, automate, and operate Red Hat Hybrid OpenShift platforms across cloud and on-prem. Implement GitOps, CI/CD, monitoring, and SRE practices; develop operators and tooling; lead incident response, on-call duties, and postmortems; mentor peers and improve platform reliability and self-service.
Summary Generated by Built In

Job Summary

The Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at scale by enabling customer self-service, making our monitoring system more sustainable, and eliminating toil through automation.

In the IT OpenShift team you will have the opportunity to influence the complex challenges of scale which are unique to Red Hat IT managed cloud platform services, while using your skills in coding, operations, and large-scale distributed system design. We develop, deploy, and maintain Red Hat's next-generation mission critical platform across hybrid cloud infrastructures.  

We are a global team operating on-premise and in the public cloud, using the latest technologies from Red Hat and beyond. Red Hat relies on teamwork and openness for its success. We learn from our failures in a blameless environment to support the continuous improvement of the team. At Red Hat, your individual contributions have more visibility than most large companies, and visibility means career opportunities and growth. Successful applicants must reside in a state where Red Hat is registered to do business.

At Red Hat, our commitment to open source innovation extends beyond our products - it’s embedded in how we work and grow. Red Hatters embrace change – especially in our fast-moving technological landscape – and have a strong growth mindset. That's why we encourage our teams to proactively, thoughtfully, and ethically use AI to simplify their workflows, cut complexity, and boost efficiency. This empowers our associates to focus on higher-impact work, creating smart, more innovative solutions that solve our customers' most pressing challenges.

What You Will Do

  • Design, build, and manage our large scale infrastructure and platform services, including public cloud, private cloud, and datacenter-based

  • Automate cloud infrastructure through use of technologies (e.g. auto scaling, load balancing, etc.), scripting (python and golang), monitoring and alerting solutions (e.g. Splunk, Splunk IM, Prometheus, Grafana, Catchpoint, DataDog etc)

  • Design, develop, and become expert in IT’s Red Hat OpenShift offerings by leveraging emerging industry standards

  • Build & support standardized CI/CD platform components using OpenShift Pipelines and Tekton, GitLab to enable multiple application deployments

  • Apply Infrastructure as Code methodologies using GitOps practices with ArgoCD for declarative platform management

  • Breakdown complex engineering efforts into consumable chunks while working with teams to understand deliverables

  • Design and development of software like Kubernetes operators, webhooks, cli-tools

  • Implement and maintain intelligent infrastructure and application monitoring designed to enable application engineering teams

  • Ensure the production environment is operating in accordance with established procedures and best practices

  • Lead escalation support for high severity and critical platform-impacting events

  • Provide feedback around bugs and feature improvements to the various Red Hat Product Engineering teams

  • Design software tests and lead peer reviews to increase the quality of our codebase

  • Help and develop peers’ capabilities through knowledge sharing, mentoring, and collaboration

  • Participate in a regular on-call schedule, supporting the operation needs of our tenants

  • Drive sustainable incident response and lead blameless postmortems

  • Work within a small agile team to develop and improve SRE methodologies, support your peers, plan and self-improve

What You Will Bring

  • 5+ years of experience operating production services on Kubernetes/OpenShift

  • 3+ years of programming experience in Python, Go

  • 2+ years of experience of using cloud providers and technologies (Google, Azure, Amazon, etc.)

  • Hands-on experience with Kubernetes/OpenShift, Linux AWS

  • Experience with GitOps workflows for managing infrastructure or application configuration

  • Solid understanding of Linux systems administration (RHEL/Fedora preferred)

  • Understanding of standard networking (TCP/IP, DNS, HTTP/TLS) and authentication protocols (LDAP)

  • Comfort with incident response, on-call responsibilities

  • Ability to work independently with minimal supervision while keeping the team informed

  • Knowledge of SRE principles — SLOs, error budgets, toil measurement

The Following Are Considered A Plus:

  • Contributions to open source projects

  • Experience with the Operator SDK or building Kubernetes operators

#LI-JS1

The salary range for this position is $118,600.00 - $195,680.00. Actual offer will be based on your qualifications.

Pay Transparency

Red Hat determines compensation based on several factors including but not limited to job location, experience, applicable skills and training, external market value, and internal pay equity. Annual salary is one component of Red Hat’s compensation package. This position may also be eligible for bonus, commission, and/or equity. For positions with Remote-US locations, the actual salary range for the position may differ based on location but will be commensurate with job duties and relevant work experience. 

About Red Hat

Red Hat is the world’s leading provider of enterprise open source software solutions, using a community-powered approach to deliver high-performing Linux, cloud, container, and Kubernetes technologies. Spread across 40+ countries, our associates work flexibly across work environments, from in-office, to office-flex, to fully remote, depending on the requirements of their role. Red Hatters are encouraged to bring their best ideas, no matter their title or tenure. We're a leader in open source because of our open and inclusive environment. We hire creative, passionate people ready to contribute their ideas, help solve complex problems, and make an impact.

Benefits
●    Comprehensive medical, dental, and vision coverage
●    Flexible Spending Account - healthcare and dependent care
●    Health Savings Account - high deductible medical plan
●    Retirement 401(k) with employer match
●    Paid time off and holidays
●    Paid parental leave plans for all new parents
●    Leave benefits including disability, paid family medical leave, and paid military leave
●    Additional benefits including employee stock purchase plan, family planning reimbursement, tuition reimbursement, transportation expense account, employee assistance program, and more! 

Note: These benefits are only applicable to full time, permanent associates at Red Hat located in the United States. 

Inclusion at Red Hat
Red Hat’s culture is built on the open source principles of transparency, collaboration, and inclusion, where the best ideas can come from anywhere and anyone. When this is realized, it empowers people from different backgrounds, perspectives, and experiences to come together to share ideas, challenge the status quo, and drive innovation. Our aspiration is that everyone experiences this culture with equal opportunity and access, and that all voices are not only heard but also celebrated. We hope you will join our celebration, and we welcome and encourage applicants from all the beautiful dimensions that compose our global village.

Equal Opportunity Policy (EEO)
Red Hat is proud to be an equal opportunity workplace and an affirmative action employer. We review applications for employment without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, citizenship, age, veteran status, genetic information, physical or mental disability, medical condition, marital status, or any other basis prohibited by law.


Red Hat does not seek or accept unsolicited resumes or CVs from recruitment agencies. We are not responsible for, and will not pay, any fees, commissions, or any other payment related to unsolicited resumes or CVs except as required in a written contract between Red Hat and the recruitment agency or party requesting payment of a fee.
Red Hat supports individuals with disabilities and provides reasonable accommodations to job applicants. If you need assistance completing our online job application, email [email protected]. General inquiries, such as those regarding the status of a job application, will not receive a reply. 

Skills Required

  • 5+ years operating production services on Kubernetes/OpenShift
  • 3+ years programming experience in Python and Go
  • 2+ years using cloud providers and technologies (Google, Azure, Amazon)
  • Hands-on experience with Kubernetes/OpenShift and Linux (RHEL/Fedora preferred)
  • Experience with GitOps workflows (ArgoCD) for managing infrastructure or application configuration
  • Experience building and supporting CI/CD platform components using OpenShift Pipelines, Tekton, and GitLab
  • Experience with monitoring and alerting solutions (Splunk, Prometheus, Grafana, Catchpoint, DataDog, Splunk IM)
  • Solid understanding of Linux systems administration
  • Understanding of standard networking (TCP/IP, DNS, HTTP/TLS) and authentication protocols (LDAP)
  • Comfort with incident response, on-call responsibilities, and SRE principles (SLOs, error budgets, toil measurement)
  • Ability to work independently with minimal supervision and collaborate within an agile team
  • Contributions to open source projects
  • Experience with the Operator SDK or building Kubernetes operators
  • Design and development of software such as Kubernetes operators, webhooks, and CLI tools

Red Hat Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Red Hat and has not been reviewed or approved by Red Hat.

  • Healthcare Strength Healthcare coverage is presented as comprehensive, spanning medical, dental, and vision along with life and disability coverage. Access to HSA/FSA options and broadly positive reception of health benefits support the view that healthcare is a core strength.
  • Leave & Time Off Breadth Time-off offerings are described as generous, with substantial PTO for new hires plus additional recharge days and an end-of-year shutdown for many non-critical roles. Paid volunteer time, holidays, sick days, and supportive expectations around taking time off reinforce the breadth of leave benefits.
  • Strong & Reliable Incentives The rewards package includes performance bonuses and a recurring quarterly bonus program tied to company and individual performance. Availability of ESPP participation further adds to incentive pathways beyond base pay.

Red Hat Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Raleigh, NC
20,000 Employees
Year Founded: 1993

What We Do

At Red Hat, we connect an innovative community of customers, partners, and contributors to deliver an open source stack of trusted, high-performing solutions. We offer cloud, Linux, middleware, storage, and virtualization technologies, together with award-winning global customer support, consulting, and implementation services. Red Hat is a rapidly growing company supporting more than 90% of Fortune 500 companies.

Why Work With Us

Red Hatters freely exchange different viewpoints, contribute ideas, and solve problems together. Our love of collaboration, accountability, a sense of community, and a measure of autonomy combine to create a powerful force that fosters innovation and makes Red Hat a great place to work.

Gallery

Gallery

Similar Jobs

Microsoft Logo Microsoft

Senior Site Reliability Engineer

Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
In-Office or Remote
2 Locations
206870 Employees
120K-261K Annually

Bank of America Logo Bank of America

Senior Site Reliability Engineer

Big Data • Fintech • Mobile • Payments • Financial Services • Data Privacy
In-Office
3 Locations
208000 Employees
153K-192K Annually

Akamai Technologies Logo Akamai Technologies

Senior Site Reliability Engineer

Cloud • Security • Software • Cybersecurity
In-Office or Remote
2 Locations
10285 Employees
121K-219K Annually

Akamai Technologies Logo Akamai Technologies

Senior Site Reliability Engineer

Cloud • Security • Software • Cybersecurity
In-Office or Remote
2 Locations
10285 Employees
121K-219K Annually

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account