Principal Site Reliability Engineer - Azure Red Hat OpenShift

Posted 2 Days Ago
Be an Early Applicant
5 Locations
Remote
Senior level
Cloud • Information Technology • Internet of Things • Software • Consulting • Infrastructure as a Service (IaaS) • Automation
Creating better technology the open source way
The Role
The Principal Site Reliability Engineer will develop and optimize OpenShift services, manage large-scale distributed systems, mentor engineers, and enhance monitoring systems, participating in incident response and continuous improvement processes.
Summary Generated by Built In

The Red Hat Site Reliability Engineering (SRE) team is looking for a Principal  Site Reliability Engineer to join us. In this role, you will develop, scale, and operate our OpenShift managed cloud services. OpenShift is Red Hat’s enterprise Kubernetes distribution. As an SRE you will contribute to running OpenShift at scale by enabling customer self-service, making our monitoring system more sustainable, and eliminating work through automation.

On the SRE team, you will have the opportunity to influence the complex challenges of scale which are unique to Red Hat Managed Cloud Services, while using your skills in coding, operations, and large-scale distributed system design.

What you will do
  • Lead the development of code and automation scripts to optimize the scalability, reliability, and performance of services.

  • Conduct thorough code reviews and implement best practices in software development to maintain a high-quality codebase.

  • Mentor and guide junior engineers, fostering a culture of continuous learning and improvement within the team.

  • Design and implement advanced monitoring and alerting systems to proactively detect and resolve issues.

  • Coordinate and lead complex incident response procedures, ensuring timely resolution and thorough postmortems.

  • Interface with internal stakeholders and external cloud providers to architect, build, and maintain fault-tolerant systems.

  • Manage large-scale, distributed systems, focusing on minimizing downtime and improving system resilience.

  • Participate in an on-call rotation and provide leadership during critical incidents to ensure 24x7x365 production support.

  • Lead the continuous enhancement of the SRE team's processes, tools, and methodologies to support the evolving needs of the service.

What you will bring
  • 5+ years of software engineering experience using object-oriented languages; Golang is preferred

  • Extensive experience managing Linux-based systems in a public cloud such as AWS, GCP, or Azure

  • Proficient experience with enterprise systems monitoring; knowledge of Prometheus is preferred

  • Extensive experience with enterprise configuration management such as Ansible, Puppet, or Chef

  • Proficient experience delivering hosted cloud services

  • 1+ year experience with container-related technologies like Docker or Kubernetes

  • Experience delivering hosted cloud services

  • Experience with containers on Linux

  • Solid understanding of standard TCP/IP networking and common protocols like DNS and HTTP

  • Good verbal and written communication skills in English

About Red Hat

Red Hat is the world’s leading provider of enterprise open source software solutions, using a community-powered approach to deliver high-performing Linux, cloud, container, and Kubernetes technologies. Spread across 40+ countries, our associates work flexibly across work environments, from in-office, to office-flex, to fully remote, depending on the requirements of their role. Red Hatters are encouraged to bring their best ideas, no matter their title or tenure. We're a leader in open source because of our open and inclusive environment. We hire creative, passionate people ready to contribute their ideas, help solve complex problems, and make an impact.

Inclusion at Red Hat
Red Hat’s culture is built on the open source principles of transparency, collaboration, and inclusion, where the best ideas can come from anywhere and anyone. When this is realized, it empowers people from different backgrounds, perspectives, and experiences to come together to share ideas, challenge the status quo, and drive innovation. Our aspiration is that everyone experiences this culture with equal opportunity and access, and that all voices are not only heard but also celebrated. We hope you will join our celebration, and we welcome and encourage applicants from all the beautiful dimensions that compose our global village.

Equal Opportunity Policy (EEO)
Red Hat is proud to be an equal opportunity workplace and an affirmative action employer. We review applications for employment without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, citizenship, age, veteran status, genetic information, physical or mental disability, medical condition, marital status, or any other basis prohibited by law.


Red Hat does not seek or accept unsolicited resumes or CVs from recruitment agencies. We are not responsible for, and will not pay, any fees, commissions, or any other payment related to unsolicited resumes or CVs except as required in a written contract between Red Hat and the recruitment agency or party requesting payment of a fee.

Red Hat supports individuals with disabilities and provides reasonable accommodations to job applicants. If you need assistance completing our online job application, email [email protected]. General inquiries, such as those regarding the status of a job application, will not receive a reply.

Top Skills

Ansible
AWS
Azure
Chef
Docker
GCP
Go
Kubernetes
Linux
Prometheus
Puppet
Tcp/Ip
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Raleigh, NC
20,000 Employees
Year Founded: 1993

What We Do

At Red Hat, we connect an innovative community of customers, partners, and contributors to deliver an open source stack of trusted, high-performing solutions.

We offer cloud, Linux, middleware, storage, and virtualization technologies, together with award-winning global customer support, consulting, and implementation services. Red Hat is a rapidly growing company supporting more than 90% of Fortune 500 companies.

Why Work With Us

Red Hatters freely exchange different viewpoints, contribute ideas, and solve problems together. Our love of collaboration, accountability, a sense of community, and a measure of autonomy combine to create a powerful force that fosters innovation and makes Red Hat a great place to work.

Gallery

Gallery

Similar Jobs

GitLab Logo GitLab

Account Executive

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
28 Locations

GitLab Logo GitLab

Senior Product Design Manager, AI

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
30 Locations
155K-240K Annually

ServiceNow Logo ServiceNow

Architect

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Remote or Hybrid
Milan, ITA

Workiva Logo Workiva

Sales Manager

Artificial Intelligence • Cloud • Fintech • Professional Services • Software • Analytics • Financial Services
Remote
Italy

Similar Companies Hiring

Standard Template Labs Thumbnail
Software • Information Technology • Artificial Intelligence
New York, NY
10 Employees
PRIMA Thumbnail
Travel • Software • Marketing Tech • Hospitality • eCommerce
US
15 Employees
Rain Thumbnail
Web3 • Payments • Infrastructure as a Service (IaaS) • Fintech • Financial Services • Cryptocurrency • Blockchain
New York, NY
40 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account