Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
Hiring Remotely in România
Remote
Senior level
Information Technology • Marketing Tech • Social Media
The Role
Operate and scale GoDaddy’s OpenStack-based global hosting infrastructure across compute, networking, and storage. Lead migration tooling, validation, and rollback strategies; automate operational tasks with Python and Puppet; improve observability; participate in on-call and incident response; conduct blameless post-incident reviews; review designs, document systems, mentor junior SREs, and support AI-assisted operational workflows.
Summary Generated by Built In

Location Details: 

At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​

Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings.  

About The Team....
Global Compute runs Optimised Hosting, GoDaddy's global platform for all customer hosting products. Squad R is the engineering team responsible for operating, scaling, and continuously improving the OpenStack-based clouds that power that platform. We treat reliability as an engineering problem: we automate toil away, we plan capacity ahead of demand, and we instrument everything so that we understand our systems before they surprise us. As an SRE III on the team, you'll be a senior technical contributor who others lean on for the hard problems.

What you'll get to do...

  • Operate and scale GoDaddy's cloud infrastructure, including our OpenStack-based hosting platform. You'll troubleshoot and improve services spanning compute, networking, and storage in large-scale production environments.
  • Drive the OpenStack migration. Help move customer hosting workloads onto the platform safely — designing and executing migration tooling, validation, and rollback strategies that protect customer experience.
  • Work within a large-scale global hosting environment supporting thousands of servers and customer workloads across multiple regions.
  • Eliminate toil through automation. Build and maintain automation in Python and Puppet to replace manual operational work. Treat repeated manual effort as a bug to be fixed.
  • Strengthen observability. Improve monitoring, alerting, and dashboards so that signal reaches the right engineer at the right time, and so that we can reason about system behavior from data.
  • Participate in on-call and incident response. Take a fair share of the on-call rotation, lead incident response when you're the responder, and drive the blameless post-incident process that turns failures into permanent fixes.
  • Raise the engineering bar. Review peers' code and designs, document systems and runbooks clearly, and mentor SRE I/II engineers.
  • Contribute to the AI/MCP initiative. Help bring AI-assisted workflows and internal MCP tooling into our operations so internal customers can resolve problems and incidents faster.
  • Participate in a shared on-call rotation (after onboarding) and help lead incident response activities when needed.

Your experience should include...

  • 5+ years in SRE, infrastructure, platform, or systems engineering roles operating production systems at scale.
  • Strong Linux systems fundamentals — networking, storage, processes, performance troubleshooting.
  • Proficiency in Python for automation and tooling (writing maintainable, tested code — not just scripts).
  • Hands-on experience operating distributed systems and diagnosing issues across service boundaries.
  • Experience with infrastructure-as-code and configuration management (Puppet, Ansible, or equivalent).
  • Comfort owning production reliability: on-call experience, incident response, and a track record of reducing toil through automation.
  • Clear written and verbal communication — able to document systems, write post-incident reviews, and collaborate across a distributed, remote team.

You might also have...

  • Experience operating OpenStack (Nova, Neutron, Ceph) or comparable cloud infrastructure platforms
  • Experience with Docker, Kolla, or containerised infrastructure
  • Experience supporting large distributed systems at scale
  • Experience operating Ceph or other software-defined storage at scale.
  • Exposure to OpenStack migration or cloud-migration programs.
  • Interest in applying AI/LLM tooling to operational workflows.

We encourage you to apply even if your experience or skillset doesn’t align perfectly with every requirement. We value a wide range of backgrounds and transferable skills, and we are excited to support learning and growth.

We've got your back...  We offer a range of total rewards that may include paid time off, retirement savings (e.g., 401k, pension schemes), bonus/incentive eligibility, equity grants, participation in our employee stock purchase plan, competitive health benefits, and other family-friendly benefits including parental leave. GoDaddy’s benefits vary based on individual role and location and can be reviewed in more detail during the interview process.

About us...  GoDaddy is empowering everyday entrepreneurs around the world by providing the help and tools to succeed online, making opportunity more inclusive for all. GoDaddy is the place people come to name their idea, build a professional website, attract customers, sell their products and services, and manage their work. Our mission is to give our customers the tools, insights, and people to transform their ideas and personal initiative into success. To learn more about the company, visit About Us. 

At GoDaddy, we know diverse teams build better products—period. Our people and culture reflect and celebrate that sense of diversity and inclusion in ideas, experiences and perspectives. But we also know that’s not enough to build true equity and belonging in our communities. That’s why we prioritize integrating diversity, equity, inclusion and belonging principles into the core of how we work every day—focusing not only on our employee experience, but also our customer experience and operations. It’s the best way to serve our mission of empowering entrepreneurs everywhere, and making opportunity more inclusive for all. To read more about these commitments, as well as our representation and pay equity data, check out our Diversity and Pay Parity annual report which can be found on our Diversity Careers page.

We also embrace our diverse culture and offer a range of Employee Resource Groups (Culture). Have a side hustle? No problem. We love entrepreneurs! Most importantly, come as you are and make your own way. 

GoDaddy is proud to be an equal opportunity employer. GoDaddy will consider for employment qualified applicants with criminal histories in a manner consistent with local and federal requirements.Refer to our full EEO policy.

Our recruiting team is available to assist you in completing your application. If they could be helpful, please reach out to [email protected]. 

GoDaddy doesn’t accept unsolicited resumes from recruiters or employment agencies.

Skills Required

  • 5+ years of experience in SRE, infrastructure, platform, or systems engineering roles operating production systems at scale
  • Strong Linux systems fundamentals, including networking, storage, processes, and performance troubleshooting
  • Proficiency in Python for maintainable, tested automation and tooling
  • Hands-on experience operating distributed systems and diagnosing issues across service boundaries
  • Experience with infrastructure-as-code and configuration management, such as Puppet, Ansible, or equivalent
  • On-call and incident response experience, with a track record of reducing toil through automation
  • Clear written and verbal communication skills for documentation, post-incident reviews, and distributed-team collaboration
  • Experience operating OpenStack, including Nova, Neutron, and Ceph, or comparable cloud infrastructure platforms
  • Experience with Docker, Kolla, or containerized infrastructure
  • Experience supporting large distributed systems at scale
  • Experience operating Ceph or other software-defined storage at scale
  • Exposure to OpenStack migration or cloud-migration programs
  • Interest in applying AI/LLM tooling to operational workflows

GoDaddy Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about GoDaddy and has not been reviewed or approved by GoDaddy.

  • Healthcare Strength Comprehensive medical, dental, and vision coverage is paired with disability and life insurance, fertility support, an Employee Assistance Program, and wellness resources like gym discounts and virtual therapy. HSAs/FSAs and preventive care options further bolster the package.
  • Leave & Time Off Breadth Unlimited vacation, paid parental leave, adoption assistance, childcare subsidies, paid holidays and sick time, and volunteer time are available. Sabbaticals and wellness days provide additional flexibility for time away.
  • Equity Value & Accessibility RSUs, company stock grants, and an Employee Stock Purchase Program augment cash pay. Engineering compensation includes stock components that increase total rewards.

GoDaddy Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Tempe, AZ
10,000 Employees
Year Founded: 1997

What We Do

GoDaddy is empowering everyday entrepreneurs around the world by providing all of the help and tools to succeed online. GoDaddy is the place people come to name their idea, build a professional website, attract customers, sell their products and services, and manage their work. Our mission is to give our customers the tools, insights and the people to transform their ideas and personal initiative into success. To learn more about the company, visit About Us (https://aboutus.godaddy.net/about-us/overview/default.aspx.)

Why Work With Us

We’re the world’s largest web services platform. Our mission is to make opportunity more inclusive for all and fuel a new generation of entrepreneurial endeavors — commercial, civic, creative. Join our diverse collective of 9k+ employees across 47 global locations.

Gallery

Gallery

Similar Jobs

In-Office or Remote
6 Locations
174 Employees

ScorePlay Logo ScorePlay

Senior Platform Engineer

Artificial Intelligence • Digital Media • Software • Sports
In-Office or Remote
27 Locations
70 Employees

P2P.org Logo P2P.org

Site Reliability Engineer

Information Technology
Remote
30 Locations
179 Employees

GoDaddy Logo GoDaddy

Senior Site Reliability Engineer

Information Technology • Marketing Tech • Social Media
Remote
România
10000 Employees

Similar Companies Hiring

PRIMA Thumbnail
Travel • Software • Marketing Tech • Hospitality • eCommerce
US
15 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account