Staff Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
3 Locations
Hybrid
181K-263K Annually
Expert/Leader
Big Data • Cloud • Marketing Tech • Social Impact • Software
LiveRamp makes it safe and easy for companies to use data effectively—and needs brilliant people to make it happen.
The Role
Defines organization-wide SRE strategy and reliability standards; architects globally distributed infrastructure; leads Kubernetes, cloud, database, automation, observability, FinOps, and production readiness initiatives. Serves as the final escalation point for major incidents, drives postmortems and architecture reviews, mentors staff engineers, influences cross-functional technical decisions, and supports due diligence for acquisitions and partnerships.
Summary Generated by Built In

LiveRamp is shaping the future of responsible data collaboration between the world’s leading brands, retailers, financial services providers, and healthcare innovators. As consumers embrace new AI-driven experiences, the LiveRamp data collaboration network exponentially expands the breadth and accuracy of the data on which marketing AI capabilities operate, powering deeper customer insight and measurable performance on a global scale. LiveRamp is headquartered in San Francisco, California, with offices worldwide. Learn more at LiveRamp.com.

The Global SRE team is responsible for owning and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability engineering across LiveRamp's global infrastructure. This is a senior individual contributor role with organization-wide scope—you will define and own the SRE strategy, influence product and platform architecture decisions, and raise the engineering bar across multiple teams and regions.

 

You Will:

  • Define and own the SRE strategy across the organization—SLOs/SLAs, error budgets, and operational excellence frameworks

  • Oversee automation of critical areas to mitigate risk and align with engineering priorities

  • Develop and own some of the most complex software infrastructure spanning multiple products and services

  • Drive engineering-wide system design, automation, and performance optimization standards

  • Lead distributed systems architecture reviews and kickoffs across engineering teams

  • Drive high-quality API and interface designs across multiple teams

  • Drive overall architecture improvements across multiple products and services

  • Shape product and service design vision inside engineering, anticipating the unexpressed needs of internal teams

  • Understand global industry and market trends and apply them to deliver superior infrastructure solutions

  • Maintain a complete view of LiveRamp products and how SRE OKRs support the product roadmap

  • Contribute technical due diligence to M&A evaluations of potential acquisitions and partnerships

  • Serve as the escalation point of last resort for high-impact production incidents globally, leading postmortems with org-wide action items

  • Establish and enforce production readiness standards across engineering

  • Champion FinOps strategy across Kubernetes, cloud resources, and database infrastructure

  • Mentor Staff Engineers and provide technical feedback and guidance

  • Hold peers accountable for on-time, quality delivery

  • Represent LiveRamp's best interests in the broader technology ecosystem

 

About you:

  • B.S./M.S. in Computer Science, Software Engineering, or equivalent

  • 10+ years in SRE, production engineering, or platform engineering; 3+ years at senior or staff level

  • Expert in Infrastructure as Code (Terraform) at scale across multi-environment, multi-team setups

  • Proven experience designing and operating highly available, globally distributed systems

  • Deep Kubernetes expertise: internals, autoscaling, multi-tenant workload management, and rightsizing

  • Advanced experience with real-time and NoSQL databases (SingleStore, ScyllaDB, Cassandra, DynamoDB)

  • Strong proficiency in Python and/or Go; able to build production-grade internal tooling adopted across teams

  • Expertise in observability engineering—SLOs, SLI pipelines, and high-signal alerting systems

  • Deep FinOps expertise: cost attribution, reserved capacity strategy, and cloud cost governance at scale

  • Experience maturing CI/CD platforms (Jenkins, CircleCI, or equivalent) for multiple engineering teams

  • Strong cloud security background: IAM, network segmentation, secrets management, SOC 2 / ISO 27001 (GCP and/or AWS)

  • Peer-recognized expert in a field relevant to LiveRamp's technology ecosystem

  • Exceptional communicator across both technical and executive audiences

  • Proven ability to lead without authority and influence across engineering organizations

 

Nice to Have:

  • Experience building or operating multi-region active-active architectures

  • Contributions to open source observability or infrastructure tooling

  • Experience with chaos engineering frameworks (Gremlin, Chaos Monkey, or equivalent)

  • Prior experience in a staff-plus IC role or as a technical lead for a global SRE organization

  • Familiarity with LLMs and AI-assisted development workflows, including tools such as Claude Code; experience applying agentic software development patterns to automate infrastructure tasks, incident response, or operational toil

The approximate annual base compensation range is $181,000 to $263,000. The actual offer, reflecting the total compensation package and benefits, will be determined by a number of factors including the applicant's experience, knowledge, skills, and abilities, geography, as well as internal equity among our team.

 

We use automated decision systems (ADS) as part of our recruitment and hiring process. If you require an accommodation or believe that the use of an ADS may create a barrier to your application or participation in the hiring process due to a disability or other protected characteristic, please let us know. We are committed to providing reasonable accommodations and ensuring an equitable hiring experience for all candidates.

To all recruitment agencies: LiveRamp does not accept agency resumes. Please do not forward resumes to our jobs alias, LiveRamp employees or any other company location. LiveRamp is not responsible for any fees related to unsolicited resumes.

Skills Required

  • B.S. or M.S. in Computer Science, Software Engineering, or equivalent
  • 10+ years of experience in SRE, production engineering, or platform engineering
  • 3+ years of experience at senior or staff level
  • Expertise in Infrastructure as Code using Terraform at scale across multi-environment, multi-team setups
  • Experience designing and operating highly available, globally distributed systems
  • Deep Kubernetes expertise, including internals, autoscaling, multi-tenant workload management, and rightsizing
  • Advanced experience with real-time and NoSQL databases, including SingleStore, ScyllaDB, Cassandra, or DynamoDB
  • Strong proficiency in Python and/or Go, with ability to build production-grade internal tooling
  • Expertise in observability engineering, including SLOs, SLI pipelines, and high-signal alerting systems
  • Deep FinOps expertise, including cost attribution, reserved capacity strategy, and cloud cost governance
  • Experience maturing CI/CD platforms such as Jenkins or CircleCI for multiple engineering teams
  • Strong cloud security background in IAM, network segmentation, secrets management, SOC 2, and ISO 27001 across GCP and/or AWS
  • Peer-recognized expertise in a field relevant to LiveRamp's technology ecosystem
  • Exceptional communication skills across technical and executive audiences
  • Ability to lead without authority and influence across engineering organizations
  • Experience building or operating multi-region active-active architectures
  • Contributions to open-source observability or infrastructure tooling
  • Experience with chaos engineering frameworks such as Gremlin or Chaos Monkey
  • Prior staff-plus individual contributor experience or technical leadership of a global SRE organization
  • Familiarity with LLMs and AI-assisted development workflows, including Claude Code and agentic software development patterns

LiveRamp Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about LiveRamp and has not been reviewed or approved by LiveRamp.

  • Retirement Support — Retirement programs include a dollar‑for‑dollar 401(k) match up to 6% with no vesting. Feedback suggests this level of matching is stronger than many peers and a clear financial pillar.
  • Healthcare Strength — Health coverage offers multiple employer‑verified medical, dental, and vision options, alongside disability coverage and other core protections. Feedback suggests the breadth and depth of plan choices are well‑regarded in the U.S.
  • Parental & Family Support — Paid parental bonding leave for all new parents is available, with additional resources such as backup care and family‑forming support. Feedback suggests these offerings materially support caregiving needs and work‑life balance.

LiveRamp Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA
1,190 Employees
Year Founded: 2011

What We Do

Those who want to make a lasting impact in all that they do will find a home at LiveRamp—an inclusive, collaborative environment where exceptional talent is nurtured and championed. If you love collaborating with great people to solve complex problems and champion innovative ideas, view our career opportunities and consider joining our team!

Similar Jobs

Hybrid
5 Locations
6000 Employees
194K-267K Annually

AlphaSense Logo AlphaSense

Site Reliability Engineer

Artificial Intelligence • Fintech • Machine Learning • Natural Language Processing • Business Intelligence
Remote or Hybrid
United States
2000 Employees
150K-225K Annually

MongoDB Logo MongoDB

Site Reliability Engineer

Big Data • Cloud • Software • Database
Easy Apply
Remote or Hybrid
10 Locations
5550 Employees
127K-249K Annually

MongoDB Logo MongoDB

Site Reliability Engineer

Big Data • Cloud • Software • Database
Easy Apply
Remote or Hybrid
6 Locations
5550 Employees
126K-248K Annually

Similar Companies Hiring

Ford Energy Thumbnail
Automotive • Software • Energy • Utilities • Manufacturing • Renewable Energy
US
55 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
70 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account