Staff Platform Engineer

Reposted 2 Days Ago
Be an Early Applicant
San Francisco, CA, USA
In-Office
198K-331K Annually
Senior level
Software
The Role
Lead cross-team platform initiatives to improve developer experience, reliability, security, and cost. Design AI-augmented platform tooling, IaC standards for Kubernetes/AWS/GCP, evolve CI/CD (Argo/GitHub Actions), drive observability (Datadog/Amplitude), own incident response and postmortems, mentor engineers, and translate product needs into platform investments.
Summary Generated by Built In

Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com.

As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do.

Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive.

Amplitude's Cloud Platform team builds the systems that every Amplitude engineer relies on every day to ship code — and we're rebuilding them for the AI era. As a Staff Platform Engineer, you'll set technical direction for the platform across teams, lead our highest-complexity and highest-leverage initiatives, and shape a platform where AI agents are first-class users alongside humans: kicking off deploys, opening pull requests against infrastructure, and triaging incidents, so a single engineer can get the throughput of a team.

You'll operate across team boundaries — partnering with product engineering, fellow Staff+ engineers, and engineering leadership to make Kubernetes and cloud infrastructure effortless across the entire engineering org. You'll build the self-service automation, shared standards, and scalable AWS and GCP infrastructure that let dozens of product teams ship faster, safer, and with less cognitive load — and you'll multiply the engineers around you while you do it.

Key Responsibilities
  • Set technical direction — shape platform and domain-level technical strategy that improves developer experience, reliability, security, and cost, and lead the high-complexity, cross-cutting initiatives that deliver it with measurable impact for the organization.
  • Drive clarity through ambiguity. Take on the most loosely-defined problems, validate the critical assumptions early, and create alignment with stakeholders across teams so others can move quickly and confidently — driving cross-team decisions to a timely close and escalating when needed.
  • Build the AI-augmented platform. Design org-wide tooling, guardrails, and policy-as-code that help every engineer get more out of AI-assisted development — infra primitives an LLM can safely reason about and PR against, automated review, and standards that hold as AI changes how code gets written.
  • Own Infrastructure-as-Code standards for Kubernetes, AWS, and GCP using Terraform, Helm, Kustomize, and emerging tooling — setting the patterns other teams adopt and making the platform consumable enough that humans and agents can safely extend it.
  • Evolve our CI/CD backbone (Argo CD / Workflows / Rollouts, GitHub Actions) into shared, well-reasoned standards that make deploys faster, safer, and easier to reason about across the engineering org.
  • Instrument and operate. Drive observability with Datadog and Amplitude, set and raise SLOs for the team, own the dashboards, and drive improvements to the shared services and dependencies that move them.
  • Serve as an escalation point for complex, cross-system P0s, lead incident response, and turn postmortems into durable platform improvements that prevent whole classes of failure.
  • Identify causes, not symptoms. Find the architectural debt that slows multiple teams and build strategies to address it, and recognize when a system must evolve to meet new requirements, be re-platformed, or be deprecated and removed.
  • Find and de-risk high-leverage bets — spot the opportunities others miss, write the proposals, get buy-in from stakeholders, and build POCs to de-risk them before committing the team.
  • Multiply the team. Mentor engineers and remove single points of dependency (including yourself) by simplifying systems others can own, raise the hiring bar and help refine hiring practices, and help the team get more leverage out of AI-assisted development.
  • Connect the platform to customers and the business by translating customer and product needs into platform investments on the roadmap, and instrumenting the feedback loops that keep the platform aligned with what teams actually need.
What We're Looking For
  • 8+ years of experience in software engineering, DevOps, or Site Reliability Engineering, with deep hands-on time in cloud infrastructure.
  • Bachelors Degree in Computer Engineering or related field 
  • A track record of leading high-complexity, cross-team infrastructure initiatives that delivered measurable improvements in reliability, developer productivity, performance, or cost.
  • Deep production experience operating Kubernetes (EKS, GKE, AKS, or on-prem) and containerized applications at meaningful scale.
  • Strong programming ability in at least one language (Golang or Python preferred) and fluency with IaC tooling (Terraform).
  • Solid command of AWS core services (EC2, EKS, IAM, VPC, ALB, S3) — GCP and/or Azure experience a plus — and networking/security fundamentals.
  • Deep familiarity with GitOps workflows and the CNCF ecosystem (Argo, Helm, Backstage, Envoy, and friends), and a point of view on where they're heading.
  • A history of setting technical standards and patterns that other teams adopted — influencing outcomes well beyond your own keyboard, often without direct authority.
  • Curiosity and conviction about AI as a force multiplier in infrastructure work — whether that's using AI-assisted development tools to ship faster or building platforms and tooling that help your teammates get more leverage from AI in their day-to-day work.
  • Strong communicator — you can break down complex topics for varied audiences, drive cross-team decisions to a timely close, and default to collaborative problem solving.
  • Comfort navigating ambiguity by validating assumptions early and using data to de-risk technical decisions — with the judgment to say no, reshape scope, and protect the team's focus when the expected impact no longer justifies the work.
  • A pragmatic, business-aligned engineering mindset, a bias for action, and a habit of continuous learning and knowledge sharing.
This role is eligible for equity, benefits and other forms of compensation.
Based on Colorado law, the following details are for individuals who will work for Amplitude in Colorado. Colorado range: $198,000 - $299,000 total target cash (inclusive of bonus or commission)
Based on legislation in New York City, the following details are for individuals who will work for Amplitude in New York City. New York City salary range: $220,000 - $331,000 total target cash (inclusive of bonus or commission)
Based on legislation in California, the following details are for individuals who will work for Amplitude in San Francisco Bay Area of California. Salary range: $220,000 - $331,000 total target cash (inclusive of bonus or commission)
Based on legislation in California, the following details are for individuals who will work for Amplitude in California outside of the San Francisco Bay Area. California salary range: $198,000 - $299,000 total target cash (inclusive of bonus or commission)
Based on legislation in Washington state, the following details are for individuals who will work for Amplitude in Washington state. Washington salary range: $198,000 - $299,000 total target cash (inclusive of bonus or commission)
Based on legislation in Washington state, the following details are for individuals who will work for Amplitude in Washington only: unlimited PTO, 10 to 13 holidays annually (will vary), medical dental and vision PPO and CDHP plans. Finally, a company sponsored 401(k) retirement plan.

By applying for this job, you acknowledge that Amplitude processes your personal data in accordance with the Amplitude Applicant Privacy Notice.

Staying Safe - Protect Yourself From Recruitment Fraud
We are aware of individuals and entities fraudulently representing themselves as Amplitude recruiters and/or hiring managers. Amplitude will never ask for financial information or payment, or for personal information such as bank account number or social security number during the job application or interview process. Any emails from the Amplitude recruiting team will come from an @amplitude.com email address. You can learn more about how to protect yourself from these types of fraud by referring to this article. Please exercise caution and cease communications if something feels suspicious about your interactions.

Skills Required

  • 8+ years in software engineering, DevOps, or Site Reliability Engineering with deep cloud infrastructure experience
  • Bachelor's degree in Computer Engineering or related field
  • Proven track record leading high-complexity, cross-team infrastructure initiatives with measurable impact
  • Production experience operating Kubernetes (EKS, GKE, AKS, or on-prem) and containerized applications at scale
  • Strong programming ability in at least one language
  • Experience with Golang or Python
  • Fluency with Infrastructure-as-Code tooling (Terraform)
  • Solid command of AWS core services (EC2, EKS, IAM, VPC, ALB, S3)
  • GCP and/or Azure experience
  • Deep familiarity with GitOps workflows and CNCF ecosystem (Argo, Helm, Backstage, Envoy)
  • Experience evolving CI/CD backbones (Argo CD/Workflows/Rollouts, GitHub Actions)
  • Observability and SLO experience using Datadog and Amplitude; owning dashboards and metrics
  • Experience leading incident response, postmortems, and driving durable platform improvements
  • History of setting technical standards and influencing cross-team adoption without direct authority
  • Strong communication skills; able to drive cross-team decisions and translate product needs into platform investments

Amplitude Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Amplitude and has not been reviewed or approved by Amplitude.

  • Fair & Transparent Compensation Pay is considered competitive and fair across many roles and levels. Public compensation ranges and stability over time reinforce a perception of above‑average total compensation.
  • Parental & Family Support Parental leave for birthing and non‑birthing parents is paired with fertility support via Carrot, adoption and surrogacy assistance, and back‑up childcare. These programs are consistently highlighted across official materials and recent job postings.
  • Wellbeing & Lifestyle Benefits Global access to Modern Health, a monthly Lifestyle Spending Account, and select One Medical availability support whole‑person wellness. Quarterly learning stipends and flexible benefits enhance everyday support.

Amplitude Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA
505 Employees
Year Founded: 2012

What We Do

Amplitude is the Digital Optimization System. Powered by the proprietary Amplitude Behavioral Graph, the Digital Optimization System enables organizations to see and predict which combination of features and actions translate to business outcomes – from loyalty to lifetime value – and intelligently adapt each experience in real-time based on these insights. Amplitude is the brain behind more than 45,000 digital products at over 1,000 enterprise customers and 23 of the Fortune 100, helping them innovate faster and smarter by answering the strategic question: "How do our digital products drive our business?"

Similar Jobs

2K Logo 2K

Platform Engineer

Gaming • Information Technology • Mobile • Software • Esports
Hybrid
Los Angeles, CA, USA
4200 Employees

ServiceNow Logo ServiceNow

Machine Learning Engineer

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Mountain View, CA, USA
29000 Employees

ServiceNow Logo ServiceNow

Software Engineer

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Santa Clara, CA, USA
29000 Employees
176K-308K Annually

ServiceNow Logo ServiceNow

Machine Learning Engineer

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Mountain View, CA, USA
29000 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account