We're a fast-moving scale-up and our infrastructure needs to keep pace. That’s why we're looking for a Senior Site Reliability Engineer to own the reliability, automation, and security of our GCP platform, and to set the bar for how we run production as we grow.
You'll spend most of your time building and operating systems, but you'll also mentor more junior engineers and shape the practices the rest of the team works by. If you like having real ownership, moving quickly without breaking things, and turning manual toil into automation, you'll fit right in.
Responsibilities:
- Own the reliability of our production systems on GCP: define SLOs/SLIs, drive down incidents, and lead blameless postmortems.
- Build the observability stack (Prometheus, Grafana, and related tooling) so we catch problems before customers do.
- Design and maintain infrastructure as code with Terraform, no click-ops.
- Operate and scale our Kubernetes / GKE workloads.
- Build and harden CI/CD pipelines so teams can ship safely and often.
- Relentlessly automate manual work; if it's done twice by hand, it's a candidate for automation.
- Bake security into the platform: IAM, secrets management, network policy, and least-privilege by default.
- Help us meet compliance requirements as we mature, and make the secure path the easy path for other engineers.
- Own cloud cost visibility and efficiency: right-sizing, capacity planning, and scaling strategy as the company grows.
Requirements:
- 5+ years in infrastructure, SRE, DevOps, or platform engineering, including production ownership of cloud systems.
- Strong hands-on experience with GCP.
- Deep experience with Terraform.
- Production experience running Kubernetes / GKE.
- Proven track record building and maintaining CI/CD pipelines.
- Solid grasp of observability practices and tooling (Prometheus, Grafana, etc.).
- A bias toward automation and a security-conscious mindset.
- The communication skills to mentor others and influence how the team works.
- Experience scaling infrastructure in a startup or high-growth environment.
- Scripting/programming beyond config (Go, Python, etc.).
- Service mesh, GitOps (e.g., ArgoCD/Flux), or progressive delivery experience.
Recruitment Process:
- Screening call with Doriane (30 min)
- Hiring Manager interview (45 min)
- Technical onsite Interview - System Design (60 min)
- Technical Interview - Problem Solving (60 min)
- Leadership Interview (30 min)
- Fit Interview (30 min)
Skills Required
- 5+ years in infrastructure, SRE, DevOps, or platform engineering with production ownership of cloud systems
- Strong hands-on experience with GCP
- Deep experience with Terraform
- Production experience running Kubernetes / GKE
- Proven track record building and maintaining CI/CD pipelines
- Solid grasp of observability practices and tooling (Prometheus, Grafana, etc.)
- Bias toward automation and a security-conscious mindset (IAM, secrets management, network policy, least-privilege)
- Communication skills to mentor others and influence team practices
- Experience scaling infrastructure in a startup or high-growth environment
- Scripting/programming beyond config (Go, Python)
- Service mesh, GitOps (ArgoCD/Flux), or progressive delivery experience
What We Do
Founded in 2020, Alice & Bob secured $30 M funding in Series A. Today, we're recognized as one of the leaders in Quantum Computing. Our cat qubits are error corrected by design, reducing hardware requirements by up to 200 times compared to other platforms enabling fault-tolerant quantum computing at scale. Our headquarters is located in Paris, with offices in Boston, MA. We are 80+ smart and dedicated innovators, coming from 17+ countries and growing by the day. We are committed to one mission: build a useful quantum computer. Check open positions: bit.ly/joinusatAB







