Monitoring & Observability Engineer

Posted 3 Days Ago
Be an Early Applicant
Reading, Berkshire, England, GBR
In-Office
Mid level
Quantum Computing
The Role
Monitor live quantum cryogenic systems, triage alerts, provide first-line incident response, develop dashboards and alerting strategies, analyse incidents for root causes, collaborate to improve observability tooling, document processes, and automate operational responses to maximise uptime and reliability.
Summary Generated by Built In

At OQC, we aren’t just theorising about the future; we’re building it. Born from a philosophy of bold innovation, we’ve successfully transitioned quantum computing from an academic dream into a commercial reality. The most exciting thing is that we’re just getting started and we’ve recently closed our £260 million Series C funding round – the largest fundraise ever completed by a quantum computing company in Europe.

The Purpose

As a Monitoring and Observability Engineer, you'll help keep OQC's live quantum computing systems running at their best. By improving system visibility, developing intelligent monitoring solutions and driving operational improvements, you'll play a key role in maximising uptime and enabling world-class live quantum computing services.

The Role

Working within our Live Services team, you'll monitor the health and performance of our live cryogenic systems, respond to operational incidents and continuously improve our observability capabilities. You'll collaborate across engineering, operations, software and reliability teams to develop dashboards, refine alerting strategies and automate operational responses that improve reliability and reduce downtime.

What You'll Be Working On
  • Monitor live systems, triage operational alerts and provide first-line incident response to maximise system availability.
  • Develop and maintain dashboards that deliver clear visibility into system health, trends and operational performance.
  • Analyse recurring incidents and monitoring data to identify root causes, improve alert quality and drive preventative action.
  • Collaborate with Reliability Engineering, Software and Operations teams to enhance observability tooling and operational processes.
  • Define monitoring thresholds, alerting strategies and operational metrics that support reliable live quantum computing services.
  • Produce documentation covering monitoring processes, incident response, escalation procedures and operational handovers.
  • Contribute to automation and continuous improvement initiatives that reduce manual intervention and improve service reliability.
What We're Looking For
  • Experience working in monitoring, observability or operational support for complex technical systems.
  • Experience working with dashboards, telemetry, alarms, logs or time-series data.
  • Experience supporting uptime, incident response or on-call operational environments.
  • Strong analytical and problem-solving skills with the ability to identify trends, anomalies and operational risks.
  • Experience developing software in Python using modern development practices.
  • Experience creating and maintaining operational dashboards.
  • Excellent communication skills with the ability to collaborate effectively across technical teams.
  • Degree, HNC/HND, apprenticeship or equivalent experience in Engineering, Physics, Computer Science, Controls, Instrumentation, Data or a related technical discipline.
  • Willingness to participate in an on-call rota and travel internationally when required.
The 'Nice-to-Haves'
  • Experience with Grafana Labs or similar observability platforms.
  • Experience monitoring high-availability technical systems.
  • Understanding of incident management, root cause analysis, reliability engineering or Site Reliability Engineering (SRE) principles.
  • Experience with HTTP/REST APIs, Git, containerisation, Kubernetes or Infrastructure as Code.
  • Continuous improvement, ownership or leadership experience.
Why Join OQC

You will join a world-class team at the forefront of the next computational era. We offer a culture of bold innovation, the chance to work with unique lab infrastructure, and the opportunity to see your work redefine the limits of computation.

Learn more about our benefits and positive work culture here: https://oqc.tech/company/careers-at-oqc/

Skills Required

  • Experience working in monitoring, observability or operational support for complex technical systems.
  • Experience working with dashboards, telemetry, alarms, logs or time-series data.
  • Experience supporting uptime, incident response or on-call operational environments.
  • Strong analytical and problem-solving skills with ability to identify trends, anomalies and operational risks.
  • Experience developing software in Python using modern development practices.
  • Experience creating and maintaining operational dashboards.
  • Excellent communication skills with ability to collaborate effectively across technical teams.
  • Degree, HNC/HND, apprenticeship or equivalent experience in Engineering, Physics, Computer Science, Controls, Instrumentation, Data or related technical discipline.
  • Willingness to participate in an on-call rota and travel internationally when required.
  • Experience with Grafana Labs or similar observability platforms.
  • Experience monitoring high-availability technical systems.
  • Understanding of incident management, root cause analysis, reliability engineering or Site Reliability Engineering (SRE) principles.
  • Experience with HTTP/REST APIs, Git, containerisation, Kubernetes or Infrastructure as Code.
  • Continuous improvement, ownership or leadership experience.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Reading
129 Employees
Year Founded: 2017

What We Do

Quantum computing is poised to reshape our world by addressing the complex challenges we face today. We deliver enterprise-ready quantum solutions that will empower humanity with quantum capabilities, paving the way for a brighter future.

Similar Jobs

Expedia Group Logo Expedia Group

Android Engineer

AdTech • eCommerce • Information Technology • Travel • Generative AI
Hybrid
London, England, GBR
16000 Employees
300K-300K Annually
Hybrid
London, Greater London, England, GBR
289097 Employees
Hybrid
London, Greater London, England, GBR
289097 Employees
Hybrid
Bournemouth, Dorset, England, GBR
289097 Employees

Similar Companies Hiring

Mastercard Thumbnail
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Purchase, NY
38800 Employees
HRL Laboratories Thumbnail
Artificial Intelligence • Hardware • Software • Nanotechnology • Semiconductor • Quantum Computing • Defense
Malibu, CA
850 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account