Head of Platform Reliability

Posted 12 Days Ago
Be an Early Applicant
New York, NY, USA
Hybrid
165K-210K Annually
Expert/Leader
Software • Industrial • Automation • Manufacturing
The Role
Own the reliability of production blockchain infrastructure supporting regulated institutional money. Define service levels, error-budget policies, release authority, and operational standards. Build and lead a site reliability or production engineering team, including a sustainable continuous-coverage rotation. Diagnose stateful chain-platform failures involving infrastructure and protocol behavior. Required experience includes production reliability leadership, practical error-budget management, continuous on-call coverage, Kubernetes, Terraform, and hands-on observability troubleshooting.
Summary Generated by Built In
Company Description

We're partnering with Quant, a global leader in digital transformation and technology solutions, seeking a Head of Platform Reliability to join their team.

About Quant:

Almost all the money in the economy is commercial bank money. Deposits, sitting on bank balance sheets, moving through payment systems designed decades before anyone had a reason to make money programmable. Almost none of it moves on chain.

Quant builds the infrastructure that changes that. Our technology lets a bank issue, move and settle its own money on programmable rails while staying connected to the systems it already runs on. That constraint is the reason this has taken as long as it has.

Central banks and commercial banks have built on our platform, including work on the digital pound and the digital euro. We are now building our team in New York.

 

Job Description

Head of Platform Reliability

Chain infrastructure carrying regulated money, in production, for institutions that hold it to the highest possible standard.

You would own its reliability. That means defining the service levels, setting the error budget policy, and holding the authority to stop a release when the budget is spent — a call that stands regardless of seniority.

You would build and lead the engineering team that runs it, on a rotation you design. We would rather you designed it properly than inherited something and patched it.

Chain platforms fail differently from the systems most reliability engineers have run. State is expensive, restarts are not a strategy, and a surprising amount of what looks like infrastructure turns out to be protocol behaviour. If that sounds interesting rather than daunting, we should talk.

Qualifications

You will need

  • To have led site reliability or production engineering somewhere failure had consequences.
  • To have run error budgets in practice — used one to stop a release, and defended the call
  • Honest experience of continuous coverage — what burns people out, what does not, and the difference between a rotation on paper and one people can live with
  • Kubernetes, Terraform, and an observability stack you have actually debugged, rather than configured.

Useful, not essential

  • Blockchain nodes in production.
  • Banking, payments, or somewhere else an auditor asked you to prove something.

Additional Information

There is more to this than we can put in an advertisement. If it sounds like your kind of problem, apply and we will tell you the rest on a call.

Skills Required

  • Experience leading site reliability or production engineering in an environment where failures had significant consequences
  • Practical experience running error budgets, including stopping and defending a release decision
  • Experience designing and managing sustainable continuous-coverage or on-call rotations
  • Hands-on experience with Kubernetes
  • Hands-on experience with Terraform
  • Hands-on experience debugging an observability stack
  • Production experience operating blockchain nodes
  • Experience in banking, payments, or an environment with audit requirements
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
4,000 Employees
Year Founded: 2005

What We Do

Cielo Projects (operating as Velotic™) is a leading independent industrial software company providing data‑driven solutions that improve manufacturing efficiency, productivity, and operational insight. Serving customers across manufacturing, oil & gas, utilities, and infrastructure, the company leverages a portfolio anchored by Proficy, Kepware, and ThingWorx to support global industrial operators, with a primary focus on AI‑enabled manufacturing and industrial software.

Similar Jobs

Dynatrace Logo Dynatrace

Senior Mainframe Developer - HLASM

Artificial Intelligence • Big Data • Cloud • Information Technology • Software • Big Data Analytics • Automation
Remote or Hybrid
United States
5600 Employees
161K-241K Annually

Morningstar Logo Morningstar

Assistant Vice President, Marketing Automation & Analytics

Artificial Intelligence • Big Data • Enterprise Web • Fintech • Software • Financial Services
Hybrid
New York, NY, USA
11500 Employees
95K-300K Annually

Morningstar Logo Morningstar

Marketing Manager

Artificial Intelligence • Big Data • Enterprise Web • Fintech • Software • Financial Services
Hybrid
New York, NY, USA
11500 Employees
100K-500K Annually
Hybrid
New York, NY, USA
289097 Employees

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account