Head of Platform Reliability
- Full-time
- Position Type: Permanent
- Compensation: USD 165000 - USD 210000 - yearly
Company Description
We're partnering with Quant, a global leader in digital transformation and technology solutions, seeking a Head of Platform Reliability to join their team.
About Quant:
Almost all the money in the economy is commercial bank money, and almost none of it moves on chain. Quant is a leading provider of programmable money infrastructure, and our technology changes that: it lets banks issue, move and settle tokenized deposits around the clock, automatically and securely, while staying connected to the systems they already run on. It is already deployed in regulated environments with central and commercial banks in the UK and around the world, including work on the digital pound and the digital euro.
The Clearing House, whose networks clear and settle more than $2 trillion a day, has now selected Quant to power its On-Chain Money Initiative: a new interoperable payments network that will let financial institutions of all sizes clear and settle tokenized deposits, with connectivity to the RTP® and CHIPS® networks. Our technology provides the network's interoperability, orchestration and transaction-management layer, enabling payments that settle immediately and transactions that trigger automatically once agreed conditions are met.
Through our Tokenized Deposits-as-a-Service solution, banks of any size can join without building new infrastructure themselves. The network is expected to become available to participating institutions in the first half of 2027, with use cases across corporate treasury, liquidity management, cross-border payments and digital asset settlement. We are building the New York team that will run it.
Job Description
Head of Platform Reliability
The Clearing House's On-Chain Money Initiative will let US financial institutions clear and settle tokenized deposits on chain, connected to the RTP® and CHIPS® networks. Quant powers it, and we need someone to own its reliability ahead of an expected launch in the first half of 2027.
That means the full production environment (OP Stack, a private Besu L1, Paladin nodes, the Overledger Gateway, the Settlement Bridge and the Audit Store), the service levels and error-budget policy, and the authority to stop a release when the budget is spent. You would also build and lead the team that runs it, on a three-shift rotation you design and take part in yourself.
Chain platforms fail differently: state is expensive, restarts are not a strategy, and much of what looks like infrastructure is protocol behavior. If that sounds interesting rather than daunting, we should talk.
Qualifications
You will need
- To have led site reliability or production engineering somewhere failure had consequences.
- To have run error budgets in practice: used one to stop a release, and defended the call.
- Honest experience of 24x7, multi-shift coverage. What burns people out, what does not, and the difference between a rotation on paper and one people can live with.
- Kubernetes, Terraform, and an observability stack you have actually debugged rather than configured.
Useful, not essential
- Blockchain nodes in production, particularly Besu or another Ethereum client.
- Banking, payments, or somewhere else an auditor asked you to prove something.
Additional Information
There is more to this than we can put in an advertisement. If it sounds like your kind of problem, apply and we will tell you the rest on a call.
By clicking the link above or any third-party link within this posting, you are leaving this site and going to a third-party website where the third-party website's terms and privacy policy apply