devopsroles

Platform Reliability Technical Lead

Cielo Projects Open · verified Sep 29, 2026
Hybrid New York Full-time Senior Site Reliability

At a glance

Senior Site Reliability role at Cielo Projects. New York · full-time · $170,000–$220,000 base.

SalaryStated by Cielo Projects
$170,000 – $220,000
Base salary for this role, in USD per year, as published in the posting.
This role, from Cielo ProjectsOur estimate for the market: $192,650–$242,500Median $214,775
In line with the market median of $214,775 for senior site reliability roles in New York.
Worked out from 12 postings that stated pay, median $214,775. Last updated Sep 29, 2026.
RoleSite Reliability
SenioritySenior
LocationNew York
WorkplaceHybrid
EmploymentFull-time
Base pay$170,000–$220,000
PostedSep 25, 2026 · 4d ago
Role brief

devopsroles summary, based on the employer's posting.

What you'll do

  • Take the final escalation for production issues across the OP Stack, private Besu L1, Overledger Gateway and Settlement Bridge
  • Co-own service levels and release decisions, helping decide when changes are ready to go live
  • Keep runbooks current, make alerts actionable and turn postmortems into concrete changes
  • Join the team's on-call rotation while remaining in a hands-on technical role

What you bring

  • Deep experience in reliability or production engineering, with a track record as the person others call for difficult escalations
  • Strong Kubernetes, Terraform and Helm skills, including the ability to debug issues in those systems
  • Experience with distributed systems where state must be handled carefully, and clear written documentation of how platforms work

Who this fits

This senior individual-contributor role suits an experienced reliability engineer who wants to stay technical rather than move into management. You’ll be part of the on-call rotation and work in a hybrid role based in New York.

From the employer

We're partnering with Quant, a global leader in digital transformation and technology solutions, seeking a Platform Reliability Technical Lead to join their team.

About Quant:

Almost all the money in the economy is commercial bank money, and almost none of it moves on chain. Quant is a leading provider of programmable money infrastructure, and our technology changes that: it lets banks issue, move and settle tokenized deposits around the clock, automatically and securely, while staying connected to the systems they already run on. It is already deployed in regulated environments with central and commercial banks in the UK and around the world, including work on the digital pound and the digital euro.

The Clearing House, whose networks clear and settle more than $2 trillion a day, has now selected Quant to power its On-Chain Money Initiative: a new interoperable payments network that will let financial institutions of all sizes clear and settle tokenized deposits, with connectivity to the RTP® and CHIPS® networks. Our technology provides the network's interoperability, orchestration and transaction-management layer, enabling payments that settle immediately and transactions that trigger automatically once agreed conditions are met.

Through our Tokenized Deposits-as-a-Service solution, banks of any size can join without building new infrastructure themselves. The network is expected to become available to participating institutions in the first half of 2027, with use cases across corporate treasury, liquidity management, cross-border payments and digital asset settlement. We are building the New York team that will run it.

Platform Reliability Technical Lead

When the engineer on shift cannot work out why the chain is doing what it is doing, they call you.

The chain is the network behind The Clearing House's On-Chain Money Initiative, expected to launch in the first half of 2027. You would be the final escalation across the production estate, from the OP Stack and private Besu L1 to the Overledger Gateway and Settlement Bridge. You would co-own service levels and release gating, and set how the team works: current runbooks, alerts that mean something, postmortems that change something.

This role exists so that someone excellent can stay technical rather than move into management. You would still rotate through on-call alongside the team.

You will need

  • Serious depth in reliability or production engineering. You are already the person others escalate to.
  • Kubernetes, Terraform and Helm at the level where you debug them.
  • Distributed systems where state matters and restarting is not a strategy.
  • To write things down. If the only record of how the platform works is in your head, nothing has been fixed.

Useful, not essential

  • Blockchain node operations, particularly Besu or another Ethereum client.
  • You have read consensus client source to explain a production behavior.

There is more to this than we can put in an advertisement. If it sounds like your kind of problem, apply and we will tell you the rest on a call.