Site Reliability Engineering Lead
ZenithYou'll be redirected to the original listing.
Description
Tags: Web3 Jobs • Web3 Full Time Jobs • Web3 Engineering Jobs • Web3 SRE Jobs • Cryptocurrency Developer Jobs • Blockchain Remote Jobs
About Zenith
Zenith is the first Ethereum extension for the Canton Network, enabling participants to launch fully Ethereum-compatible execution environments that integrate directly with Canton's privacy, compliance, and settlement infrastructure. With Zenith EVM, teams deploy their Ethereum applications onto Canton using familiar tooling — Solidity, Hardhat, Foundry, and MetaMask — while Zenith Stack lets institutions run a fully customized blockchain environment without building or maintaining the infrastructure themselves. Engineered for high-frequency institutional finance, these environments deliver 100,000+ TPS and sub-second finality, giving institutions and builders the foundation for interoperable, compliant, and scalable financial systems on Canton. Support for further environments, beginning with a Zenith SVM for Solana applications, is on the way.
Operating at the intersection of DeFi and TradFi, Zenith connects Ethereum's builders and liquidity with the institutional capital flowing through Canton — bringing consumer blockchains and institutional-grade ledgers together to modernize global finance.
Our team combines decades of experience across blockchain infrastructure, capital markets, and venture innovation. Zenith is a registered Super Validator within Canton's Global Synchronizer — one of 50+ Super Validators securing the network — extending the ecosystem to deliver institutional trust, compliance, and scalability to the next generation of digital assets.
What you’ll do
This is a high-impact, hands-on SRE leadership role for a senior engineer who can own Zenith’s infrastructure systems today and build the operational foundation for a world-class SRE team tomorrow. You will lead all reliability, performance, and operational functions across Zenith’s internal systems, Zenith Stack environments, and validator infrastructure.
Ideally, you have built and are currently managing systems that simply cannot have any downtime. We are looking for someone who can initially build our system reliability single-handedly and later on lead a team to ensure our clients are serviced at all times and that the system is further refined.
In this role, the following aspects are essential:
- Own production reliability: Ensure strong operational reliability and resilience practices while partnering with engineering teams responsible for software quality.
- Manage and evolve cloud environments: Build, automate, and maintain scalable cloud systems supporting multi-layered blockchain execution and off-chain runtimes.
- Build SRE processes: Define and implement incident response, observability, monitoring, alerting, on-call rotation, runbooks, and reliability KPIs.
- Lead operational excellence: Establish systems for versioning, release management, configuration, secrets, environment orchestration, and reproducibility.
- Enable client Zenith Stack deployments: Work closely with engineering and product teams to support the deployment, scaling, and reliability of client-tailored Zenith Stacks.
- Collaborate on protocol operations: Partner with core engineering to support validator operations, network participation, upgrades, and secure infrastructure.
- Drive automation: Identify manual inefficiencies and replace them with robust automation for …
Related remote jobs