Skip to main content
Telecommut

Manager, Site Reliability Engineering

Teli Labs

Full time Posted: 3 hours ago Software Development

Hiring from: Canada

Back to jobs

New

Manager, Site Reliability Engineering

Vancouver, BC

LayerZero

The Future is Omnichain.

Founded in 2021, LayerZero’s vision is to create a community of cross-chain developers, building dApps that are no longer constrained by individual blockchain capabilities. With LayerZero's simple, generic messaging protocol, builders will develop cross-chain dApps designed to unify the power of individual blockchains.

We are funded by the best investors in the world including:

a16z, Sequoia, PayPal, Binance Ventures, Coinbase Ventures, Uniswap Labs, Circle Ventures, Delphi Digital, and many more.

About The Role

At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering, dedicated to crafting and maintaining large-scale, resilient systems. Our goal is to ensure that all LayerZero services — ranging from critical internal systems to those external users interact with — are reliable, meet the uptime expectations of our users, and continuously evolve at a swift pace. Our SRE professionals will monitor our system's capacity and performance to uphold these standards.

As Manager of SRE, you'll lead a team of engineers responsible for the reliability, performance, and scalability of our blockchain node infrastructure and platform services — while staying technically sharp enough to guide architecture decisions and jump into critical incidents. You'll balance people leadership with hands-on technical judgment, shaping how the team works, grows, and scales alongside LayerZero.

What You'll Do

  • Lead and develop a team of SREs — setting technical direction, growth plans, and performance expectations.
  • Own the reliability strategy for blockchain node infrastructure across a variety of DLTs, including SLOs, capacity planning, and incident response.
  • Partner with Engineering leadership and Product/Platform teams to align reliability investments with business priorities.
  • Drive infrastructure-as-code practices, with a focus on Kubernetes and Helm at scale.
  • Establish and continuously improve on-call structure, incident detection/triage automation, and postmortem culture.
  • Stay hands-on: review designs, dig into complex incidents, and set the technical bar for the team.

About You

  • Bachelor's degree in Computer Science, similar technical field of study, or equivalent practical experience.
  • 6+ years in SRE, DevOps, or infrastructure engineering, including 2+ years directly managing or leading a technical team.
  • Deep familiarity with blockchain node infrastructure (validator/full/archive nodes, RPC optimization, etc.).
  • Strong proficiency in TypeScript or Golang, with the judgment to know when to write code vs. delegate.
  • Advanced knowledge of Unix/Linux internals and distributed systems / high-availability design.
  • 3+ years running Kubernetes in production, including Helm chart authoring at scale.
  • Track record building or scaling an on-call/incident response process.
  • Excellent communication skills — able to represent the team to leadership and hire/retain strong engineers.

Equal Opportunity Employer

LayerZero Labs is committed to fostering a diverse and inclusive workplace. LayerZero Labs is an equal opportunity employer and does not discriminate on the basis of race, national origin, religion, gender, gender identity, sexual orientation, marital status, protected veteran status, disability, age, or any other legally protected status.

Create a Job Alert

Interested in building your career at LayerZero Labs? Get future opportunities sent straight to your email.

Apply for this job

indicates a required field

First Name*

Last Name*

Preferred First Name

Email*

Phone

Country

Phone

Resume/CV

Attach

Enter manually

Accepted file types: pdf, doc, docx, txt, rtf

Cover Letter

Attach

Enter manually

Accepted file types: pdf, doc, docx, txt, rtf

Are you able to work onsite at our Vancouver, BC office? If not currently in Vancouver, are you open to relocating?*

Select...

LinkedIn Profile

Website

Telegram

How to apply

To apply for this job you need to login. If you don't have an account yet, please register.

Post a resume

Similar jobs

Laborer

Treasure Chest Casino

Boyd Gaming Corporation has been successful in gaming jurisdiction in which we operate in the United States and is one of the premier casino entertainment companies in the United States. Never content to rest upon our successes, we will continue...

Full time Posted: 2 hours ago Hiring from: United States

Overview Metropolis is an artificial intelligence company for the real world. We use computer vision to enable checkout-free parking experiences. So there’s no fumbling with tickets, machines, apps, or credit cards. You just “drive in and drive out.” We are...

Full time Posted: 2 hours ago Hiring from: United States

About Xanadu Xanadu’s mission is to build quantum computers that are useful and available to people everywhere. At Xanadu, we are learners, innovators, researchers, collaborators and problem solvers. We are creating something that has never been built before. What we...

Full time Posted: 3 hours ago Hiring from: Canada

About Clutch We’re on a mission to reinvent the way people buy, sell, and own cars. Are you game? Clutch is Canada’s largest online used car retailer, delivering a seamless, hassle-free car-buying experience to drivers everywhere. Customers can browse hundreds...

Full time Posted: 3 hours ago Hiring from: Canada