Manager, Site Reliability Engineering
Layerzero Labs
LayerZero The Future is Omnichain.
Founded in 2021, LayerZero’s vision is to create a community of cross-chain developers, building dApps that are no longer constrained by individual blockchain capabilities. With LayerZero's simple, generic messaging protocol, builders will develop cross-chain dApps designed to unify the power of individual blockchains.
We are funded by the best investors in the world including:
a16z, Sequoia, PayPal, Binance Ventures, Coinbase Ventures, Uniswap Labs, Circle Ventures, Delphi Digital, and many more.
ABOUT THE ROLE
At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering, dedicated to crafting and maintaining large-scale, resilient systems. Our goal is to ensure that all LayerZero services — ranging from critical internal systems to those external users interact with — are reliable, meet the uptime expectations of our users, and continuously evolve at a swift pace. Our SRE professionals will monitor our system's capacity and performance to uphold these standards.
As Manager of SRE, you'll lead a team of engineers responsible for the reliability, performance, and scalability of our blockchain node infrastructure and platform services — while staying technically sharp enough to guide architecture decisions and jump into critical incidents. You'll balance people leadership with hands-on technical judgment, shaping how the team works, grows, and scales alongside LayerZero.
WHAT YOU'LL DO
- Lead and develop a team of SREs — setting technical direction, growth plans, and performance expectations.
- Own the reliability strategy for blockchain node infrastructure across a variety of DLTs, including SLOs, capacity planning, and incident response.
- Partner with Engineering leadership and Product/Platform teams to align reliability investments with business priorities.
- Drive infrastructure-as-code practices, with a focus on Kubernetes and Helm at scale.
- Establish and continuously improve on-call structure, incident detection/triage automation, and postmortem culture.
- Stay hands-on: review designs, dig into complex incidents, and set the technical bar for the team.
**ABOUT YOU
**- Bachelor's degree in Computer Science, similar technical field of study, or equivalent practical experience.
- 6+ years in SRE, DevOps, or infrastructure engineering, including 2+ years directly managing or leading a technical team.
- Deep familiarity with blockchain node infrastructure (validator/full/archive nodes, RPC optimization, etc.).
- Strong proficiency in TypeScript or Golang, with the judgment to know when to write code vs. delegate.
- Advanced knowledge of Unix/Linux internals and distributed systems / high-availability design.
- 3+ years running Kubernetes in production, including Helm chart authoring at scale.
- Track record building or scaling an on-call/incident response process.
- Excellent communication skills — able to represent the team to leadership and hire/retain strong engineers.
LayerZero Labs is committed to fostering a diverse and inclusive workplace. LayerZero Labs is an equal opportunity employer and does not discriminate on the basis of race, national origin, religion, gender, gender identity, sexual orientation, marital status, protected veteran status, disability, age, or any other legally protected status.
- ...Site Reliability Engineering Manager Home Based - APAC; Home based - EMEA Canonical is a leading provider of open-source software and operating systems for global enterprise and technology markets. Our platform, Ubuntu, is very widely used in breakthrough enterprise...SuggestedWork at officeLocal areaRemote workWork from homeWorldwide
$100 per hour
...team of platform-focused SRE engineers, providing technical mentorship... ...development, and performance management while fostering a culture of... ..., and operate their services reliably Manage observability... ...regular team and company off-sites throughout the year as well as...SuggestedWork experience placementLive inLocal areaRemote workWork from homeHome officeFlexible hours- ...Leading a team of engineers, the full-time remote Site Reliability Engineering Manager will oversee the development of platforms and infrastructure to ensure reliable and scalable services, while fostering a culture of automation and continuous improvement. Key responsibilities...SuggestedFull timeRemote work
- ...flex cards, and member engagement solutions. We partner with managed care organizations to provide innovative healthcare... ...Location: Remote (US-based candidates only) Manager, Site Reliability Engineering (SRE) Position Overview We are seeking a Manager, Site...SuggestedRemote workFlexible hours
$139.7k - $232.9k
...Wilmington Center Wilmington, DE, location with the flexibility to work from home one day per week Overview The Site Reliability Engineering (SRE) Manager leads teams responsible for the reliability, availability, performance, and operational excellence of critical...SuggestedFull timeWork from home1 day per week$276.1k - $311.4k
...defines your career! The role As SRE Manager, you'll build the Vehicle Software SRE... ...defining its charter, hiring its founding engineers, establishing the operating model, and... ...creating the technical strategy that makes reliability a first-class property of the software...Permanent employmentFull timeWork at officeWork from home$106.7k - $177.9k
...Banking organization and help drive the reliability, resiliency, and performance of the... ...depend on every day. As a Senior Software Engineer, Site Reliability Engineering (SRE), you... ...verification efforts.• Utilize source code management tools to manage and deploy...Permanent employmentFull timeWork experience placement$81.5k - $141.3k
...in more than 160 countries.JOB DESCRIPTION:Position Title: Site Reliability Engineer IITeam: CRM DevOps Employment Type: Full-TimeAbout the... ...Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We are seeking a highly skilled and mission-driven...Remote workShift work$96.8k - $145.2k
...-thinking organization, apply now.We are currently seeking a Site Reliability Engineer (Onsite Hybrid) to join our team in Plano, Texas (US-TX), United States (US).Job Responsibilities Include: Own and manage observability using New Relic (APM, infrastructure monitoring...Temporary workWork at officeRemote workFlexible hours- ...The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also...Full timeWork experience placementRemote work
- As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing AI, HPC, Kubernetes, and enterprise infrastructure environments. This role transforms newly installed hardware into production...Work at officeImmediate startWorldwideShift work
$115.5k - $164.8k
...company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will... ...previously required human intervention with reliable, tested automation. You will also... ...years of applicable experience.Experience managing cloud platforms such as Azure, AWS, or similar...Work experience placementWork at officeRemote work$87.12k - $151.25k
...thinking organization, apply now.We are currently seeking a Site Reliability Engineer- REMOTE - Onsite Training to join our team in Memphis,... ...experience in troubleshooting using APM (Application Performance Management) tools like Datadog and Dynatrace.4+ years' experience...Full timeTemporary workWork experience placementWork at officeRemote workFlexible hours- Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer... ...release decisions, prioritize reliability initiatives, and manage operational risk.Design and maintain observability...Local areaRemote workFlexible hoursShift work
$130k - $180k
...building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference... ...and AI R&D. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. This is a remote position...Temporary workImmediate startRemote work$262k - $364k
...infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and... ...Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines... ...SRE team, you’ll have the opportunity to manage the complex challenges of scale which are...- ...Senior Site Reliability Engineer Bentley Systems Location Exton or Philadelphia, PA (Hybrid - 3 times a week in-office) Position Summary... ...business operations for our users. Responsibilities Manage, implement, and improve automation (CI/CD Infrastructure)...Casual workWork at officeWorldwide
$140k - $180k
...Learn more at Opportunity We’re looking for a Senior Site Reliability Engineer to help build and scale a high-impact SRE function. You’ll... ...~ Deep understanding of observability, incident management, and system performance ~ Proficiency in at least one...Work experience placementLocal areaRemote workVisa sponsorshipWork visa- ...customers rely on us in the moments that matter. Engineering delivers on that promise. The Senior Site Reliability Engineer is responsible for ensuring our SaaS... ...SRE’s at DFIN take on availability, performance, managing change, monitoring, response and are guardians...Work experience placementRemote workFlexible hours
- ...Site Reliability Engineer Company: GitLab Work Type: Remote Employment: Full Time Location: CA, US Seniority: Senior Level Technologies: Terraform, Ansible, Kubernetes, Go, Ruby, Jsonnet, Prometheus, ELK, Grafana Requirements: Senior-level SRE with strong Terraform/IaC...Full timeRemote work
- ...encourage you to apply. The Role As a Senior Platform Engineer, you are a champion for DevOps and SRE culture and industry... ...met. \n What You Will Be Doing Improving production reliability and system resilience within an SRE scoped team Championing...Remote workFlexible hours
- ...Engineering, Product, Design, and Marketing Engineering Compensation ~ Zone 1 Base... ...collaborative workspaces, Mail’s inbox management, and Go, the proactive AI assistant that... ...for building software to ensure the reliability of our back-end systems, working with engineers...WorldwideHome officeFlexible hours
$141.8k - $195k
...we give customers the choice, control, and flexibility to manage and analyze telemetry for both humans and agents, so they can... ...You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all...Temporary workRemote work$166.3k - $238.3k
...technology that simply works. The SRE Engineering Enablement org for our Network Platform... ..., build tools, code review, artifact management, CI, education, and documentation. Our... ...engineers at Cisco.YOUR IMPACT As a Lead Site Reliability Engineer, you will architect, build,...Full timeTemporary workLocal areaRemote workFlexible hours- ...0m series C round. SUMMARY We're hiring a Senior Site Reliability Engineer to design, implement, and maintain our cloud and on-premise... ...best practices, and high-availability standards. ~ Manage and optimize Kubernetes cluster environments (Helm, ArgoCD,...Local areaRemote work
- ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis...Full timeRemote work
$180k - $230k
...re looking for a Senior SRE to own the reliability, scalability, and observability of our... ...ll work closely with platform and data engineering to keep high-throughput, data-intensive... ...toil — deployment pipelines, capacity management, self-healing systems Partner with engineering...Work at officeLocal areaImmediate startRemote work3 days per week- ...A senior Site Reliability Engineer will join an established infrastructure function responsible for highly available, security-conscious cloud... ...core component of infrastructure provisioning and lifecycle management. The position blends hands-on production engineering with...Full timeRemote work
- ...the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-... ...technical ownership and customer-facing engineering: you'll define how we measure... ...updated within the restricted environment.Manage containerized services (Docker,...Full timeContract workRemote workFlexible hours
$125k - $250k
...reimagining how developers build reliable, scalable, event-driven... ...and monitor SLIs/SLOs and help manage error budgets across the platform... ...Partner closely with engineering teams to improve system resiliency... ...5+ years of experience in Site Reliability Engineering, DevOps...Full timeImmediate startRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
- site reliability engineer Remote
- site reliability engineer remote Remote
- site reliability engineer sre Remote
- site safety Remote
- website coordinator Remote
- on-site clinical research associate (traveling/remote) Remote
- site services specialist Remote
- on site coordinator Remote
- construction site safety Remote
- junior website developer Remote



