Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$276.1k - $311.4k

Lindus Health

About us

Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems.

Our vision is to create autonomy that propels the world forward. Our intelligent, mapless, and hardware-agnostic AI products are designed for automakers, accelerating the transition from assisted to automated driving.
In our fast-paced environment big problems ignite us—we embrace uncertainty, leaning into complex challenges to unlock groundbreaking solutions. We aim high and stay humble in our pursuit of excellence, constantly learning and evolving as we pave the way for a smarter, safer future.

At Wayve, your contributions matter. We value diversity, embrace new perspectives, and foster an inclusive work environment; we back each other to deliver impact.

Make Wayve the experience that defines your career!

The role

As SRE Manager, you'll build the Vehicle Software SRE team from the ground up — defining its charter, hiring its founding engineers, establishing the operating model, and creating the technical strategy that makes reliability a first-class property of the software running on our vehicles.

You'll work in a production environment unlike most: a globally distributed fleet of autonomous vehicles operating at the intersection of software, hardware, networking, sensors, and the physical world. Failures are often intermittent, hard to reproduce, and distributed across ownership boundaries. You'll move reliability upstream — from reactive field support to prevention through architecture, automation, observability, and disciplined production readiness.

You'll embed your team within Vehicle Software, partnering with product teams who retain ownership of what they build while your team provides the reliability engineering, standards, and leverage that help them operate fleet-critical software safely at scale. You'll stay hands-on throughout — writing code, reviewing critical designs, and leading the investigations that matter most.

The systems you help harden will connect Wayve's AI to physical vehicles and underpin the transition from engineering fleets to commercial operations. Few engineering leadership roles offer this combination of zero-to-one team building, deep systems work, and direct influence on the safety and scalability of autonomous mobility.

Key Responsibilities
  • Team building & leadership : Build and lead a new SRE team from the ground up, staying hands-on as a player-coach on the team's most consequential work.
  • Reliability strategy: Own technical direction for vehicle software reliability across deployment, service health, telemetry, and diagnostics; define what production-ready means at Wayve.
  • Production readiness: Define SLIs, SLOs, and error budgets for fleet-critical workflows; drive release criteria, automated gates, rollback strategies, and fault-injection practices across Vehicle Software.
  • Observability & tooling: Design and implement the observability and automation that shortens the path from vehicle symptom to root cause, cuts the manual toil between failure and fix, and shapes systems for robustness, recoverability, and debuggability.
  • Incident response: Lead investigations into complex failures, ensure every incident produces a durable fix, and strengthen on-call practices and escalation paths across service-owning teams.
  • Mentorship & communication: Mentor engineers and emerging leaders, and give senior leadership the clarity on reliability health, risks, and investment they need to make good decisions.
About you

Essential

  • 8+ years building and operating production software systems with strong depth in SRE, production engineering, platform engineering, embedded systems, or robotics, and a recent track record of writing production-quality code and leading architecture reviews across Linux-based, distributed, or hardware-software systems.
  • 3+ years in people leadership with a track record of hiring, coaching, and growing engineers across levels while staying actively engaged in coding, design, and code review; experience forming a new team or capability from scratch is a strong plus.
  • Proven experience with SLOs, error budgets, production-readiness standards, observability, incident management, postmortems, and toil-reduction programmes, with measurable outcomes to show for it.
  • A track record of turning ambiguous, cross-functional problems into clear ownership, sequenced plans, and reliable delivery without relying on formal authority.
  • Hands-on experience building production software, automation, and diagnostic tooling in C++, Rust, Python, or Go, with familiarity with CI/CD, release systems, telemetry pipelines, and modern observability tooling.
  • Calm and structured during incidents, with clear communication across software, hardware, operations, product, and executive stakeholders; a leadership approach grounded in ownership, blameless learning, and autonomy with accountability.

Desirable

  • Experience with autonomous vehicles, robotics, automotive software, safety-critical systems, or cyber-physical products deployed into uncontrolled real-world environments.
  • Knowledge of onboard compute, sensors, middleware, vehicle networks, OTA deployment, data capture and offload, or multi-variant hardware-software integration.
  • Experience improving reliability through the transition from research and prototype systems to commercial, multi-market operations.
  • Experience partnering with fleet operations, field engineering, validation, hardware, or external vehicle and technology partners.
What success looks like
  • In the first 90 days, you have aligned the charter and ownership model, established a reliability baseline, mapped the highest-risk systems and workflows, defined the hiring plan, and personally contributed design, code, or tooling to deliver early reliability wins.
  • Within six months, the founding team is operating effectively inside priority Vehicle Software domains, with clearer service ownership, stronger observability, improved runbooks and escalation paths, and initial automated prevention and release controls in production; you remain a trusted technical contributor through design, coding, and code review.
  • Over 6-12 months, repeat incidents, mean time to detect, mean time to recover, and unowned escalations are trending down, while deployment success, vehicle readiness, data-collection reliability, and safe fleet availability are improving against the baseline.
  • Vehicle Software teams are increasingly able to own and operate their systems reliably, with SRE acting as a force multiplier rather than a permanent support queue.

This is a full-time role based in our office in Sunnyvale. At Wayve we want the best of all worlds so we operate a hybrid working policy that combines time together in our offices and workshops to fuel innovation, culture, relationships and learning, and time spent working from home.

The reasonably estimated salary for this role ranges from $276,100 - $311,400 plus a competitive equity package. Actual compensation is based on the candidate's skills, qualifications, and experience.

Wayve is committed to creating an inclusive interview experience. If you require any accommodations or adjustments to participate fully in our interview process, please let us know.

We understand that everyone has a unique set of skills and experiences and that not everyone will meet all of the requirements listed above. If you’re passionate about self-driving cars and think you have what it takes to make a positive impact on the world, we encourage you to apply.

At Wayve we're committed to creating a diverse, fair and respectful culture that is inclusive of everyone based on their unique skills and perspectives, and regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, veteran status, pregnancy or related condition (including breastfeeding) or any other basis as protected by applicable law.

For more information visit Careers at Wayve.

To learn more about what drives us, visit Values at Wayve

For US candidates only, please visit E-Verify Notice and Participation and Right to Work

DISCLAIMER: We will not ask about marriage or pregnancy, care responsibilities or disabilities in any of our job adverts or interviews. However, we do look to capture information about care responsibilities, and disabilities among other diversity information as part of an optional DEI Monitoring form to help us identify areas of improvement in our hiring process and ensure that the process is inclusive and non-discriminatory.

#J-18808-Ljbffr
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Sunnyvale, CA vacancy
  • $170k - $200k

    We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance,... 
    Suggested
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    2 days ago
  • $152k - $241.5k

     ...infrastructure platforms for automated host lifecycle management, fleet reliability/auto-healing, E2E observability or data-driven operations (...  ...languages such as Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through design reviews,... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $148k - $235.75k

     ...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change...  ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"... 
    Suggested
    Night shift

    Forward Networks

    Santa Clara, CA
    2 hours ago
  • $145k - $175k

     ...straightforward communication and clinical domain expertise, Commence cuts straight to better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability, scalability, and operational health of our mission-critical healthcare data... 
    Suggested
    Full time
    Remote work

    GrabJobs

    Santa Clara, CA
    2 days ago
  •  ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless... 

    TechDigital Group

    Santa Clara, CA
    5 days ago
  •  ...Site Reliability Engineer, Data Platform - USDS Responsibilities Engage in and improve the whole lifecycle of service, from inception and design, through to deployment, operation and refinement. Ensure reliable, fault-tolerant, efficiently scalable and cost-effective data... 

    Tik Tok

    Mountain View, CA
    5 days ago
  •  ...Google is seeking a Senior Engineering Manager for Collaboration SRE to lead a multi-site engineering organization across Sunnyvale and Zurich. You will own the...  ..., and drive high-impact projects that improve reliability, performance, and scalability of #J-18808-Ljbffr

    Socket

    Sunnyvale, CA
    8 hours ago
  •  ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation... 

    Oracle

    Santa Clara, CA
    4 days ago
  •  ...in Cupertino, California, invites an experienced CDN Solutions Engineer to join the Content Delivery Network Solutions team. You will...  ...and collaborate with engineering groups across Apple to ensure reliable delivery at scale. The ideal candidate has 4+ years in CDNs and... 

    Apple

    Cupertino, CA
    1 day ago
  •  ...ServiceNow in Santa Clara, CA, seeks a Staff Software Engineer – SRE & AIOps to drive infrastructure automation, resilience, and toil...  ...for global engineering teams. Embedded within the Site Reliability & Database Engineering organization, you will architect SRE tooling... 

    ServiceNow

    Santa Clara, CA
    4 days ago
  •  ...function to support one of the world’s fastest-growing AI inference services, powered by the Wafer-Scale Engine (WSE). This team will help deliver world-class, ultra-reliable inference infrastructure for leading model builders such as OpenAI and other frontier labs.As a... 
    Shift work

    CEREBRAS SYSTEMS INC.

    Sunnyvale, CA
    1 day ago
  •  ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable... 

    Epic Games

    Sunnyvale, CA
    5 days ago
  •  ...Overview Title: Site Reliability Engineer SRE – ML platform Location: Austin, TX or Sunnyvale, CA Employment type: Full-time • Seniority: Mid-Senior level • ONLY W2 Responsibilities Continuous Deployment using GitHub Actions, Flux, Kustomize Design and implement cloud... 
    Full time

    Saransh

    Sunnyvale, CA
    5 days ago
  • $128k - $216k

     ...consumers to one another millions of times a day - quickly, reliably, and securely. Any time you swipe your credit card, pay...  ...a global scale, come make a difference at Fiserv. Sr. Site Reliability Engineer About Clover Clover is a pioneer in the fintech space... 
    Worldwide

    BentoBox

    Sunnyvale, CA
    5 days ago
  • $65 - $85 per hour

     ...Site Reliability Engineer Sustainable Talent is partnering with a global leader who's been transforming computer graphics, PC gaming, and accelerated computing for over 25 years. We are looking for a Site Reliability Engineer to support our client's team based out of... 
    Full time
    Contract work
    Worldwide

    Sustainable Talent

    Santa Clara, CA
    2 days ago
  • $168k - $270.25k

     ...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 

    Google

    Sunnyvale, CA
    2 hours ago
  • $101k - $161k

     ...excellence has earned us several prestigious awards, such as Best Engineering Team, Best Company for Diversity, Compensation, and Work-...  ...we do.Job DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s CloudVision-as-a-... 

    Arista Networks

    Santa Clara, CA
    3 days ago
  • $255.7k - $300k

     ...designs from peers, providing feedback to ensure best practices in reliability, security, and efficiency.Triage and resolve complex system...  ...execution of software development initiatives.Mentor other engineers and contribute to the engineering community through documentation... 
    Full time

    Google

    Sunnyvale, CA
    2 days ago
  • $262k - $364k

     ...services within the AViD ecosystem have reliability and uptime appropriate to users' needs with...  ...capacity and performance.Build creative engineering solutions to operations and...  ...changing circumstances in a strategic way.Site Reliability Engineering (SRE) combines software... 

    Google

    Mountain View, CA
    2 hours ago
  • $222k - $300.5k

     ...possible.Job OverviewAbout the TeamIntuit's Infrastructure and Site Reliability organization owns the operational backbone that keeps...  ...hundreds of millions of customers. The Fintech Platform Systems Engineering team builds and operates the AWS-based infrastructure, resiliency... 
    Worldwide
    Shift work

    Intuit

    Mountain View, CA
    3 days ago
  • $248k - $396.75k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production systems with exceptional efficiency, resilience, and availability. It combines software and systems engineering practices with... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...design by customizing MES tool per business needs Education Requirements, Ideal Experience: Associate’s degree in Industrial Engineering or IT related field Minimum of 0-3 years’ relevant experience Experience in C#, Delphi desired Knowledge of the... 
    Work at office

    Foxconn Industrial Internet - FII

    Sunnyvale, CA
    28 days ago
  •  ...of Huobi globe spanning infrastructure. •       Work with engineering teams to make sure new features and changes are deployed quickly...  .... •       Constantly improve our system performance and reliability through better tools, process and monitoring system. •... 
    Worldwide

    Cryptoware Technologies Inc

    Santa Clara, CA
    28 days ago
  • $160k - $240k

     ...one another millions of times a day - quickly, reliably, and securely. Any time you swipe your credit...  ...come make a difference at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our global team in... 
    Full time

    Fiserv

    Sunnyvale, CA
    20 days ago
  • $168k - $270.25k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline to design, build and maintain large scale production systems with high efficiency and availability using the combination of software and systems engineering practices. This is a highly specialized... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $217.57k - $260k

     ...job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA;...  ...which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    Shift work

    ID.me

    Mountain View, CA
    4 days ago
  • $207k - $300k

    Lead a team of Software/Systems Engineers on projects for users and be directly responsible for uptime.Own end-to-end availability...  ...Science or Engineering.1 year of people management experience. Site Reliability Engineering (SRE) combines software and systems engineering... 

    Google

    Mountain View, CA
    4 days ago
  • $255.7k - $300k

    Lead a team of engineers to maintain service uptime while managing global on-call rotations...  ...improve operational practices to drive reliability, maintainability, and stakeholder alignment...  ...or in a Manager, Software Engineer, Site Reliability Engineering-related occupation... 
    Full time
    Work at office

    Google

    Sunnyvale, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!