Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer - Site Reliability Engineering

$140k - $230k
Full-time

Zoox Inc.

Zoox is seeking a Site Reliability Engineer to help ensure the availability, performance, and resilience of the services that power the development and operation of our autonomous vehicles. In this role, you will own the full lifecycle of our services—from designing fault-tolerant, maintainable systems to deploying, operating, and continuously improving them in production. As a robotics company, Zoox embraces automation at every layer of our infrastructure, and you’ll help drive that ethos forward. You’ll work hands-on with systems that process massive volumes of data and support compute-intensive pipelines running on both CPUs and GPUs. 

In this role, you will:



  • Architect and optimize scalable systems: You will design, implement, and continuously improve highly reliable infrastructure, directly impacting the success and safety of Zoox's autonomous vehicle platform.



  • Build proactive monitoring solutions: You will develop advanced monitoring, alerting, and reporting tools to ensure potential issues are identified and resolved before they affect production.



  • Collaborate across engineering: You will partner closely with software engineering teams to elevate our system architecture, streamline deployment processes, and drive automation initiatives.



  • Lead incident resolution: You will conduct thorough root cause analyses on production issues and rapidly deploy corrective actions to maintain a resilient and stable environment.



  • Ensure business continuity: You will safeguard the company's operations by designing and implementing robust disaster recovery plans to keep the Zoox fleet running smoothly under any circumstances.


Qualifications



  • SRE & Distributed Systems Experience: 5+ years of experience in site reliability engineering or a similar role, with a strong, objective background in managing large-scale distributed systems.



  • Cloud & Infrastructure as Code (IaC): Proven experience operating within major cloud platforms (AWS, GCP, or Azure) and utilizing IaC tools like Terraform, Ansible, Salt, or CloudFormation.



  • Container Orchestration: Technical expertise in deploying, managing, and scaling systems using container orchestration technologies such as Kubernetes.



  • Core Infrastructure Knowledge: Deep, foundational understanding of networking protocols, storage solutions, and database technologies.



  • Programming Proficiency: Strong, demonstrable programming and scripting skills in languages such as Python, Go, C/C++, or Java.


Bonus Qualifications



  • Experience in the automotive or autonomous vehicle industry.



  • Knowledge of security best practices and compliance requirements.


$140,000 - $230,000 a year

About Zoox

Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We’re looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team.

Accommodations

If you need an accommodation to participate in the application or interview process please reach out to [email protected] or your assigned recruiter.

A Final Note:

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Software Engineer - Site Reliability Engineering in Remote vacancy
  • $96k - $163k

     ...realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work...  ...our developers during the application build phase in software run principles that include operational design, automation... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    19 hours ago
  • $96k - $163k

     ...realize their greatest potential. Title and Summary Senior Site Reliability Engineer, Performance Engineering Senior Site Reliability...  ...suites from LoadRunner to Gatling or BlazeMeter. • Strong software performance analysis, bottleneck identification, scaling,... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    19 hours ago
  • $76k - $127k

     ...realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site...  .... Support and improve the CI/CD pipeline for software promotion, enforcing validation and operational gating.... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours
    Early shift

    Mastercard

    O Fallon, MO
    19 hours ago
  •  ...their greatest potential. Title and Summary Senior Site Reliability Engineer Overview-The ProCOM team is looking for a Site...  ...• Support the application CI/CD pipeline for promoting software into higher environments through validation and operational... 
    Suggested
    Full time
    Part time
    Immediate start
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    19 hours ago
  • $76k - $127k

     ...governments realize their greatest potential. Title and Summary Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we...  ...our developers during the application build phase in software run principles that include operational design, automation... 
    Suggested
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    19 hours ago
  •  ...realize their greatest potential. Title and Summary Lead Site Reliability Engineer Job Description Summary Overview: Who is...  ...support our developers during the application build phase in software run principals that includes operational design, automation... 
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    19 hours ago
  • $122k - $207k

     ...realize their greatest potential. Title and Summary Manager, Site Reliability Engineering Who is Mastercard? At Mastercard technology, we work...  ...support platform stability through monitoring. We support software run principals that includes change implementation,... 
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    19 hours ago
  • $62k - $141k

    Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether...  ...in network engineering, systems administration, or software development, if you have a passion for making... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Chantilly, Loudoun County, VA
    4 days ago
  •  ...ensuring the availability, scalability, and reliability of systems and applications.What will...  ...or CloudFormation.Mentor junior engineers and provide technical guidance.Stay up-...  ...and Kubernetes.Experience in application software developmentCompany BenefitsCompetitive... 
    Work at office
    Remote work

    Interactive Brokers

    Greenwich, CT
    4 days ago
  • $146k - $194k

     ...focused on positioning Anduril as a lead provider of specialized engineering and products for Intelligence Community (IC) customers. We...  ...pressing national security requirements.ABOUT THE JOBAs a Site Reliability Engineer, your primary mission is to ensure the health,... 
    Full time
    Work experience placement
    Immediate start
    Remote work

    Anduril Industries

    Reston, VA
    3 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range...  ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...  ...if youHave 6+ years of experience in software development and operating distributed systemsAre... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Austin, TX
    2 days ago
  • $158.5k - $172k

     ...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will...  ...high-impact position driving continuous reliability, deep system optimization, and...  ...ensuring fast, secure, and friction-free software delivery workflows.Secure and Standardize... 
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    New York, NY
    19 hours ago
  •  ...exceptional talent to help us stay at the cutting edge. As a DevOps Site Reliability Engineer, you’ll have the chance to contribute to the continuous...  ...by providing a non-smoking environment. Reynolds and Reynolds is an equal opportunity employer.Category: Software Development
    Work from home
    2 days per week

    Reynolds & Reynolds

    Tallahassee, FL
    4 days ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  • $90k - $180k

     ...people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale,...  .... Identify and eliminate performance bottlenecks in software and infrastructure, ensuring low-latency, high-throughput,... 
    Remote work

    Abbott

    Sunnyvale, CA
    2 days ago
  •  ...we’re shaping the future of software delivery. From generative AI...  ...platforms to advanced release engineering practices, our teams are redefining...  ...preferred Exposure to reliability engineering concepts such as...  ...KC1#GMFjobsAbout The Role:The Site Reliability Engineer under the... 
    Work experience placement
    H1b
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours
    Shift work
    2 days per week

    GM Financial

    Arlington, TX
    3 days ago
  • $130k - $150k

     ...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are...  ...on-premises and cloud environments. This role blends software engineering and operations practices to reduce manual toil... 
    Work at office
    Work from home
    3 days per week

    CRA International

    Boston, MA
    4 days ago
  • $15k

     ...office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to...  ...mechanisms when off-the-shelf ones won't doHelp software and research teams design policies around fair cluster usage... 
    Work at office
    Local area
    Remote work

    The Voleon Group

    Berkeley, CA
    3 days ago
  •  ...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of...  ...MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment... 
    Full time
    Work at office
    Local area
    Remote work

    MUFG

    Jersey City, NJ
    4 days ago
  •  ...home day is currently Tuesday.Engineering at Lambda is responsible for...  ...plane services and dataplane software running on SmartNICsDevelop tooling...  ...teams to improve service reliability and deployment...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    2 days ago
  • $118.6k - $195.68k

     ...SummaryThe Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat...  ...teams to understand deliverablesDesign and development of software like Kubernetes operators, webhooks, cli-toolsImplement... 
    Permanent employment
    Full time
    Contract work
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Red Hat

    Raleigh, NC
    3 days ago
  • $130k - $180k

     ...belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a...  ...growth in its flagship cloud product. We’re seeking senior software and systems engineers specializing in reliability and platform... 
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to Friday
    Flexible hours

    Imanage

    Chicago, IL
    2 days ago
  • $115.5k - $164.8k

     ...with our ecosystem of devices and cloud software. Like our products, we work better together...  ...where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend...  ...required human intervention with reliable, tested automation. You will also participate... 
    Work experience placement
    Work at office
    Remote work

    Axon

    Washington DC
    2 days ago
  • $102.1k - $202.2k

     ...yearEmployment type: Full-TimeWork site: 0 days / week in-office -...  ...ContributorTravel: Less than 25%Profession: Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...workloads. As a Site Reliability Engineer II, you will take ownership of reliability... 
    Ongoing contract
    Work at office
    Local area

    Microsoft

    Redmond, WA
    4 days ago
  • $105.6k - $145.2k

    Architect the Future as our Site Reliability Engineer!Are you ready to take your skills to the next level as a self-motivated and enthusiastic Site Reliability Engineer with hands-on experience supporting multiple connected Cloud-based products? Trimble is a global technology... 
    Ongoing contract
    Full time
    Work at office
    Local area
    Worldwide

    Trimble Navigation

    Westminster, CO
    1 day ago
  • $108.08k - $172.5k

    Work with development and platform engineering teams to migrate and maintain applications in Google Cloud. Apply Observability concepts...  ...and manage toil. Work with application teams to ensure designed software solutions meet non-functional requirements, including... 
    Full time
    Remote work
    Worldwide

    CME- Group

    Chicago, IL
    16 hours ago
  • $165k - $190k

     ...DevOps / SRE TeamThe DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-...  ...security platformAddress complex challenges around scalability, reliability, observability, and cost efficiencyCollaborate with Engineering... 
    Work from home

    Obsidian Security

    Palo Alto, CA
    2 days ago
  • $80k - $133k

    Job Family:Software Development & SupportTravel Required:Up to 10%Clearance Required:Ability to Obtain Public TrustWhat You Will Do:Collaborate...  ...Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud... 
    Permanent employment
    Full time
    Contract work
    Remote work
    Flexible hours

    Guidehouse

    McLean, VA
    2 days ago
  • $230k - $250k

    GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is... 
    Remote work

    Govcio

    Arlington, VA
    19 hours ago
  • $91.7k - $163.7k

     ...technology systems in accordance with modern design standards.As a Site Reliability Engineer (SRE), you will play a key role in ensuring the...  ...bridge the gap between development and operations, applying software engineering practices to solve operational challenges. Your... 
    Minimum wage
    Full time
    Work experience placement
    Local area
    Remote work

    UnitedHealth Group

    Eden Prairie, MN
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer - Site Reliability Engineering. Be the first to apply!