Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

Latitude AI

Senior Site Reliability Engineer

Latitude AI is building the future of Ford's autonomy roadmap to make travel safer, less stressful, and more enjoyable for everyone. Bringing this vision to scale, our fully in-house developed hands-free ADAS platform will debut on Ford's all-new Universal Electric Vehicle in 2027.

When you join the Latitude team, you'll work alongside leading experts across machine learning and robotics, cloud platforms, mapping, sensors and compute systems, test operations, systems and safety engineering – all dedicated to redefining the relationship between people and their vehicles for millions of customers.

As a Ford Motor Company subsidiary, we operate independently to develop automated driving technology at the speed of a technology startup. Latitude is headquartered in Pittsburgh with engineering centers in Dearborn, Mich., and Palo Alto, Calif.

Meet the team:

As a Site Reliability Engineer on the team, you will be responsible for helping to build and run these mission critical systems. Through the implementation of monitoring and automation, you will constantly ensure the health, reliability, scalability, and performance of the platforms.

The Site Reliability team interacts with engineering teams including ingest/data processing, mapping, labeling, triage, machine learning (detection, prediction, tracking), motion planning/control, offline simulation, and release/deployment teams to provide uniform service observability and incident response.

What you'll do:

  • Build monitoring to ensure our platform is healthy and its reliability measurable
  • Build alerting and a set of runbooks to enable faster detection and remediation of platform issues
  • Debug complex issues that may combine multiple components of the stack and ensure proper fixes are implemented to prevent these issues from happening again
  • Participate in an on-call rotation and culture of continuous improvement through blameless postmortems
  • Design and implement components of the platform to enable features that make the work of our customers possible, simpler and more efficient
  • Build Kubernetes controllers to automate operations

What you'll need to succeed:

  • Bachelor's degree in Computer Engineering, Computer Science, Electrical Engineering, Robotics or a related field and 4+ years of relevant experience (or Master's degree and 2+ years of relevant experience, or PhD)
  • Fundamental understanding of Linux operating system internals, TCP/IP networking, and storage subsystems
  • Hands on development in Go or Python to create robust software that can run reliably in production
  • Strong experience scaling and securing services in the cloud (AWS, GCP) or cloud native environments
  • Experience using infrastructure-as-code principles to automate the creation of infrastructure resources (e.g. Terraform, CloudFormation)
  • Experience authoring and maintaining Kubernetes Controllers in Go
  • Experience running Kubernetes and related core components in a large-scale, production environment
  • Experience with metrics (e.g. Prometheus), logging (e.g. Elasticsearch, Loki) and tracing (e.g. Jaeger, Tempo) systems
  • Understanding of engineering design limitations and ability to provide guidance to teams to scale their services to achieve desired performance within budget
  • A focus on increasing service reliability through defining and adhering to SLOs
  • Strong communication skills and the ability to work effectively in a diverse and distributed team

What we offer you:

  • Competitive compensation packages
  • High-quality individual and family medical, dental, and vision insurance
  • Health savings account with available employer match
  • Employer-matched 401(k) retirement plan with immediate vesting
  • Employer-paid group term life insurance and the option to elect voluntary life insurance
  • Paid parental leave
  • Paid medical leave
  • Unlimited vacation
  • 15 paid holidays
  • Daily lunches, snacks, and beverages available in all office locations
  • Pre-tax spending accounts for healthcare and dependent care expenses
  • Pre-tax commuter benefits
  • Monthly wellness stipend
  • Adoption/Surrogacy support program
  • Backup child and elder care program
  • Professional development reimbursement
  • Employee assistance program
  • Discounted programs that include legal services, identity theft protection, pet insurance, and more
  • Company and team bonding outlets: employee resource groups, quarterly team activity stipend, and wellness initiatives
Vacancy posted 3 hours ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Palo Alto, CA vacancy
  •  ...professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of...  ...and position yourself among the top echelon in site reliability.  As a Senior Lead Site Reliability Engineer at JPMorgan Chase within... 
    Senior

    J.P. Morgan

    Palo Alto, CA
    5 days ago
  •  ...The Role We're looking for a Senior Site Reliability Engineer to own the reliability, scalability, and operational excellence of the production systems that power Nectar's platform. We run high-volume data ingestion pipelines and real-time AI agents on top of a fast... 
    Senior
    Remote work

    Nectar Social

    Palo Alto, CA
    4 days ago
  •  ...Site Reliability Engineer There are NO limits to your career: come shape the future and be part of a truly unique global culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of software... 
    Senior
    Immediate start
    Remote work
    Worldwide

    OutSystems

    Menlo Park, CA
    2 hours ago
  • $137.77k - $194.59k

     ...distributed team of roughly 80 scientists and engineers building and operating Rubin's petascale...  ...Your role: \n You will own the reliability and robustness of Rubin Observatory's...  ...nature of this position, SLAC is open to on-site, hybrid, and remote work options. \n \... 
    Senior
    Remote work
    Flexible hours
    Night shift

    Stanford University

    Menlo Park, CA
    22 hours ago
  • $150k - $175k

     ...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed... 
    Senior
    Remote work

    ASAPP

    Mountain View, CA
    1 day ago
  •  ...Senior Site Reliability Engineer LeanData helps the world's fastest-growing companies automate, simplify, and accelerate revenue. We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly... 
    Senior
    Full time
    Work at office
    Flexible hours
    2 days per week

    LeanData

    Santa Clara, CA
    22 hours ago
  • $180k - $260k

     ...effortless integration into customers' logistics operations. About the role We are seeking an experienced Senior/Staff Site Reliability Engineer to support the operation, monitoring, and scaling of our growing fleet of autonomous vehicles. In this role, you will... 
    Senior
    Odd job
    Work at office
    Remote work

    Gatik AI

    Santa Clara, CA
    3 days ago
  •  ...strong customer and partner networks. About the Role This role leads reliability strategy and architectural improvements across infrastructure, GPU systems, observability, ML Ops and IT Ops. Mentor engineers, manage high‑severity incidents, and drive SLO governance. You... 
    Senior
    Full time

    Gruve

    Redwood City, CA
    1 day ago
  •  ...Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will... 
    Senior

    Kody

    Palo Alto, CA
    28 days ago
  • $151.6k - $245.3k

     ...Site Reliability Engineer Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture... 

    Palo Alto Networks

    Palo Alto, CA
    22 hours ago
  • $90k - $180k

     ...medicines. Our 115,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: About the Role This Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.... 
    Senior
    Full time
    Remote work
    Shift work

    Abbott

    Sunnyvale, CA
    3 days ago
  • $189k - $232k

     ...Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck.... 
    Full time
    Work at office
    Local area
    2 days per week

    Earnin

    Mountain View, CA
    2 days ago
  •  ...Site Reliability Engineer III There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability... 
    Work at office

    Chase

    Palo Alto, CA
    21 hours ago
  • $170k - $250k

     ...Site Reliability Engineer (SRE) Location: San Francisco, CA / Palo Alto, CA Company Stage of Funding: Growth-Stage AI Infrastructure Company ($80M Raised) Office Type: Onsite (4 Days Per Week) Salary: $170,000–$250,000 + Competitive Equity We're representing a rapidly... 
    Work at office
    Visa sponsorship
    Flexible hours

    Recruiting from Scratch

    Palo Alto, CA
    1 day ago
  • $170k - $230k

     ...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community,... 
    Work at office
    Local area
    1 day per week

    Mithril

    Palo Alto, CA
    2 days ago
  • $165k - $190k

     ...Site Reliability Engineer Palo Alto, California, USA Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast... 
    Work from home
    Flexible hours

    Obsidian Security

    Palo Alto, CA
    2 days ago
  • $217.57k - $260k

     ...job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA;...  ...which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    Shift work

    ID.me

    Mountain View, CA
    22 hours ago
  •  ...Lead Site Reliability Engineer Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within... 

    Chase

    Palo Alto, CA
    1 day ago
  • $200k - $260k

     ...Lead Site Reliability Engineer Glean is the Work AI platform that helps everyone work smarter with AI. What began as the industry's most advanced...  ...practical experience. ~8+ years of experience in a senior-level role within Site Reliability Engineering or similar role... 
    Work at office
    Home office
    Flexible hours

    Softbank Investment Advisers

    Mountain View, CA
    1 day ago
  • $252k - $308k

     ...Staff Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck... 
    Full time
    Work at office
    2 days per week

    Earnin

    Mountain View, CA
    4 days ago
  • $100k - $200k

    OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about... 
    Full time

    OPPO

    Palo Alto, CA
    2 days ago
  •  ...that's more connected, more intelligent, more sustainable for everyone. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that underpins the build pipelines that allow our companies to... 
    Full time
    Contract work
    Local area

    Rivian and Volkswagen Group Technologies

    Palo Alto, CA
    1 day ago
  •  ...keep the world running. Location: 5 on-site days a week in Sunnyvale, CA Headquarters. Our Team's Vision: Our Engineering team is shaping the future of cybersecurity...  ...: We are looking for an experienced Senior Site Reliability Engineer (SRE) with a strong background... 
    Senior
    Work experience placement
    Immediate start

    Illumio

    Sunnyvale, CA
    2 days ago
  •  ...Job Title 12+ years in platform engineering, SRE, or DevOps. Experience with HPC clusters (Slurm, PBS, Grid Engine). Cloud infrastructure expertise (GCP/AWS preferred). Proficiency with Terraform, Ansible, Prometheus, Grafana, ELK. Strong Linux administration... 
    Senior

    Saxon Global

    Mountain View, CA
    4 days ago
  • # Senior Software Engineer, AI PlatformMeta## Job DescriptionMeta AI is looking for experienced Senior Software Engineers to contribute to our AI...  ...401(k) matching.- Generous paid time off and holidays.- On-site amenities and perks.- Opportunity to shape the future of AI... 
    Senior

    Aibreakingwire

    Menlo Park, CA
    3 days ago
  • $184k - $276k

     ...expertise in modern IT ecosystems? We're seeking a Sr. IT Systems Engineer passionate about driving automation, maturing enterprise...  ...IAM), and architecting scalable, secure SaaS integrations. This senior individual contributor role serves as a key technical resource,... 
    Senior
    Temporary work

    SpaceXAI

    Palo Alto, CA
    22 hours ago
  • $75 - $80 per hour

     ...Orchestration, Mesos, Docker Swarm). Experience with streaming or queuing systems such as Kafka, ActiveMq, etc. Experience engineering, operating, troubleshooting and testing SaaS/Cloud services. Excellent troubleshooting, critical thinking, and data analysis skills... 
    Senior

    Cynet Systems

    Palo Alto, CA
    1 day ago
  • $126k - $248k

     ...About the Role We’re looking for a Senior Engineer to help build the next-generation inference platform that supports embedding models...  ...integration with Atlas, and contribute to a platform designed for reliability, performance, and ease of use. We're looking to speak with... 
    Senior
    Local area
    Flexible hours

    United States Digital Space LLC

    Palo Alto, CA
    22 hours ago
  • $186k - $232.5k

     ...future that’s more connected, more intelligent, more sustainable for everyone. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that underpins the build pipelines that allow our companies to... 
    Full time
    Local area

    Rivian and Volkswagen Group Technologies

    Palo Alto, CA
    1 day ago
  • $163.2k - $220.8k

     ...allow exceptional opportunities for professional achievement and career growth. Wilson Sonsini is seeking an experienced Senior Software Developer to join our Innovations team. The software developers on our team are the primary contributors to Neuron on both... 
    Senior
    Remote work
    Worldwide

    Wilson Sonsini Goodrich and Rosati

    Palo Alto, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!