Senior Site Reliability Engineer
Latitude AI
Senior Site Reliability Engineer
Latitude AI is building the future of Ford's autonomy roadmap to make travel safer, less stressful, and more enjoyable for everyone. Bringing this vision to scale, our fully in-house developed hands-free ADAS platform will debut on Ford's all-new Universal Electric Vehicle in 2027.
When you join the Latitude team, you'll work alongside leading experts across machine learning and robotics, cloud platforms, mapping, sensors and compute systems, test operations, systems and safety engineering – all dedicated to redefining the relationship between people and their vehicles for millions of customers.
As a Ford Motor Company subsidiary, we operate independently to develop automated driving technology at the speed of a technology startup. Latitude is headquartered in Pittsburgh with engineering centers in Dearborn, Mich., and Palo Alto, Calif.
Meet the team:
As a Site Reliability Engineer on the team, you will be responsible for helping to build and run these mission critical systems. Through the implementation of monitoring and automation, you will constantly ensure the health, reliability, scalability, and performance of the platforms.
The Site Reliability team interacts with engineering teams including ingest/data processing, mapping, labeling, triage, machine learning (detection, prediction, tracking), motion planning/control, offline simulation, and release/deployment teams to provide uniform service observability and incident response.
What you'll do:
- Build monitoring to ensure our platform is healthy and its reliability measurable
- Build alerting and a set of runbooks to enable faster detection and remediation of platform issues
- Debug complex issues that may combine multiple components of the stack and ensure proper fixes are implemented to prevent these issues from happening again
- Participate in an on-call rotation and culture of continuous improvement through blameless postmortems
- Design and implement components of the platform to enable features that make the work of our customers possible, simpler and more efficient
- Build Kubernetes controllers to automate operations
What you'll need to succeed:
- Bachelor's degree in Computer Engineering, Computer Science, Electrical Engineering, Robotics or a related field and 4+ years of relevant experience (or Master's degree and 2+ years of relevant experience, or PhD)
- Fundamental understanding of Linux operating system internals, TCP/IP networking, and storage subsystems
- Hands on development in Go or Python to create robust software that can run reliably in production
- Strong experience scaling and securing services in the cloud (AWS, GCP) or cloud native environments
- Experience using infrastructure-as-code principles to automate the creation of infrastructure resources (e.g. Terraform, CloudFormation)
- Experience authoring and maintaining Kubernetes Controllers in Go
- Experience running Kubernetes and related core components in a large-scale, production environment
- Experience with metrics (e.g. Prometheus), logging (e.g. Elasticsearch, Loki) and tracing (e.g. Jaeger, Tempo) systems
- Understanding of engineering design limitations and ability to provide guidance to teams to scale their services to achieve desired performance within budget
- A focus on increasing service reliability through defining and adhering to SLOs
- Strong communication skills and the ability to work effectively in a diverse and distributed team
What we offer you:
- Competitive compensation packages
- High-quality individual and family medical, dental, and vision insurance
- Health savings account with available employer match
- Employer-matched 401(k) retirement plan with immediate vesting
- Employer-paid group term life insurance and the option to elect voluntary life insurance
- Paid parental leave
- Paid medical leave
- Unlimited vacation
- 15 paid holidays
- Daily lunches, snacks, and beverages available in all office locations
- Pre-tax spending accounts for healthcare and dependent care expenses
- Pre-tax commuter benefits
- Monthly wellness stipend
- Adoption/Surrogacy support program
- Backup child and elder care program
- Professional development reimbursement
- Employee assistance program
- Discounted programs that include legal services, identity theft protection, pet insurance, and more
- Company and team bonding outlets: employee resource groups, quarterly team activity stipend, and wellness initiatives
- ...professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of... ...and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase within...Senior
- ...The Role We're looking for a Senior Site Reliability Engineer to own the reliability, scalability, and operational excellence of the production systems that power Nectar's platform. We run high-volume data ingestion pipelines and real-time AI agents on top of a fast...SeniorRemote work
- ...Site Reliability Engineer There are NO limits to your career: come shape the future and be part of a truly unique global culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of software...SeniorImmediate startRemote workWorldwide
$137.77k - $194.59k
...distributed team of roughly 80 scientists and engineers building and operating Rubin's petascale... ...Your role: \n You will own the reliability and robustness of Rubin Observatory's... ...nature of this position, SLAC is open to on-site, hybrid, and remote work options. \n \...SeniorRemote workFlexible hoursNight shift$150k - $175k
...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed...SeniorRemote work- ...Senior Site Reliability Engineer LeanData helps the world's fastest-growing companies automate, simplify, and accelerate revenue. We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly...SeniorFull timeWork at officeFlexible hours2 days per week
$180k - $260k
...effortless integration into customers' logistics operations. About the role We are seeking an experienced Senior/Staff Site Reliability Engineer to support the operation, monitoring, and scaling of our growing fleet of autonomous vehicles. In this role, you will...SeniorOdd jobWork at officeRemote work- ...strong customer and partner networks. About the Role This role leads reliability strategy and architectural improvements across infrastructure, GPU systems, observability, ML Ops and IT Ops. Mentor engineers, manage high‑severity incidents, and drive SLO governance. You...SeniorFull time
- ...Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will...Senior
$151.6k - $245.3k
...Site Reliability Engineer Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture...$90k - $180k
...medicines. Our 115,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: About the Role This Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division....SeniorFull timeRemote workShift work$189k - $232k
...Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck....Full timeWork at officeLocal area2 days per week- ...Site Reliability Engineer III There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability...Work at office
$170k - $250k
...Site Reliability Engineer (SRE) Location: San Francisco, CA / Palo Alto, CA Company Stage of Funding: Growth-Stage AI Infrastructure Company ($80M Raised) Office Type: Onsite (4 Days Per Week) Salary: $170,000–$250,000 + Competitive Equity We're representing a rapidly...Work at officeVisa sponsorshipFlexible hours$170k - $230k
...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community,...Work at officeLocal area1 day per week$165k - $190k
...Site Reliability Engineer Palo Alto, California, USA Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast...Work from homeFlexible hours$217.57k - $260k
...job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA;... ...which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure...Full timeTemporary workWork at officeRemote workFlexible hoursShift work- ...Lead Site Reliability Engineer Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within...
$200k - $260k
...Lead Site Reliability Engineer Glean is the Work AI platform that helps everyone work smarter with AI. What began as the industry's most advanced... ...practical experience. ~8+ years of experience in a senior-level role within Site Reliability Engineering or similar role...Work at officeHome officeFlexible hours$252k - $308k
...Staff Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck...Full timeWork at office2 days per week$100k - $200k
OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about...Full time- ...that's more connected, more intelligent, more sustainable for everyone. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that underpins the build pipelines that allow our companies to...Full timeContract workLocal area
- ...keep the world running. Location: 5 on-site days a week in Sunnyvale, CA Headquarters. Our Team's Vision: Our Engineering team is shaping the future of cybersecurity... ...: We are looking for an experienced Senior Site Reliability Engineer (SRE) with a strong background...SeniorWork experience placementImmediate start
- ...Job Title 12+ years in platform engineering, SRE, or DevOps. Experience with HPC clusters (Slurm, PBS, Grid Engine). Cloud infrastructure expertise (GCP/AWS preferred). Proficiency with Terraform, Ansible, Prometheus, Grafana, ELK. Strong Linux administration...Senior
- # Senior Software Engineer, AI PlatformMeta## Job DescriptionMeta AI is looking for experienced Senior Software Engineers to contribute to our AI... ...401(k) matching.- Generous paid time off and holidays.- On-site amenities and perks.- Opportunity to shape the future of AI...Senior
$184k - $276k
...expertise in modern IT ecosystems? We're seeking a Sr. IT Systems Engineer passionate about driving automation, maturing enterprise... ...IAM), and architecting scalable, secure SaaS integrations. This senior individual contributor role serves as a key technical resource,...SeniorTemporary work$75 - $80 per hour
...Orchestration, Mesos, Docker Swarm). Experience with streaming or queuing systems such as Kafka, ActiveMq, etc. Experience engineering, operating, troubleshooting and testing SaaS/Cloud services. Excellent troubleshooting, critical thinking, and data analysis skills...Senior$126k - $248k
...About the Role We’re looking for a Senior Engineer to help build the next-generation inference platform that supports embedding models... ...integration with Atlas, and contribute to a platform designed for reliability, performance, and ease of use. We're looking to speak with...SeniorLocal areaFlexible hours$186k - $232.5k
...future that’s more connected, more intelligent, more sustainable for everyone. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that underpins the build pipelines that allow our companies to...Full timeLocal area$163.2k - $220.8k
...allow exceptional opportunities for professional achievement and career growth. Wilson Sonsini is seeking an experienced Senior Software Developer to join our Innovations team. The software developers on our team are the primary contributors to Neuron on both...SeniorRemote workWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Palo Alto, CA
- site reliability engineer sre Palo Alto, CA
- senior app developer Palo Alto, CA
- senior international account manager Palo Alto, CA
- senior magento developer Palo Alto, CA
- sr marketing manager Palo Alto, CA
- sr technical product manager Palo Alto, CA
- senior manager pmo Palo Alto, CA
- senior accountant part time Palo Alto, CA
- senior electrical designer Palo Alto, CA


