Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer, Fleet Infrastructure

$166k - $220k

Anduril Industries

Anduril Industries is a defense technology company with a mission to transform U.S. and allied military capabilities with advanced technology. By bringing the expertise, technology, and business model of the 21st century’s most innovative companies to the defense industry, Anduril is changing how military systems are designed, built and sold. Anduril’s family of systems is powered by Lattice OS, an AI-powered operating system that turns thousands of data streams into a realtime, 3D command and control center. As the world enters an era of strategic competition, Anduril is committed to bringing cutting-edge autonomy, AI, computer vision, sensor fusion, and networking technology to the military in months, not years.ABOUT THE TEAMAnduril products are designed to operate in high stakes environments, in some cases with life-and-death consequences for success or failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services & infrastructure across the fleet need to be able to be observed & monitored so that our users can trust in their ability to perform their expected functions. The Observability team builds and operates Anduril’s production telemetry plane, aggregating signals from the edge to cloud on our core ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril’s production telemetry systems for ground & cloud nodes. This includes architecting and designing our core observability plane for high-volume, high-availability telemetry ingestion, and operate/manage our central monitoring and alerting infrastructure. You’ll also build for both cloud-connected as well as off-line/airgapped environments - and ensure a seamless user experience transitioning between both. You’ll also collaborate closely with the Robotics Data Foundation & Fleet Management teams (which own our vehicle telemetry infrastructure as well as our software delivery plane for edge systems) to ensure that monitoring and alerting are first-class primitives supported in both of those products. A critical aspect of this role is determining our monitoring and alerting posture for the core observability system, as well as interfacing with other SRE and software teams across the business to ensure their infrastructure is seamlessly integrated.REQUIRED QUALIFICATIONSBuild & operate a robust, high-availability core observability plane — We run Anduril’s centralized observability system for 100s of environments. Engineering teams depend on these systems daily to remain operationally responsible systems - these systems must gracefully scale to meet demand and meet strict uptime requirements.Eagerly engage with our customers — We partner closely with engineering teams to understand their needs, gaps of our systems, and proactively engage with them so that our systems continually improve. We actively work to close gaps early on, and maintain a close working relationship with internal customers.Enable metrics-driven engineering and seamless investigation of production incidents — the ultimate mandate of the Observability team is to enable our organization to make data-driven decisions in real-time. This also extends to providing the substrate for agentic workflows for root-cause analysis, and enabling users to synthesize information across services and environments when triaging and actioning issues.PREFERRED QUALIFICATIONS8+ years of experience as a Site Reliability Engineer / related role supporting production systemsBachelors degree in Computer Science, Computer Engineering, Electrical Engineering, or related field (or equivalent experience)Familiarity with common container orchestration systems and cloud infrastructure (Docker, Kubernetes, experience developing and deploying systems on AWS, GCP, or Azure). Experience with infrastructure components of the observability stack is a plus (ClickHouse/ClickStack, Victoria Metrics, Prometheus, Grafana, ELK stack, etc).Experience serving as part of an on-call rotation to support high-availability systemsUser-focused engineering mindset - ability to translate user needs into technical solutions while balancing user experience with engineering constraintsU.S. Person status is required as this position needs to access export controlled dataUS Salary Range$166,000—$220,000 USDThe salary range for this role is an estimate based on a wide range of compensation factors, inclusive of base salary only. Actual salary offer may vary based on (but not limited to) work experience, education and/or training, critical skills, and/or business considerations. Highly competitive equity grants are included in the majority of full time offers; and are considered part of Anduril's total compensation package. Additionally, Anduril offers top-tier benefits for full-time employees, including: BenefitsAt Anduril, we invest in our people. Our comprehensive, competitive benefits package (available at little to no cost to employees) ensures you’re supported in health, recovery, and whatever comes next. For more information, Explore Our Benefits.Protecting Yourself from Recruitment ScamsAnduril is committed to maintaining the integrity of our Talent acquisition process and the security of our candidates. We've observed a rise in sophisticated phishing and fraudulent schemes where individuals impersonate Anduril representatives, luring job seekers with false interviews or job offers. These scammers often attempt to extract payment or sensitive personal information.To ensure your safety and help you navigate your job search with confidence, please keep the following critical points in mind:No Financial Requests: Anduril will never solicit payment or demand personal financial details (such as banking information, credit card numbers, or social security numbers) at any stage of our hiring process. Our legitimate recruitment is entirely free for candidates.Please always verify communications:Direct from Anduril: If you receive an email from one of our recruiters, it will only come from an @anduril.com address.Via Agency Partner: If contacted by a recruiting agency for an Anduril role, their email will clearly identify their agency. If you suspect any suspicious activity, please verify the agency's authenticity by reaching out to View email address on click.appcast.io. Exercise Caution with Unsolicited Outreach: If you receive any communication that appears suspicious, contains grammatical errors, or makes unusual requests, do not engage. Always confirm the sender's email domain is @anduril.com before providing any personal information or clicking on links.What to Do If You Suspect Fraud: Should you encounter any questionable or fraudulent outreach claiming to be from Anduril, please report it immediately to View email address on click.appcast.io. Your proactive caution is invaluable in protecting your personal information and upholding the security and trustworthiness of our recruitment efforts.Data PrivacyTo view Anduril's candidate data privacy policy, please visit . By submitting your application, you consent to Anduril Industries using a third-party service provider to conduct pre-employment risk, integrity, and due diligence screening and assessing potential risks as part of your application process. This third-party service provider provides risk-intelligence services that may include analysis of sanctions and watchlists, adverse media, public-record information, and other lawful open-source or commercial data sources. This third-party service provider does not act as a consumer reporting agency. Use of this provider helps to ensure compliance with applicable laws and protect technology, intellectual property, and organizational security.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer, Fleet Infrastructure in Washington DC vacancy
  • $176k - $276k

    Cloud Foundations Reliability (CFR) is part of NVIDIA’s Global Network Infrastructure (GNI) organization. We deploy, integrate...  ...We are looking for a hands-on senior engineer to own the lifecycle and...  ...large, multi-region Kubernetes fleets, including fleet-wide upgrades... 
    Senior
    Fleet
    Full time
    Remote work
    Weekend work

    Nvidia

    Washington DC
    10 hours ago
  • $160k - $185k

     ...skilled and motivated Sr. Infrastructure Engineer to join our Hardware...  ...performant and reliable infrastructure solutions...  ...the guidance of more senior engineers. Help...  ...Work closely with our Fleet Operations Team, enabling...  ...in cloud operations, site reliability... 
    Senior
    Fleet
    Full time
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    Coreweave

    Washington DC
    14 hours ago
  • $126.89k - $166.13k

     ...at College Park, and the infrastructure underneath that work is a...  ...build.We are looking for a Senior Systems and Network Engineer to own the classical infrastructure...  ...at our College Park site. This means provisioning...  ...tooling to manage a growing fleet of servers without... 
    Senior
    Fleet
    Permanent employment
    Contract work
    Work at office
    Remote work

    IONQ

    College Park, MD
    10 hours ago
  • $150k - $180k

     ...are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale...  ...the mission- and business-critical infrastructure that powers Umbra's systems. In this role...  ....This position is based on-site in either our Arlington, VA office, Reston... 
    Senior
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work
    Worldwide

    Umbra

    Arlington, VA
    1 day ago
  • $166k - $220k

     ...expectations. Our systems integration engineers internalize the nuances of each deployment...  ....ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our...  ...in the development of Kubernetes cloud infrastructure, DevOps, CI/CD and improving the developer... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    1 day ago
  • $207k - $284.9k

     ...AI by building the trusted, neutral infrastructure that enables organizations to...  ...mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from...  ...let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours
    Day shift

    Okta

    Washington DC
    4 days ago
  • $175k - $250k

     ...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to...  ...ensuring scalability, performance, and reliability across environments. What You’ll Do Design... 
    Senior
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    Washington DC
    1 day ago
  •  ...Azure, Oracle, Cassandra, SQL Server, My SQL and Mongo DB Seniority level Seniority level Mid-Senior level Employment type Employment...  ...new job is posted. Sign in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259,000.00 1 day ago... 
    Senior
    Contract work
    Remote work

    Signature IT World Inc

    Washington DC
    1 day ago
  • $149.4k - $202k

     ...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC...  ...of cloud native systems. Our SREs don’t just manage infrastructure; they build it using Infrastructure as Code (IaC), monitor... 
    Senior
    Remote work

    Noctua Technology

    Washington DC
    3 days ago
  • $140k - $205k

     ...Senior Technology Site Reliability Engineer Cooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operationsteam. Position summary: The Senior Technology Site Reliability Engineer("SRE") is responsible for ensuring the reliability... 
    Senior
    Full time
    Temporary work
    Work at office
    Flexible hours
    Weekend work

    Cooley

    Washington DC
    14 hours ago
  • $86k - $148k

     ...Position Summary We’re looking for a Senior Engineer to lead complex initiatives and elevate...  ...focused on vulnerability management, infrastructure updates, and compliance requirements....  ...solutions, and improve system reliability. Security-First Mindset: Experienced... 
    Senior
    Work at office
    Immediate start
    Remote work
    Flexible hours

    GrabJobs

    Washington DC
    4 days ago
  • $102.16k - $155k

     ...network architecture, reverse engineering, software and hardware...  ...synthetic training environments to fleet sustainment, environmental...  ...efforts across IT infrastructure, cybersecurity, physical facilities...  ...$169,413What you will doThe Senior Network Engineer is responsible... 
    Senior
    Fleet
    Full time
    Local area
    Worldwide

    HII Mission Technologies Division

    Springfield, VA
    4 days ago
  • $85.19k - $185k

     ...cybersecurity, network architecture, reverse engineering, software and hardware development...  ...and synthetic training environments to fleet sustainment, environmental remediation...  ...Security and Modernization efforts across IT infrastructure, cybersecurity, physical facilities,... 
    Senior
    Fleet
    Full time
    Work experience placement
    Local area
    Worldwide

    HII Mission Technologies Division

    Springfield, VA
    2 days ago
  • $145k - $200k

     ...Palantir’s core production infrastructure — 100s of K8s clusters —...  ...footprint or large. As a Senior Software Engineer on Substrate, you will design...  ...and operating the entire fleet of K8s clusters with zero...  ...architecture, design patterns, reliability and scaling) of new and... 
    Senior
    Fleet
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    2 days ago
  •  ...Title: Senior Embodied AI Engineer About Us: UnitX builds the world's leading physical AI systems...  ...the virtual proving grounds for our fleet, but also design, train, and deploy...  ...precision, low latency, and reliability. Basic Qualifications Education... 
    Senior
    Fleet
    Full time

    Unitx

    Washington DC
    14 hours ago
  • $191k - $253k

     ...business line at Anduril, and relies upon our fleet of autonomous vehicles, Lattice operating...  ...Anduril Cyber is hiring a Software Engineer focused on embedded and mission software...  ...products, and mission workflows cleanly and reliably.Incorporate AI/LLM-enabled workflow... 
    Senior
    Fleet
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    3 days ago
  •  ...diagnostics and ensuring quality repair standards within the fleet operation. This senior role influences uptime performance and resolves high-...  ...peers. Join us to drive efficient robotic deliveries and improve reliability through innovative solutions. #J-18808-Ljbffr... 
    Senior
    Fleet

    Industrious Ventures

    Washington DC
    3 days ago
  •  ...Service Technician III to lead technical efforts within their fleet repair operations in Washington, DC. This role requires advanced...  ...troubleshooting and coordination of daily repairs, ensuring quality and reliability standards are met across the technician team. With 4-6+ years... 
    Senior
    Fleet

    Serve Robotics

    Washington DC
    4 days ago
  •  ...business line at Anduril, and relies upon our fleet of autonomous vehicles, Lattice operating...  ...Anduril Cyber is hiring a Software Engineer focused on embedded and mission software...  ...products, and mission workflows cleanly and reliably. Incorporate AI/LLM-enabled workflow... 
    Senior
    Fleet

    PVH (Tommy Hilfiger/Calvin Klein)

    Washington DC
    3 days ago
  •  ...satellite operator, providing reliable and secure satellite-...  ...years.  Backed by a legacy of engineering excellence, reliability and industry...  ...s state-of-the-art Satellite fleet consists of 12 GEO satellites...  ...LinkedIn or visit As our Senior NMS developer, you will be responsible... 
    Senior
    Fleet
    Work at office
    Worldwide

    Telesat

    Bethesda, MD
    10 days ago
  •  ...Palantir’s core production infrastructure — 100s of K8s clusters —...  ...footprint or large. As a Senior Software Engineer on Substrate, you will design...  ...and operating the entire fleet of K8s clusters with zero...  ...architecture, design patterns, reliability and scaling) of new and... 
    Senior
    Fleet

    hackajob

    Washington DC
    4 days ago
  •  ...Senior Platform Engineer Washington, DC – onsite US citizenship required...  ...secure, innovative hybrid infrastructure for one of the nation's...  ...standards • Manage Kubernetes fleet operations across AWS EKS,...  ...and improve platform reliability across both cloud and on-premises... 
    Senior
    Fleet
    Contract work
    Local area

    System One

    Washington DC
    a month ago
  •  ...and forecasting activities. Analyze supply chain, sustainment, engineering, manpower, and operational processes. Develop analytical products...  ...Qualification Prior USCG, DHS, DoD, logistics, maintenance, or fleet support experience. Experience supporting strategic planning... 
    Senior
    Fleet
    For contractors

    Strategic Resilience Group LLC

    Washington DC
    1 day ago
  • $102.3k - $209.5k

     ...guardrails. Enhances automation for fleet operations; improves CI/CD,...  ...paradigms to maintain reliability and availability. Work closely with platform engineering, operations, firmware development...  ...brings together the data, infrastructure, applications, and expertise... 
    Senior
    Fleet
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Washington DC
    4 days ago
  •  ...Senior Platform Engineer Washington, DC – onsite US citizenship required...  ...secure, innovative hybrid infrastructure for one of the nation's...  ...standards • Manage Kubernetes fleet operations across AWS EKS,...  ...and improve platform reliability across both cloud and on-premises... 
    Senior
    Fleet
    Contract work
    Local area

    System One

    Washington DC
    a month ago
  • $166k - $220k

     ...Anduril, and relies upon our fleet of autonomous vehicles, Lattice...  ...Cyber is hiring a Software Engineer focused on embedded and mission...  .... The role is aimed at a senior hands-on engineer who can connect...  ...while balancing performance, reliability, manufacturability, and... 
    Senior
    Fleet
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    1 day ago
  • $165k - $205k

     ...Rumble Cloud is seeking a Senior Platform Engineer (Security) to help operate...  ...and continuously harden the infrastructure that powers our public...  ...hardening across our Linux fleet and cloud control plane, track...  ...the platform meets both reliability and security expectations... 
    Senior
    Fleet

    Rumble

    Washington DC
    12 days ago
  • $185k - $230k

    As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services... 
    Senior
    Full time
    Local area
    Immediate start

    MetroStar Systems

    Washington DC
    1 day ago
  • $253.9k - $298.7k

     ...managing a validator fleet that represents...  ...security, compliance, and reliability. The Role We are looking for a Senior Staff Software Engineer to serve as Coinbase...  ...and build critical infrastructure for validator...  ...compatible with this site click here to download... 
    Senior
    Fleet
    Local area

    Coinbase

    Washington DC
    1 day ago
  • $165k - $230k

     ...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink...  ...both on-premises and in the cloudDeploy and manage core infrastructure such as databases, monitoring and storageClosely collaborate... 
    Senior
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer, Fleet Infrastructure. Be the first to apply!