Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$112k - $149k

Anduril

Anduril Industries is a defense technology company with a mission to transform U.S. and allied military capabilities with advanced technology. By bringing the expertise, technology, and business model of the 21st century’s most innovative companies to the defense industry, Anduril is changing how military systems are designed, built and sold. Anduril’s family of systems is powered by Lattice OS, an AI-powered operating system that turns thousands of data streams into a realtime, 3D command and control center. As the world enters an era of strategic competition, Anduril is committed to bringing cutting-edge autonomy, AI, computer vision, sensor fusion, and networking technology to the military in months, not years.About the Team The Imaging team builds and fields state-of-the-art camera and sensor systems deployed to solve real security challenges for the United States and its allies. We work across the full stack, from bare-metal hardware and firmware to networked services and cloud integrations, and we own our systems all the way through to the field. When something breaks in a deployed environment, we fix it. About the Role We're looking for a Site Reliability Engineer to join the Imaging team. This is not a product development role, and it isn't a traditional cloud-SRE role either. You are the frontline for keeping fielded imaging systems alive: the person field personnel and customer-support escalations turn to when a deployed system isn't behaving. You'll be the second SRE on a small, high-trust team, working directly with our lead, with real room to shape how Imaging reliability and support operate as we scale. The work is varied and often ambiguous. You might spend a morning triaging a networking failure on a system being stood up at a remote site, an afternoon walking a field operator through a sensor calibration over the phone, and the back half of a week hardening a runbook so the next person never has to solve that problem live again. When a fielded system misbehaves, the root cause could be anywhere in the stack--and you're the one who narrows it down. If you get energy from diagnosing live systems under pressure, being the reason a deployment stays up, and turning recurring fire-drills into durable fixes, this role is built for you. What You'll Do Own fielded system reliability. You are responsible for the health and uptime of deployed imaging systems. When issues arise--whether the root cause is in the network, the calibration, an upgrade, or the sensor hardware itself--you triage, diagnose, and drive resolution. Run point on escalations. You'll be first and second line of response for issues coming through our support channels and Anduril's customer-support pipeline (Tier 0 to Mission Success / Product Operations to SRE), acting as the deep-expertise backstop the rest of the funnel escalates to. Turn fires into runbooks. Recurring issues shouldn't be solved live twice. You'll build and maintain runbooks, diagnostics, and self-service tooling that reduce repeat problems and shrink the support load over time. Hold the boundary with engineering. You own everything short of a code fix. When an issue turns out to be a genuine software defect, you'll cleanly reproduce it, document it, and hand it off to the Mission Software Engineers--protecting their focus by keeping frontline support where it belongs. Feed reliability back into the product. You'll turn what you learn in the field into signals that make our systems more supportable--better observability, safer upgrades, more graceful failure. Travel expected approximately 15% of the time for field support and deployment windows. What We're Looking For Required 3+ years in SRE, DevOps, field/systems engineering, or production support of deployed hardware/software systems--with real ownership after systems ship, not just standing them up Strong Linux fundamentals, including comfort troubleshooting real networking issues (IP, routing, VPNs, connectivity in constrained or field environments) Demonstrated ability to diagnose and resolve issues across system boundaries (networking, services, hardware interaction) without always having full visibility into every component Comfortable owning a structured on-call rotation, including scheduled after-hours and weekend coverage Strong written and verbal communication skills, including the ability to run a remote troubleshooting session with a non-technical operator and document what happened Eligibility to obtain and maintain a U.S. Secret clearance Preferred Experience supporting fielded or deployed systems, not just development environments Experience with fielded hardware or sensor systems (EO/IR, optical, or similar), including familiarity with sensor calibration Scripting for diagnostics and automation (Python, Bash, or similar) Familiarity with Nix or NixOS--uncommon, but valuable on our stack Familiarity with systemd service management and observability practices on Linux Familiarity with incident tooling (PagerDuty or equivalent) What Makes Someone Successful Here The people who thrive in this role share a few traits that don't always show up on a resume: They close loops. When they pick up a problem they own it until it's resolved, not until it's someone else's problem. They stay calm under pressure. A system down in the field with an operator waiting is a high-stress moment. They work the problem methodically instead of reacting. They communicate proactively. They don't wait to be asked for a status update. They document what they find, share what they learn, and keep the people who need to know informed. They're comfortable with ambiguity at the stack boundary. When a device behaves unexpectedly in the field, the root cause could be anywhere. They have a methodology for narrowing it down systematically rather than guessing or waiting for someone else to have the answer. They know where the line is. They own frontline support fully, and they know when an issue is a true code-level defect that belongs with engineering--handing it off cleanly rather than either sitting on it or tossing it over the wall half-formed.US Salary Range$112,000—$149,000 USDThe salary range for this role is an estimate based on a wide range of compensation factors, inclusive of base salary only. Actual salary offer may vary based on (but not limited to) work experience, education and/or training, critical skills, and/or business considerations. Highly competitive equity grants are included in the majority of full time offers; and are considered part of Anduril's total compensation package. Additionally, Anduril offers top-tier benefits for full-time employees, including: Benefits At Anduril, we invest in our people. Our comprehensive, competitive benefits package (available at little to no cost to employees) ensures you’re supported in health, recovery, and whatever comes next. For more information, Explore Our Benefits.

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Waltham, MA vacancy
  •  ...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient... 
    Suggested
    Ongoing contract
    Full time
    Temporary work
    Work experience placement
    Worldwide

    SS&C Technologies

    Waltham, MA
    1 day ago
  • $40 - $45.78 per hour

     ...Job Description Job Description Site Reliability Engineer 1 Job Details Site Reliability Engineer 1 (Contract) Location: Waltham, MA 02451 (Hybrid) Duration: 10/22/2025 to 4/03/2026 Team: Campaign Core RD US Key Responsibilities: Deploy and manage... 
    Suggested
    Hourly pay
    Contract work

    Cypress HCM

    Waltham, MA
    22 days ago
  • $134.25k - $214.8k

     ...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance... 
    Suggested
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Axon

    Boston, MA
    1 day ago
  •  ...and best in class outcomesVisionary in future focused problem-solvingExceptional in execution and impactThe RoleAs a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to deliver... 
    Suggested
    Full time
    Flexible hours

    Proofpoint

    Boston, MA
    1 day ago
  • $134.25k - $214.8k

     ...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed...  ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,... 
    Suggested
    Work experience placement
    Work at office
    Remote work

    Axon

    Boston, MA
    4 days ago
  • $160k - $200k

     ...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident... 
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interface

    Somerville, MA
    2 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    3 days ago
  • $130k - $150k

     ...systems and hybrid infrastructure, meaning experience with cloud technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are reliable, scalable, and performant across on-premises and cloud... 
    Work at office
    Work from home
    3 days per week

    CRA International

    Boston, MA
    5 days ago
  • $127k - $249k

     ...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas...  ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    4 days ago
  • $160k - $200k

    Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware... 
    Local area
    Remote work

    QuEra Computing

    Boston, MA
    2 days ago
  • $135k - $165k

     ...Own assigned problem statements, estimate work, balance effort and risk, and drive tasks to completion with guidance from senior engineers. Collaborate across engineering teams and communicate progress, blockers, and outcomes to technical and non-technical stakeholders... 
    Full time

    Xometry

    Boston, MA
    17 days ago
  •  ...reviews across the platform. Support platform architecture decisions and technology standardization under the guidance of senior engineers. Design and build a Kubernetes-based platform for automated lifecycle management of core infrastructure. Collaborate with... 
    Full time
    Work at office
    Remote work

    Axon

    Boston, MA
    2 days ago
  •  ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the...  ...ideal candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based... 
    Full time

    Federal Reserve Bank of Boston

    Boston, MA
    3 days ago
  • Role Description & ResponsibilitiesOur DELMIAWORKS team is looking for a Pre-sales Solution Engineer. The preferred location for this role is on the East Coast.*Please note: This is for a future opening. By submitting your application, you will be added to our talent pool... 

    DASSAULT SYSTÈMES

    Waltham, MA
    2 days ago
  •  ...DescriptionTekWissen provides a unique portfolio of innovative capabilities that seamlessly combines clients insights, strategy, design, software engineering, and systems integration. Job DescriptionWe are currently seeking a talented DB2 Systems Programmer for a 1 YEAR project in... 

    TekWissen

    Waltham, MA
    1 day ago
  • $86.9k - $198k

     ...specifications makes you an integral part of delivering a customer-focused engineering solution.As an IT Systems Engineer on our team, you’ll have the...  ...total benefits by visiting the Resource page on our Careers site and reviewing Our Employee Benefits page.Salary at Booz Allen... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Lexington, MA
    1 day ago
  • $127k - $220k

     ..., MAHybridFull Time$127k - $220k Job Description AI Solutions Engineer Location: Waltham, MAWork Schedule: Hybrid, 3 Days Onsite Per...  ...beyond proof-of-concept and into production environments where reliability, scalability, and measurable business impact matter. This position... 
    Full time
    Flexible hours
    3 days per week

    Motion Recruitment

    Waltham, MA
    2 days ago
  • $100k - $300k

     ...society to the next level.Tutor Intelligence builds software to enable ordinary robots to achieve extraordinary things. As a platform engineer, your work lies at the center of this challenge, orchestrating real time robot code, machine learning systems, data labeling... 
    Full time
    Work at office
    Shift work

    Tutor Intelligence

    Watertown, MA
    1 day ago
  • $114k - $184k

    DescriptionJob Title: Platform SW Engineer III The Elevator Pitch  Are you looking for a role that has meaning and purpose on a team that...  ...a platform strategy so future product lines share a common, reliable software foundation. Actively contribute to architecture and design... 
    Full time
    Work at office
    Local area
    Remote work
    Flexible hours

    Evolv Technology

    Waltham, MA
    3 days ago
  •  ...organization in the insurance industry is seeking a Senior Software Engineer, Platform Engineering to serve as the primary Platform...  ...technical direction Champion engineering excellence, operational reliability, and continuous improvement across the development lifecycle... 
    Local area
    Flexible hours

    Motion Recruitment

    Waltham, MA
    2 days ago
  • $100k - $140k

     ...we transition to a large-scale research fleet, the Controls Reliability Team serves as a Technical Force Multiplier. In this role, you...  ...will be the bridge between high-volume field operations and Engineering, resolving complex system failures and building the diagnostic... 
    Full time

    Boston Dynamics

    Waltham, MA
    1 day ago
  • $100k - $300k

     ...the next level.Tutor Intelligence builds software to enable ordinary robots to achieve extraordinary things. As a robotics software engineer, your work lies at the center of this challenge, orchestrating real time robot code, optimization systems for motion planning,... 
    Full time
    Work at office
    Flexible hours
    Shift work

    Tutor Intelligence

    Watertown, MA
    1 day ago
  • $100k - $300k

     ...effectorsExperience integrating and programming robot armsMechanical design of mounts and stands.About Our Roles & TitlesAt Tutor, we believe great engineers and researchers are defined by what they build and the impact they have — not where they sit in an org chart or what title they... 
    Full time
    Work at office
    Shift work

    Tutor Intelligence

    Watertown, MA
    1 day ago
  • $198k - $300k

    We're looking for a Senior Engineering Manager to lead our ML Platform Team - a growing team responsible for the foundational infrastructure...  ..., particularly in the near term as the team growsDrive reliability, performance, and cost efficiency across distributed training... 
    Full time
    Temporary work
    Work at office

    Boston Dynamics

    Waltham, MA
    5 days ago
  •  ...'s largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionDirector of Engineering AI/Data ScienceLocation(s): Waltham, MAGet To Know the TeamYou'll be joining a collaborative, fast-moving team of data scientists... 
    Ongoing contract
    Full time
    Casual work
    Flexible hours

    SS&C Technologies

    Waltham, MA
    2 days ago
  • $127k - $220k

    Waltham, MAHybridFull Time$127k - $220k Job Description Software Engineer, AI Applications Location: Waltham, MAWork Schedule: Hybrid, 3...  ...and monitoring processes to improve AI performance and reliability Collaborate with Product and Engineering teams to design new features... 
    Full time
    Flexible hours
    3 days per week

    Motion Recruitment

    Waltham, MA
    2 days ago
  • $196.35k - $292.6k

    Job SummaryAs a Software Engineer, you will play a key role in delivering an enterprise‑class NetApp Software Defined Storage (SDS) product...  ..., SREs, and Product Managers. You will contribute to scalable, reliable storage systems that power mission-critical cloud workloads,... 
    Local area

    NetApp

    Waltham, MA
    2 days ago
  • $69.3k - $158k

    Software EngineerThe Opportunity:As a software engineer, you know that great software isn’t just built—it’s crafted. It’s more than an elegant...  ...our total benefits by visiting the Resource page on our Careers site and reviewing Our Employee Benefits page.Salary at Booz Allen is... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Lexington, MA
    3 days ago
  • $163.8k - $257.4k

     ...ll be joining a highly interdisciplinary team transforming the core of Zoominfo’s business. Collaborating with PMs, designers, and engineers you will rapidly drive innovation by building intelligent systems that power cutting edge user experiences, create valuable... 
    Worldwide

    ZoomInfo

    Waltham, MA
    2 days ago
  • $166k - $220k

     ...RoleWe're looking for a Mission Software Engineer to join the Imaging Interfaces team. This...  ...environments. These services need to be reliable, observable, and maintainable by engineers...  ...teams; sometimes remotely, sometimes on-site. You don't need to know everything, but you... 
    Full time
    Work experience placement
    Immediate start
    Remote work
    Day shift

    Anduril Industries

    Lexington, MA
    19 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!