Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$170k - $219k

RADAR

Job Description

Job Description

ABOUT US

E-commerce got real-time data infrastructure decades ago. Physical stores still have not. RADAR is changing that.

RADAR is building the data infrastructure layer for the physical world, starting with retail. Our hardware-enabled SaaS platform uses proprietary overhead sensors, software, and AI-powered analytics to locate every product in a store, continuously, down to the fixture. We are deployed across 1,500+ stores with retailers including American Eagle Outfitters and Old Navy, processing tens of billions of real-world events every day, delivering 99%+ accuracy in complex, noisy environments - at fleet scale.

RADAR is one of the best-funded companies in retail technology, backed by a recent Series B financing at a $1 billion valuation. Inventory accuracy is only the beginning. We believe RADAR can become foundational infrastructure for the physical economy, powering new AI-driven commerce experiences across retail and beyond.

Join us if you want to work on a large, unsolved, technically challenging problem with an ambitious team building category-defining technology. OUR VALUES
  • Mission-Driven: We're transforming retail with cutting-edge technology and building something that truly matters.
  • Collaborative Team: We thrive on curiosity, shared goals, and solving complex problems together.
  • High Impact: You'll make meaningful contributions from day one and help shape the future of our product and company.
  • Clear Communication: We value honesty, humility, and respectful dialogue—everyone's voice matters.
  • Balanced Lives: We work hard, but not at the expense of well-being. We respect time, boundaries, and life outside of work.
  • Diverse Perspectives: We believe better ideas come from diverse backgrounds, experiences, and viewpoints.
  • Empathy-Driven Design: We build with deep respect for our end users, listening closely to their feedback and needs.
ABOUT THE JOB

RADAR runs data infrastructure across 1,600+ live retail stores, processing tens of billions of real-world events every day. We're hiring a Site Reliability Engineer to own the reliability of that system end to end — leading incident response, running day-to-day NOC operations, and building the observability foundation that lets us catch issues before they hit a store floor. You'll be the steady hand during a live incident, and the engineer making sure there are fewer of them to begin with.

Responsibilities:
  • Own the incident management lifecycle end to end: detection, triage, escalation, communication, resolution, and postmortem for production incidents.
  • Act as Incident Commander for high severity incidents, coordinating across engineering, support, and leadership until resolution.
  • Run day-to-day NOC (Network Operations Center) operations, including 24/7 shift coverage, escalation matrices, and shift handover protocols.
  • Coach and mentor NOC analysts on triage discipline and escalation judgment, and own NOC KPIs like response time and escalation accuracy.
  • Design and maintain observability pipelines across metrics, logs, and traces, and define SLIs/SLOs with engineering and product.
  • Build dashboards and alert that surface true signal from our sensor and platform data, cutting down on noise and alert fatigue.
  • Facilitate blameless postmortems and root cause analysis, and track corrective actions through to closure.
  • Maintain on-call rotations, runbooks, and escalation policies, and report on MTTA/MTTR/MTBF trends to leadership.
ABOUT YOU Required :
  • You have 5+ years of experience in Site Reliability Engineering, DevOps, Infrastructure, or Production Operations, with direct incident response and on-call experience.
  • You have experience running or actively contributing to a NOC, including shift scheduling, escalation processes, and performance metrics.
  • You have strong hands-on experience with observability tooling (Prometheus, Grafana, Datadog, New Relic, Splunk, ELK, OpenTelemetry, or similar).
  • You have a solid understanding of SLIs, SLOs, SLAs, and error budgets, and how to use them to drive prioritization.
  • You have hands-on release engineering experience, including CI/CD pipelines, deployment automation, and safe rollout practices like canary releases, feature flags, and automated rollbacks.
  • You are proficient in at least one scripting or programming language (Python, Go, Bash, etc.).
  • You have experience with infrastructure-as-code tools (Terraform, Ansible).
  • You have experience with cloud platforms (AWS, GCP, or Azure) and container orchestration (Kubernetes, Docker).
  • You are a clear, direct communicator who stays calm and organized under pressure during live incidents.
Preferred:
  • You have experience building or scaling a NOC from the ground up.
  • You have a background in distributed systems architecture and microservices troubleshooting.
  • You have familiarity with chaos engineering and resilience testing.
  • You have a certification such as AWS Certified SysOps Administrator, Google Professional Cloud DevOps Engineer, or ITIL.
WHAT YOU'LL DO In your first 30 days, you will:
  • Learn RADAR's mission, technology stack and core values.
  • Complete onboarding and security compliance training.
  • Shadow the NOC across shifts and review recent incident history and open postmortem action items.
In your first 60 days, you will:
  • Take on-call as primary or secondary responder for at least one service area.
  • Tune or consolidate at least one high-volume, low-signal alert source, and audit existing observability coverage for major gaps.
  • Draft or update runbooks for the top recurring incident types, and instrument one under-monitored service.
In your first 90 days, you will:
  • Lead Incident Commander duties for high severity incidents, including full postmortem facilitation.
  • Deliver a reliability report on incident trends, NOC KPIs, and observability maturity gaps.
  • Present a roadmap for the next 2–3 quarters covering NOC process, Automation, SLO definitions, and observability investment.

At RADAR, your base pay is one part of your total compensation package. The expected base salary range for this position is $170,000 - $219,000 . Individual pay is determined by work location and additional factors,  including job-related skills, experience and relevant education or training. You will also be eligible to receive other benefits including: equity, comprehensive medical and dental coverage, life and disability benefits, 401k plan,  flexible time off, and paid parental leave. The pay range listed for this position is a good faith and reasonable estimate of the range of possible base compensation at the time of posting. 

Research has shown that women & underrepresented minorities are more likely to read lists of requirements and consider themselves unqualified if they don't meet every single one. This list represents what we're ideally looking for, but everyone has unique strengths & weaknesses, and we hire for strength & potential, not lack of weakness.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Seattle, WA vacancy
  •  ...thousands of customers depend on every day. We're hiring a senior, hands-on engineer to own the reliability, availability, security, and performance of that...  ...you are: ~5+ years of hands-on Cloud Operations and Site Reliability Engineering, operating production-scale SaaS... 
    Senior
    Full time

    MangoApps

    Seattle, WA
    1 day ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper). As... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Seattle, WA
    4 days ago
  • $134.25k - $214.8k

     ...upload. Every piece of digital evidence. Every chain of custody log that holds up in court. That's us.Axon's Platform team is the engine behind what hundreds of thousands of officers rely on every day. We're one of the world's largest blob storage customers, ingesting... 
    Senior
    Work experience placement
    Work at office
    Remote work

    Axon

    Seattle, WA
    1 day ago
  • $55k - $151.47k

     ...ApplicableSpecialismIFS - Internal Firm Services - OtherManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our... 
    Senior
    Full time
    H1b

    PwC

    Seattle, WA
    2 days ago
  • $170k - $219k

     ...Site Reliability Engineer RADAR runs data infrastructure across 1,600+ live retail stores, processing tens of billions of real-world events every day. We're hiring a Site Reliability Engineer to own the reliability of that system end to end — leading incident response... 
    Senior
    Flexible hours
    Shift work
    Night shift

    Radar

    Seattle, WA
    1 day ago
  •  ...system and process health, performance, and reliability, including alerting quality, dashboards,...  ...while aligning stakeholders across IT, engineering and partner teams. Proactively...  ...experience. ~5+ years of experience in site reliability engineering, infrastructure... 
    Senior

    Blink Health

    Seattle, WA
    1 day ago
  • $160k - $250k

     ...DevOps And Systems Engineer Hive is the leading provider of cloud-based AI solutions to understand, search, and generate content...  ...machine learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS offering... 
    Senior

    Hive

    Seattle, WA
    1 day ago
  •  ...team is responsible for the reliability, scalability, and efficiency...  ...building features, but about engineering the resilience and performance...  ...maintain system stability.As a Site Reliability Engineer, you...  ...practices while working alongside senior engineers to solve... 
    Senior

    TikTok

    Seattle, WA
    4 days ago
  • $232k - $319k

     ...to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and...  ...enabled with self-serviceAccelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    Bellevue, WA
    22 hours ago
  • $160k - $200k

     ...change and achieving remarkable growth in a rapidly evolving industry. Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management across Cognitiv. Our immediate challenge is to scale... 
    Senior
    Work at office
    Immediate start
    Remote work
    Work from home

    Cognitiv

    Bellevue, WA
    2 days ago
  • $190k - $250k

    Sr Manager, Site Reliability Engineering About Invoca Invoca is an AI-powered revenue execution platform that brings together marketing, commerce...  ...view on what great looks like. This position reports to the Senior Director, Infrastructure & Security. What you'll do Lead... 
    Senior
    Currently hiring
    Local area
    Remote work
    Flexible hours

    Invoca

    Seattle, WA
    1 day ago
  • $134.25k - $214.8k

     ...real change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability... 
    Senior
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Axon

    Seattle, WA
    2 days ago
  •  ...This is an engineering-first Senior SRE role. We’re looking for senior engineers who have: Built and shipped significant backend systems...  ...end-to-end in production (design → launch → on-call → reliability improvements) Led incident response and driven durable follow... 
    Senior

    Practice by Numbers

    Bellevue, WA
    2 days ago
  •  ...certification), ISO 27001:2005 Information Security Management System (ISMS), and CMMI-DEV Level 3. Job Description Sr. Site Reliability Engineer Location – Seattle, WA Duration – 12 months Interview – in-person if local or Phone + Skype Minimum... 
    Senior
    Local area
    Worldwide

    Comtech LLC

    Seattle, WA
    1 day ago
  •  ...Senior Site Reliability Engineer (SRE) Location: Seattle, hybrid - 2 times a week in the office Job Type: Full-time, direct hire Industry: High-Growth Technology / SaaS About the Role We are seeking a highly skilled Senior Site Reliability Engineer... 
    Full time
    Work at office

    TalentDome Staffing

    Seattle, WA
    1 day ago
  • $143k - $194k

     ...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental...  ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and... 
    Full time
    Temporary work
    Work experience placement
    Immediate start

    Anduril Industries

    Seattle, WA
    2 days ago
  • $151.2k - $204.6k

    Would you like to be an engineer who builds the systems that power advertising...  ...every day, where latency, reliability, and quality translate directly...  ...customer trust.We are looking for a Senior System Development Engineer, operating as a Site Reliability Engineer, to raise... 
    Flexible hours

    Amazon

    Seattle, WA
    22 hours ago
  •  ...The RoleThis hybrid role combines the hands-on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE).The ideal candidate has a strong technical foundation, thrives in a... 
    Full time
    Local area

    F5 Networks

    Seattle, WA
    15 hours ago
  • Company DescriptionComtech LLC is a woman-owned small business focused on delivering end-to-end solutions and products. Since 1998, we have successfully serviced enterprises across the public and private sectors, and the Department of Defense. Our services span all aspects...

    Comtech

    Seattle, WA
    4 days ago
  • $184.3k - $264.95k

     ...Principal Site Reliability Engineer At UKG, the work you do matters. The code you ship, the decisions you make, and the care you show a customer all add up to real impact. Today, tens of millions of workers start and end their days with our workforce operating platform... 

    UKG, Inc.

    Seattle, WA
    4 hours ago
  • $152k - $241.5k

     ...the world.Join the Simulation Software team at NVIDIA as a Senior System Software Engineer! This role offers an outstanding opportunity to work on...  ...development by enabling Chips Simulation as a trusted and reliable virtual platform.What you will be doing:Drive early... 
    Senior
    Full time

    Nvidia

    Seattle, WA
    22 hours ago
  •  ...where you can push the limits of what's possible.As a Lead Software Engineer at JPMorganChase within the Enterprise Technology,...  .... These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare... 

    JP Morgan Chase

    Seattle, WA
    22 hours ago
  • $204k - $306k

     ...all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity, from...  ...in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    Okta

    Bellevue, WA
    2 days ago
  • $165k - $206k

     ...coordinate support and resolve platform issues across CPU/radio SoCs, MCU/PIC, NPU/GPU, and peripheral devices.Support hardware engineering teams with deep technical debugging and contribute to OS/platform modernization efforts.What You’ll NeedBasic Qualifications:Bachelor... 
    Senior
    Full time
    Temporary work
    Work at office
    Immediate start
    Visa sponsorship
    Work visa

    Sonos

    Seattle, WA
    22 hours ago
  • $194k - $267k

     ...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Bellevue, WA
    2 days ago
  • $194k - $267k

     ...do something more than once, automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Bellevue, WA
    3 days ago
  •  ...future together. We are responsible for the reliability of all the company's major data warehouse products, services, and query engines. We serve business needs across domains...  ...practices, and emerging technologies related to site reliability and infrastructure engineering.... 

    TikTok

    Seattle, WA
    4 days ago
  •  ...Engineering, Product, Design, and Marketing Engineering Compensation ~ Zone 1 Base Pay: $214K – $260K Superhuman offers...  ...role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them... 
    Worldwide
    Home office
    Flexible hours

    Superhuman

    Seattle, WA
    2 days ago
  • $95k - $134k

     ...disrupting the industry it helped build. Job Application Deadline: 10/31/2026 The Opportunity DAT is looking for a Site Reliability Engineer to join our SRE platform team. This position will work hybrid in Seattle, WA or Portland, OR Candidate profile DAT... 
    Temporary work
    For contractors
    Work experience placement
    Work at office
    Local area
    Immediate start
    Flexible hours

    DAT Freight & Analytics

    Seattle, WA
    2 days ago
  •  ...on one unified cloud. One cloud for compute, inference, and agents. Role Overview We are seeking a skilled Site Reliability Engineer to join the GMI Global Infrastructure team. This role is hands-on and critical to ensuring the stability, efficiency, and... 

    GMI Cloud

    Seattle, WA
    5 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!