Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Sr. Site Reliability Engineer

Full-time

Veeamsoftware

Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in both data resilience and data security posture management, Veeam is built for the convergence of identity, data, security, and AI risk. Headquartered in Seattle with offices in more than 30 countries, Veeam protects over 550,000 customers worldwide, who trust Veeam to keep their businesses running. Join us as we go fearlessly forward together, growing, learning, and making a real impact for some of the world’s biggest brands. About The Role Veeam is building a global SRE function to support the Veeam Data Cloud, our new SaaS platform. This role focuses on our Government and Sovereign Cloud environment. Due to clearance and access requirements, this team operates with restricted access to GOV infrastructure. That means you'll be part of a small team responsible for the full platform stack — including all VDC workloads. You won't always be able to hand off problems to other teams; you need to understand the entire architecture well enough to own it. You'll need to get up to speed on the platform quickly, often by reading code, docs, and architecture artifacts rather than getting direct access to environments from day one. This is a ground-up role — you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going forward. What You'll Do Discovery & Documentation - Get up to speed on the full platform — all VDC workloads, dependencies, and risk areas. Much of this will happen through code, docs, and conversations rather than direct environment access. - Work with SMEs across the org to fill knowledge gaps and build onboarding material for the team. - Write and maintain runbooks, architecture docs, and operational guides. Reliability & Incident Response - Design infrastructure for high availability and fault tolerance on Azure (including Azure Government). - Define SLIs, SLOs, and error budgets where none exist today. - Run incident response and blameless postmortems.

Turn incidents into improvements. - Identify reliability risks across modern and legacy workloads and build practical remediation plans that work within compliance constraints. Observability - Close observability gaps — define instrumentation requirements and drive implementation. - Set alerting, telemetry, and monitoring standards with partner teams. - Build automation to reduce toil and support fleet management. - Participate in on-call rotations. Infrastructure & Delivery - Work with IaC , CI/CD, deployment automation, and config management — including in air-gapped or compliance-restricted environments. - Build and maintain testing, canary deployment, and release validation pipelines. - Integrate chaos engineering and monitoring tools, adapting choices to meet regulatory requirements. Collaboration - Work across product, platform, security, legal, compliance, and operations teams. - Own problems end-to-end — identify gaps, drive solutions, don't wait for direction. - Mentor other engineers and help spread SRE practices across the org. Technologies we work with - Microsoft TFS, Azure DevOps, Git, BitBucket - Azure (Entra ID, API Management, Cosmos Db, Storage services, Azure Functions, static website hosting, Azure security, etc. ) - IaC tools (Azure ARM templates, AWS CloudFormation, Terraform, the Serverless Framework, etc. ) - Observability (Azure Monitor, AppInsights, Elastic Stack) What You'll Bring - 7+ years in Software Engineering, with 3+ years in SRE, Platform Engineering, or similar — across multi-service platforms, not just single-service environments. - Experience with Government or Sovereign Cloud (e. g. , Azure Government, AWS GovCloud).

Vacancy posted 12 days ago
Similar jobs that could be interesting for youBased on the Sr. Site Reliability Engineer in Remote vacancy
  • $87.12k - $151.25k

     ...to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a Digital Site Reliability Sr Engineer - Remote to join our team in Memphis, Tennessee (US-TN), United States (US).Digital Site Reliability Senior EngineerWe are... 
    Senior
    Temporary work
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Memphis, TN
    1 day ago
  • $134.25k - $214.8k

     ...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance... 
    Senior
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Axon

    Seattle, WA
    4 days ago
  •  ...About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure... 
    Senior
    Remote work

    MeridianLink

    United States
    1 day ago
  •  ...consumers and companies, alikeKlover’s engineering team powers one of the fastest-growing fintech...  ...-grade systems that prioritize reliability, security, and performance, and that integrate...  ...the RoleAs a Senior/Staff Site Reliability Engineer, you will play a critical... 
    Senior
    Work at office
    Immediate start
    Remote work

    Attain Data

    Chicago, IL
    4 days ago
  • $160k - $185k

     ...in their fitness journey and revolutionized the industry along the way. And we’re just getting started!OverviewThe Sr. Manager, Site Reliability Engineering (SRE) leads the strategy, execution, and continuous improvement of reliability, availability, and performance... 
    Senior
    Work at office
    Local area
    Remote work
    Work from home

    Planet Fitness

    Hampton, NH
    1 day ago
  • $107.8k - $162k

     ...an expectation of a minimum of three days per week working in the office and flexibility to work remotely on the remaining days. On-site expectations may evolve over time to support business needs, with clear communication provided in advance. Job Description Operates... 
    Senior
    Work at office
    Local area
    Remote work
    3 days per week

    Green Dot

    Los Angeles, CA
    4 days ago
  •  ...Sr Site Reliability Engineer (SRE) SigNoz is an open-source observability platform that helps modern engineering teams monitor, debug, and optimize their applications with deep visibility into metrics, traces, and logs — all in one place. We're built natively on OpenTelemetry... 
    Senior
    Remote work

    SigNoz

    United States
    5 days ago
  •  ...infrastructure that companies around the world depend on to keep their engineering teams focused on what matters most - their own product....  ...that everyone can appreciate. About the Role: As a Site Reliability Engineer, you will play a critical role in ensuring the... 
    Senior
    Remote work
    Flexible hours

    AuthZed

    Canada, KY
    3 days ago
  •  ...Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has been at the forefront of online real estate...  ...confidence through expert guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization,... 
    Senior
    Work at office
    Local area

    realtor.com

    Austin, TX
    3 days ago
  •  ...straightforward communication and clinical domain expertise, Commence cuts straight to better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability, scalability, and operational health of our mission-critical healthcare data... 
    Senior
    Remote work

    Commence Corporation

    United States
    8 hours ago
  •  ...Financial Software company Position: SRE Engineer Location: Hybrid remote/Buffalo Grove/Chicago Comp: Solid...  ...; coach teams on using them to balance feature velocity with reliability and communicate system health to stakeholders. • Lead alert... 
    Senior
    Remote work

    The Judge Group

    Wheeling, IL
    4 days ago
  •  ...Job Description The Opportunity: Versant's Sports & Entertainment Digital Products division is seeking a Senior Site Reliability Engineer to help drive the reliability, scalability, and usability of internal developer platforms, tooling, and engineering workflows... 
    Senior
    Local area
    Remote work
    Worldwide

    Versant

    United States
    1 day ago
  • $140k - $170k

     ...and intelligence. If you want to push yourself and reshape a $200B+ market, we're excited to talk to you! What will the Site Reliability Engineer do? We're looking for a Senior Site Reliability Engineer who's passionate about building and maintaining reliable,... 
    Senior
    Full time
    Immediate start
    Remote work
    Visa sponsorship
    Flexible hours

    Coterie Applications, Inc.

    United States
    3 days ago
  • $151.5k - $252.5k

     ...artifacts rather than getting direct access to environments from day one. This is a ground-up role - you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going... 
    Senior
    Base plus commission
    Local area
    Remote work
    Worldwide

    Veeam Software

    United States
    4 days ago
  • $165k - $225k

     ...demanding AI workloads with enterprise-grade reliability and compliance. Your Role: You will...  ...core. Working closely with our systems engineers, network engineers, and platform...  ...custom Kubernetes networking solutions with SR-IOV for high-performance GPU interconnects... 
    Senior
    Remote work
    Flexible hours

    Moonlite

    Chicago, IL
    19 days ago
  • $110k - $155k

     ...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is... 
    Senior
    Contract work
    Work at office
    Work from home
    Flexible hours

    Vertafore

    Denver, CO
    27 days ago
  • $120k - $170k

    Sr. Manager/Manager Site Reliability Engineering Join to apply for the Sr. Manager/Manager Site Reliability Engineering role at Aritzia Sr. Manager/Manager Site Reliability Engineering 1 day ago Be among the first 25 applicants Join to apply for the Sr. Manager/Manager... 
    Senior
    Full time
    Work at office
    Remote work
    Flexible hours

    Aritzia

    Seattle, WA
    4 days ago
  •  ...: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8 to...  ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will have... 
    Senior
    Remote work

    SRI Tech

    Plano, TX
    4 days ago
  • $210k - $230k

    GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation... 
    Senior
    Currently hiring
    Remote work

    Govcio

    Arlington, VA
    1 day ago
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 
    Senior

    Google

    Sunnyvale, CA
    4 days ago
  •  ...home day is currently Tuesday.Engineering at Lambda is responsible for...  ...networking teams to improve service reliability and deployment...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering,...  ...virtualization technologies, SR-IOV, and DPDKUnderstanding of... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    1 day ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    RELX Group

    Buford, GA
    8 hours ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  • $119.8k - $234.7k

     ...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual...  ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...workloads. As a Senior Site Reliability Engineer, you will lead reliability improvements... 
    Senior
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Redmond, WA
    8 hours ago
  •  ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE... 
    Senior
    Full time
    Remote work

    CyberArk

    United States
    1 day ago
  • $15k

     ...packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage engineering skills... 
    Senior
    Work at office
    Local area
    Remote work

    The Voleon Group

    Berkeley, CA
    2 days ago
  • The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also... 
    Senior
    Full time
    Work experience placement
    Remote work

    Shutterstock

    New York, NY
    8 hours ago
  • $127k - $249k

    The TeamPlatform Engineering sits within SRE and builds the core infrastructure powering MongoDB...  ...plays a pivotal role in engineering the reliable, globally connected, multi-cloud network...  ...are seeking a talented Senior Site Reliability Engineer (SRE) with a strong... 
    Senior
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    4 days ago
  • $158.5k - $172k

     ...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and...  .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology... 
    Senior
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    New York, NY
    4 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Seattle, WA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Sr. Site Reliability Engineer. Be the first to apply!