Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$109.8k - $183k

Veeam Software

Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in both data resilience and data security posture management, Veeam is built for the convergence of identity, data, security, and AI risk. Headquartered in Seattle with offices in more than 30 countries, Veeam protects over 550,000 customers worldwide, who trust Veeam to keep their businesses running. Join us as we go fearlessly forward together, growing, learning, and making a real impact for some of the world’s biggest brands.

About The Role

Veeam is building a global SRE function to support the Veeam Data Cloud, our new SaaS platform. This role focuses on our Government and Sovereign Cloud environment.

Due to clearance and access requirements, this team operates with restricted access to GOV infrastructure. That means you'll be part of a small team responsible for the full platform stack — including all VDC workloads. You won't always be able to hand off problems to other teams; you need to understand the entire architecture well enough to own it. You'll need to get up to speed on the platform quickly, often by reading code, docs, and architecture artifacts rather than getting direct access to environments from day one.

This is a ground-up role — you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going forward.

What You'll Do
Discovery & Documentation
  • Get up to speed on the full platform — all VDC workloads, dependencies, and risk areas. Much of this will happen through code, docs, and conversations rather than direct environment access.

  • Work with SMEs across the org to fill knowledge gaps and build onboarding material for the team.

  • Write and maintain runbooks, architecture docs, and operational guides.

Reliability & Incident Response
  • Design infrastructure for high availability and fault tolerance on Azure (including Azure Government).

  • Define SLIs, SLOs, and error budgets where none exist today.

  • Run incident response and blameless postmortems. Turn incidents into improvements.

  • Identify reliability risks across modern and legacy workloads and build practical remediation plans that work within compliance constraints.

Observability
  • Close observability gaps — define instrumentation requirements and drive implementation.

  • Set alerting, telemetry, and monitoring standards with partner teams.

  • Build automation to reduce toil and support fleet management.

  • Participate in on-call rotations.

Infrastructure & Delivery
  • Work with IaC, CI/CD, deployment automation, and config management — including in air-gapped or compliance-restricted environments.

  • Build and maintain testing, canary deployment, and release validation pipelines.

  • Integrate chaos engineering and monitoring tools, adapting choices to meet regulatory requirements.

Collaboration
  • Work across product, platform, security, legal, compliance, and operations teams.

  • Own problems end-to-end — identify gaps, drive solutions, don't wait for direction.

  • Mentor other engineers and help spread SRE practices across the org.

Technologies we work with
  • Microsoft TFS, Azure DevOps, Git, BitBucket
  • Azure (Entra ID, API Management, Cosmos Db, Storage services, Azure Functions, static website hosting, Azure security, etc.) 
  • IaC tools (Azure ARM templates, AWS CloudFormation, Terraform, the Serverless Framework, etc.) 
  • Observability (Azure Monitor, AppInsights, Elastic Stack) 
What You'll Bring
  • 7+ years in Software Engineering, with 3+ years in SRE, Platform Engineering, or similar — across multi-service platforms, not just single-service environments.

  • Experience with Government or Sovereign Cloud (e.g., Azure Government, AWS GovCloud).

  • Experience in regulated compliance environments — government (FedRAMP, CMMC, IL2/IL4/IL5), financial (PCI-DSS, SOX), or healthcare (HIPAA, HITRUST). You understand how compliance shapes architecture and operations.

  • Strong experience building and running production services on cloud infrastructure (Azure preferred, including Azure Government).

  • Able to learn large, complex platforms quickly with limited guidance — comfortable building understanding from code, docs, and architecture artifacts when direct environment access is restricted.

  • Can investigate systems independently and produce clear docs, risk assessments, and improvement plans.

  • Comfortable working across teams — engineering, product, security, compliance, operations.

  • Programming skills in one or more of: TypeScript/JS, Go, Java, C#, or similar.

  • Experience with monitoring and observability tools (e.g., Prometheus, Grafana, OpenTelemetry, ELK stack).

  • Experience with IaC (Terraform, Terragrunt, Pulumi) and container orchestration (Kubernetes).

  • Experience with CI/CD and GitOps tooling — GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, FluxCD, or Dagger.

  • Solid grasp of distributed systems, networking, and cloud-native architecture.

  • Clear written and verbal communication skills 

Bonus Skills
  • Experience on B2B SaaS platforms in regulated or government markets.

  • Background in chaos engineering, resilience testing, or performance/load testing.

  • Have built an SRE or reliability function from scratch before.

  • Experience across mixed environments — modern cloud-native and older legacy systems.

  • Familiar with AI-first development workflows — using LLM-powered tools for infrastructure automation, code generation, and documentation.

Why Join?
  • Build the GOV reliability practice from day one — your decisions will shape how this team works.

  • Help define SRE at Veeam across a globally distributed engineering org.

  • Work with strong teams across product, cloud engineering, security, and compliance.

  • Professional development resources including mentorship, training, and volunteer days.

  • Competitive compensation and benefits.

#LI-Remote 

#LI-RW1

What you'll get

  • Unlimited paid time off, 12 paid holidays including 4 global VeeaMe Days for self-care and 24 paid volunteer hours annually through Veeam Cares
  • Paid parental leave: 8 weeks for all parents, 16 weeks for birthing parents
  • Medical, dental, and vision coverage starting on your first day
  • Mental health support, therapy sessions, and digital wellness tools via our Employee Assistance Program
  • 401(k) retirement plan with company matching contributions
  • Fertility, adoption, and surrogacy support through Maven, plus paid volunteer time
  • AirVet: 24/7 virtual veterinary care at no cost
  • Legal services, identity protection, and supplemental health insurance options
  • Tax-advantaged spending accounts for healthcare, dependent care, and commuting
  • Opportunities to learn and grow through on-demand libraries (LinkedIn Learning, O’Reilly), mentoring, workshops, and learning events like our annual Global Day of Learning

Compensation Transparency

Veeam is committed to pay transparency and equitable compensation. For this role, the compensation range below reflects the expected total target compensation (TTC), inclusive of base pay and a competitive performance-based bonus. For roles with a commission plan, the compensation range represents On Target Earnings (OTE), which includes base salary plus variable commission. When determining compensation, Veeam takes into consideration factors such as experience, education, skills, and geographic zone. Offers are typically made below the midpoint of the range.

In addition to compensation, Veeam provides a comprehensive benefits package, including health coverage, retirement plans, and unlimited time off.

U.S. Geographic Zones & Compensation Ranges (TTC / OTE) Zone 1: San Francisco Bay Area, New York City Boroughs$151,500—$252,500 USDZone 2: Washington, California (excluding San Francisco Bay Area)$138,900—$231,400 USDZone 3: Texas, Illinois, North Carolina, Colorado, Massachusetts, Pennsylvania, Virginia, Oregon, Nevada, Hawaii, New York (excluding NYC boroughs); Sales roles located in Georgia, Ohio, and Arizona$126,300—$210,400 USDZone 4: All other US locations$109,800—$183,000 USD

Veeam Software is an equal opportunity employer and does not tolerate discrimination in any form on the basis of race, color, religion, gender, age, national origin, citizenship, disability, veteran status or any other classification protected by federal, state or local law. All your information will be kept confidential.

Personal data collected during the recruitment process will be processed in accordance with our Recruiting Privacy Notice, which explains how your information is collected, used, and handled in connection with hiring activities. By applying for this position, you consent to this processing. 

By submitting your application, you confirm that the information provided, including any supporting documents, is complete and accurate to the best of your knowledge. Any misrepresentation, omission, or falsification may result in disqualification from consideration or, if discovered after employment begins, termination of employment.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in United States vacancy
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change...  ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"... 
    Suggested
    Night shift

    Forward Networks

    Santa Clara, CA
    5 days ago
  • We are looking for an experienced Site Reliability Engineer (SRE) to strengthen observability and operational resilience across a Microsoft Azure environment. This long-term Contract role will work closely with DevOps and engineering teams to establish monitoring standards... 
    Suggested
    Long term contract

    Robert Half

    Maumee, OH
    1 day ago
  • $158.5k - $172k

     ...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and...  .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology... 
    Suggested
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    New York, NY
    5 days ago
  • $148.5k - $223.9k

     ...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations,... 
    Suggested
    Full time
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    4 days ago
  • $141k - $216.6k

     ...—it means helping shape the future of emergency response and building a safer, more connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability, and operational excellence of our Unified Call (UC) platform—the mission-critical... 
    Suggested
    Work experience placement
    Work at office

    Axon

    New York, NY
    3 days ago
  • $182.8k - $247.3k

     ...mission to develop education for our half a billion (and growing!) learners around the world.About the role...As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed systems... 
    Work experience placement

    Duolingo

    Pittsburgh, PA
    2 days ago
  •  ...selected candidate for this role to work on site in the specified location(s).The Client...  ...team is responsible for ensuring the reliability, scalability, and operational excellence...  ...around the clock. As a Site Reliability Engineer, you will partner across application engineering... 
    Full time
    Work at office

    The Charles Schwab Corporation

    Austin, TX
    17 hours ago
  • Senior Site Reliability EngineerLocation: Exton or Philadelphia, PA (Hybrid - 3 times a week in-office)Position SummaryAre you ready to start...  ...looking for you!We are looking for a Senior Site Reliability Engineer to take on the responsibility of automating cloud-based... 
    Casual work
    Work at office
    Worldwide

    Bentley Systems

    Philadelphia, PA
    3 days ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and broad business problems with simple and straightforward solutions... 

    JP Morgan Chase

    Plano, TX
    4 days ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Corporate and Investment Banking team, you will solve complex and broad business problems... 

    JP Morgan Chase

    Orem, UT
    5 days ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    RELX Group

    Atlanta, GA
    1 day ago
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure... 

    Alembic

    San Francisco, CA
    5 days ago
  • The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams... 
    Full time
    Work at office
    Local area

    Castleton Commodities International

    Stamford, CT
    5 days ago
  • $143k - $191k

     ...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental...  ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and... 
    Full time
    Temporary work
    Work experience placement
    Immediate start

    Anduril Industries

    Seattle, WA
    5 days ago
  • Job ID: 28091697Reference Number: 26-00413Title: Site Reliability EngineerLocation: Iselin, NJ, 08830Posted Date: 2026-04-27Contact: Deepak...  ...Phone: (***) ***-****Company: HAN Staffing As a Site Reliability Engineer at JPMorgan Chase within the Commercial & Investment Banking,... 

    HAN Staffing

    Iselin, NJ
    1 day ago
  •  ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises... 
    Permanent employment
    Full time
    Part time
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Monday to Friday
    Shift work
    Day shift

    Truist

    Raleigh, NC
    5 days ago
  • Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through...  ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform...  ...vendor resources Willingness to work on-site at stated location in the job openingDepartment... 
    Contract work
    For contractors
    Work experience placement

    Cedent Consulting

    Dallas, TX
    5 days ago
  • $167.7k - $245.2k

     ...very effective.We’re looking for talented engineers with a software or operations background...  ...development teams to ensure the reliability, performance and security of our infrastructure...  ...insurance. Please see the Cisco careers site to discover more benefits and perks.... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    New York, NY
    4 days ago
  • $160k - $180k

     ...expertise, and world-class customer satisfaction. The Platform Engineering group at CentralReach builds the underlying technologies that...  ...in Software Engineering to drive adoption of modern reliability practices like SLOs, error budget policies, actionable alerts... 
    Full time
    Worldwide

    CentralReach

    Holmdel, NJ
    17 hours ago
  •  ...and responsible for ensuring the availability, scalability, and reliability of systems and applications.What will be your responsibilities...  ...using tools like Terraform or CloudFormation.Mentor junior engineers and provide technical guidance.Stay up-to-date with industry trends... 
    Work at office
    Remote work

    Interactive Brokers

    Greenwich, CT
    4 days ago
  • $45 - $85 per hour

    DescriptionThe Resy Site Reliability Engineering groups goal is to ensure Resy Customers can always use the service reliably. We're looking for engineers to be part of an empowered, self-organizing group, with the opportunity to use modern languages and tools and to operate... 
    Contract work
    Temporary work

    TEKsystems

    Phoenix, AZ
    5 days ago
  •  ...importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).As a Site Reliability Engineer supporting the Cashiering organization, you will play a critical role in ensuring the stability,... 
    Full time
    Work at office

    The Charles Schwab Corporation

    Southlake, TX
    17 hours ago
  •  ...and foster a dynamic work environment where new ideas thrive. Are you ready to join our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the reliability, scalability, and performance of our Azure-based SaaS... 
    Temporary work
    Casual work
    Worldwide

    TeamViewer

    Austin, TX
    1 day ago
  • $134.25k - $214.8k

     ...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed...  ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,... 
    Work experience placement
    Work at office
    Remote work

    Axon

    Seattle, WA
    1 day ago
  • $150k - $180k

     ...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business...  ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or... 
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work
    Worldwide

    Umbra

    Arlington, VA
    5 days ago
  • $96.8k - $145.2k

     ...If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a Site Reliability Engineer (Onsite Hybrid) to join our team in Plano, Texas (US-TX), United States (US).Job Responsibilities Include: Own and manage... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Plano, TX
    2 days ago
  • Reliability Engineering Design, implement, and operate scalable, resilient, and highly available systems on Google Cloud Platform. Improve service...  ...Skills, and Abilities Three or more years of experience in Site Reliability Engineering, platform engineering, DevOps, cloud... 
    Remote work

    Patterson-UTI

    Houston, TX
    4 days ago
  • $87.1k - $157.45k

     ...throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across...  ...to come in and help us build systems that stay reliable when things get complicated.We need a Site Reliability Engineer who has experience building, deploying... 
    Full time
    Work from home
    Flexible hours

    Leidos

    Chantilly, Loudoun County, VA
    1 day ago
  • $80k - $140k

    Job DescriptionRBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and... 
    Full time
    Flexible hours
    Shift work

    Royal Bank of Canada

    Minneapolis, MN
    1 day ago
  • $165k - $270k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most... 
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Redmond, WA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!