Sr. Site Reliability Engineer
$151.5k - $252.5kVeeam Software
Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in both data resilience and data security posture management, Veeam is built for the convergence of identity, data, security, and AI risk. Headquartered in Seattle with offices in more than 30 countries, Veeam protects over 550,000 customers worldwide, who trust Veeam to keep their businesses running. Join us as we go fearlessly forward together, growing, learning, and making a real impact for some of the world's biggest brands.
About The Role Veeam is building a global SRE function to support the Veeam Data Cloud, our new SaaS platform. This role focuses on our Government and Sovereign Cloud environment. Due to clearance and access requirements, this team operates with restricted access to GOV infrastructure. That means you'll be part of a small team responsible for the full platform stack - including all VDC workloads. You won't always be able to hand off problems to other teams; you need to understand the entire architecture well enough to own it. You'll need to get up to speed on the platform quickly, often by reading code, docs, and architecture artifacts rather than getting direct access to environments from day one. This is a ground-up role - you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going forward.What You'll Do
Discovery & Documentation
- Get up to speed on the full platform - all VDC workloads, dependencies, and risk areas. Much of this will happen through code, docs, and conversations rather than direct environment access.
- Work with SMEs across the org to fill knowledge gaps and build onboarding material for the team.
- Write and maintain runbooks, architecture docs, and operational guides.
- Design infrastructure for high availability and fault tolerance on Azure (including Azure Government).
- Define SLIs, SLOs, and error budgets where none exist today.
- Run incident response and blameless postmortems. Turn incidents into improvements.
- Identify reliability risks across modern and legacy workloads and build practical remediation plans that work within compliance constraints.
- Close observability gaps - define instrumentation requirements and drive implementation.
- Set alerting, telemetry, and monitoring standards with partner teams.
- Build automation to reduce toil and support fleet management.
- Participate in on-call rotations.
- Work with IaC, CI/CD, deployment automation, and config management - including in air-gapped or compliance-restricted environments.
- Build and maintain testing, canary deployment, and release validation pipelines.
- Integrate chaos engineering and monitoring tools, adapting choices to meet regulatory requirements.
- Work across product, platform, security, legal, compliance, and operations teams.
- Own problems end-to-end - identify gaps, drive solutions, don't wait for direction.
- Mentor other engineers and help spread SRE practices across the org.
- Microsoft TFS, Azure DevOps, Git, BitBucket
- Azure (Entra ID, API Management, Cosmos Db, Storage services, Azure Functions, static website hosting, Azure security, etc.)
- IaC tools (Azure ARM templates, AWS CloudFormation, Terraform, the Serverless Framework, etc.)
- Observability (Azure Monitor, AppInsights, Elastic Stack)
- 7+ years in Software Engineering, with 3+ years in SRE, Platform Engineering, or similar - across multi-service platforms, not just single-service environments.
- Experience with Government or Sovereign Cloud (e.g., Azure Government, AWS GovCloud).
- Experience in regulated compliance environments - government (FedRAMP, CMMC, IL2/IL4/IL5), financial (PCI-DSS, SOX), or healthcare (HIPAA, HITRUST). You understand how compliance shapes architecture and operations.
- Strong experience building and running production services on cloud infrastructure (Azure preferred, including Azure Government).
- Able to learn large, complex platforms quickly with limited guidance - comfortable building understanding from code, docs, and architecture artifacts when direct environment access is restricted.
- Can investigate systems independently and produce clear docs, risk assessments, and improvement plans.
- Comfortable working across teams - engineering, product, security, compliance, operations.
- Programming skills in one or more of: TypeScript/JS, Go, Java, C#, or similar.
- Experience with monitoring and observability tools (e.g., Prometheus, Grafana, OpenTelemetry, ELK stack).
- Experience with IaC (Terraform, Terragrunt, Pulumi) and container orchestration (Kubernetes).
- Experience with CI/CD and GitOps tooling - GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, FluxCD, or Dagger.
- Solid grasp of distributed systems, networking, and cloud-native architecture.
- Clear written and verbal communication skills
- Experience on B2B SaaS platforms in regulated or government markets.
- Background in chaos engineering, resilience testing, or performance/load testing.
- Have built an SRE or reliability function from scratch before.
- Experience across mixed environments - modern cloud-native and older legacy systems.
- Familiar with AI-first development workflows - using LLM-powered tools for infrastructure automation, code generation, and documentation.
- Build the GOV reliability practice from day one - your decisions will shape how this team works.
- Help define SRE at Veeam across a globally distributed engineering org.
- Work with strong teams across product, cloud engineering, security, and compliance.
- Professional development resources including mentorship, training, and volunteer days.
- Competitive compensation and benefits.
#LI-RW1 What you'll get
- Unlimited paid time off, 12 paid holidays including 4 global VeeaMe Days for self-care and 24 paid volunteer hours annually through Veeam Cares
- Paid parental leave: 8 weeks for all parents, 16 weeks for birthing parents
- Medical, dental, and vision coverage starting on your first day
- Mental health support, therapy sessions, and digital wellness tools via our Employee Assistance Program
- 401(k) retirement plan with company matching contributions
- Fertility, adoption, and surrogacy support through Maven, plus paid volunteer time
- AirVet: 24/7 virtual veterinary care at no cost
- Legal services, identity protection, and supplemental health insurance options
- Tax-advantaged spending accounts for healthcare, dependent care, and commuting
- Opportunities to learn and grow through on-demand libraries (LinkedIn Learning, O'Reilly), mentoring, workshops, and learning events like our annual Global Day of Learning
Compensation Transparency Veeam is committed to pay transparency and equitable compensation. For this role, the compensation range below reflects the expected total target compensation (TTC), inclusive of base pay and a competitive performance-based bonus. For roles with a commission plan, the compensation range represents On Target Earnings (OTE), which includes base salary plus variable commission. When determining compensation, Veeam takes into consideration factors such as experience, education, skills, and geographic zone. Offers are typically made below the midpoint of the range. In addition to compensation, Veeam provides a comprehensive benefits package, including health coverage, retirement plans, and unlimited time off. U.S. Geographic Zones & Compensation Ranges (TTC / OTE) Zone 1: San Francisco Bay Area, New York City Boroughs $151,500-$252,500 USD Zone 2: Washington, California (excluding San Francisco Bay Area) $138,900-$231,400 USD Zone 3: Texas, Illinois, North Carolina, Colorado, Massachusetts, Pennsylvania, Virginia, Oregon, Nevada, Hawaii, New York (excluding NYC boroughs); Sales roles located in Georgia, Ohio, and Arizona $126,300-$210,400 USD Zone 4: All other US locations $109,800-$183,000 USD Veeam Software is an equal opportunity employer and does not tolerate discrimination in any form on the basis of race, color, religion, gender, age, national origin, citizenship, disability, veteran status or any other classification protected by federal, state or local law. All your information will be kept confidential. Personal data collected during the recruitment process will be processed in accordance with our Recruiting Privacy Notice, which explains how your information is collected, used, and handled in connection with hiring activities. By applying for this position, you consent to this processing.
By submitting your application, you confirm that the information provided, including any supporting documents, is complete and accurate to the best of your knowledge. Any misrepresentation, omission, or falsification may result in disqualification from consideration or, if discovered after employment begins, termination of employment.
- ...thousands of companies. Join us as we help people all over the world thrive at work.Location: Salt Lake City, UTAs a Senior Site Reliability Engineer, you will help define the future of reliability for our world-class employee recognition platform. You'll leverage...SeniorFull timeShift work
- Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has been at the forefront of online real estate... ...confidence through expert guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization,...SeniorWork at officeLocal area
$138.4k - $173k
...infrastructure as well as help improve the reliability, quality of services and overall... ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability... ...about our locations by visiting our site.Compensation & BenefitsThe base salary that...SeniorFull timeFlexible hours- ...selected candidate for this role to work on site in the specified location(s).Schwab... ...their money by delivering innovative and reliable technology solutions that support investing... ...Within the Bank Platform Operations and Engineering organization, you will help ensure the...SeniorFull timeWork at office
$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...SeniorFull timeLocal areaImmediate start$130k - $140k
...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient...SeniorOngoing contractFull timeTemporary workWork experience placement$165k - $265k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy...SeniorPermanent employmentTemporary workWorldwideWeekend work$134.25k - $214.8k
...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance...SeniorWork at officeRemote workFlexible hours$165k - $280k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s...SeniorPermanent employmentTemporary workWorldwideWeekend work$139.76k - $287.75k
...their business.We are seeking a Senior Site ReliabilityEngineer to help operate, scale... ...will be instrumental in advancing the reliability, scalability, automation, observability,... ...The ideal candidate is a highly hands-on engineer with strong production experience and a...SeniorWork at officeLocal areaRelocationRelocation package$165k - $230k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts...SeniorPermanent employmentTemporary workImmediate startWeekend work$87.12k - $151.25k
...to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a Digital Site Reliability Sr Engineer - Remote to join our team in Memphis, Tennessee (US-TN), United States (US).Digital Site Reliability Senior EngineerWe are...SeniorTemporary workWork at officeRemote workFlexible hours$84.24k - $142.48k
OverviewJoin us to work collaboratively with our talented team of dynamic and passionate engineers to deliver capabilities that enable our customers to make a difference. You'll deploy and operate ArcGIS Velocity and ArcGIS Workflow Manager SaaS solutions. You will also...SeniorWorldwideFlexible hours- ...consumers and companies, alikeKlover’s engineering team powers one of the fastest-growing fintech... ...-grade systems that prioritize reliability, security, and performance, and that integrate... ...the RoleAs a Senior/Staff Site Reliability Engineer, you will play a critical...SeniorWork at officeImmediate startRemote work
$160k - $185k
...in their fitness journey and revolutionized the industry along the way. And we’re just getting started!OverviewThe Sr. Manager, Site Reliability Engineering (SRE) leads the strategy, execution, and continuous improvement of reliability, availability, and performance...SeniorWork at officeLocal areaRemote workWork from home- ...Site Reliability Engineer As a Site Reliability Engineer, you will play a critical role in ensuring the reliability, availability, and performance of our systems. You will be responsible for designing, implementing, and maintaining scalable infrastructure solutions...SeniorRemote workFlexible hours
- ...Sr Site Reliability Engineer (SRE) SigNoz is an open-source observability platform that helps modern engineering teams monitor, debug, and optimize their applications with deep visibility into metrics, traces, and logs — all in one place. We're built natively on OpenTelemetry...SeniorRemote work
$117.63k - $176.44k
...you will be responsible for ensuring the reliability, scalability, and performance of our data systems. Working closely with data engineers and other operation sub-teams, you will manage... ...and benefits summary on our careers site for more details.EducationBachelor's DegreeWhile...SeniorFull time$125k - $145k
...General information Press space or enter keys to toggle section visibility Job Title Sr. Site Reliability Engineer Functional Area Teammate - Information Technology City Remote Work Location Type...SeniorFull timeWork experience placementLocal areaRemote workFlexible hoursShift work- ...to grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise. The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers...SeniorWork experience placementRemote workFlexible hours
- Role Description We're looking for an SRE to own the reliability, scalability, and operability of the SigNoz cloud platform. You'll keep... ...Benefits ~Work on a globally used open-source project that engineers actually love. ~Huge scope and ownership — your work directly...SeniorFull timeRemote work
$140k - $215k
...operate at the intersection of our Core Platform and Embedded Reliability charters: building the foundational libraries, services, and... ...product group depends on, while embedding directly with product engineering teams and their leadership to drive reliability outcomes at...SeniorFull timeWork experience placementWork at officeLocal area2 days per week3 days per week$140k - $170k
...and intelligence. If you want to push yourself and reshape a $200B+ market, we're excited to talk to you! What will the Site Reliability Engineer do? We're looking for a Senior Site Reliability Engineer who's passionate about building and maintaining reliable,...SeniorFull timeImmediate startRemote workVisa sponsorshipFlexible hours$165k - $230k
...actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. HARDWARE / INFRASTRUCTURE SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink,...SeniorPermanent employmentTemporary workWork at officeWorldwideMonday to FridayWeekend work- ...security to responsibly propel the global lottery industry ever forward. Position Summary We are looking for a skilled Site Reliability Engineer (SRE) to enhance the stability, performance, and reliability of our production systems. The SRE will work closely with...SeniorPermanent employmentWork experience placementLocal area
$150k
...Job Title: Senior Site Reliability Engineer (SRE) Location: Cincinnati, OH (Hybrid – 3 days onsite/week) Type: Direct Hire Compensation: $150,000 Range (based on experience) We’re seeking an experienced Senior Site Reliability Engineer to join a small...SeniorFull timeLocal area3 days per week$130k - $160k
...Description NBCU is looking for creative engineers willing to learn from the current process... .../existing systems; propose & deploy more reliable scalable solutions Responsible for... ...monitoring deliverables to improve site reliability Evaluate new software releases...SeniorFull timeWork at officeLocal areaRotating shift$120k - $200k
Sr Site Reliability Engineer (Prisma Access) 2 days ago Be among the first 25 applicants Job Description This role requires US Citizenship. Your Career Palo Alto Networks runs a large infrastructure and is one of the biggest GCP customers. As a Principal SRE, you'll...SeniorRotating shift$120k - $170k
Sr. Manager/Manager Site Reliability Engineering Join to apply for the Sr. Manager/Manager Site Reliability Engineering role at Aritzia Sr. Manager/Manager Site Reliability Engineering 1 day ago Be among the first 25 applicants Join to apply for the Sr. Manager/Manager...SeniorFull timeWork at officeRemote workFlexible hours- Sr. Director, Site Reliability and Platform Engineering Optomi, in partnership with our premier client in the Technology industry, is seeking a Senior Director to join our client’s SRE and DevOps team in Tacoma, WA, reporting to the Vice President, Engineering. In this...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr. Site Reliability Engineer. Be the first to apply!
- site reliability engineer United States
- lead site reliability engineer United States
- site reliability engineering manager United States
- site reliability engineer remote United States
- site reliability engineer sre United States
- senior maintenance supervisor United States
- senior lead project manager United States
- senior robotics software engineer United States
- senior firewall engineer United States
- senior devops engineer remote United States



