Sr. Site Reliability Engineer
Veeamsoftware
Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in both data resilience and data security posture management, Veeam is built for the convergence of identity, data, security, and AI risk. Headquartered in Seattle with offices in more than 30 countries, Veeam protects over 550,000 customers worldwide, who trust Veeam to keep their businesses running. Join us as we go fearlessly forward together, growing, learning, and making a real impact for some of the world’s biggest brands. About The Role Veeam is building a global SRE function to support the Veeam Data Cloud, our new SaaS platform. This role focuses on our Government and Sovereign Cloud environment. Due to clearance and access requirements, this team operates with restricted access to GOV infrastructure. That means you'll be part of a small team responsible for the full platform stack — including all VDC workloads. You won't always be able to hand off problems to other teams; you need to understand the entire architecture well enough to own it. You'll need to get up to speed on the platform quickly, often by reading code, docs, and architecture artifacts rather than getting direct access to environments from day one. This is a ground-up role — you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going forward. What You'll Do Discovery & Documentation - Get up to speed on the full platform — all VDC workloads, dependencies, and risk areas. Much of this will happen through code, docs, and conversations rather than direct environment access. - Work with SMEs across the org to fill knowledge gaps and build onboarding material for the team. - Write and maintain runbooks, architecture docs, and operational guides. Reliability & Incident Response - Design infrastructure for high availability and fault tolerance on Azure (including Azure Government). - Define SLIs, SLOs, and error budgets where none exist today. - Run incident response and blameless postmortems.
Turn incidents into improvements. - Identify reliability risks across modern and legacy workloads and build practical remediation plans that work within compliance constraints. Observability - Close observability gaps — define instrumentation requirements and drive implementation. - Set alerting, telemetry, and monitoring standards with partner teams. - Build automation to reduce toil and support fleet management. - Participate in on-call rotations. Infrastructure & Delivery - Work with IaC , CI/CD, deployment automation, and config management — including in air-gapped or compliance-restricted environments. - Build and maintain testing, canary deployment, and release validation pipelines. - Integrate chaos engineering and monitoring tools, adapting choices to meet regulatory requirements. Collaboration - Work across product, platform, security, legal, compliance, and operations teams. - Own problems end-to-end — identify gaps, drive solutions, don't wait for direction. - Mentor other engineers and help spread SRE practices across the org. Technologies we work with - Microsoft TFS, Azure DevOps, Git, BitBucket - Azure (Entra ID, API Management, Cosmos Db, Storage services, Azure Functions, static website hosting, Azure security, etc. ) - IaC tools (Azure ARM templates, AWS CloudFormation, Terraform, the Serverless Framework, etc. ) - Observability (Azure Monitor, AppInsights, Elastic Stack) What You'll Bring - 7+ years in Software Engineering, with 3+ years in SRE, Platform Engineering, or similar — across multi-service platforms, not just single-service environments. - Experience with Government or Sovereign Cloud (e. g. , Azure Government, AWS GovCloud).
$87.12k - $151.25k
...to be part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a Digital Site Reliability Sr Engineer - Remote to join our team in Memphis, Tennessee (US-TN), United States (US).Digital Site Reliability Senior EngineerWe are...SeniorTemporary workWork at officeRemote workFlexible hours$134.25k - $214.8k
...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance...SeniorWork experience placementWork at officeRemote workFlexible hours- ...About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure...SeniorRemote work
- ...consumers and companies, alikeKlover’s engineering team powers one of the fastest-growing fintech... ...-grade systems that prioritize reliability, security, and performance, and that integrate... ...the RoleAs a Senior/Staff Site Reliability Engineer, you will play a critical...SeniorWork at officeImmediate startRemote work
$160k - $185k
...in their fitness journey and revolutionized the industry along the way. And we’re just getting started!OverviewThe Sr. Manager, Site Reliability Engineering (SRE) leads the strategy, execution, and continuous improvement of reliability, availability, and performance...SeniorWork at officeLocal areaRemote workWork from home$107.8k - $162k
...an expectation of a minimum of three days per week working in the office and flexibility to work remotely on the remaining days. On-site expectations may evolve over time to support business needs, with clear communication provided in advance. Job Description Operates...SeniorWork at officeLocal areaRemote work3 days per week- ...Sr Site Reliability Engineer (SRE) SigNoz is an open-source observability platform that helps modern engineering teams monitor, debug, and optimize their applications with deep visibility into metrics, traces, and logs — all in one place. We're built natively on OpenTelemetry...SeniorRemote work
- ...infrastructure that companies around the world depend on to keep their engineering teams focused on what matters most - their own product.... ...that everyone can appreciate. About the Role: As a Site Reliability Engineer, you will play a critical role in ensuring the...SeniorRemote workFlexible hours
- ...Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has been at the forefront of online real estate... ...confidence through expert guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization,...SeniorWork at officeLocal area
- ...straightforward communication and clinical domain expertise, Commence cuts straight to better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability, scalability, and operational health of our mission-critical healthcare data...SeniorRemote work
- ...Financial Software company Position: SRE Engineer Location: Hybrid remote/Buffalo Grove/Chicago Comp: Solid... ...; coach teams on using them to balance feature velocity with reliability and communicate system health to stakeholders. • Lead alert...SeniorRemote work
- ...Job Description The Opportunity: Versant's Sports & Entertainment Digital Products division is seeking a Senior Site Reliability Engineer to help drive the reliability, scalability, and usability of internal developer platforms, tooling, and engineering workflows...SeniorLocal areaRemote workWorldwide
$140k - $170k
...and intelligence. If you want to push yourself and reshape a $200B+ market, we're excited to talk to you! What will the Site Reliability Engineer do? We're looking for a Senior Site Reliability Engineer who's passionate about building and maintaining reliable,...SeniorFull timeImmediate startRemote workVisa sponsorshipFlexible hours$151.5k - $252.5k
...artifacts rather than getting direct access to environments from day one. This is a ground-up role - you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going...SeniorBase plus commissionLocal areaRemote workWorldwide$165k - $225k
...demanding AI workloads with enterprise-grade reliability and compliance. Your Role: You will... ...core. Working closely with our systems engineers, network engineers, and platform... ...custom Kubernetes networking solutions with SR-IOV for high-performance GPU interconnects...SeniorRemote workFlexible hours$110k - $155k
...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is...SeniorContract workWork at officeWork from homeFlexible hours$120k - $170k
Sr. Manager/Manager Site Reliability Engineering Join to apply for the Sr. Manager/Manager Site Reliability Engineering role at Aritzia Sr. Manager/Manager Site Reliability Engineering 1 day ago Be among the first 25 applicants Join to apply for the Sr. Manager/Manager...SeniorFull timeWork at officeRemote workFlexible hours- ...: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8 to... ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will have...SeniorRemote work
$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...SeniorCurrently hiringRemote work$174k - $252k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...Senior- ...home day is currently Tuesday.Engineering at Lambda is responsible for... ...networking teams to improve service reliability and deployment... ...rotationYouHave 5+ years of experience in Site Reliability Engineering,... ...virtualization technologies, SR-IOV, and DPDKUnderstanding of...SeniorWork at officeLocal areaWork from homeFlexible hours
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SeniorFull timeWork at officeLocal areaRemote workWork from home- ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering... ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or...SeniorWork at officeLocal areaWork from homeFlexible hours
$119.8k - $234.7k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...workloads. As a Senior Site Reliability Engineer, you will lead reliability improvements...SeniorOngoing contractLocal area3 days per week- ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE...SeniorFull timeRemote work
$15k
...packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage engineering skills...SeniorWork at officeLocal areaRemote work- The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also...SeniorFull timeWork experience placementRemote work
$127k - $249k
The TeamPlatform Engineering sits within SRE and builds the core infrastructure powering MongoDB... ...plays a pivotal role in engineering the reliable, globally connected, multi-cloud network... ...are seeking a talented Senior Site Reliability Engineer (SRE) with a strong...SeniorLocal areaRemote workWorldwideFlexible hours$158.5k - $172k
...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and... .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology...SeniorFull timeTemporary workWork at officeFlexible hours3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr. Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Remote
- site reliability engineer Remote
- site reliability engineer remote Remote
- senior operations technician Remote
- senior operations associate Remote
- senior cloud service delivery manager Remote
- senior it service manager Remote
- senior chief engineer Remote
- sr operations manager Remote
- senior account director Remote


