Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$151.5k - $252.5k
Full-time

Veeam Software

Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in both data resilience and data security posture management, Veeam is built for the convergence of identity, data, security, and AI risk. Headquartered in Seattle with offices in more than 30 countries, Veeam protects over 550,000 customers worldwide, who trust Veeam to keep their businesses running. Join us as we go fearlessly forward together, growing, learning, and making a real impact for some of the world’s biggest brands. About The Role Veeam is building a global SRE function to support the Veeam Data Cloud, our new SaaS platform. This role focuses on our Government and Sovereign Cloud environment. Due to clearance and access requirements, this team operates with restricted access to GOV infrastructure. That means you'll be part of a small team responsible for the full platform stack — including all VDC workloads. You won't always be able to hand off problems to other teams; you need to understand the entire architecture well enough to own it. You'll need to get up to speed on the platform quickly, often by reading code, docs, and architecture artifacts rather than getting direct access to environments from day one. This is a ground-up role — you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going forward. What You'll Do Discovery & Documentation Get up to speed on the full platform — all VDC workloads, dependencies, and risk areas. Much of this will happen through code, docs, and conversations rather than direct environment access. Work with SMEs across the org to fill knowledge gaps and build onboarding material for the team. Write and maintain runbooks, architecture docs, and operational guides. Reliability & Incident Response Design infrastructure for high availability and fault tolerance on Azure (including Azure Government). Define SLIs, SLOs, and error budgets where none exist today. Run incident response and blameless postmortems. Turn incidents into improvements. Identify reliability risks across modern and legacy workloads and build practical remediation plans that work within compliance constraints. Observability Close observability gaps — define instrumentation requirements and drive implementation. Set alerting, telemetry, and monitoring standards with partner teams. Build automation to reduce toil and support fleet management. Participate in on-call rotations. Infrastructure & Delivery Work with IaC, CI/CD, deployment automation, and config management — including in air-gapped or compliance-restricted environments. Build and maintain testing, canary deployment, and release validation pipelines. Integrate chaos engineering and monitoring tools, adapting choices to meet regulatory requirements. Collaboration Work across product, platform, security, legal, compliance, and operations teams. Own problems end-to-end — identify gaps, drive solutions, don't wait for direction. Mentor other engineers and help spread SRE practices across the org. Technologies we work with Microsoft TFS, Azure DevOps, Git, BitBucket Azure (Entra ID, API Management, Cosmos Db, Storage services, Azure Functions, static website hosting, Azure security, etc.) IaC tools (Azure ARM templates, AWS CloudFormation, Terraform, the Serverless Framework, etc.) Observability (Azure Monitor, AppInsights, Elastic Stack) What You'll Bring 7+ years in Software Engineering, with 3+ years in SRE, Platform Engineering, or similar — across multi-service platforms, not just single-service environments. Experience with Government or Sovereign Cloud (e.g., Azure Government, AWS GovCloud). Experience in regulated compliance environments — government (FedRAMP, CMMC, IL2/IL4/IL5), financial (PCI-DSS, SOX), or healthcare (HIPAA, HITRUST). You understand how compliance shapes architecture and operations. Strong experience building and running production services on cloud infrastructure (Azure preferred, including Azure Government). Able to learn large, complex platforms quickly with limited guidance — comfortable building understanding from code, docs, and architecture artifacts when direct environment access is restricted. Can investigate systems independently and produce clear docs, risk assessments, and improvement plans. Comfortable working across teams — engineering, product, security, compliance, operations. Programming skills in one or more of: TypeScript/JS, Go, Java, C#, or similar. Experience with monitoring and observability tools (e.g., Prometheus, Grafana, OpenTelemetry, ELK stack). Experience with IaC (Terraform, Terragrunt, Pulumi) and container orchestration (Kubernetes). Experience with CI/CD and GitOps tooling — GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, FluxCD, or Dagger. Solid grasp of distributed systems, networking, and cloud-native architecture. Clear written and verbal communication skills Bonus Skills Experience on B2B SaaS platforms in regulated or government markets. Background in chaos engineering, resilience testing, or performance/load testing. Have built an SRE or reliability function from scratch before. Experience across mixed environments — modern cloud-native and older legacy systems. Familiar with AI-first development workflows — using LLM-powered tools for infrastructure automation, code generation, and documentation. Why Join? Build the GOV reliability practice from day one — your decisions will shape how this team works. Help define SRE at Veeam across a globally distributed engineering org. Work with strong teams across product, cloud engineering, security, and compliance. Professional development resources including mentorship, training, and volunteer days. Competitive compensation and benefits. #LI-Remote

#LI-RW1

What you'll get Unlimited paid time off, 12 paid holidays including 4 global VeeaMe Days for self-care and 24 paid volunteer hours annually through Veeam Cares Paid parental leave: 8 weeks for all parents, 16 weeks for birthing parents Medical, dental, and vision coverage starting on your first day Mental health support, therapy sessions, and digital wellness tools via our Employee Assistance Program 401(k) retirement plan with company matching contributions Fertility, adoption, and surrogacy support through Maven, plus paid volunteer time AirVet: 24/7 virtual veterinary care at no cost Legal services, identity protection, and supplemental health insurance options Tax-advantaged spending accounts for healthcare, dependent care, and commuting Opportunities to learn and grow through on-demand libraries (LinkedIn Learning, O’Reilly), mentoring, workshops, and learning events like our annual Global Day of Learning Compensation Transparency Veeam is committed to pay transparency and equitable compensation. For this role, the compensation range below reflects the expected total target compensation (TTC), inclusive of base pay and a competitive performance-based bonus. For roles with a commission plan, the compensation range represents On Target Earnings (OTE), which includes base salary plus variable commission. When determining compensation, Veeam takes into consideration factors such as experience, education, skills, and geographic zone. Offers are typically made below the midpoint of the range. In addition to compensation, Veeam provides a comprehensive benefits package, including health coverage, retirement plans, and unlimited time off. U.S. Geographic Zones & Compensation Ranges (TTC / OTE) Zone 1: San Francisco Bay Area, New York City Boroughs

$151,500—$252,500 USD

Zone 2: Washington, California (excluding San Francisco Bay Area)

$138,900—$231,400 USD

Zone 3: Texas, Illinois, North Carolina, Colorado, Massachusetts, Pennsylvania, Virginia, Oregon, Nevada, Hawaii, New York (excluding NYC boroughs); Sales roles located in Georgia, Ohio, and Arizona

$126,300—$210,400 USD

Zone 4: All other US locations

$109,800—$183,000 USD

Veeam Software is an equal opportunity employer and does not tolerate discrimination in any form on the basis of race, color, religion, gender, age, national origin, citizenship, disability, veteran status or any other classification protected by federal, state or local law. All your information will be kept confidential. Personal data collected during the recruitment process will be processed in accordance with our Recruiting Privacy Notice, which explains how your information is collected, used, and handled in connection with hiring activities. By applying for this position, you consent to this processing. By submitting your application, you confirm that the information provided, including any supporting documents, is complete and accurate to the best of your knowledge. Any misrepresentation, omission, or falsification may result in disqualification from consideration or, if discovered after employment begins, termination of employment.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in United States vacancy
  • $189k - $283.6k

     ...the SRE team, you will proactively and reactively improve the reliability of Block's platform and critical infrastructure. You are metrics...  ...of accountability * A strong desire to perform and grow as an engineer * 5+ years of software development experience Technologies... 
    Suggested
    Full time
    Local area
    Remote work
    Relocation package
    Flexible hours
    Shift work

    Block

    California
    4 days ago
  • $130k - $145k

     ...and your desire to team up with some of the best and brightest in technology and entertainment. The Role The Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus... 
    Suggested
    Full time
    Local area
    Worldwide
    Flexible hours

    AXS

    Los Angeles, CA
    4 days ago
  • $120k - $175k

     ...of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible.... 
    Suggested
    Full time
    Remote work
    Work visa
    Flexible hours

    PrizePicks

    United States
    7 hours ago
  • $146.4k

     ...Communications group. A cross-functional engineering team that develops the distributed...  ...network. Our system supports fast and reliable configuration of Akamai's global network...  ...various metadata systems. As a Senior Site Reliability Engineer, you will be responsible... 
    Suggested
    Full time
    Work experience placement
    Work at office

    Akamai

    United States
    1 day ago
  • $140k - $165k

     ...deployed services and infrastructure components, empowering cloud engineering teams to move fast without sacrificing stability is essential...  ...CI/CD pipelines to ensure cloud software changes are deployed reliably and efficiently. Own and manage developer self-service... 
    Suggested
    Full time
    Remote work

    Flock

    United States
    4 days ago
  • $76k - $127k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex... 
    Full time
    Part time
    Worldwide
    Flexible hours
    Early shift

    Mastercard

    O Fallon, MO
    1 day ago
  • $118.6k - $195.68k

    About the Job The Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at... 
    Permanent employment
    Full time
    Contract work
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Red River

    North Carolina
    7 hours ago
  • $166k - $220k

     ...networking technology to the military in months, not years. ABOUT THE TEAM We are seeking a highly skilled and mission-driven Site Reliability Engineer (SRE) to join our Mission Autonomy team. In this critical role, you will be responsible for ensuring the reliability,... 
    Full time
    Work experience placement
    Immediate start
    Remote work

    Anduril Industries

    Costa Mesa, CA
    3 days ago
  • $76k - $127k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive... 
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    14 hours ago
  •  ...services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview-The ProCOM team is looking for a Site Reliability Engineering (SRE) who can help us solve problems, build our... 
    Full time
    Part time
    Immediate start
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    14 hours ago
  • $96k - $163k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that... 
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    14 hours ago
  • $100k - $200k

     ...backed by top tier investors. Our lean, world-class team of engineers and operators is applying a first-principles approach to...  ...culture of urgency, accountability and transparency. DevOps / Site Reliability Engineer We are seeking a highly capable DevOps / Site... 
    Full time
    Weekend work

    General Matter

    Los Angeles, CA
    7 hours ago
  • $235k - $275k

     ...Inc. as one of the most innovative and fastest-growing technology companies in the country. \n Role Summary As a Staff Site Reliability Engineer at Filevine, you are the senior technical authority on the SRE team and a strategic partner to engineering leadership.... 
    Permanent employment
    Full time
    Temporary work

    Filevine

    United States
    3 days ago
  • $217.57k - $260k

     ...job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA;...  ...which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    Shift work

    ID.me

    Mountain View, CA
    7 hours ago
  • $195k - $257.5k

     ...work environment where new ideas are encouraged and everyone is a stakeholder. What you’ll be responsible for: As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and operate the infrastructure that powers our blockchain platform at... 
    Full time
    Remote work
    Flexible hours

    Circle

    United States
    7 hours ago
  • $197.3k - $313.7k

     ...ensure you are not duplicating efforts. Job Category Software Engineering Job Details About Salesforce Salesforce is the #1 AI CRM,...  ...and you are the future of Salesforce. Job Title: Director, Site Reliability Engineering Location: New York, NY; San Francisco, CA;... 
    Full time
    Immediate start

    Salesforce

    San Francisco, CA
    7 hours ago
  • $139.7k - $232.9k

     ...designing, implementing, and continuously improving highly reliable, scalable, and resilient platform solutions across the enterprise. Operates as a subject matter expert (SME) in Site Reliability Engineering, driving reliability engineering practices, operational excellence... 
    Full time
    Work experience placement

    M&T Bank

    Buffalo, NY
    7 hours ago
  • $96k - $163k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer, Performance Engineering Senior Site Reliability Engineer, Performance Engineering Payment Optimization unifies... 
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    14 hours ago
  • Role Description We are looking for a Site Reliability Engineer (SRE) who is passionate about infrastructure reliability, automation, and building scalable production systems. ~Own and improve production infrastructure reliability and stability ~Prepare, execute, and... 
    Full time
    Remote work

    Social Discovery Group

    Remote
    5 days ago
  • $152.5k - $205k

    Role Description As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind critical digital-assets, AI, and application workloads. You will bring an engineering mindset to production... 
    Full time

    Circle

    Remote
    19 hours ago
  • Role Description We’re looking for a Senior Platform Engineer to design, build, and operate the core services that power Optura’s AI...  ...systems end-to-end, from model and agent orchestration to routing, reliability, and observability. You will partner closely with product and... 
    Full time
    Remote work

    Optura

    Remote
    7 days ago
  • $54k - $150k

    Role Description As Senior Site Reliability Engineer for Remote Build, you'll own the operational excellence and infrastructure strategy that makes Build's platform reliable, performant, and safe for customers. You'll report to the Engineering Manager and work closely with... 
    Full time
    Local area
    Immediate start
    Remote work
    Home office
    Flexible hours

    Referral Board

    Remote
    14 hours ago
  • $190k - $240k

    Role Description As a Sr. Site Reliability Engineer (SRE) at ICD, you will play a critical role in ensuring the reliability and seamless operation of our global platform and AWS infrastructure to create scalable and highly reliable software systems. Job Responsibilities... 
    Full time
    Work at office
    Immediate start
    Flexible hours

    Tradeweb

    Remote
    7 days ago
  • $54k - $150k

    Role Description As Senior Site Reliability Engineer for Remote Build, you'll own the operational excellence and infrastructure strategy that makes Build's platform reliable, performant, and safe for customers. You'll report to the Engineering Manager and work closely with... 
    Full time
    Local area
    Remote work
    Home office
    Flexible hours

    Remote

    Remote
    1 day ago
  • Role Description Stack AV Site Reliability Engineers are responsible for enabling and ensuring our production systems meet their service-level objectives. Through the implementation of centralized observability and automation, the SRE team constantly ensures the health... 
    Full time

    Stack AV

    Remote
    3 days ago
  • Role Description The Senior Site Reliability Engineer is a technical leader responsible for architecting the reliability strategy for large-scale, distributed government systems. You will lead the implementation of the SRE framework, driving the adoption of SLO-based management... 
    Contract work
    Remote work

    Arctiq

    Remote
    6 days ago
  • $137.9k - $221.4k

     ...for someone to lead development aspects of the Infrastructure engineering team at ServiceTitan. You must have a strong background in...  ...leadership and strong architectural thought process. Our Site Reliability and Infrastructure Engineering team is an investment by Cloud... 
    Full time
    Immediate start
    Flexible hours

    ServiceTitan

    Remote
    4 days ago
  • Role Description As a Site Reliability Engineer on the Central AI team, you will help Health Catalyst engineer teams adopt AI responsibly and effectively. You bring deep experience solutioning and implementing AI systems, and you use that expertise to evaluate architectures... 
    Full time

    Health Catalyst

    Remote
    5 days ago
  • Role Description We are expanding our Site Reliability Engineering (SRE) team and seeking a highly skilled and passionate Senior SRE to join us. As a member of our growing SRE function, you will play a critical role in ensuring the reliability, scalability, and performance... 
    Full time
    Temporary work

    QAD, Inc.

    Remote
    7 days ago
  • Role Description We’re looking for a Senior Site Reliability Engineer who takes ownership seriously — someone who designs for reliability, ships the automation, and stands behind it in production. You’ll work across cloud-native infrastructure on systems that process millions... 
    Full time

    CertifyOS

    Remote
    7 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!