Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer (SRE)

Dental Intelligence

About The Role:

We're looking for a Senior Site Reliability Engineer to help us mature and scale the infrastructure behind our multi-cloud SaaS platform. Most of our footprint runs on Microsoft Azure, built from the ground up around cloud architecture principles: autoscaling App Service and Container Apps workloads, VM Scale Sets, and Kafka-based event streaming form the backbone. We also have a smaller, well-run presence on AWS, and a Windows-based on-premises component that sits at the core of our product and integrates with our cloud environment. This role is primarily focused on Azure, where the biggest opportunity for impact lives, and you'll work across the full multi-cloud picture. We'd like to see the on-prem and cloud sides operate as a more unified, well-integrated system than they do today, and that integration work is part of what makes this role interesting.

This is a high-impact role for someone who thinks in terms of systems rather than tickets, and who treats infrastructure like software. You'll have significant ownership over how our cloud infrastructure is architected, provisioned, secured, and operated going forward. If you get energized by taking a fast-growing environment and giving it real architectural rigor (consistent patterns, full infrastructure-as-code coverage, sane permission models, and cost discipline), this role was built for you.

We're looking for someone who wants to build the operating model, establish the standards, and lead the transformation. You should bring genuine software engineering discipline to how that infrastructure work gets done: everything in git, everything reviewed, everything automated. No snowflakes, no manual changes made "just this once."

Location: This is a fully remote role available to candidates located in U.S. states where Dental Intelligence currently has employees.

What You'll do:

  • Own the reliability, scalability, and security posture of our Azure environment end-to-end
  • Lead the effort to bring our infrastructure fully under Terraform-managed IaC, replacing manual and ad-hoc provisioning with repeatable, version-controlled deployments
  • Treat infrastructure code like production software: everything lives in git, changes go through pull requests and peer review, and modules are tested before they ship
  • Define and implement a coherent Azure architecture strategy, including resource organization, naming and tagging standards, subscription and management group hierarchy, and network topology
  • Redesign and enforce least-privilege access and permission boundaries across Azure RBAC, Entra ID, and service principals
  • Identify and eliminate wasteful or redundant resource provisioning, and build cost visibility and accountability into how infrastructure is deployed
  • Build CI/CD pipelines for infrastructure changes so that plan/apply, validation, and policy checks are automated rather than run by hand
  • Build monitoring, alerting, and observability practices (Azure Monitor, Log Analytics, App Insights, or equivalent) that give the team real signal
  • Manage and modernize the Windows-based on-prem component that sits at the core of our product, and work to integrate it more tightly with our Azure environment
  • Bring the same IaC and automation discipline to bear on our AWS footprint as needed, keeping it as clean and well-run as it is today
  • Drive incident response, postmortems, and reliability engineering practices such as SLOs, error budgets, and capacity planning
  • Partner closely with engineering teams to bake reliability, security, and IaC discipline into the software delivery lifecycle
  • Mentor other engineers on cloud, IaC, and software engineering best practices, raising the bar across the team

What You Bring:

  • ​​6+ years in SRE, DevOps, or infrastructure engineering roles, with deep, hands-on Azure experience
  • Strong, demonstrable Terraform expertise is required. You should be comfortable designing module structures, managing state, and using Terraform to manage complex, multi-resource environments from scratch
  • A genuine cloud mentality: you default to automation, reproducibility, and IaC over manual changes, and you get uncomfortable when infrastructure can't be traced back to code
  • A real software engineering mindset applied to infrastructure: fluency with git workflows (branching, PRs, code review), a bias toward automating anything done more than once, and discomfort with manual, undocumented changes
  • Experience with Windows Server administration, IIS, Active Directory/Entra ID, and Windows-based application stacks, which are critical for the on-prem component at the core of our product
  • Solid understanding of Azure networking (VNets, peering, private endpoints, NSGs), identity (Entra ID, RBAC, service principals and managed identities), and cost management tooling
  • Working knowledge of AWS core services (IAM, VPC, EC2/ECS, etc.) sufficient to support and extend an already well-architected environment
  • Experience integrating on-premises infrastructure with cloud environments (VPN/ExpressRoute, hybrid identity, hybrid networking)
  • A track record of introducing structure and standards into environments that grew organically or quickly; you enjoy imposing order, not just maintaining it
  • Strong scripting ability (PowerShell and/or Bash/Python)
  • Excellent judgment around security and access control; you think in terms of least privilege by default
  • Comfort operating with a high degree of autonomy and ownership, including making architectural calls and driving them through to adoption

Nice to Have:

  • Familiarity with CI/CD tooling (Azure DevOps) for infrastructure pipelines
  • Prior experience leading a cloud environment from an ungoverned state to a well-architected one
  • Relevant certifications (Azure Solutions Architect, Azure Administrator, HashiCorp Terraform Associate)

What You'll Love About Dental Intelligence:

  • Flexible Paid Time Off + 10 company-wide paid holidays
  • Competitive Medical, Dental & vision offerings, with buy up plan options, AND we match your HSA contributions.
  • Fully Paid Parental Leave
  • 401K Retirement savings plan with company match up to 5.5% of your earnings+ unlimited access to personal financial advisors.
  • Learning & Development Reimbursement program
  • Company paid Life, Disability & AD&D
  • Mental Health support programs, Cellphone & Gym membership Discounts, Corporate Sundance Passes, and more!

Please Note: All offers of employment are contingent upon successful completion of a background check, which may include verification of education, employment history, and other credentials. By applying, you confirm that all information you have provided is accurate and complete to the best of your knowledge, and you understand that misrepresentation may result in disqualification or termination.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer (SRE) in United States vacancy
  •  ...TX** Our Opportunity: We are looking for a skilled engineer with disciplines that incorporate aspects of software...  ...including AI/ML-driven approaches to observability and reliability. What you’ll do: • Evangelize SRE mindset and solve problems through systematization.... 
    Senior

    Mindlance

    Austin, TX
    1 day ago
  • $149.4k - $202k

     ...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing... 
    Senior
    Remote work

    Noctua Technology

    Washington DC
    3 days ago
  • $70k - $150k

     ...Address:4100 Gordon Baker RoadJob Family Group:TechnologyWe are seeking a highly motivated and technically strong Senior Lead, Site Reliability Engineering (SRE) to provide technical leadership within the DCOE Reliability Engineering organization. This role is critical to... 
    Senior
    Full time
    Contract work
    Part time

    BMO Bank

    Toronto, OH
    4 days ago
  • $140k - $180k

     ...accelerated growth in the AI-driven world. Learn more at Opportunity We’re looking for a Senior Site Reliability Engineer to help build and scale a high-impact SRE function. You’ll be a technical leader on a team responsible for improving system reliability,... 
    Senior
    Work experience placement
    Local area
    Remote work
    Visa sponsorship
    Work visa

    UJET

    United States
    1 day ago
  •  ...Senior Site Reliability Engineer At Swile, we believe that good products can help reduce friction in daily professional life and boost employee satisfaction...  .... Your role as a Senior Site Reliability Engineer (SRE) centers around creatively solving problems, ensuring a... 
    Senior
    Remote work

    Swile

    United States
    4 days ago
  • Role Summary The Senior Site Reliability Engineer (SRE) is a hands-on role responsible for the availability, performance, and end-to-end observability of QSR digital platforms across Mobile (iOS/Android), Web, and POS systems. This role is part of the Observability... 
    Senior
    Flexible hours

    Donato Technologies, Inc

    Washington DC
    1 day ago
  •  ...high-growth company. The Role Nium is looking for a Senior Manager, Site Reliability Engineering to lead the teams responsible for the availability,...  ...with finance and engineering leadership.  Represent SRE in executive reviews, translating technical risk and system... 
    Senior
    Full time
    Work at office
    Local area
    Worldwide
    Flexible hours
    3 days per week

    Nium

    Remote
    7 days ago
  •  ...Role: Senior Site Reliability Engineer (SRE) Cloud & Kubernetes Location: Atlanta, GA (Onsite) Contract Role Summary: Lead the reliability, scalability, security, and operational excellence of customer-facing platforms across Azure, GCP, and Kubernetes... 
    Senior
    Contract work

    Noblesoft Technologies

    Atlanta, GA
    1 day ago
  • $175k - $215k

     ...enhance these exciting experiences.Sr. Manager, Site Reliability Engineer provides strategic leadership across multiple SRE teams and their managers, ensuring alignment...  ...driving resilience and innovation. Influences senior internal and external stakeholders to secure funding... 
    Senior

    Disney Interactive

    Orlando, FL
    15 hours ago
  •  ...Job Title Location Remote - United States Job Category Information Technology, Platform Engineering, Site Reliability Engineering Industry Computer Software, SaaS, National Security Employee Type FT Exempt Manage Others No Minimum Experience 5 Years... 
    Senior
    Remote work

    CenCore

    United States
    15 hours ago
  • $165k - $225k

     ...demanding AI workloads with enterprise-grade reliability and compliance. Your Role: You will...  ...core. Working closely with our systems engineers, network engineers, and platform...  ...Requirements Experience: 5+ years in SRE, DevOps, or infrastructure engineering roles... 
    Senior
    Remote work
    Flexible hours

    Moonlite

    Chicago, IL
    a month ago
  •  ...security to responsibly propel the global lottery industry ever forward. Position Summary We are looking for a skilled Site Reliability Engineer (SRE) to enhance the stability, performance, and reliability of our production systems. The SRE will work closely with... 
    Senior
    Permanent employment
    Work experience placement
    Local area

    Scientific Games

    Alpharetta, GA
    1 day ago
  •  ...Sr Application Performance and Observability Engineer At Sequoia Connect, we are a Talent-First Technology Ecosystem that redefines how elite professionals interact with the global digital landscape. We move beyond traditional models to act as a catalyst for the top... 
    Senior
    Remote work
    Worldwide

    Sequoia Connect

    United States
    4 days ago
  • $120k - $175k

     ...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's... 
    Senior
    Full time
    Remote work
    Work visa
    Flexible hours

    AEG Presents

    Atlanta, GA
    1 day ago
  •  ...GEICO is seeking a Senior SRE Software Engineer to design, build, and operate high-performance distributed platforms with zero-downtime reliability. You will own incident management tooling, lead on-call operations, and help scale automation across complex systems.... 
    Senior

    Jobleads-US

    Richardson, TX
    2 days ago
  • $175k - $215k

     ...Sr. Manager, Site Reliability Engineer At Disney, we're storytellers. We make the impossible possible...  ...strategic leadership for multiple SRE teams, fostering a culture of reliability...  ...skills, with experience influencing senior stakeholders and driving cross-functional... 
    Senior
    Local area

    The Walt Disney Studios

    Orlando, FL
    3 days ago
  • $152k - $241.5k

     ...infrastructure for AI workloads. We are looking for Software Engineers with SRE or Production Engineering experience who have worked hands-...  ...provisioning through repair.Experience managing production reliability through on-call duties, incident response, observability,... 
    Senior
    Permanent employment
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $300k

     ...experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and...  .... Skills / Must Have: ~7+ years of experience in SRE, DevOps, or Infrastructure Engineering roles supporting... 
    Senior
    Permanent employment
    San Francisco, CA
    more than 2 months ago
  •  ...Job Description Job Description Job Title: Senior AWS Site Reliability Engineer (SRE) Location: Birmingham, Alabama Type: Contract To Hire Work Model: Onsite – onsite Hours: 40.0 Security Clearance: Overview Responsibilities Implement and improve... 
    Senior
    Contract work
    Local area

    System One

    Columbia, SC
    20 days ago
  • $500 per month

     ...Ireland, Spain or Portugal. Key engineering and product teams are based in these locales...  .... The Role We’re hiring a Senior Site Reliability Engineer to join our Platform team...  ..., and make an impact. As an SRE at Maze, you will: Build, operate... 
    Senior
    Remote job
    Full time
    Flexible hours

    Maze

    Remote
    7 days ago
  • $135k - $155k

     ...global manufacturing capacity.Xometry is seeking a Site Reliability Engineer II to join our Site Reliability Engineering (SRE) Organization. In this role as an individual...  ...drive them to completion with guidance from senior engineers.Write clean, efficient, and well-documented... 
    Flexible hours

    Thomas

    Denver, CO
    2 days ago
  • $70.8k - $131.4k

     ...DescriptionThomson Reuters is strengthening its Site Reliability Engineering capability to help engineering and...  ...ResponsibilitiesSupport and maintain SRE operational tooling, including...  ...generated documentation, with guidance from senior engineers and service owners.Assist... 
    Full time
    Work at office
    Local area
    Flexible hours

    Thomson Reuters

    Eagan, MN
    3 days ago
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability... 
    Contract work

    2T Consulting

    Atlanta, GA
    a month ago
  • $207k - $300k

     ...by pushing for changes that improve reliability and velocity.Practice sustainable...  ...Master's degree in Computer Science or Engineering.Experience mentoring engineers and...  ...across cross-functional teams.Site Reliability Engineering (SRE) combines software and systems engineering... 

    Google

    New York, NY
    3 days ago
  •  ...We’re seeking a highly skilled Site Reliability Engineer (SRE) to join our engineering team and help ensure the reliability, scalability, and performance of our systems. As an SRE, you’ll blend software engineering with systems engineering to build and maintain resilient... 
    Temporary work
    Interim role
    Remote work
    Flexible hours

    OutSolve - Beyond Compliance

    Mission, KS
    1 day ago
  •  ...Period of performance: Up to 2 years in duration MUST HAVES: Minimum of 8 years of experience as a Site Reliability Engineer with a strong understanding of SRE principles for highly scalable and reliable systems Possess a bachelor's degree Experience working... 
    Local area
    Relocation package
    3 days per week

    Beyond SOF

    Vienna, VA
    15 hours ago
  • $100k - $200k

     ...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about... 
    Full time

    OPPO

    Palo Alto, CA
    3 days ago
  •  ...arc of the patient journey. The Opportunity: Machine Learning Engineer Patients count on our platform 24/7. You'll build and...  ...certificate lifecycles in line with HIPAA. What You Bring 5+years SRE/DevOps experience running production workloads on AWS, GCP or Azure... 

    Tala Health

    Eastern, KY
    15 hours ago
  • $100k - $180k

     ...Site Reliability Engineer (SRE) - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established... 
    Full time
    H1b
    Local area
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    Edison, NJ
    2 days ago
  •  ...Job Title:  Site Reliability Engineer (Azure Government & Infrastructure) Pay Type : SALARIED EXEMPT  Location:  Remote Citizenship Requirement...  ...Role/Responsibilities The Site Reliability Engineer (SRE) for Azure Government & Infrastructure plays a critical role... 
    Full time
    Remote work
    Monday to Friday

    Quzara LLC

    United States
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer (SRE). Be the first to apply!