Senior Site Reliability Engineer
$168k - $200kDatavant
Datavant is the data collaboration platform trusted for healthcare. Guided by our mission to make the world’s health data secure, accessible and actionable, we provide critical data solutions for organizations across the healthcare ecosystem - including providers, health plans, researchers, and life sciences companies. From fulfilling a single patient’s request for their medical records to powering the AI revolution in healthcare, Datavanters are building the future of how data is connected and used to improve health.
By joining Datavant today, you’re stepping onto a driven and highly collaborative team that is passionate about creating transformative change in healthcare.
What We’re Looking For
We’re looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You’ll be at the forefront of building and operating a resilient, observable, and scalable platform that enables mission-critical data and ML workloads across our organization.
This role is ideal for someone who combines a strong SRE mindset with deep cloud infrastructure and data platform experience . You're comfortable operating at scale in a complex, hybrid cloud environment and can architect systems that balance velocity, safety, and cost. You’ll work closely with Data & ML Engineers, Data Scientists, Analysts, and App Engineering teams to build a modern data platform that is secure, self-service, and production-grade.
What You Will Do
Operate and Improve Databricks and Snowflake : Own Databricks & Snowflake platforms lifecycle—including automation, workspace governance, job orchestration, and cost optimization.
Design for Reliability : Architect resilient, scalable, and secure infrastructure across cloud environments. Drive initiatives around failover, autoscaling, chaos testing, and capacity planning.
Advance Observability : Build and maintain platform-wide monitoring, alerting, and logging infrastructure using Datadog and other open tooling. Define and enforce SLOs/SLAs for critical services.
Drive CI/CD for Data & ML : Automate deployments of data pipelines, ML workflows, and infra components using GitHub Actions , Terraform, and related IaC tooling.
Enable Data Flow Across Platforms : Build patterns and tooling to support inter- and intra-cloud data movement across systems like Snowflake, S3, Delta Lake, and Kafka.
Champion Event-Driven Architectures : Leverage cloud-native tools like EventBridge , SNS/SQS, and Lambda to build loosely coupled, scalable data systems.
Collaborate Across Teams : Serve as the SRE and platform partner for teams across the organization, ensuring the platform meets the needs of analytics, data science, and product use cases.
Contribute to Strategy : Influence engineering-wide decisions on data platform architecture , ML enablement , and data product strategy .
What You Need to Succeed
6+ years in SRE, platform engineering, or DevOps roles supporting data-intensive or ML-powered applications.
AI-native working style: daily use of Claude Code, Cursor, Copilot, or equivalent, with views on how they make a team faster.
Hands-on Databricks experience , including workspace setup, cluster/job management, and integration with CI/CD and data orchestration tools. Experience with Snowflake as well.
Deep understanding of cloud-native infrastructure on AWS (or similar), including VPCs, IAM, event-driven patterns, and serverless compute.
Proven expertise with observability tools (especially Datadog) and architecting platform-wide logging and monitoring solutions.
Strong command of CI/CD tooling , especially GitHub Actions , infrastructure-as-code (Terraform), and deployment automation for data systems.
Working knowledge in shell scripting and Python.
Experience building and supporting highly available, fault-tolerant systems .
Excellent communication and collaboration skills; able to work effectively across teams.
What Helps You Stand Out
DevSecOps mindset : Familiarity with implementing security best practices in IaC, CI/CD, secret management, and audit logging.
Experience with ML infrastructure tooling such as MLflow, Feature Stores, and GPU workload orchestration.
Strong experience in both Databricks and Snowflake in a large scale production lakehouse with cross-warehouse interoperability, e.g. Iceberg v3, Glue, etc.
Background in compliance-aware architecture (e.g., HIPAA, SOC 2) or regulated industries.
Familiarity with multi-cloud or hybrid cloud data environments ; experience with Azure.
Contributions to open-source infrastructure, SRE, or observability tools.
We are committed to building a diverse team of Datavanters who are all responsible for stewarding a high-performance culture in which all Datavanters belong and thrive. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status.
At Datavant our total rewards strategy powers a high-growth, high-performance, health technology company that rewards our employees for transforming health care through creating industry-defining data logistics products and services.
The range posted is for a given job title, which can include multiple levels. Individual rates for the same job title may differ based on their level, responsibilities, skills, and experience for a specific job.
The estimated total cash compensation range for this role is:
$168,000—$200,000 USD
To ensure the safety of patients and staff, many of our clients require post-offer health screenings and proof and/or completion of various vaccinations such as the flu shot, Tdap, COVID-19, etc. Any requests to be exempted from these requirements will be reviewed by Datavant Human Resources and determined on a case-by-case basis. Depending on the state in which you will be working, exemptions may be available on the basis of disability, medical contraindications to the vaccine or any of its components, pregnancy or pregnancy-related medical conditions, and/or religion.
This job is not eligible for employment sponsorship.
Datavant is committed to a work environment free from job discrimination. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status. To learn more about our commitment, please review our EEO Commitment Statement here ( . Know Your Rights ( , explore the resources available through the EEOC for more information regarding your legal rights and protections. In addition, Datavant does not and will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay.
At the end of this application, you will find a set of voluntary demographic questions. If you choose to respond, your answers will be anonymous and will help us identify areas for improvement in our recruitment process. (We can only see aggregate responses, not individual ones. In fact, we aren’t even able to see whether you’ve responded.) Responding is entirely optional and will not affect your application or hiring process in any way.
Datavant is committed to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need an accommodation while seeking employment, please request it here, ( by selecting the ‘Interview Accommodation Request’ category. You will need your requisition ID when submitting your request, you can find instructions for locating it here ( . Requests for reasonable accommodations will be reviewed on a case-by-case basis.
For more information about how we collect and use your data, please review our Privacy Policy ( .
- As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing AI, HPC, Kubernetes, and enterprise infrastructure environments. This role transforms newly installed hardware into production...SeniorWork at officeImmediate startWorldwideShift work
- ...Site Reliability Engineer We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You'll partner with engineers and data scientists to build, automate, and...Senior
$120k - $175k
...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's...SeniorFull timeRemote workWork visaFlexible hours- ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront of Cloud and Big Data technology. In this role you will establish yourself as a technical leader by exposing yourself to...Senior
- ...Senior Site Reliability Engineer Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems...SeniorWorldwide
- ...professionalism. We are seeking an experienced AWS solution design engineer/architect to join our infrastructure cloud team. The... ...features efficiently and confidently them into production. As Senior SRE, you will be responsible for providing leadership, design and...Senior
$136.2k - $214.01k
...outcomes Visionary in future focused problem-solving Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to...SeniorFull timeFlexible hours$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SeniorTemporary workImmediate startFlexible hoursShift work$149.8k - $241.5k
...efficiency, accelerate time-to-value, and deliver better customer experiences. About The Role We're looking for a Senior Site Reliability Engineer who's passionate about building reliable, scalable infrastructure that helps developers ship better software faster....SeniorWork at officeLocal areaRemote workWork from homeWorldwideHome officeFlexible hours$178.13k - $205.4k
...Bachelor's degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5)... ...websites that are not Workday Careers. Please be aware of sites that may ask for you to input your data in connection with a job...SeniorWork at officeRemote workFlexible hours$123.4k - $222.53k
...! Ready grow your career as part of the Uncarrier journey at T-Mobile? Our team is searching for our next Sr. Site Reliability Engineer to strengthen the reliability and resilience of the systems powering T-Mobile's payment platforms, enabling faster, safer...SeniorFull timeTemporary workPart timeWork experience placementLocal areaFlexible hours$169.3k - $304.7k
...in building and maintaining fast, efficient, scalable, and reliable routing software and infrastructure that is responsible... ...growth and stability of our global platform. As a Principal Site Reliability Engineer - Network, you will be responsible for: Architecting,...Work experience placementWork at office- ...of America)Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and... ...platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability...Contract work
- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...Senior
- ...configure the monitoring and alerting metrics so the support engineers can proactively and timely validate, troubleshoot and... ...availability critical application components. • 1+ Years in Site Reliability Engineering organization preferred • Overall 4-6years of experience...Work experience placement
- ...availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual processes using Python, Ruby, Unix Shell (bash,...Immediate start
$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term potential and is located in Atlanta, GA (Onsite). Please review the job description below and contact me ASAP if you are interested...Contract workLocal areaImmediate start- ...Technical Support Specialist In Site Reliability Engineering (Sre) Mandatory skills: Scripting and programming languages like Python, Java, Ruby. Cloud and infrastructure management – AWS, Google cloud and Azure is a plus- CI/CD Automation, Database Management. The...
$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office- ...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's... ...is founder-led, profitable, and growing. We are hiring a Site Reliability Engineer Our goal is to perfect enterprise infrastructure DevOps...Work at officeLocal areaRemote workWork from homeWorldwide
- ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it...Full timeLive inWork at office
- ...data, and human expertise. We deliver faster, smarter, more reliable insights to insurance carriers and single-family rental... ...at scale, you're in the right place. The Role As a Site Reliability Engineer, you'll be responsible for the availability, scalability,...Full timeFlexible hours
$100k - $120k
...OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying...Full timeTemporary workWork experience placementFlexible hours- ...Job Purpose At Intercontinental Exchange (NYSE:ICE), we engineer technology, exchanges and clearing houses that connect companies... ...-oriented people to join our team. We are seeking a Site Reliability Engineer to bring 3+ years of hands-on experience to our SRE...
- ...Join to apply for the Site Reliability Engineer role at Motion Recruitment Join to apply for the Site Reliability Engineer role at... ...and security enhancements. Posted By: VMS Sourcing Seniority level ~ Seniority level Mid-Senior level Employment...Contract workWorldwide
$130k - $145k
...Back Site Reliability Engineer Cloud/Infrastructure Atlanta , GA Sep 2, 2026 Site Reliability Engineer Atlanta, GA / Hybrid Blu Omega is seeking a Site Reliability Engineer to support a federal program focused on enterprise cloud modernization. This role operates...Temporary work$130k - $150k
....00/yr Overview: We are seeking a highly skilled Site Reliability Engineer (SRE) to join our team and help build and maintain scalable... ...software development lifecycle and CI/CD principles Seniority level ~ Seniority level Mid-Senior level Employment...Full timeRemote work- ...We have an immediate need for a Senior Release Train Engineer for a contract assignment located in Carmel, Indiana . The Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train (ART) by steering it to success and navigating the complexity...SeniorContract workWork at officeImmediate start
$145k - $160k
...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives critical to our multi-region disaster recovery roadmap. You will architect and implement robust observability pipelines, ensure deep...Temporary workRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer remote Atlanta, GA
- site reliability engineer sre Atlanta, GA
- senior living director Atlanta, GA
- senior php developer remote Atlanta, GA
- senior manager customer operations Atlanta, GA
- senior java developer Atlanta, GA
- senior software engineer ruby on rails Atlanta, GA
- sr finance manager Atlanta, GA
- sr marketing manager Atlanta, GA




