Senior Site Reliability Engineer
$168k - $200kDatavant
Datavant is the data collaboration platform trusted for healthcare. Guided by our mission to make the world's health data secure, accessible and actionable, we provide critical data solutions for organizations across the healthcare ecosystem - including providers, health plans, researchers, and life sciences companies. From fulfilling a single patient's request for their medical records to powering the AI revolution in healthcare, Datavanters are building the future of how data is connected and used to improve health.
By joining Datavant today, you're stepping onto a driven and highly collaborative team that is passionate about creating transformative change in healthcare.
What We're Looking For
We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the forefront of building and operating a resilient, observable, and scalable platform that enables mission-critical data and ML workloads across our organization.
This role is ideal for someone who combines a strong SRE mindset with deep cloud infrastructure and data platform experience . You're comfortable operating at scale in a complex, hybrid cloud environment and can architect systems that balance velocity, safety, and cost. You'll work closely with Data & ML Engineers, Data Scientists, Analysts, and App Engineering teams to build a modern data platform that is secure, self-service, and production-grade.
What You Will Do
Operate and Improve Databricks and Snowflake : Own Databricks & Snowflake platforms lifecycle-including automation, workspace governance, job orchestration, and cost optimization.
Design for Reliability : Architect resilient, scalable, and secure infrastructure across cloud environments. Drive initiatives around failover, autoscaling, chaos testing, and capacity planning.
Advance Observability : Build and maintain platform-wide monitoring, alerting, and logging infrastructure using Datadog and other open tooling. Define and enforce SLOs/SLAs for critical services.
Drive CI/CD for Data & ML : Automate deployments of data pipelines, ML workflows, and infra components using GitHub Actions , Terraform, and related IaC tooling.
Enable Data Flow Across Platforms : Build patterns and tooling to support inter- and intra-cloud data movement across systems like Snowflake, S3, Delta Lake, and Kafka.
Champion Event-Driven Architectures : Leverage cloud-native tools like EventBridge , SNS/SQS, and Lambda to build loosely coupled, scalable data systems.
Collaborate Across Teams : Serve as the SRE and platform partner for teams across the organization, ensuring the platform meets the needs of analytics, data science, and product use cases.
Contribute to Strategy : Influence engineering-wide decisions on data platform architecture , ML enablement , and data product strategy .
What You Need to Succeed
6+ years in SRE, platform engineering, or DevOps roles supporting data-intensive or ML-powered applications.
AI-native working style: daily use of Claude Code, Cursor, Copilot, or equivalent, with views on how they make a team faster.
Hands-on Databricks experience , including workspace setup, cluster/job management, and integration with CI/CD and data orchestration tools. Experience with Snowflake as well.
Deep understanding of cloud-native infrastructure on AWS (or similar), including VPCs, IAM, event-driven patterns, and serverless compute.
Proven expertise with observability tools (especially Datadog) and architecting platform-wide logging and monitoring solutions.
Strong command of CI/CD tooling , especially GitHub Actions , infrastructure-as-code (Terraform), and deployment automation for data systems.
Working knowledge in shell scripting and Python.
Experience building and supporting highly available, fault-tolerant systems .
Excellent communication and collaboration skills; able to work effectively across teams.
What Helps You Stand Out
DevSecOps mindset : Familiarity with implementing security best practices in IaC, CI/CD, secret management, and audit logging.
Experience with ML infrastructure tooling such as MLflow, Feature Stores, and GPU workload orchestration.
Strong experience in both Databricks and Snowflake in a large scale production lakehouse with cross-warehouse interoperability, e.g. Iceberg v3, Glue, etc.
Background in compliance-aware architecture (e.g., HIPAA, SOC 2) or regulated industries.
Familiarity with multi-cloud or hybrid cloud data environments ; experience with Azure.
Contributions to open-source infrastructure, SRE, or observability tools.
We are committed to building a diverse team of Datavanters who are all responsible for stewarding a high-performance culture in which all Datavanters belong and thrive. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status.
At Datavant our total rewards strategy powers a high-growth, high-performance, health technology company that rewards our employees for transforming health care through creating industry-defining data logistics products and services.
The range posted is for a given job title, which can include multiple levels. Individual rates for the same job title may differ based on their level, responsibilities, skills, and experience for a specific job.
The estimated total cash compensation range for this role is:
$168,000-$200,000 USD
To ensure the safety of patients and staff, many of our clients require post-offer health screenings and proof and/or completion of various vaccinations such as the flu shot, Tdap, COVID-19, etc. Any requests to be exempted from these requirements will be reviewed by Datavant Human Resources and determined on a case-by-case basis. Depending on the state in which you will be working, exemptions may be available on the basis of disability, medical contraindications to the vaccine or any of its components, pregnancy or pregnancy-related medical conditions, and/or religion.
This job is not eligible for employment sponsorship.
Datavant is committed to a work environment free from job discrimination. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status. To learn more about our commitment, please review our EEO Commitment Statement here ( . Know Your Rights ( , explore the resources available through the EEOC for more information regarding your legal rights and protections. In addition, Datavant does not and will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay.
At the end of this application, you will find a set of voluntary demographic questions. If you choose to respond, your answers will be anonymous and will help us identify areas for improvement in our recruitment process. (We can only see aggregate responses, not individual ones. In fact, we aren't even able to see whether you've responded.) Responding is entirely optional and will not affect your application or hiring process in any way.
Datavant is committed to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need an accommodation while seeking employment, please request it here, ( by selecting the 'Interview Accommodation Request' category. You will need your requisition ID when submitting your request, you can find instructions for locating it here ( . Requests for reasonable accommodations will be reviewed on a case-by-case basis.
For more information about how we collect and use your data, please review our Privacy Policy ( .
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SeniorFull timeWork at officeLocal areaRemote workWork from home- ...Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation...SeniorFull timeWorldwideFlexible hours
- ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...vision, audio, and emerging multimodal architectures — fast, reliable, and effortless to deploy at massive scale. To deliver on that...Senior
- ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront of Cloud and Big Data technology. In this role you will establish yourself as a technical leader by exposing yourself to...Senior
$120k - $175k
...Requirements: We need at least 5 years of experience as a reliability-focused engineer in a fast-moving, rapidly expanding enterprise setting.... ...to qualified remote applicants anywhere in the U.S. This Senior Site Reliability Engineer role offers a typical salary range...SeniorFull timeRemote workVisa sponsorshipFlexible hours$141k - $208k
...be a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability...SeniorLocal areaRemote workHome officeFlexible hours$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SeniorTemporary workImmediate startFlexible hoursShift work- ...development, testing, implementation, and operation of secure, scalable, resilient, and highly available software platforms using Site Reliability Engineering and AI-native engineering practices. The engineer collaborates across technical teams to build and support cloud-native,...SeniorFull timeTemporary workPart timeWork experience placementLocal areaFlexible hours
- ...Senior Site Reliability Engineer Atlanta, Georgia Who We Are QGenda is redefining healthcare workforce management everywhere care is delivered. We're on a mission to empower the healthcare industry to better onboarding, deploy, and manage their workforce. Over...SeniorPermanent employmentFull timeWork at officeRemote workWork from homeWork visa
$178.13k - $205.4k
...Bachelor's degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5)... ...websites that are not Workday Careers. Please be aware of sites that may ask for you to input your data in connection with a job...SeniorWork at officeRemote workFlexible hours- OneTrust is seeking a Senior Software Engineer - SRE in Atlanta to join the Software Engineering team. You will design, implement, and maintain... ...focusing on observability, automated incident response, and reliable deployments. You will collaborate with cross-functional...Senior
- OneTrust is seeking a Senior Software Engineer in Atlanta, Georgia. The role involves designing and maintaining a reliable application platform, collaborating with engineering teams, and enhancing customer experiences through observability tools. The ideal candidate will...Senior
- The Challenge We’re looking for a Senior Software Engineer that will report to the Development Manager / R&D Head. In this role you will be part... .... Manage and enforce error budgets to balance system reliability with product feature velocity. Improve alert quality by reducing...SeniorLocal areaFlexible hours
$35 - $45 per hour
DescriptionKforce has a client that is seeking a remote Site Reliability Engineer to join their team.Summary:The team consists of systems that can track lead management, job management and sales management. It is built on Salesforce but underpinned by a lot of Java/API'...Remote work$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying...Full timeTemporary workWork experience placementFlexible hours$138.1k - $198.2k
...more intuitive with technology that simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments... ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$152.13k - $162.13k
...challenge the status-quo.Unum is changing, and we’re excited about what’s next. Join us.General Summary:Unum Group seeks Site Reliability Engineers in Atlanta, GA.Applicants who are interested in this position may apply at (Ref #66753) for consideration.Design, build,...Full timeTemporary workWork at officeRemote work- T-Mobile USA, Inc. is seeking an experienced Site Reliability Engineer to lead reliability across cloud platforms and enterprise applications. You will own monitoring, incident response, CI/CD, and automation while guiding AI-enabled development practices and cross-team...Senior
- T-Mobile USA, Inc. seeks a Site Reliability Engineer lead to design, build, and operate secure, scalable software platforms using SRE and AI-native practices across cloud-native and distributed systems. You will mentor the SRE team, drive Terraform-based infrastructure,...Senior
- ...speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage.About The Role:The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services...SeniorWeekly payPermanent employmentFull timeWork at office
$101.5k - $169.1k
...include an incentive program.Job DescriptionThe Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train... ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures. Maintain...SeniorFull timeWork at officeRemote workVisa sponsorshipFlexible hours- ...of America)Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and... ...platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
$167.7k - $245.2k
...Cisco Meraki, we are responsible for building and growing the cloud that supports these customers and their networks. As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze...Permanent employmentFull timeTemporary workLocal areaFlexible hours- ...Site Reliability Engineer (SRE) Location: Atlanta, GA 30303 Duration of the project: 12 Months Strong expertise in Ansible with an SRE background Ability to review and test GitLab Duo generated code CICD pipelines (GitLab, GitHub Actions) Infrastructure automation...
- ...Technical Support Specialist In Site Reliability Engineering (Sre) Mandatory skills: Scripting and programming languages like Python, Java, Ruby. Cloud and infrastructure management – AWS, Google cloud and Azure is a plus- CI/CD Automation, Database Management. The...
$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term potential and is located in Atlanta, GA (Onsite). Please review the job description below and contact me ASAP if you are interested...Contract workLocal areaImmediate start- ...availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual processes using Python, Ruby, Unix Shell (bash,...Immediate start
- Job description Snowflake SRE JD Your Role Accountabilities Primarily responsible for administrating Snowflake environments on AWS Identify, tune, and fix the performance issues on priority. Diagnose and troubleshoot Snowflake related errors and work with team to raise...
$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...Full timeTemporary workWork experience placementRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer remote Atlanta, GA
- site reliability engineer sre Atlanta, GA
- site reliability engineer Atlanta, GA
- senior operations associate Atlanta, GA
- senior safety specialist Atlanta, GA
- senior technology project manager Atlanta, GA
- remote senior business analyst Atlanta, GA
- senior director fp&a Atlanta, GA
- senior manager clinical operations Atlanta, GA
- senior supervisor Atlanta, GA


