Senior Site Reliability Engineer
$168k - $200kDatavant
Datavant is the data collaboration platform trusted for healthcare. Guided by our mission to make the world’s health data secure, accessible and actionable, we provide critical data solutions for organizations across the healthcare ecosystem - including providers, health plans, researchers, and life sciences companies. From fulfilling a single patient’s request for their medical records to powering the AI revolution in healthcare, Datavanters are building the future of how data is connected and used to improve health.
By joining Datavant today, you’re stepping onto a driven and highly collaborative team that is passionate about creating transformative change in healthcare.
What We’re Looking For
We’re looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You’ll be at the forefront of building and operating a resilient, observable, and scalable platform that enables mission-critical data and ML workloads across our organization.
This role is ideal for someone who combines a strong SRE mindset with deep cloud infrastructure and data platform experience . You're comfortable operating at scale in a complex, hybrid cloud environment and can architect systems that balance velocity, safety, and cost. You’ll work closely with Data & ML Engineers, Data Scientists, Analysts, and App Engineering teams to build a modern data platform that is secure, self-service, and production-grade.
What You Will Do
Operate and Improve Databricks and Snowflake : Own Databricks & Snowflake platforms lifecycle—including automation, workspace governance, job orchestration, and cost optimization.
Design for Reliability : Architect resilient, scalable, and secure infrastructure across cloud environments. Drive initiatives around failover, autoscaling, chaos testing, and capacity planning.
Advance Observability : Build and maintain platform-wide monitoring, alerting, and logging infrastructure using Datadog and other open tooling. Define and enforce SLOs/SLAs for critical services.
Drive CI/CD for Data & ML : Automate deployments of data pipelines, ML workflows, and infra components using GitHub Actions , Terraform, and related IaC tooling.
Enable Data Flow Across Platforms : Build patterns and tooling to support inter- and intra-cloud data movement across systems like Snowflake, S3, Delta Lake, and Kafka.
Champion Event-Driven Architectures : Leverage cloud-native tools like EventBridge , SNS/SQS, and Lambda to build loosely coupled, scalable data systems.
Collaborate Across Teams : Serve as the SRE and platform partner for teams across the organization, ensuring the platform meets the needs of analytics, data science, and product use cases.
Contribute to Strategy : Influence engineering-wide decisions on data platform architecture , ML enablement , and data product strategy .
What You Need to Succeed
6+ years in SRE, platform engineering, or DevOps roles supporting data-intensive or ML-powered applications.
AI-native working style: daily use of Claude Code, Cursor, Copilot, or equivalent, with views on how they make a team faster.
Hands-on Databricks experience , including workspace setup, cluster/job management, and integration with CI/CD and data orchestration tools. Experience with Snowflake as well.
Deep understanding of cloud-native infrastructure on AWS (or similar), including VPCs, IAM, event-driven patterns, and serverless compute.
Proven expertise with observability tools (especially Datadog) and architecting platform-wide logging and monitoring solutions.
Strong command of CI/CD tooling , especially GitHub Actions , infrastructure-as-code (Terraform), and deployment automation for data systems.
Working knowledge in shell scripting and Python.
Experience building and supporting highly available, fault-tolerant systems .
Excellent communication and collaboration skills; able to work effectively across teams.
What Helps You Stand Out
DevSecOps mindset : Familiarity with implementing security best practices in IaC, CI/CD, secret management, and audit logging.
Experience with ML infrastructure tooling such as MLflow, Feature Stores, and GPU workload orchestration.
Strong experience in both Databricks and Snowflake in a large scale production lakehouse with cross-warehouse interoperability, e.g. Iceberg v3, Glue, etc.
Background in compliance-aware architecture (e.g., HIPAA, SOC 2) or regulated industries.
Familiarity with multi-cloud or hybrid cloud data environments ; experience with Azure.
Contributions to open-source infrastructure, SRE, or observability tools.
We are committed to building a diverse team of Datavanters who are all responsible for stewarding a high-performance culture in which all Datavanters belong and thrive. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status.
At Datavant our total rewards strategy powers a high-growth, high-performance, health technology company that rewards our employees for transforming health care through creating industry-defining data logistics products and services.
The range posted is for a given job title, which can include multiple levels. Individual rates for the same job title may differ based on their level, responsibilities, skills, and experience for a specific job.
The estimated total cash compensation range for this role is:
$168,000—$200,000 USD
To ensure the safety of patients and staff, many of our clients require post-offer health screenings and proof and/or completion of various vaccinations such as the flu shot, Tdap, COVID-19, etc. Any requests to be exempted from these requirements will be reviewed by Datavant Human Resources and determined on a case-by-case basis. Depending on the state in which you will be working, exemptions may be available on the basis of disability, medical contraindications to the vaccine or any of its components, pregnancy or pregnancy-related medical conditions, and/or religion.
This job is not eligible for employment sponsorship.
Datavant is committed to a work environment free from job discrimination. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status. To learn more about our commitment, please review our EEO Commitment Statement here ( . Know Your Rights ( , explore the resources available through the EEOC for more information regarding your legal rights and protections. In addition, Datavant does not and will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay.
At the end of this application, you will find a set of voluntary demographic questions. If you choose to respond, your answers will be anonymous and will help us identify areas for improvement in our recruitment process. (We can only see aggregate responses, not individual ones. In fact, we aren’t even able to see whether you’ve responded.) Responding is entirely optional and will not affect your application or hiring process in any way.
Datavant is committed to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need an accommodation while seeking employment, please request it here, ( by selecting the ‘Interview Accommodation Request’ category. You will need your requisition ID when submitting your request, you can find instructions for locating it here ( . Requests for reasonable accommodations will be reviewed on a case-by-case basis.
For more information about how we collect and use your data, please review our Privacy Policy ( .
$160k - $200k
...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$134.25k - $214.8k
...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed... ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,...SeniorWork experience placementWork at officeRemote work$127k - $249k
Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that... ...maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering...SeniorLocal areaWorldwideFlexible hours$166k - $220k
...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &... ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril...SeniorFull timeWork experience placementImmediate start$128k - $160k
...step at a time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Senior Site Reliability Engineer, you’ll build and scale the critical Kubernetes infrastructure that powers our platforms and services. You’ll solve...SeniorFull timeImmediate start$160k - $200k
...Senior Site Reliability Engineer This role is located in Somerville, MA - We are a hybrid work environment and are in the office 3+ days/per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce with...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week$140k - $210.9k
...position will be primarily on-site with residency commutable to... ...DevOps backgrounds or software engineering backgrounds (e.g., Java... ...interest in operating and improving reliability of distributed production... ...Responsibilities As a Senior Engineer of the SRE / Production...SeniorFull timeTemporary workPart timeWork at officeShift work$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SeniorWork experience placementWork at office$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SeniorTemporary workImmediate startFlexible hoursShift work- ...Information Technology group delivers secure, reliable technology solutions that enable... ...You Will Have in This RoleAs a Senior Application Support Engineer, you will help power DTCC's global... ...processing and settlement.Leveraging Site Reliability Engineering (SRE) principles...SeniorRemote workFlexible hours
- ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the... ...candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based...SeniorFull time
$60 - $70 per hour
...Job Description- 100% REMOTE! This DevOps Automation Engineer role sits within a platform operations team and focuses on supporting... ...a global, regulated MedTech context. The position emphasizes site reliability engineering and platform operations over CI/CD-heavy...SeniorContract workTemporary workRemote work$139k - $257.55k
...Individual Contributor The Challenge The Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe...SeniorTemporary workLocal areaRemote workRelocation- The Depository Trust & Clearing Corporation (DTCC) seeks a Senior Application Support Engineer to ensure reliability and performance of its critical trade processing platforms. You will apply SRE principles, drive automation, and partner with global teams to support AWS...Senior
$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...career mentoring and performance coaching from an assigned senior colleague. Additional leadership and collaboration opportunities...Work at officeWork from home3 days per week$160k - $200k
Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware...Local areaRemote work- ...mission-critical industries, helping partners move more quickly and reliably from algorithm to silicon. Our platform accelerates deployment... .... The Roles We are looking for an experienced software engineer to help us build a new generation of transpilation tools...SeniorFull timeRemote workRelocation packageFlexible hours
$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed...
- ...efforts and improve strategic decision-making. As a member of our engineering team, you'll be working closely with other team members and our... ...engineering teams can work and the tools they useLocation: on-site in BostonWe believe that it takes a diverse team to build the...Senior
- ...ISEE is seeking an experienced Senior Software Engineer to join our team. The ideal candidate has several years of work experience, and have worked on complex, performance-critical code-bases. Role responsibilities include: - Support full software development life...SeniorFull timeWork experience placement
$108k - $209k
...seeking an experienced, creative, and talented Principal / Senior Software Engineer. The ideal candidate will have a strong background in software... .... Leverage AWS cloud infrastructure to build scalable, reliable, and efficient applications and AI-powered services. Uphold...Senior$160k - $225k
...Staff Site Reliability Engineer Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients. Our products speed up workflows in areas from target identification and clinical development to market access and precision medicine in the...$148k - $185k
...the future together. The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure... ...into clear, actionable insights that help teams and senior leaders make better decisions about reliability, risk, and...Full timeImmediate start$150k - $215k
...individuals optimize their health, fitness, and recovery. As a Senior Software Engineer on the AI team, you will play a key role in building and... ...is ideal for an engineer who is passionate about building reliable, scalable applications and thrives in a fast-paced,...SeniorFull timeWork at officeRelocation$130k - $140k
...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient...SeniorOngoing contractFull timeTemporary workWork experience placement$150k - $195k
...empowers members to perform at a higher level through a deeper understanding of their bodies and daily lives.WHOOP is seeking a Senior Reliability Engineer to lead the charge in ensuring our hardware products deliver a consistent, high-reliability experience for members. In...SeniorFull timeWork at officeRelocation$138k - $252k
...scaled autonomy Solicit and incorporate feedback from end users of the APIs and implementations, and collaborate with adjacent engineering teams to help make the product vision a reality Develop and improve our APIs for commanding and controlling teams of...SeniorFull timeWork experience placementLocal areaRelocation packageFlexible hours$140k - $160k
...About this role: Pickle is seeking a dynamic, driven Senior Software Engineer, Navigation, to enhance the speed and safety of our autonomous... ...algorithms and capable of optimizing for performance and reliability. Detail-oriented, but with a system-level mindset....SeniorFull timeWork at office3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Boston, MA
- site reliability engineer sre Boston, MA
- srs distribution Boston, MA
- senior operations coordinator Boston, MA
- senior associate architect Boston, MA
- senior dynamics crm developer Boston, MA
- senior application security Boston, MA
- senior account director Boston, MA
- sr hr business partner Boston, MA
- senior supervisor Boston, MA


