Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$168k - $200k

Datavant

Datavant is the data collaboration platform trusted for healthcare. Guided by our mission to make the world’s health data secure, accessible and actionable, we provide critical data solutions for organizations across the healthcare ecosystem - including providers, health plans, researchers, and life sciences companies. From fulfilling a single patient’s request for their medical records to powering the AI revolution in healthcare, Datavanters are building the future of how data is connected and used to improve health.

By joining Datavant today, you’re stepping onto a driven and highly collaborative team that is passionate about creating transformative change in healthcare.

What We’re Looking For

We’re looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You’ll be at the forefront of building and operating a resilient, observable, and scalable platform that enables mission-critical data and ML workloads across our organization.

This role is ideal for someone who combines a strong SRE mindset with deep cloud infrastructure and data platform experience . You're comfortable operating at scale in a complex, hybrid cloud environment and can architect systems that balance velocity, safety, and cost. You’ll work closely with Data & ML Engineers, Data Scientists, Analysts, and App Engineering teams to build a modern data platform that is secure, self-service, and production-grade.

What You Will Do

  • Operate and Improve Databricks and Snowflake : Own Databricks & Snowflake platforms lifecycle—including automation, workspace governance, job orchestration, and cost optimization.

  • Design for Reliability : Architect resilient, scalable, and secure infrastructure across cloud environments. Drive initiatives around failover, autoscaling, chaos testing, and capacity planning.

  • Advance Observability : Build and maintain platform-wide monitoring, alerting, and logging infrastructure using Datadog and other open tooling. Define and enforce SLOs/SLAs for critical services.

  • Drive CI/CD for Data & ML : Automate deployments of data pipelines, ML workflows, and infra components using GitHub Actions , Terraform, and related IaC tooling.

  • Enable Data Flow Across Platforms : Build patterns and tooling to support inter- and intra-cloud data movement across systems like Snowflake, S3, Delta Lake, and Kafka.

  • Champion Event-Driven Architectures : Leverage cloud-native tools like EventBridge , SNS/SQS, and Lambda to build loosely coupled, scalable data systems.

  • Collaborate Across Teams : Serve as the SRE and platform partner for teams across the organization, ensuring the platform meets the needs of analytics, data science, and product use cases.

  • Contribute to Strategy : Influence engineering-wide decisions on data platform architecture , ML enablement , and data product strategy .

What You Need to Succeed

  • 6+ years in SRE, platform engineering, or DevOps roles supporting data-intensive or ML-powered applications.

  • AI-native working style: daily use of Claude Code, Cursor, Copilot, or equivalent, with views on how they make a team faster.

  • Hands-on Databricks experience , including workspace setup, cluster/job management, and integration with CI/CD and data orchestration tools. Experience with Snowflake as well.

  • Deep understanding of cloud-native infrastructure on AWS (or similar), including VPCs, IAM, event-driven patterns, and serverless compute.

  • Proven expertise with observability tools (especially Datadog) and architecting platform-wide logging and monitoring solutions.

  • Strong command of CI/CD tooling , especially GitHub Actions , infrastructure-as-code (Terraform), and deployment automation for data systems.

  • Working knowledge in shell scripting and Python.

  • Experience building and supporting highly available, fault-tolerant systems .

  • Excellent communication and collaboration skills; able to work effectively across teams.

What Helps You Stand Out

  • DevSecOps mindset : Familiarity with implementing security best practices in IaC, CI/CD, secret management, and audit logging.

  • Experience with ML infrastructure tooling such as MLflow, Feature Stores, and GPU workload orchestration.

  • Strong experience in both Databricks and Snowflake in a large scale production lakehouse with cross-warehouse interoperability, e.g. Iceberg v3, Glue, etc.

  • Background in compliance-aware architecture (e.g., HIPAA, SOC 2) or regulated industries.

  • Familiarity with multi-cloud or hybrid cloud data environments ; experience with Azure.

  • Contributions to open-source infrastructure, SRE, or observability tools.

We are committed to building a diverse team of Datavanters who are all responsible for stewarding a high-performance culture in which all Datavanters belong and thrive. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status.

At Datavant our total rewards strategy powers a high-growth, high-performance, health technology company that rewards our employees for transforming health care through creating industry-defining data logistics products and services.

The range posted is for a given job title, which can include multiple levels. Individual rates for the same job title may differ based on their level, responsibilities, skills, and experience for a specific job.

The estimated total cash compensation range for this role is:

$168,000—$200,000 USD

To ensure the safety of patients and staff, many of our clients require post-offer health screenings and proof and/or completion of various vaccinations such as the flu shot, Tdap, COVID-19, etc. Any requests to be exempted from these requirements will be reviewed by Datavant Human Resources and determined on a case-by-case basis. Depending on the state in which you will be working, exemptions may be available on the basis of disability, medical contraindications to the vaccine or any of its components, pregnancy or pregnancy-related medical conditions, and/or religion.

This job is not eligible for employment sponsorship.

Datavant is committed to a work environment free from job discrimination. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status. To learn more about our commitment, please review our EEO Commitment Statement here ( . Know Your Rights ( , explore the resources available through the EEOC for more information regarding your legal rights and protections. In addition, Datavant does not and will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay.

At the end of this application, you will find a set of voluntary demographic questions. If you choose to respond, your answers will be anonymous and will help us identify areas for improvement in our recruitment process. (We can only see aggregate responses, not individual ones. In fact, we aren’t even able to see whether you’ve responded.) Responding is entirely optional and will not affect your application or hiring process in any way.

Datavant is committed to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need an accommodation while seeking employment, please request it here, ( by selecting the ‘Interview Accommodation Request’ category. You will need your requisition ID when submitting your request, you can find instructions for locating it here ( . Requests for reasonable accommodations will be reviewed on a case-by-case basis.

For more information about how we collect and use your data, please review our Privacy Policy ( .

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Boston, MA vacancy
  • $160k - $200k

     ...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident... 
    Senior
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interface

    Somerville, MA
    4 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    5 days ago
  • $134.25k - $214.8k

     ...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed...  ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,... 
    Senior
    Work experience placement
    Work at office
    Remote work

    Axon

    Boston, MA
    1 day ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that...  ...maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering... 
    Senior
    Local area
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    5 days ago
  • $166k - $220k

     ...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &...  ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Boston, MA
    5 days ago
  • $128k - $160k

     ...step at a time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Senior Site Reliability Engineer, you’ll build and scale the critical Kubernetes infrastructure that powers our platforms and services. You’ll solve... 
    Senior
    Full time
    Immediate start

    DraftKings Inc.

    Boston, MA
    2 days ago
  • $160k - $200k

     ...Senior Site Reliability Engineer This role is located in Somerville, MA - We are a hybrid work environment and are in the office 3+ days/per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce with... 
    Senior
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interfaces

    Somerville, MA
    3 days ago
  • $140k - $210.9k

     ...position will be primarily on-site with residency commutable to...  ...DevOps backgrounds or software engineering backgrounds (e.g., Java...  ...interest in operating and improving reliability of distributed production...  ...Responsibilities As a Senior Engineer of the SRE / Production... 
    Senior
    Full time
    Temporary work
    Part time
    Work at office
    Shift work

    Federal Reserve Bank

    Boston, MA
    4 days ago
  • $121.4k - $218.6k

     ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner...  ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling... 
    Senior
    Work experience placement
    Work at office

    Akamai

    Boston, MA
    1 day ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection... 
    Senior
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Boston, MA
    4 days ago
  •  ...Information Technology group delivers secure, reliable technology solutions that enable...  ...You Will Have in This RoleAs a Senior Application Support Engineer, you will help power DTCC's global...  ...processing and settlement.Leveraging Site Reliability Engineering (SRE) principles... 
    Senior
    Remote work
    Flexible hours

    DTCC- The Depository Trust & Clearing Corporation

    Boston, MA
    5 days ago
  •  ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the...  ...candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based... 
    Senior
    Full time

    Federal Reserve Bank of Boston

    Boston, MA
    5 days ago
  • $60 - $70 per hour

     ...Job Description- 100% REMOTE! This DevOps Automation Engineer role sits within a platform operations team and focuses on supporting...  ...a global, regulated MedTech context. The position emphasizes site reliability engineering and platform operations over CI/CD-heavy... 
    Senior
    Contract work
    Temporary work
    Remote work

    Actalent

    Boston, MA
    3 days ago
  • $139k - $257.55k

     ...Individual Contributor The Challenge The Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe... 
    Senior
    Temporary work
    Local area
    Remote work
    Relocation

    Adobe

    Waltham, MA
    3 days ago
  • The Depository Trust & Clearing Corporation (DTCC) seeks a Senior Application Support Engineer to ensure reliability and performance of its critical trade processing platforms. You will apply SRE principles, drive automation, and partner with global teams to support AWS... 
    Senior

    The Depository Trust & Clearing Corporation (DTCC)

    Boston, MA
    5 days ago
  • $130k - $150k

     ...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are...  ...career mentoring and performance coaching from an assigned senior colleague. Additional leadership and collaboration opportunities... 
    Work at office
    Work from home
    3 days per week

    CRA International

    Boston, MA
    2 days ago
  • $160k - $200k

    Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware... 
    Local area
    Remote work

    QuEra Computing

    Boston, MA
    4 days ago
  •  ...mission-critical industries, helping partners move more quickly and reliably from algorithm to silicon. Our platform accelerates deployment...  .... The Roles We are looking for an experienced software engineer to help us build a new generation of transpilation tools... 
    Senior
    Full time
    Remote work
    Relocation package
    Flexible hours

    Code Metal

    Boston, MA
    1 day ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Cambridge, MA
    1 day ago
  •  ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed... 

    Watershed Informatics

    Cambridge, MA
    5 days ago
  •  ...efforts and improve strategic decision-making. As a member of our engineering team, you'll be working closely with other team members and our...  ...engineering teams can work and the tools they useLocation: on-site in BostonWe believe that it takes a diverse team to build the... 
    Senior

    Roberts Recruiting

    Boston, MA
    4 days ago
  •  ...ISEE is seeking an experienced Senior Software Engineer to join our team. The ideal candidate has several years of work experience, and have worked on complex, performance-critical code-bases. Role responsibilities include: - Support full software development life... 
    Senior
    Full time
    Work experience placement

    Isee

    Cambridge, MA
    1 day ago
  • $108k - $209k

     ...seeking an experienced, creative, and talented Principal / Senior Software Engineer. The ideal candidate will have a strong background in software...  .... Leverage AWS cloud infrastructure to build scalable, reliable, and efficient applications and AI-powered services. Uphold... 
    Senior

    Seres Therapeutics

    Cambridge, MA
    2 days ago
  • $160k - $225k

     ...Staff Site Reliability Engineer Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients. Our products speed up workflows in areas from target identification and clinical development to market access and precision medicine in the... 

    Manifold

    Cambridge, MA
    2 days ago
  • $148k - $185k

     ...the future together. The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure...  ...into clear, actionable insights that help teams and senior leaders make better decisions about reliability, risk, and... 
    Full time
    Immediate start

    DraftKings

    Boston, MA
    9 hours ago
  • $150k - $215k

     ...individuals optimize their health, fitness, and recovery. As a Senior Software Engineer on the AI team, you will play a key role in building and...  ...is ideal for an engineer who is passionate about building reliable, scalable applications and thrives in a fast-paced,... 
    Senior
    Full time
    Work at office
    Relocation

    Whoop

    Boston, MA
    1 day ago
  • $130k - $140k

     ...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient... 
    Senior
    Ongoing contract
    Full time
    Temporary work
    Work experience placement

    SS&C Technologies

    Waltham, MA
    3 days ago
  • $150k - $195k

     ...empowers members to perform at a higher level through a deeper understanding of their bodies and daily lives.WHOOP is seeking a Senior Reliability Engineer to lead the charge in ensuring our hardware products deliver a consistent, high-reliability experience for members. In... 
    Senior
    Full time
    Work at office
    Relocation

    WHOOP

    Boston, MA
    1 day ago
  • $138k - $252k

     ...scaled autonomy Solicit and incorporate feedback from end users of the APIs and implementations, and collaborate with adjacent engineering teams to help make the product vision a reality Develop and improve our APIs for commanding and controlling teams of... 
    Senior
    Full time
    Work experience placement
    Local area
    Relocation package
    Flexible hours

    Anduril Industries

    Boston, MA
    1 day ago
  • $140k - $160k

     ...About this role: Pickle is seeking a dynamic, driven Senior Software Engineer, Navigation, to enhance the speed and safety of our autonomous...  ...algorithms and capable of optimizing for performance and reliability. Detail-oriented, but with a system-level mindset.... 
    Senior
    Full time
    Work at office
    3 days per week

    Pickle Robot Company

    Charlestown, MA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!