Senior Site Reliability Engineer
$168k - $200kDatavant
Datavant is the data collaboration platform trusted for healthcare. Guided by our mission to make the world’s health data secure, accessible and actionable, we provide critical data solutions for organizations across the healthcare ecosystem - including providers, health plans, researchers, and life sciences companies. From fulfilling a single patient’s request for their medical records to powering the AI revolution in healthcare, Datavanters are building the future of how data is connected and used to improve health.
By joining Datavant today, you’re stepping onto a driven and highly collaborative team that is passionate about creating transformative change in healthcare.
What We’re Looking For
We’re looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You’ll be at the forefront of building and operating a resilient, observable, and scalable platform that enables mission-critical data and ML workloads across our organization.
This role is ideal for someone who combines a strong SRE mindset with deep cloud infrastructure and data platform experience . You're comfortable operating at scale in a complex, hybrid cloud environment and can architect systems that balance velocity, safety, and cost. You’ll work closely with Data & ML Engineers, Data Scientists, Analysts, and App Engineering teams to build a modern data platform that is secure, self-service, and production-grade.
What You Will Do
Operate and Improve Databricks and Snowflake : Own Databricks & Snowflake platforms lifecycle—including automation, workspace governance, job orchestration, and cost optimization.
Design for Reliability : Architect resilient, scalable, and secure infrastructure across cloud environments. Drive initiatives around failover, autoscaling, chaos testing, and capacity planning.
Advance Observability : Build and maintain platform-wide monitoring, alerting, and logging infrastructure using Datadog and other open tooling. Define and enforce SLOs/SLAs for critical services.
Drive CI/CD for Data & ML : Automate deployments of data pipelines, ML workflows, and infra components using GitHub Actions , Terraform, and related IaC tooling.
Enable Data Flow Across Platforms : Build patterns and tooling to support inter- and intra-cloud data movement across systems like Snowflake, S3, Delta Lake, and Kafka.
Champion Event-Driven Architectures : Leverage cloud-native tools like EventBridge , SNS/SQS, and Lambda to build loosely coupled, scalable data systems.
Collaborate Across Teams : Serve as the SRE and platform partner for teams across the organization, ensuring the platform meets the needs of analytics, data science, and product use cases.
Contribute to Strategy : Influence engineering-wide decisions on data platform architecture , ML enablement , and data product strategy .
What You Need to Succeed
6+ years in SRE, platform engineering, or DevOps roles supporting data-intensive or ML-powered applications.
AI-native working style: daily use of Claude Code, Cursor, Copilot, or equivalent, with views on how they make a team faster.
Hands-on Databricks experience , including workspace setup, cluster/job management, and integration with CI/CD and data orchestration tools. Experience with Snowflake as well.
Deep understanding of cloud-native infrastructure on AWS (or similar), including VPCs, IAM, event-driven patterns, and serverless compute.
Proven expertise with observability tools (especially Datadog) and architecting platform-wide logging and monitoring solutions.
Strong command of CI/CD tooling , especially GitHub Actions , infrastructure-as-code (Terraform), and deployment automation for data systems.
Working knowledge in shell scripting and Python.
Experience building and supporting highly available, fault-tolerant systems .
Excellent communication and collaboration skills; able to work effectively across teams.
What Helps You Stand Out
DevSecOps mindset : Familiarity with implementing security best practices in IaC, CI/CD, secret management, and audit logging.
Experience with ML infrastructure tooling such as MLflow, Feature Stores, and GPU workload orchestration.
Strong experience in both Databricks and Snowflake in a large scale production lakehouse with cross-warehouse interoperability, e.g. Iceberg v3, Glue, etc.
Background in compliance-aware architecture (e.g., HIPAA, SOC 2) or regulated industries.
Familiarity with multi-cloud or hybrid cloud data environments ; experience with Azure.
Contributions to open-source infrastructure, SRE, or observability tools.
We are committed to building a diverse team of Datavanters who are all responsible for stewarding a high-performance culture in which all Datavanters belong and thrive. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status.
At Datavant our total rewards strategy powers a high-growth, high-performance, health technology company that rewards our employees for transforming health care through creating industry-defining data logistics products and services.
The range posted is for a given job title, which can include multiple levels. Individual rates for the same job title may differ based on their level, responsibilities, skills, and experience for a specific job.
The estimated total cash compensation range for this role is:
$168,000—$200,000 USD
To ensure the safety of patients and staff, many of our clients require post-offer health screenings and proof and/or completion of various vaccinations such as the flu shot, Tdap, COVID-19, etc. Any requests to be exempted from these requirements will be reviewed by Datavant Human Resources and determined on a case-by-case basis. Depending on the state in which you will be working, exemptions may be available on the basis of disability, medical contraindications to the vaccine or any of its components, pregnancy or pregnancy-related medical conditions, and/or religion.
This job is not eligible for employment sponsorship.
Datavant is committed to a work environment free from job discrimination. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status. To learn more about our commitment, please review our EEO Commitment Statement here ( . Know Your Rights ( , explore the resources available through the EEOC for more information regarding your legal rights and protections. In addition, Datavant does not and will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay.
At the end of this application, you will find a set of voluntary demographic questions. If you choose to respond, your answers will be anonymous and will help us identify areas for improvement in our recruitment process. (We can only see aggregate responses, not individual ones. In fact, we aren’t even able to see whether you’ve responded.) Responding is entirely optional and will not affect your application or hiring process in any way.
Datavant is committed to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need an accommodation while seeking employment, please request it here, ( by selecting the ‘Interview Accommodation Request’ category. You will need your requisition ID when submitting your request, you can find instructions for locating it here ( . Requests for reasonable accommodations will be reviewed on a case-by-case basis.
For more information about how we collect and use your data, please review our Privacy Policy ( .
- ...and best in class outcomesVisionary in future focused problem-solvingExceptional in execution and impactThe RoleAs a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to deliver...SeniorFull timeFlexible hours
$160k - $200k
...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$134.25k - $214.8k
...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed... ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,...SeniorWork experience placementWork at officeRemote work$127k - $249k
Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that... ...maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering...SeniorLocal areaWorldwideFlexible hours- ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the... ...candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based...SeniorFull time
$160k - $200k
...promQL Key Responsibilities: Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coach Perform incident...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week$185.5k - $232k
...Senior Site Reliability Engineer New York, NY; Boston, MA; San Francisco, CA About Formation Bio Formation Bio is a tech and AI driven pharma company differentiated by radically more efficient drug development. Advancements in AI and drug discovery are creating...SeniorWork experience placementWork at officeLocal areaRelocation3 days per week$55k - $151.47k
...ApplicableSpecialismIFS - Internal Firm Services - OtherManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...SeniorFull timeH1b$140k - $210.9k
...position will be primarily on-site with residency commutable to... ...DevOps backgrounds or software engineering backgrounds (e.g., Java... ...interest in operating and improving reliability of distributed production... ...Responsibilities As a Senior Engineer of the SRE / Production...SeniorFull timeTemporary workPart timeWork at officeShift work$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SeniorWork experience placementWork at office$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SeniorTemporary workImmediate startFlexible hoursShift work- ...Information Technology group delivers secure, reliable technology solutions that enable... ...You Will Have in This RoleAs a Senior Application Support Engineer, you will help power DTCC's global... ...processing and settlement.Leveraging Site Reliability Engineering (SRE) principles...SeniorRemote workFlexible hours
$146.25k - $225k
...dedicated to solving complex problems and making a huge impact. We are expanding our team and recruiting for a skilled Senior/Staff Site Reliability Engineer focused on designing, building, and operating our on-prem/cloud environment.The OpportunityYou will advance the...SeniorLocal areaRemote workRelocation package$118.3k - $147.9k
...CMT is looking for a Senior Site Reliability Engineer, SecOps to help us change the world. CMT has helped protect over 65 million drivers and prevent over 126,000 crashes worldwide. We build AI to solve some of the most difficult challenges in mobility — understanding...SeniorFull timeTemporary workSummer workWork from homeWorldwideFlexible hours$134.25k - $214.8k
...real change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability...SeniorWork experience placementWork at officeRemote workFlexible hours$139k - $257.55k
...Individual Contributor The Challenge The Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe...SeniorTemporary workLocal areaRemote workRelocation$200k - $250k
...The Crown Is Yours As a Principal Site Reliability Engie r , you'll shape the long-term... ...cloud and on-premise platforms, helping engineering teams build, deploy, and operate highly... ...and operational excellence. Mentor senior engineers, influence technical strategy...Full timeImmediate start$169.3k - $304.7k
...in building and maintaining fast, efficient, scalable, and reliable routing software and infrastructure that is responsible... ...growth and stability of our global platform. As a Principal Site Reliability Engineer - Network, you will be responsible for: Architecting...Work experience placementWork at office$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...career mentoring and performance coaching from an assigned senior colleague. Additional leadership and collaboration opportunities...Work at officeWork from home3 days per week$90k - $110k
...SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocations: Boston/Waltham, MA | HybridGet To Know us:SS... ...out talented candidates for the position ofSite Reliability Engineer. This role is based out of one of our Boston-area offices (Boston...Ongoing contractFull timeCasual workWork at officeWorldwideFlexible hours- ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed...
$160k - $200k
Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware...Local areaRemote work$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- Senior DevOps Engineer - Release EngineeringLocation: Boston/Marlborough, MA — Hybrid, 3 days per weekRole... ...influence software delivery speed, reliability, developer productivity, and release... ...Collaborative workspaces Free on-site fitness centers and stocked kitchens in...SeniorFor contractorsWork at officeRemote workFlexible hours
- ...mission-critical industries, helping partners move more quickly and reliably from algorithm to silicon. Our platform accelerates deployment... .... The Roles We are looking for an experienced software engineer to help us build a new generation of transpilation tools...SeniorFull timeRemote workRelocation packageFlexible hours
$165k - $206k
...coordinate support and resolve platform issues across CPU/radio SoCs, MCU/PIC, NPU/GPU, and peripheral devices.Support hardware engineering teams with deep technical debugging and contribute to OS/platform modernization efforts.What You’ll NeedBasic Qualifications:Bachelor...SeniorFull timeTemporary workWork at officeImmediate startVisa sponsorshipWork visa$160k - $225k
...Staff Site Reliability Engineer Cambridge, MA Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients. Our products speed up workflows in areas from target identification and clinical development to market access and precision...$105.79k - $141.05k
...delivers on-demand networking at scale. As Lead SRE, you'll own the reliability of that platform — partnering with operations teams and... ..., and automation, and you'll coordinate across architecture, engineering, and systems development organizations to measurably improve...Temporary workRemote work$148k - $185k
...the future together. The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure... ...into clear, actionable insights that help teams and senior leaders make better decisions about reliability, risk, and...Full timeImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Boston, MA
- site reliability engineer Boston, MA
- site reliability engineer remote Boston, MA
- senior operations technician Boston, MA
- senior operations associate Boston, MA
- senior cloud service delivery manager Boston, MA
- senior it service manager Boston, MA
- senior chief engineer Boston, MA
- sr operations manager Boston, MA
- senior account director Boston, MA


