Site Reliability Engineer
$100k - $130kDatavant
Datavant is the data collaboration platform trusted for healthcare. Guided by our mission to make the world’s health data secure, accessible and actionable, we provide critical data solutions for organizations across the healthcare ecosystem - including providers, health plans, researchers, and life sciences companies. From fulfilling a single patient’s request for their medical records to powering the AI revolution in healthcare, Datavanters are building the future of how data is connected and used to improve health.
By joining Datavant today, you’re stepping onto a driven and highly collaborative team that is passionate about creating transformative change in healthcare.
We are seeking a Site Reliability Engineer to play a key role in designing, optimizing, and securing the underlying cloud infrastructure that powers our organization’s cloud environments. As we continue to transform, consolidate, and evolve our cloud environments, this role will be instrumental in architecting scalable, secure, automated, and resilient infrastructure, ensuring seamless workload migration, and integrating new cloud environments into our ecosystem.
With a focus on core cloud infrastructure, including networking, identity and access management, security, storage, compute, and cross-cloud integrations, you will collaborate closely with peers in security, development, and other areas of the platform engineering team to ensure our cloud environment adheres to best practices, governance frameworks, and automation-first principles.
What You Will Do
Reliability & Technical Ownership
Lead the design and implementation of reliability improvements across assigned services, with minimal guidance
Identify systemic inefficiencies in architecture, implementation, and operational process and drive solutions
Implement customized solutions to complex operational problems derived from technical requirements
Review code, systems, and configuration with a focus on efficiency gains, optimization, and best practices and hold peers to those standards
Own SLO/SLI definitions for assigned services and drive teams toward meeting and improving those targets
Lead incident response, facilitate postmortems, and ensure action items result in durable reliability improvements
M&A Integration & Environment Consolidation
Support the integration of newly acquired cloud environments into Datavant's existing infrastructure, ensuring reliability, security, and operational consistency from day one
Contribute to the implementation of hybrid-cloud and cross-cloud connectivity strategies that ensure interoperability across a growing multi-cloud footprint
Help maintain and expand the modular network security edge, keeping it flexible enough to absorb additional environments as M&A activity requires
Drive standardization of infrastructure and operational practices across consolidated environments, reducing fragmentation and toil
Partner with security and platform engineering teams to ensure newly integrated environments adhere to governance frameworks and automation-first principles
Service Delivery & Process Improvement
With limited guidance, develop tools and processes to improve team service delivery including scaling, resiliency, efficiency, visibility, quality, and operations management
Analyze service delivery data and team feedback to drive meaningful improvements to development processes
Participate in and help evolve on-call practices, runbooks, and alerting strategy for assigned teams
Address communication gaps and produce clear documentation of process changes and technical standards
Collaboration & Mentorship
Teach and lead more junior engineers on team processes, technical implementations, and SRE best practices
Accept and promote sound engineering decisions including those that weren't your own and build alignment around them
Collaborate cross-functionally with development, security, and platform engineering teams to embed reliability thinking into the software development lifecycle
Automation & Tooling
Build and maintain Infrastructure as Code (Terraform, Ansible) to support scalable, repeatable, and secure deployments
Develop and improve CI/CD pipelines, automated testing, and deployment tooling
Implement policy-based automation (AWS SCPs, Azure Policy) to maintain governance across cloud environments
Extend observability coverage through instrumentation, dashboards, and alerting using Datadog, CloudWatch, and Azure Log Analytics
Leverage AI tools and agents to accelerate and improve daily engineering workflows
What You Need to Succeed
5+ years of experience in site reliability engineering, DevOps, or platform/infrastructure engineering
Strong expertise in cloud infrastructure, including networking, security, compute, storage, and IAM
Experience supporting workload migrations and integrating new cloud environments into existing architectures
Hands-on experience with Infrastructure as Code and automation (Terraform and Ansible)
Strong security knowledge, including IAM, encryption, network security, and compliance frameworks (SOC2, HITRUST, NIST)
Demonstrated ability to solve complex operational problems independently and drive solutions end-to-end
Strong proficiency in at least one systems language (Python, Go, or similar) and comfort across multiple languages and configuration formats
Proven ability to conduct meaningful code and system reviews not just for correctness, but for efficiency and architectural soundness
Strong communication skills, including the ability to document technical decisions, address gaps in understanding, and build team alignment
Experience leveraging AI agents to accelerate daily workload
Competency with the following technologies:
Compute: AWS EC2, Azure VMs, Kubernetes / containerized workloads
Network security: AWS (VPC, TGW, Peering, SG, ALB, NLB); Azure (VNets, NSG, AGW); VPN
Observability: Datadog, CloudWatch, Azure Log Analytics instrumentation, dashboards, alerting
Automation: Terraform, Ansible, GitHub Actions or equivalent CI/CD tooling
Cloud platforms: Hands-on experience in AWS and/or Azure; familiarity with GCP
IAM: AWS IAM, Azure RBAC, least-privilege access patterns
Core OS: Linux (required), Windows (helpful); DNS / IPAM
What Helps You Stand Out
Experience with multi-cloud environments (AWS, Azure, GCP) and multi-account cloud governance (AWS Organizations, Azure Policy, SCPs)
Background in driving environment consolidation and integration bringing order to disparate hybrid and multi-cloud architectures
Experience defining and operating against SLOs/SLIs and error budgets
Familiarity with FinOps principles and cloud cost optimization
Track record of improving on-call health reducing alert fatigue, improving runbook quality, driving down MTTR
Experience with cloud-native identity management (AWS IAM Identity Center, Entra ID, RBAC)
Knowledge of multi-region architecture and disaster recovery strategies
We are committed to building a diverse team of Datavanters who are all responsible for stewarding a high-performance culture in which all Datavanters belong and thrive. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status.
At Datavant our total rewards strategy powers a high-growth, high-performance, health technology company that rewards our employees for transforming health care through creating industry-defining data logistics products and services.
The range posted is for a given job title, which can include multiple levels. Individual rates for the same job title may differ based on their level, responsibilities, skills, and experience for a specific job.
The estimated total cash compensation range for this role is:
$100,000—$130,000 USD
To ensure the safety of patients and staff, many of our clients require post-offer health screenings and proof and/or completion of various vaccinations such as the flu shot, Tdap, COVID-19, etc. Any requests to be exempted from these requirements will be reviewed by Datavant Human Resources and determined on a case-by-case basis. Depending on the state in which you will be working, exemptions may be available on the basis of disability, medical contraindications to the vaccine or any of its components, pregnancy or pregnancy-related medical conditions, and/or religion.
This job is not eligible for employment sponsorship.
Datavant is committed to a work environment free from job discrimination. We are proud to be an Equal Employment Opportunity employer and all qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, disability, veteran status, or other legally protected status. To learn more about our commitment, please review our EEO Commitment Statement here ( . Know Your Rights ( , explore the resources available through the EEOC for more information regarding your legal rights and protections. In addition, Datavant does not and will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay.
At the end of this application, you will find a set of voluntary demographic questions. If you choose to respond, your answers will be anonymous and will help us identify areas for improvement in our recruitment process. (We can only see aggregate responses, not individual ones. In fact, we aren’t even able to see whether you’ve responded.) Responding is entirely optional and will not affect your application or hiring process in any way.
Datavant is committed to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need an accommodation while seeking employment, please request it here, ( by selecting the ‘Interview Accommodation Request’ category. You will need your requisition ID when submitting your request, you can find instructions for locating it here ( . Requests for reasonable accommodations will be reviewed on a case-by-case basis.
For more information about how we collect and use your data, please review our Privacy Policy ( .
$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SuggestedTemporary workImmediate startFlexible hoursShift work$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...SuggestedWork experience placementWork at office$105.79k - $141.05k
...shape the future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...SuggestedTemporary workRemote work$84.9k - $209.5k
...Job Description As a Principal Site Reliability Engineer (IC4), you will be responsible for designing, building, and operating highly available, scalable, secure, and resilient cloud services. You will combine software engineering with infrastructure expertise to improve...SuggestedTemporary workFlexible hours$51.9 per hour
...OVERVIEW: This job is responsible for the reliability, availability, and performance of... ...operational efficiency. This role blends software engineering, clinical engineering, and security... .... Works cross-functionally with AHN site leaders and teams to navigate and to monitor...SuggestedFor contractorsLocal area$132.23k - $176.31k
...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...Full timeTemporary workRemote work$94.9k - $135.6k
...development, testing, operations, and platform teams to deliver value safely and efficiently. Cardinal Health is seeking a Release Engineer to lead iteration and release management activities supporting mission critical warehouse transformation initiatives on Program...Temporary workLocal areaImmediate startFlexible hours$103.71k - $138.28k
...demonstrated knowledge and experience in system architecture and engineering disciplines. Specific technical knowledge of enterprise level... ...Amazon Web Services. •Supports due diligence activities including site surveys, design, design review, bill of materials creation,...Temporary workRemote work$125k - $191.7k
...Job Description Hybrid: This role is categorized as hybrid/Remote Role: As a Senior Software Systems Engineer on the Software Validation team within the AV organization, you will play a critical role in leading the strategy and execution of validation efforts...Local areaRemote workWork from homeFlexible hours$102.3k - $209.5k
...software, firmware, and hardware layers, working in DevOps and incident response paradigms to maintain reliability and availability. Work closely with platform engineering, operations, firmware development, silicon/board vendors, and data center teams to support and...Temporary workImmediate startFlexible hoursShift work$92.5k - $209.5k
...services in a distributed, multi-tenant cloud environment. OCI gives engineers the opportunity to work on systems that directly power... ...supply chain, release velocity, security posture, and operational reliability. The team works on distributed build services, build...Temporary workFixed term contractFlexible hours$94.1k - $150k
...The Platform Engineer (Ops Technology Lead) is responsible for designing, implementing, and maintaining IT infrastructure platforms within the CASTLE-NET program, ensuring reliability, scalability, and security. This role supports application deployment and management...Contract workWork at office$92.5k - $209.5k
...Job Description We’re hiring a Platform Software Engineer to help build the execution layer behind OCI’s AI and GPU growth motion.... ...Object Storage, Data Flow) to move training and inference data reliably across customer and internal pipelines. • Build and extend GPU...Temporary workFlexible hours$80k
...Specific Essential Duties and Responsibilities: Provide Tier‑3 engineering support for Microsoft 365 GCC, Exchange Online, hybrid Exchange... .... Support SharePoint Online platform operations, including site collections, permissions, integrations, and collaboration...Contract work$75k - $110k
...Home (USA). Job Summary We are looking for a Customer Solutions Engineer to provide business and technical support to our Customers in cash... ...of work (SOW) detailing customer requirements. Perform on‑site presentations, product demonstrations, and observations of product...Work at officeRemote workWork from home- ...Epic is seeking a Technical Solutions Engineer to work on impactful software affecting millions worldwide. This role involves diagnosing issues, developing solutions, and managing implementations across multiple locations. The ideal candidate will hold a Bachelor's degree...WorldwideRelocationRelocation package
$100k
...Maximus is currently seeking a Cloud Platform Engineer. This is a remote position. Maximus is a trusted federal partner supporting... ...VoIP, VTC, and real-time communications systems, ensuring reliability, performance, and operational continuity. Job-Specific Minimum...Contract workRemote work$83k - $166.1k
...maintaining backend systems, and ensuring the long-term scalability, reliability, and sustainability of the platform. The ideal candidate... ...Bachelor's degree in Computer Science, Information Systems, Engineering, Healthcare Informatics, or a related field. ~ Master's...Temporary workWork experience placementImmediate startFlexible hours- ...more technical specializations. As a technical leader, they mentor others and share their knowledge within the Red River Solution Engineering and Sales teams. They provide architectural guidance for customers across the Red River product portfolio, including products and...Work experience placementWork at office
$105k - $141.75k
...re-platforming. Experience with COBOL modernization and porting application code across platforms. Familiarity with agile engineering practices like Test Automation, Test-Driven Development (TDD), Continuous Integration (CI), Continuous Delivery (CD), DevOps, and...Remote workWorldwide$105k - $141.75k
...conversion, and re-platforming. Experience with COBOL modernization and porting application code across platforms. Familiarity with agile engineering practices like Test Automation, Test-Driven Development (TDD), Continuous Integration (CI), Continuous Delivery (CD), DevOps, and...$121.4k - $218.6k
...teams to solve complex challenges? Join Our Custom Government Engineering Team! The Custom Government team operates across the full... ...Sector customers. This role partners closely with Operations and Site Reliability Engineering (SRE) teams in a highly collaborative...Work experience placementWork at office$112.3k - $140k
...looking for a stimulating and challenging career where your engineering expertise can be leveraged to create, enhance, and maintain our... ...long-term partnership, uncompromising support, and proven reliability. Major product lines include fuel dispensers, tank gauges and...Work at officeLocal areaRemote workWorldwide$127k - $183k
...Our Mission As the world’s number 1 job site*, our mission is to help people get jobs.... ...value for employers. As a Senior Software Engineer, you will design and build software that... ...-functional business partners to deliver reliable, high-quality solutions. In this role, you...Work experience placementLocal areaImmediate startRemote work$84.63k - $112.84k
Lumen is the trusted network for the AI-powered world, connecting people, data, and applications through our expansive fiber network and connected ecosystem. We enable secure, high-performance connectivity across cloud, edge, and AI workloads for enterprises, governments...Temporary workRemote workWork from home$135.2k - $306.4k
...largest technical and business challenges. Oracle Kubernetes Engine (OKE) is OCI's managed Kubernetes service. OKE enables... ...including cluster lifecycle management, orchestration, scalability, reliability, performance, automation, observability, security, and integration...Temporary workRemote workFlexible hours$118k - $178k
...Mission As the world's number 1 job site*, our mission is to help people get jobs... ...March 2025) Day to Day As a Software Engineer III on the AI Gateway & Guardrails team... ...architectural decisions, drive service reliability through SLOs and operational readiness,...Work experience placementLocal area$114.6k - $234.6k
...Description Join OCI’s Edge Security team as a Principal Software Engineer focused on building and scaling Oracle Cloud Infrastructure’s... ..., and platform teams to deliver secure, performant, and reliable services while helping define the long-term technical vision for...Temporary workFlexible hours- ...individual development ticket and regression testing with clear documentation of test results. Correspond with Software and Template Engineering, give feedback to any questions they may have, and facilitate the provision of more information if needed. Participate in...For contractorsWork experience placementWork at officeLocal areaRemote workFlexible hours
$170k - $210k
...learn more, visit franklincovey.com. Title: Senior Software Engineer Payroll Title: Sr. Software Engineer Division &... ...while maintaining our high standards for quality, security, and reliability. You’ll also be a hands-on coach and mentor, helping junior engineers...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- construction site safety Helena, MT
- on-site clinical research associate (traveling/remote) Helena, MT
- site safety Helena, MT
- historic site Helena, MT
- IT site lead Helena, MT
- site leader Helena, MT
- junior website developer Helena, MT
- official site Helena, MT
- site services specialist Helena, MT
- site reliability engineering manager


