Site Reliability Engineer
$205k - $212.5kConsensus Cloud Solutions
Consensus Cloud Solutions is a publicly traded, leading digital cloud fax and interoperability solutions organization in the United States and globally, focusing on connecting and empowering healthcare providers, payers, care teams, and technology innovators to unify multiple systems that wouldn't otherwise talk to each other. Consensus is a trailblazer in our industry and believes that data transformation will reshape the world of healthcare.
Founded over 25 years ago, Consensus leverages its technology heritage to move from simple digital documents to advanced healthcare standards (HL7/FHIR) for secure data transport, as well as Natural Language Processing (NLP) and Artificial Intelligence (AI) to convert unstructured to structured, analytics-ready data, helping users unveil information that is meaningful and actionable for better patient care.Consensus leads the industry in data exchange solutions and we're only getting started! With exciting new initiatives on the horizon, we are continuing our strategic expansion and we are looking to add to our diverse team of innovators.
Now is the ideal time to join us in our mission to solve healthcare's biggest challenges, and work collaboratively with a diverse team of like-minded self-starters and partners to accomplish it.
Consensus Cloud Solutions is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive and equitable environment for all employees. We offer many remote and hybrid career opportunities. How you will impact the organization... Reporting to the Director, Infrastructure Operations, the SRE I (Site Reliability Engineer I) with a strong DevOps focus is a key member of a team responsible for supporting the tooling, pipelines, frameworks, and other technologies that underpin the many platforms deployed within the company's infrastructure. This role blends software engineering principles with deep DevOps expertise to automate and streamline the entire software delivery lifecycle. Additionally, the SRE I role will be an expert in Infrastructure as Code (IaC), AWS Cloud infrastructure design and best practices, and CI/CD platforms and processes, driving the automation and optimization of our operations. As SRE I, they will partner closely with Engineering and Information Security peers on developing infrastructure solutions that follow established best practices and design patterns. They will also contribute to the continued development of RFCs, standards, and frameworks for IaC, automation, and supporting tools. Responsibilities also include the development of internal tooling, modules, and libraries used by technology teams to both implement new projects and maintain and enhance existing platforms with a focus on automation, resiliency, availability, scalability, and performance that meet business needs within appropriate cost constraints, primarily leveraging open-source technologies and frameworks. This position will champion a DevOps culture and practices, providing expert full-stack support to software engineering teams (Java, Python, Node, Go, etc.) by integrating DevOps methodologies into their development and deployment workflows. Strong documentation skills and the ability to mentor other team members in DevOps, IaC, CI/CD, and operational best practices are essential. The value you will deliver...
- Lead the design, development, and maintenance of secure, scalable, resilient, and cost-effective cloud infrastructure solutions on AWS through a DevOps approach, leveraging the existing IaC framework based on Python, Terraform, and Terragrunt managing AWS resources; championing IaC best practices while ensuring adherence to best practices for security, reliability, performance, cost optimization, and operational excellence.
- Design, implement, manage, and optimize robust CI/CD pipelines using tools like GitHub Actions and AWS CodePipeline for both infrastructure and applications; maintain deep expertise in GitHub.
- Design, develop, and implement new tooling, applications, and platforms to improve and upgrade the capabilities of the IaC and automation platforms, and support infrastructure.
- Provide expert DevOps-focused full-stack guidance and support to software engineering teams (using common languages such as Java, Python, Node, Go, etc.) to integrate DevOps practices, automate builds/deployments, identify/resolve reliability/performance bottlenecks, and establish comprehensive documentation.
- Champion and implement DevOps best practices across teams, fostering a culture of collaboration, automation, and continuous improvement, as well as providing mentorship and leadership in DevOps methodologies, IaC, CI/CD, and cloud technologies.
- Participate in grooming and prioritizing development efforts in extending and supporting the IaC, tooling, and infrastructure support application platforms.
- Partner with other teams across the technology group to propose and draft RFCs and standards for development of best practices and design patterns for applications and platforms using IaC and the established deployment pipelines and tooling.
- Research, propose, and implement solutions to improve and upgrade cloud-based resources, infrastructure, and systems, ensuring they are performant, efficient, and resilient.
- Initiate efforts to review and ensure existing platforms are performing resiliently, efficiently, and are cost effective.
- Monitor ticket queues, provide timely and accurate updates, and resolve feature request and development tickets.
- Monitor and respond to requests and questions in Slack channels, providing guidance to and assisting troubleshooting for developers and team members.
- Create tickets and participate in deployments following Change Management procedures.
- Participate in a 24/7 on-call rotation to respond to and resolve production incidents; lead and contribute to blameless postmortems.
- Define and manage Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
- Evaluate, implement, and support Open Source frameworks and projects (e.g., ECS/Docker, Prometheus, Grafana, ELK stack, Kafka).
- Update and maintain documentation for troubleshooting, and Methods of Procedure for deployments.
- Ensure systems and users follow security standards and follow established policy, and support audit processes.
- Light travel to summit meetings or conferences may be required.
- Perform other duties and responsibilities as required, assigned, or requested. Consensus reserves the right to add or change duties at any time.
- 6+ years hands-on experience managing and automating UNIX/Linux system environments within a DevOps context.
- 5+ years of experience in a DevOps Engineer or SRE role with a strong DevOps focus, emphasizing infrastructure automation, CI/CD pipeline development, and cloud services.
- 4+ years experience designing and implementing infrastructure as code within the AWS ecosphere using Terraform.
- Mastery in DevOps discipline and processes, including building and managing CI/CD pipelines supporting Infrastructure as Code frameworks such as Terraform, and Continuous Delivery Tools such as AWS CodePipeline, GitHub Actions, Jenkins, Git, Artifactory, etc.
- Expert-level proficiency with Terraform and Terragrunt for managing AWS infrastructure as code.
- Deep expertise in GitHub, GitHub Actions for CI/CD, including design, troubleshooting, and support.
- Strong experience with AWS Cloud services (e.g., EC2, S3, RDS, VPC, IAM, Lambda, EKS/ECS, CloudWatch) and infrastructure design best practices, applied within a DevOps model.
- Mastery of observability, monitoring, metrics and alerting at scale across regionally and globally resilient and distributed platforms leveraging common open source frameworks such as Prometheus, Thanos, OpenTelemetry, Grafana, etc.
- Experience providing DevOps-centric support for applications developed in Java, Python, and Angular, including build automation, deployment pipelines, and observability.
- Expert level proficiency in at least one scripting language (e.g., Python, Bash, Perl) and one programming language (e.g., Java, Go, Node). (Code samples and/or GitHub links to prior work desirable).
- Mastery of Containerization (Docker), and strong familiarity with the container ecosystem, especially Amazon ECS.
- Mastery in config automation tool sets such as AWS Config and/or SSM, Puppet, Ansible, Chef, etc. - Includes solid knowledge of concepts and practices surrounding such solutions.
- Hands on experience with APM tools such as Zipkin, Jaeger, OpenTelemetry, NewRelic, etc.
- Proficient with Jira, Confluence, and git toolset.
- Hands-on experience with Agile/Scrum & Waterfall process environments.
- Experience implementing and supporting a variety of Open Source frameworks and projects relevant to DevOps and SRE.
- Consistently exhibits a personal accountability to outcomes to all team members, peers, and stakeholders.
- Able to prioritize and manage multiple projects simultaneously in order to meet deadlines.
- Self-starter able to work independently with minimal supervision, and high organization and communication skills to ensure alignment with team and project goals.
- Driven to learn and stay abreast of the latest technologies and DevOps best practices.
- Strong analytical and problem-solving skills with a proactive, blameless, and detail-oriented approach.
- Excellent communication and collaboration skills, essential for fostering a DevOps culture and working effectively across teams; ability to mentor others.
- Experience with PCI, HiTrust, FedRamp/GovCloud and/or similar certification methodologies.
- Experience with migrating and educating teams to newer SDLC and DevOps concepts.
- Experience with APM/Observability and advanced DevOps/SRE concepts and methodologies.
- Proven experience mentoring team members in DevOps practices.
- Location requirements: Fully remote within the U.S. (Los Angeles or Las Vegas preferred.)
- Travel requirements: Up to 10% travel.
- Physical requirements: Must be able to sit for long periods, as well as, handle long periods of screen time.
- Technology requirements: Reliable, high speed internet.
- Eligible for sponsorship: No
- This position is contingent upon satisfactory completion of a more extensive background check through the Federal Government as a Public Trust Position. Requirements include (subject to change), but are not limited to, active U.S. Citizenship or green card holder residing in the U.S. for a minimum of 3 years and working location in the U.S.
We are not accepting agency submissions for this role. To learn more about us visit consensus.com
$76k - $127k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex...SuggestedFull timePart timeWorldwideFlexible hoursEarly shift$96k - $163k
...services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview The BizOps team is looking for a Senior Site Reliability Engineer who can help us solve problems and...SuggestedFull timePart timeWorldwideFlexible hoursShift work$76k - $127k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...SuggestedFull timePart timeWorldwideFlexible hours$96k - $163k
...services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...SuggestedFull timePart timeWorldwideFlexible hours- ...Mass General Brigham in Massachusetts is seeking an Epic Data Courier Administrator, Systems Engineer to manage data migrations and change control across Epic environments. The role sits at the center of deployment operations and collaborates with Epic analysts and application...Suggested
$184k - $264.5k
...diverse and supportive environment, where We are seeking a Site Reliability Operations Technical Lead to serve as the senior technical... ...platform teams, and provides technical leadership to site support engineers. Be responsible for the hardest issues across Active...Permanent employmentWork at officeLocal areaRelocation- ...Lambda Inc. in San Francisco is seeking a Storage Engineer to own the reliability, performance, and capacity of our production storage fleet across multiple data centers, using a software-defined data plane. You will build monitoring, dashboards, and alerting for storage...
$145k - $175k
...straightforward communication and clinical domain expertise, Commence cuts straight to better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability, scalability, and operational health of our mission-critical healthcare data...Full timeRemote work- ...The Site Reliability Engineer (SRE) / Subject Matter Expert (SME) – Computer Systems Engineer/Architect will provide senior-level reach-back expertise to support the reliability, scalability, performance, and operational resilience of the GEOMAP platform in secure cloud...Full timeContract workFor contractorsFor subcontractorRemote work
$145k - $193k
...entertainment, we want to talk to you. About the Role & Team The SRE team at PENN Entertainment is looking for a Senior Site Reliability Engineer to help build and operate the infrastructure behind a large-scale sports betting and media platform. You'll own critical...Remote work- ...organization across multiple locations in the US, South America, and India. Location: Remote (US-Based Candidates Only) Site Reliability Engineer II (SRE) Position Overview We are seeking a Site Reliability Engineer II (SRE) to join our growing Site...Remote workFlexible hoursShift workWeekday work
- ...About The Role: We're looking for a Senior Site Reliability Engineer to help us mature and scale the infrastructure behind our multi-cloud SaaS platform. Most of our footprint runs on Microsoft Azure, built from the ground up around cloud architecture principles:...Remote workFlexible hours
- ...SitusAMC in Annapolis, MD, is seeking a seasoned Cloud Reliability Engineer to lead AWS-based deployments and SaaS reliability initiatives. You will optimize CI/CD pipelines, implement IaC, and drive observability across complex microservices. Join a collaborative...Remote work
$194k - $237k
## Principal Site Reliability EngineerApplylocations: Scottsdaletime type: Full timeposted on: Posted 5 Days Agojob requisition id: REQ2026... ...sponsorship.**Overall Purpose**The Principal Site Reliability Engineer partners with development teams by designing availability and...Hourly payWork at officeImmediate startVisa sponsorshipWork visaFlexible hours- ...troubleshooting, providing rubric-based written feedback. This role requires hands-on Kubernetes expertise in EKS/GKE/AKS or self-managed clusters, with strong scripting in Go, Python, or TypeScript, and ability to document findings clearly for engineering #J-18808-Ljbffr
$350k
...with leading AI companies and infrastructure providers to build reliable, high-performance platforms supporting next-generation AI workloads. This opportunity is for a Staff Site Reliability Engineer to lead the reliability of large-scale GPU infrastructure, covering...$120k - $175k
...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible....Full timeRemote workWork visaFlexible hours$200k - $240k
...systems across all product teams. You will collaborate closely with engineering leadership, product managers, and cross-functional teams to... ...and Helm ~ Understand the importance of performant and reliable systems ~ Education - Ideally looking for a B.A. / B.S. degree...Work at officeImmediate start3 days per week- ...Site Reliability Engineer Company: Quzara Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: Azure, Terraform, Bicep, Ansible, Azure Monitor, Azure Automation, Azure Policy, Azure Site Recovery, TLS/SSL Requirements: 4+ years in SRE...Full timeRemote work
$160k - $180k
...big impact. See Arkestro in action at arkestro.com. About the Role Arkestro is hiring for a Senior SRE Engineer to manage our performance and reliability for our software platform and infrastructure. The right candidate will own and develop our infrastructural...Local areaRemote work- ...JPMorganChase in Seattle seeks a Lead Software Engineer to join the Enterprise Technology, Infrastructure Platforms team. You will act as a core technical contributor, delivering trusted, scalable technology across multiple domains, while guiding AI-assisted engineering...
$165k - $280k
...actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. SR. SITE RELIABILITY ENGINEER (STARLINK) At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s...Permanent employmentTemporary workWorldwideWeekend work- ...Nscale is seeking a Senior Site Reliability Engineer to own the reliability bar for AI infrastructure operations. You will tackle the hardest reliability challenges, mentor peers, and shape SRE practices across the platform. You will work closely with teams running...
- ...Site Reliability Engineer, Data Platform - USDS Responsibilities Engage in and improve the whole lifecycle of service, from inception and design, through to deployment, operation and refinement. Ensure reliable, fault-tolerant, efficiently scalable and cost-effective data...
- ...customers globally. As part of the ongoing investment in the reliability and modernization of these systems, client is making a multi-... ...delivery of change without impact to reliability. Integration Engineering, ensuring that client is consuming the right infrastructure...Work experience placement3 days per week
- ...Site Reliability Engineer Company: Crunchafi Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Azure, AKS, Azure Kubernetes Service, Terraform, Bicep, ARM templates, GitHub Actions, Azure DevOps, Kubernetes, Docker, App Insights...Full timeRemote work
- ...About the Role: We are looking for a Senior Site Reliability Engineer (SRE) to help modernize large-scale infrastructure and improve the reliability, scalability, and operational excellence of critical production systems. In this role, you will lead OS modernization...Remote work
$191k - $226k
...incentives to steer members to the care that helps them get healthier, faster. About the role We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of the cloud infrastructure powering Garner's products and AI/ML...Remote workWork visaFlexible hours$136.6k - $184.8k
...You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers... ...you to own them to completion. As a Senior Infrastructure Reliability Engineer you will be proactively driving the reliability risk...Work experience placementFlexible hours- ...grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise. The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers...Work experience placementRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre United States
- site reliability engineering manager United States
- site reliability engineer United States
- site reliability engineer remote United States
- site recruiter United States
- site services specialist United States
- junior website developer United States
- official site United States
- on site coordinator United States
- site leader United States

