Senior Site Reliability Engineer
$145k - $175kGrabJobs
Full-time Description At Commence, we’re the start of a new age of data-centric transformation, elevating health outcomes and powering better, more efficient process to program and patient health. We combine quality data-driven solutions that fuel answers, technology that advances performance, and clinical expertise that builds trust to create a more efficient path to quality care. With human-centered, healthcare-relevant, and value-based solutions, we create new possibilities with data. We provide proof beyond the concept and performance beyond the scope with a focus on efficiencies that transform the lives of those we serve. With a culture driven by purpose, straightforward communication and clinical domain expertise, Commence cuts straight to better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability, scalability, and operational health of our mission-critical healthcare data platform. You will bridge the gap between engineering and operations—embedding reliability as a first-class concern from architecture through deployment. This role is built for someone who thrives when systems are under pressure and who treats an outage as a problem to be engineered away permanently, not just survived. Design, implement, and own observability infrastructure including metrics, logging, tracing, and alerting across distributed systems. Define and enforce SLOs, SLIs, and error budgets in partnership with product and engineering teams. Lead incident response: triage, coordinate remediation, conduct blameless post-mortems, and drive systemic fixes. Build and maintain CI/CD pipelines that support rapid, safe delivery of changes to production. Collaborate with engineering teams on infrastructure changes; able to read, modify, and contribute to existing infrastructure-as-code (Terraform or CloudFormation). Design and operate highly available, fault-tolerant systems—including auto-scaling, failover, and disaster recovery strategies. Reduce operational toil through automation; eliminate manual processes before they become habits. Collaborate with software engineers to establish reliability-first design patterns and review architectures for operational risk. Manage Kubernetes or container orchestration environments at scale. Ensure systems meet compliance and security requirements, particularly those applicable to healthcare data (HIPAA, SOC 2). Provide technical mentorship and guidance to engineers across the organization on reliability practices. Participate in on-call rotation with a commitment to continuously reducing the need for it. Qualifications 7+ years of experience in SRE, platform engineering, or DevOps roles. Exceptional problem-solving under pressure—demonstrated track record of diagnosing complex, high-stakes system failures and building durable solutions. Deep hands-on experience with AWS services including EC2, EKS/ECS, Lambda, RDS, S3, CloudWatch, and related tooling. Familiarity with infrastructure-as-code (Terraform or CloudFormation)—able to contribute to existing configurations. Experience designing and operating distributed systems with strict availability and latency requirements. Proficiency in at least one scripting or systems language (Python, Go, Bash, or similar) for automation and tooling. Experience with container orchestration (Kubernetes, ECS) in production environments. Expertise in observability tooling (OpenSearch, Prometheus/Grafana, or equivalent). Hands-on experience with CI/CD platforms (GitHub Actions, Jenkins, CircleCI, or similar). Proven ability to define and operationalize SLOs and error budgets. Experience with relational and NoSQL databases—performance tuning, replication, and backup strategies. Strong working knowledge of networking fundamentals: DNS, load balancing, VPCs, TLS. Excellent communication skills—able to translate technical risk into business impact for non-engineering stakeholders. Additional Requirements AWS Certifications (Solutions Architect, DevOps Engineer, or SysOps Administrator). Experience in healthcare technology or other regulated industries (HIPAA, SOC 2, FedRAMP). Familiarity with chaos engineering practices and tooling. Experience with data pipeline reliability (ETL/ELT workflows, streaming systems). Exposure to AI/ML infrastructure and the reliability challenges unique to model serving. Familiarity with additional cloud platforms (Azure, Google Cloud). Contributions to open-source reliability or infrastructure tooling. *Commence' headquarters are in Virginia Beach, VA, however we are open to remote candidates in the following states: AZ, AR, CO, DE, FL, GA, IL, IN, KS, KY, MA, MD, MI, MS, MO, MT, NC, NE, NV, NY, OH, OK, PA, SC, TN, TX, VA, DC, WI, and WV* Work Environment/Physical Demands The work environment and physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. This is a remote position. While performing the duties of this job, the employee regularly works in a climate-controlled environment. Candidates must be able to sit, read, work on a computer, and watch a computer screen for extended periods of time. Occasionally required to stand, walk, use hands and fingers, kneel or crouch. Commence is an equal employment opportunity employer. All personnel processes are merit-based and applied without discrimination on the basis of race, color, religion, sex, sexual orientation, gender identity, marital status, age, disability, national or ethnic origin, military and veteran status or any other characteristic protected by applicable law. Commence.AI is committed to providing equal employment opportunities to all applicants, including individuals with disabilities. If you require a reasonable accommodation to participate in the application process due to a disability, please contact Human Resources at View phone number on click.appcast.io or View email address on click.appcast.io. Please note that unless you are requesting an accommodation, all applications must be submitted through our online application system. Salary Description $145,000-$175,000
$158.5k - $172k
...the exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate,... ...environment. This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire...SeniorFull timeTemporary workWork at officeFlexible hours3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$130k - $180k
...best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform...SeniorWork at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$127k - $249k
...Eastern or Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the... ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background....SeniorLocal areaRemote workWorldwideFlexible hours$190.8k - $267.1k
...is a unique opportunity to leave your mark on one of the most influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your knowledge of distributed systems and architecture to improve the...SeniorWork experience placementHome officeFlexible hours$125.04k - $187.56k
...services, including Finance, Legal, Sustainability, Commercial, Digital and E-commerce, Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible for ensuring the scalability, reliability, and performance of production systems through automation...SeniorFull timeWork at officeRemote workFlexible hours$130k - $165k
...Job Title: Senior Software Engineer Company: Snapsheet Job Location: USA, Remote Job Type: Full-time, direct hire Job Department: Technology Team : Site Reliability Engineering About Snapsheet: Snapsheet exists to simplify claims. We leverage...SeniorFull timeTemporary workLocal areaRemote workVisa sponsorshipWork visaFlexible hours- Get AI-powered advice on this job and more exclusive features. Direct message the job poster from Algo Capital Group Senior Site Reliability Engineer - Observability and Automation A leading high-frequency trading firm is seeking a mid to senior-level Site Reliability...SeniorFull timeWork at officeFlexible hours
$130k - $170k
Senior Site Reliability Engineer About Us Founded in 2014, we offer the industry’s first and only cloud‑based, fully‑customisable, end‑to‑end software solution to automate securities‑based lending from origination through the life of the loan. By combining thought leadership...SeniorFull timeFlexible hoursShift work$80 - $90 per hour
...LaSalle Network is hiring for a Senior Site Reliability Engineer (Compute Platform) with a leading infrastructure and platform engineering firm known for innovation and cutting-edge technological solutions. This opportunity is a remote role focused on deep infrastructure...SeniorHourly payContract workTemporary workRemote work$80 - $90 per hour
...LaSalle Network is hiring for a Senior Site Reliability Engineer (Storage Platforms) with a storage-focused, enterprise-leading company known for innovation and impactful cloud infrastructure. Join a dedicated team managing software-defined storage solutions and enterprise...SeniorHourly payContract workTemporary workRemote work- ...Hire Overview Our client is seeking a highly skilled Edge Site Reliability Engineer (Edge SRE) to lead the design, automation, and operations... ..., leadership, and cross‑functional collaboration skills. Seniority level Not Applicable Employment type Full-time Job function...SeniorFull timeContract work
- ...consumers and companies, alikeKlover’s engineering team powers one of the fastest-growing... ...-grade systems that prioritize reliability, security, and performance, and that integrate... ...the right candidateAbout the RoleAs a Senior/Staff Site Reliability Engineer, you will play a...SeniorWork at officeImmediate startRemote work
$140k - $170k
We are looking for a Senior Site Reliability Engineer to work as part of a lean, product‑focused engineering organization. This role is about building and operating reliable cloud‑based systems by writing code, automating infrastructure and delivery workflows, and reducing...SeniorFull timeWork experience placementFlexible hours$117.63k - $176.44k
...you will be responsible for ensuring the reliability, scalability, and performance of our data systems. Working closely with data engineers and other operation sub-teams, you will manage... ...and benefits summary on our careers site for more details.EducationBachelor's DegreeWhile...SeniorFull time- ...to physicians, providing critical information about the right treatments for the right patients, at the right time.The Site Reliability Engineering team works with all departments and business units to provide dependable cloud infrastructure solutions, along with support...Full time
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will solve complex and...
$108.08k - $172.5k
Work with development and platform engineering teams to migrate and maintain applications in Google Cloud. Apply Observability concepts and applications to maintain services. Monitor metrics, system health and analyze reports. Provide on-call rotation support for production...Full timeRemote workWorldwide- Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources Willingness to work on-site at stated location in the job openingDepartment...Contract workFor contractorsWork experience placement
$130k - $225k
...expectations, integrity, innovation and a willingness to challenge consensus.The Algorithmic Trading Team is looking for a Site Reliability Engineer for our Chicago office. The SRE team is critical to the success of our trading - ensuring that our production trading...Temporary workWork at officeFlexible hours- ...IL, United StatesIndustry: Trading FirmPosted: 2026-08-17Contact: Mike LaTulipEmail: ****@*****.*** Title: Site Reliability Engineer (Infrastructure & Systems)Location: Chicago, IL (Greater Metro Area)About the OpportunityJoin a premier financial technology...Local area
$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...career mentoring and performance coaching from an assigned senior colleague. Additional leadership and collaboration opportunities...Work at officeWork from home3 days per week- Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will use technology to solve...
$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying...Full timeTemporary workWork experience placementFlexible hours$100.7k - $167.8k
Job SummaryThe Site Reliability Engineer III is a pivotal architect of stability for CME Clearing & Risk. You will engineer secure, scalable, and reliable technology solutions that safeguard the global marketplace. By bridging the gap between development and operations,...Full timeWorldwide$194k - $267k
...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to...Permanent employmentWork at officeLocal areaWorldwideFlexible hours- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology team, you will solve complex and broad business problems with simple...
$132.1k - $220.1k
Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week on-site at... ...the technical roadmap of our entire SRE domain, mentoring senior engineers and setting the global standard for operational excellence...Full timeWork at officeLocal areaWorldwide2 days per week$160k - $210k
...world. What you'll do:Join our Platform Engineering team, where you'll ensure the availability... ...and mentoring engineers across reliability initiativesAnalyze, troubleshoot, and remediate... ...need:8+ years of experience in DevOps, Site Reliability Engineering, or Platform Engineering...Work at officeWorldwideMonday to FridayFlexible hours$112.5k - $187.5k
...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering... ...Site Reliability Engineer at TransUnion, you will serve as a senior technical leader and force multiplier on the SRE team....Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Chicago, IL
- site reliability engineer remote Chicago, IL
- site reliability engineer sre Chicago, IL
- senior maintenance supervisor Chicago, IL
- senior lead project manager Chicago, IL
- senior robotics software engineer Chicago, IL
- senior firewall engineer Chicago, IL
- senior devops engineer remote Chicago, IL
- senior sas administrator Chicago, IL
- senior IT manager Chicago, IL


