Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$145k - $175k

GrabJobs

Full-time Description At Commence, we’re the start of a new age of data-centric transformation, elevating health outcomes and powering better, more efficient process to program and patient health. We combine quality data-driven solutions that fuel answers, technology that advances performance, and clinical expertise that builds trust to create a more efficient path to quality care. With human-centered, healthcare-relevant, and value-based solutions, we create new possibilities with data. We provide proof beyond the concept and performance beyond the scope with a focus on efficiencies that transform the lives of those we serve. With a culture driven by purpose, straightforward communication and clinical domain expertise, Commence cuts straight to better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability, scalability, and operational health of our mission-critical healthcare data platform. You will bridge the gap between engineering and operations—embedding reliability as a first-class concern from architecture through deployment. This role is built for someone who thrives when systems are under pressure and who treats an outage as a problem to be engineered away permanently, not just survived. Design, implement, and own observability infrastructure including metrics, logging, tracing, and alerting across distributed systems. Define and enforce SLOs, SLIs, and error budgets in partnership with product and engineering teams. Lead incident response: triage, coordinate remediation, conduct blameless post-mortems, and drive systemic fixes. Build and maintain CI/CD pipelines that support rapid, safe delivery of changes to production. Collaborate with engineering teams on infrastructure changes; able to read, modify, and contribute to existing infrastructure-as-code (Terraform or CloudFormation). Design and operate highly available, fault-tolerant systems—including auto-scaling, failover, and disaster recovery strategies. Reduce operational toil through automation; eliminate manual processes before they become habits. Collaborate with software engineers to establish reliability-first design patterns and review architectures for operational risk. Manage Kubernetes or container orchestration environments at scale. Ensure systems meet compliance and security requirements, particularly those applicable to healthcare data (HIPAA, SOC 2). Provide technical mentorship and guidance to engineers across the organization on reliability practices. Participate in on-call rotation with a commitment to continuously reducing the need for it. Qualifications 7+ years of experience in SRE, platform engineering, or DevOps roles. Exceptional problem-solving under pressure—demonstrated track record of diagnosing complex, high-stakes system failures and building durable solutions. Deep hands-on experience with AWS services including EC2, EKS/ECS, Lambda, RDS, S3, CloudWatch, and related tooling. Familiarity with infrastructure-as-code (Terraform or CloudFormation)—able to contribute to existing configurations. Experience designing and operating distributed systems with strict availability and latency requirements. Proficiency in at least one scripting or systems language (Python, Go, Bash, or similar) for automation and tooling. Experience with container orchestration (Kubernetes, ECS) in production environments. Expertise in observability tooling (OpenSearch, Prometheus/Grafana, or equivalent). Hands-on experience with CI/CD platforms (GitHub Actions, Jenkins, CircleCI, or similar). Proven ability to define and operationalize SLOs and error budgets. Experience with relational and NoSQL databases—performance tuning, replication, and backup strategies. Strong working knowledge of networking fundamentals: DNS, load balancing, VPCs, TLS. Excellent communication skills—able to translate technical risk into business impact for non-engineering stakeholders. Additional Requirements AWS Certifications (Solutions Architect, DevOps Engineer, or SysOps Administrator). Experience in healthcare technology or other regulated industries (HIPAA, SOC 2, FedRAMP). Familiarity with chaos engineering practices and tooling. Experience with data pipeline reliability (ETL/ELT workflows, streaming systems). Exposure to AI/ML infrastructure and the reliability challenges unique to model serving. Familiarity with additional cloud platforms (Azure, Google Cloud). Contributions to open-source reliability or infrastructure tooling. *Commence' headquarters are in Virginia Beach, VA, however we are open to remote candidates in the following states: AZ, AR, CO, DE, FL, GA, IL, IN, KS, KY, MA, MD, MI, MS, MO, MT, NC, NE, NV, NY, OH, OK, PA, SC, TN, TX, VA, DC, WI, and WV* Work Environment/Physical Demands The work environment and physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. This is a remote position. While performing the duties of this job, the employee regularly works in a climate-controlled environment. Candidates must be able to sit, read, work on a computer, and watch a computer screen for extended periods of time. Occasionally required to stand, walk, use hands and fingers, kneel or crouch. Commence is an equal employment opportunity employer. All personnel processes are merit-based and applied without discrimination on the basis of race, color, religion, sex, sexual orientation, gender identity, marital status, age, disability, national or ethnic origin, military and veteran status or any other characteristic protected by applicable law. Commence.AI is committed to providing equal employment opportunities to all applicants, including individuals with disabilities. If you require a reasonable accommodation to participate in the application process due to a disability, please contact Human Resources at View phone number on click.appcast.io or View email address on click.appcast.io. Please note that unless you are requesting an accommodation, all applications must be submitted through our online application system. Salary Description $145,000-$175,000

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Pittsburgh, PA vacancy
  • $179.2k - $268.8k

     ...sensors and compute systems, test operations, systems and safety engineering - all dedicated to redefining the relationship between...  ...in Dearborn, Mich., and Palo Alto, Calif.Meet the team:As a Site Reliability Engineer on the team, you will be responsible for helping to... 
    Suggested
    Permanent employment
    Full time
    Work at office
    Immediate start
    Visa sponsorship

    Latitude AI

    Pittsburgh, PA
    2 days ago
  •  ...: Lovelace is the only provider of enterprise-scale context engines capable of analyzing trillions of real-time data points to create...  ...: ~ Lovelace AI is seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an SRE at... 
    Suggested
    Full time

    Lovelace Ai

    Pittsburgh, PA
    16 hours ago
  •  ...Job Description Job Description Senior Site Reliability Engineer (SRE) Location: Pittsburgh, PA / Cleveland, OH / Dallas, TX FTE Position Overview We are seeking an experienced Senior Site Reliability Engineer (SRE) to support production operations, application... 
    Suggested
    Full time
    Shift work
    Weekend work

    System One

    Pittsburgh, PA
    9 days ago
  •  ...Job Description Job Description Job Title: Senior Site Reliability Engineer Job Category: Infrastructure/Cloud Job Type: Permanent Full Time Location: Pittsburgh, Pennsylvania, United States Position Description This role will require someone onsite... 
    Suggested
    Permanent employment
    Full time
    Work at office
    Flexible hours
    Shift work
    Weekend work

    System One

    Pittsburgh, PA
    6 days ago
  • $147k - $210k

     ...product or system development code.Review code developed by other engineers and provide feedback to ensure best practices (e.g., style...  ..., and troubleshooting large-scale distributed systems. Site Reliability Engineering (SRE) is what you get when you treat operations... 
    Suggested

    Google

    Pittsburgh, PA
    3 days ago
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 

    Google

    Pittsburgh, PA
    3 days ago
  •  ...Site Reliability - Software Engineer (II)Duration: 6 months contract Location: Pittsburgh, PA (Hybrid: Tues-Thurs in the office. Mon & Friday remote)As a Site Reliability Engineer (SRE-SWE), you deliver medium sized projects from start to finish with minimal supervision... 
    Contract work
    Work at office
    Immediate start
    Remote work

    Artech

    Pittsburgh, PA
    4 days ago
  • $150k - $200k

     ...our CEO's funding announcement: . The Reliability team owns the availability, performance,...  ...enforcing reliability standards across engineering Designing incident response processes and...  ...strong ownership of production systems. As a Site Reliability Engineer on the Reliability... 
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    GrabJobs

    Pittsburgh, PA
    10 hours ago
  • $70.8k - $156.7k

    Senior Site Reliability Engineer - Local to Pittsburgh, Cleveland, or Dallas Position Description This role will require someone onsite at our client office in Cleveland, OH, Pittsburgh, PA, or Dallas, TX. Love technology? We do too. CGI is looking for a Site... 
    Work at office
    Local area
    Flexible hours
    Shift work
    Weekend work
    Pittsburgh, PA
    25 days ago
  •  ...Job Description Job Description Site Reliability Maintenance Engineer This is an exciting new position meant to be a key player in our newly created Reliability Program. The role responsible for identifying and managing reliability improvements to steel producing... 
    Full time

    Universal Stainless

    Bridgeville, PA
    a month ago
  • $151k - $297k

     ...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB's cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Pittsburgh, PA
    3 days ago
  • $93.61k - $119.72k

     ...bricas funcionen mejor. This role will support several ATS site locations within the Midwest and East. The role will be...  .... Partners with internal/external customer for engineered solutions to improve reliability and throughput. Identifies opportunities for Capital... 
    Full time
    Work at office
    Remote work
    Home office

    Advanced Technology Services

    Pittsburgh, PA
    1 day ago
  •  ...ideal opportunities. We are currently seeking a qualified Software Engineer / Autonomous Systems to join our client’s organization and...  ...collaborating with an interdisciplinary team focusing on developing reliable robotic and automation systems for a wide range of real-world... 

    OpenArc

    Pittsburgh, PA
    3 days ago
  • $100.22k - $111.18k

    Basic Qualifications Requires a Bachelor’s degree in Systems Engineering, or a related Science, Engineering, Technology or Mathematics...  ...benefitsWorkplace Options:This position is located in Pittsburgh PA. On site work is required and Hybrid/Flex work schedule is permitted.... 
    For subcontractor
    Second job
    Work at office
    Flexible hours

    General Dynamics Mission Systems

    Pittsburgh, PA
    3 days ago
  • $179.2k - $268.8k

     ...across machine learning and robotics, cloud platforms, mapping, sensors and compute systems, test operations, systems and safety engineering - all dedicated to redefining the relationship between people and their vehicles for millions of customers.As a Ford Motor Company... 
    Permanent employment
    Full time
    Work at office
    Immediate start
    Remote work
    Visa sponsorship

    Latitude AI

    Pittsburgh, PA
    4 days ago
  • POSITION TITLE: Sr Software Engineer - Transportation SystemsREPORTS TO: Sr Director - EngineeringLOCATION: American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making the best jeans-we embrace self expression, culture, optimism and connection... 
    Full time
    Part time
    Casual work
    Local area

    American Eagle Outfitters

    Pittsburgh, PA
    3 days ago
  • $179.2k - $268.8k

     ...across machine learning and robotics, cloud platforms, mapping, sensors and compute systems, test operations, systems and safety engineering - all dedicated to redefining the relationship between people and their vehicles for millions of customers.As a Ford Motor Company... 
    Permanent employment
    Full time
    Work at office
    Immediate start
    Remote work
    Visa sponsorship

    Latitude AI

    Pittsburgh, PA
    1 day ago
  •  ...system solutions designed for extreme scale, performance, and reliability across modern compute environments. Our platforms integrate cutting...  ...in the world.Role OverviewVDURA is seeking a Senior System Engineer to lead the specification, selection, and qualification of... 
    Remote work

    Panasas

    Pittsburgh, PA
    3 days ago
  •  ...UPMC Health Plan is seeking an Associate Agile Release Train Engineer (RTE) to help us continue our Business Agility journey and goals. They will lead an Agile Release Train (ART)s team of teams through the preparation, facilitation and delivery of our Program Increments... 
    Full time
    Work at office
    Work from home
    Monday to Friday

    UPMC University of Pittsburgh Medical Center

    Pittsburgh, PA
    1 day ago
  • $104.8k - $149.3k

     ...consistently providing compelling solutions and innovative technologies that improve reliability, sustainability, and performance.How will you make a difference? The Lead Systems Engineer, MCA - Will partner with a team of engineers focused on delivering software... 
    Work experience placement
    Worldwide

    Wabtec

    Pittsburgh, PA
    3 days ago
  •  ...Job Title: Senior Site Reliability Engineer Job Category: Infrastructure/Cloud Job Type: Permanent Full Time Location: Pittsburgh, Pennsylvania, United States Position Description This role will require someone onsite at our client office in Cleveland... 
    Permanent employment
    Full time
    Contract work
    Work at office
    Local area
    Flexible hours
    Shift work
    Weekend work

    System One

    Pittsburgh, PA
    8 days ago
  •  ...trains, metros, trams, maintenance services and integrated systems.More Than a Job: A MissionWe seek a passionate Safety and Reliability Engineer in Pittsburgh to drive the design and development of robust, future-proof solutions. Join dedicated teams committed to... 
    Worldwide

    ALSTOM

    Pittsburgh, PA
    2 days ago
  • $147.93k - $291.61k

     ...motion control and VD and low level actuator control. - Strong physics and control systems knowledge. - Knowledge of Systems Engineering and Verification and Validation (V&V) best practices. - Knowledge of functional safety standards (i.e. ISO 26262). - Comprehensive... 
    Full time
    Contract work
    Work at office
    Work from home
    Flexible hours

    Waabi

    Pittsburgh, PA
    17 days ago
  • $86.25k - $158.13k

     ...culture where all of our employees feel respected, valued and have an opportunity to contribute to the company’s success. As a site reliability engineer within PNC's Information Technology Group and Site Reliability Center (SRC) you will be based at one of PNC's Information... 
    Full time
    Temporary work
    Part time
    Work experience placement
    Work at office

    The PNC Financial Services Group

    Pittsburgh, PA
    3 days ago
  • $75.4k - $126.5k

     ...thinking.Woolpert is an award-winning, global leader in architecture, engineering, and geospatial services. We blend design excellence with...  ...for career growth.Position OverviewWoolpert is hiring a Site Civil Engineer to join our dynamic Land Development team. This... 
    Casual work
    Local area
    Remote work
    Flexible hours
    Night shift

    Woolpert

    Pittsburgh, PA
    16 hours ago
  • $172k - $229k

     ...forefront of the driverless revolution. We are seeking a Staff Engineer Team Lead to grow and lead our Release team, specializing in Software...  ..., and deploying the infrastructure that ensures safe and reliable software releases across our autonomous vehicle fleet. You will... 
    Work experience placement
    Work at office
    Shift work

    Motional

    Pittsburgh, PA
    3 days ago
  •  ...supervision of the Chief Technology Officer, the Lead Enterprise Systems Engineer works to analyze, plan, implement, maintain, troubleshoot, and...  ...tools and gain efficiency• Work as part of a team• Work with on-site equipmentWhat You Bring to the Team• Other duties as assigned•... 
    Full time
    For contractors
    Work at office
    Night shift

    Northwest Bank

    Bellevue, PA
    3 days ago
  •  ...Job Description Job Description Title: Lead Design Engineer Location: Onsite, Pittsburgh, PA 15213 Type: Direct-Hire/Permanent Hours: Standard business hours Start: May Overview: Join a cutting-edge lab to discover novel therapeutics that are seeking... 
    Permanent employment
    Full time

    System One

    Pittsburgh, PA
    a month ago
  • $76.92k - $183.8k

     ...get crucial goods where they need to go, and make mobility more efficient and accessible for all. We’re searching for a Software Engineer  In this role, you will Be responsible for designing, delivering, and maintaining software systems at the core of our self-... 
    Full time

    Aurora Innovation

    Pittsburgh, PA
    16 hours ago
  • $70k - $300k

     ...robots interact with the real world. We are building risk-aware, reliable, and field-ready AI systems that address the most complex...  ...hardware and software is critical. We’re looking for a Software Engineer - Mission Workflows to maintain and develop robot user workflows... 
    Permanent employment
    Full time
    Flexible hours

    Field AI

    Pittsburgh, PA
    16 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!