Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$145k - $175k

GrabJobs

Full-time Description At Commence, we’re the start of a new age of data-centric transformation, elevating health outcomes and powering better, more efficient process to program and patient health. We combine quality data-driven solutions that fuel answers, technology that advances performance, and clinical expertise that builds trust to create a more efficient path to quality care. With human-centered, healthcare-relevant, and value-based solutions, we create new possibilities with data. We provide proof beyond the concept and performance beyond the scope with a focus on efficiencies that transform the lives of those we serve. With a culture driven by purpose, straightforward communication and clinical domain expertise, Commence cuts straight to better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability, scalability, and operational health of our mission-critical healthcare data platform. You will bridge the gap between engineering and operations—embedding reliability as a first-class concern from architecture through deployment. This role is built for someone who thrives when systems are under pressure and who treats an outage as a problem to be engineered away permanently, not just survived. Design, implement, and own observability infrastructure including metrics, logging, tracing, and alerting across distributed systems. Define and enforce SLOs, SLIs, and error budgets in partnership with product and engineering teams. Lead incident response: triage, coordinate remediation, conduct blameless post-mortems, and drive systemic fixes. Build and maintain CI/CD pipelines that support rapid, safe delivery of changes to production. Collaborate with engineering teams on infrastructure changes; able to read, modify, and contribute to existing infrastructure-as-code (Terraform or CloudFormation). Design and operate highly available, fault-tolerant systems—including auto-scaling, failover, and disaster recovery strategies. Reduce operational toil through automation; eliminate manual processes before they become habits. Collaborate with software engineers to establish reliability-first design patterns and review architectures for operational risk. Manage Kubernetes or container orchestration environments at scale. Ensure systems meet compliance and security requirements, particularly those applicable to healthcare data (HIPAA, SOC 2). Provide technical mentorship and guidance to engineers across the organization on reliability practices. Participate in on-call rotation with a commitment to continuously reducing the need for it. Qualifications 7+ years of experience in SRE, platform engineering, or DevOps roles. Exceptional problem-solving under pressure—demonstrated track record of diagnosing complex, high-stakes system failures and building durable solutions. Deep hands-on experience with AWS services including EC2, EKS/ECS, Lambda, RDS, S3, CloudWatch, and related tooling. Familiarity with infrastructure-as-code (Terraform or CloudFormation)—able to contribute to existing configurations. Experience designing and operating distributed systems with strict availability and latency requirements. Proficiency in at least one scripting or systems language (Python, Go, Bash, or similar) for automation and tooling. Experience with container orchestration (Kubernetes, ECS) in production environments. Expertise in observability tooling (OpenSearch, Prometheus/Grafana, or equivalent). Hands-on experience with CI/CD platforms (GitHub Actions, Jenkins, CircleCI, or similar). Proven ability to define and operationalize SLOs and error budgets. Experience with relational and NoSQL databases—performance tuning, replication, and backup strategies. Strong working knowledge of networking fundamentals: DNS, load balancing, VPCs, TLS. Excellent communication skills—able to translate technical risk into business impact for non-engineering stakeholders. Additional Requirements AWS Certifications (Solutions Architect, DevOps Engineer, or SysOps Administrator). Experience in healthcare technology or other regulated industries (HIPAA, SOC 2, FedRAMP). Familiarity with chaos engineering practices and tooling. Experience with data pipeline reliability (ETL/ELT workflows, streaming systems). Exposure to AI/ML infrastructure and the reliability challenges unique to model serving. Familiarity with additional cloud platforms (Azure, Google Cloud). Contributions to open-source reliability or infrastructure tooling. *Commence' headquarters are in Virginia Beach, VA, however we are open to remote candidates in the following states: AZ, AR, CO, DE, FL, GA, IL, IN, KS, KY, MA, MD, MI, MS, MO, MT, NC, NE, NV, NY, OH, OK, PA, SC, TN, TX, VA, DC, WI, and WV* Work Environment/Physical Demands The work environment and physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. This is a remote position. While performing the duties of this job, the employee regularly works in a climate-controlled environment. Candidates must be able to sit, read, work on a computer, and watch a computer screen for extended periods of time. Occasionally required to stand, walk, use hands and fingers, kneel or crouch. Commence is an equal employment opportunity employer. All personnel processes are merit-based and applied without discrimination on the basis of race, color, religion, sex, sexual orientation, gender identity, marital status, age, disability, national or ethnic origin, military and veteran status or any other characteristic protected by applicable law. Commence.AI is committed to providing equal employment opportunities to all applicants, including individuals with disabilities. If you require a reasonable accommodation to participate in the application process due to a disability, please contact Human Resources at View phone number on click.appcast.io or View email address on click.appcast.io. Please note that unless you are requesting an accommodation, all applications must be submitted through our online application system. Salary Description $145,000-$175,000

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Saint Paul, MN vacancy
  • $119k - $170k

     ...the greater good, come make your next move with Zscaler. Our Engineering team built the world’s largest cloud security platform from...  ...cloud-first strategy. We’re looking for an experienced Staff Site Reliability Engineer (Federal) to join our Government Cloud team.... 
    Suggested
    Full time
    Work at office
    Local area
    Worldwide
    Night shift

    GrabJobs

    Saint Paul, MN
    4 days ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Suggested
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Saint Paul, MN
    3 days ago
  • $121.4k - $218.6k

     ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with...  ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust... 
    Suggested
    Work experience placement
    Work at office

    Akamai

    Saint Paul, MN
    3 days ago
  •  ...and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing...  ...problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes... 
    Suggested
    Work at office
    Local area
    Remote work
    Sleeping nights

    GrabJobs

    Saint Paul, MN
    1 day ago
  • $141.8k - $195k

     ...their best work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl... 
    Suggested
    Temporary work
    Remote work

    GrabJobs

    Saint Paul, MN
    3 days ago
  • $168k - $200k

     ...Senior Site Reliability Engineer Datavant is the data collaboration platform trusted for healthcare. Guided by our mission to make the world's health data secure, accessible and actionable, we provide critical data solutions for organizations across the healthcare ecosystem... 

    Minnesota Jobs

    Saint Paul, MN
    23 hours ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection... 
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Saint Paul, MN
    1 day ago
  • Job Description Job Description Mandatory 1.     Technical and Windows Application Architecture – performance evaluation and tuning of applications 2.     Cloud knowledge incl. IaaS, PaaS etc. 3.     Code as Service – experience working with containers like Docker...

    Unknown

    Saint Paul, MN
    more than 2 months ago
  •  ...Partner with software developers, platform engineers, and IT staff to improve system design,...  ...requirements, service quality, reliability, security, and compliance needs. Drive continuous...  ...Required: 8+ years of experience in Site Reliability Engineering, DevOps, Platform... 
    Work at office
    Remote work

    GrabJobs

    Saint Paul, MN
    3 days ago
  •  ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San... 
    Internship
    Work at office
    Local area
    Remote work
    Worldwide

    GrabJobs

    Saint Paul, MN
    2 days ago
  • $84.9k - $209.5k

     ...Job Description As a Principal Site Reliability Engineer (IC4), you will be responsible for designing, building, and operating highly available, scalable, secure, and resilient cloud services. You will combine software engineering with infrastructure expertise to improve... 
    Temporary work
    Flexible hours

    Oracle

    Saint Paul, MN
    4 days ago
  • $80k - $140k

    Job DescriptionRBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and... 
    Full time
    Flexible hours
    Shift work

    Royal Bank of Canada

    Minneapolis, MN
    1 day ago
  •  ...itD is seeking a Site Reliability Engineer to develop and enhance automation solutions that improve the reliability, scalability, and operational efficiency of large-scale cloud infrastructure. The ideal candidate will bring hands-on experience in site reliability engineering... 
    Work experience placement
    Remote work

    GrabJobs

    Minneapolis, MN
    1 day ago
  •  ...Join Our Team as a Site Reliability Engineer (SRE)! About Us At Energy Worldnet, Inc. (EWN), we deliver innovative technology solutions that empower our clients and support the future of the energy industry. Our SaaS platform helps organizations operate more safely, efficiently... 
    Temporary work
    Casual work
    H1b
    Work at office
    Remote work
    Home office
    Monday to Friday
    Flexible hours
    Shift work

    GrabJobs

    Minneapolis, MN
    1 day ago
  • $160k - $210k

     ...departmental collaboration and a unified sense of purpose, making teamwork a cornerstone of our success. We are looking for a Senior Site Reliability engineer to work on expanding our global footprint of datacenters and improve service management across Cognitiv. Our immediate... 
    Work at office
    Local area
    Immediate start
    Remote work

    GrabJobs

    Minneapolis, MN
    3 days ago
  • $70 - $74.07 per hour

     ...Title: Technology & Information Architectures - Performance & Reliability Engineer Location: Richfield, MN (Hybrid, Local Required) Duration:...  ...applications. Required Skills & Qualifications: 3-5+ years in Site Reliability Engineering, DevOps, or Production Support. Experience... 
    Contract work
    Work at office
    Local area

    BCforward

    Richfield, MN
    13 hours ago
  •  ...Site Reliability Engineer We are seeking a highly skilled Site Reliability Engineer (SRE) to support and enhance the reliability, scalability, and performance of enterprise applications and infrastructure. The ideal candidate will have strong experience in cloud environments... 

    eTeam

    Minneapolis, MN
    13 hours ago
  • $107.9k - $195.05k

     ...Description Leidos is seeking an experienced Release Train Engineer (RTE) to join the Air Traffic Business Area within the Homeland Sector, supporting the development of next-generation flight service and air traffic management systems. This work supports the FAA'... 
    Contract work
    Local area
    Immediate start
    Remote work

    Leidos

    Saint Paul, MN
    2 days ago
  •  ...include the following and other duties may be assigned Perform reliability activities (design verification, reliability testing, test...  ...and conducts technical reliability studies and evaluations of engineering design concepts and design of experiments (DOE) constructs Recommends... 
    Extra income
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Work visa

    ATR International

    Saint Paul, MN
    3 days ago
  • $114.6k - $234.6k

     ...creates durable fixes and preventive controls. Designs performance, reliability, and fault-tolerance improvements for drivers, services, and...  ...development lifecycle; provides guidance and coaching to engineers to drive improvements. Utilizes advanced knowledge to develop... 
    Temporary work
    Flexible hours
    Shift work

    Oracle

    Saint Paul, MN
    2 days ago
  • Minnesota Department of Health is seeking a statewide leader for overdose prevention. The role coordinates grants and contracts, provides technical assistance, and partners with public health, healthcare, and community organizations to strengthen systems and improve naloxone...

    Minnesota Department of Health

    Saint Paul, MN
    3 days ago
  •  ...We are seeking a highly skilled and motivated Lead Systems Engineer to join our team, focusing on the design, development, and optimization...  ...Engineering (MBSE) methodologies to enhance the performance, reliability, and safety of advanced hypersonic test systems. The ideal... 
    For subcontractor
    Local area

    North Wind

    Saint Paul, MN
    4 days ago
  • Company Description Sonsoft , Inc. is a USA based corporation duly organized under the laws of the Commonwealth of Georgia. Sonsoft Inc. is growing at a steady pace specializing in the fields of Software Development, Software Consultancy and Information Technology...
    Full time

    SonSoft Inc.

    Saint Paul, MN
    13 hours ago
  • $131.4k - $243.8k

     ...position title is Infrastructure Sr Mgr.* ​Mission StatementThe Observability & Middleware Platforms team within the Platform & Reliability Engineering organization empowers Securian’s technology ecosystem by delivering reliable, scalable, and automated enterprise platforms.... 
    Full time
    Work at office
    Flexible hours
    3 days per week

    Securian Financial

    Saint Paul, MN
    1 day ago
  • $140k - $200k

     ...people around the globe work on Speechify in a 100% distributed setting - Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and Google, leading PhD programs like Stanford, high growth startups... 
    Work at office
    Remote work

    Speechify

    Saint Paul, MN
    13 hours ago
  •  ...that the LLM offering is as efficient and fast as possible given the inference hardware available Maintain and optimize inference engine architecture Tune data storage configurations to optimize for scale and near real-time availability in a streaming architecture... 
    Full time
    Work at office
    Local area
    3 days per week

    Swoop Technologies

    Saint Paul, MN
    13 hours ago
  • $105.4k - $124k

    At U.S. Bank, we’re on a journey to do our best. Helping the customers and businesses we serve to make better and smarter financial decisions and enabling the communities we support to grow and succeed. We believe it takes all of us to bring our shared ambition to life,...
    Full time
    Work experience placement
    Local area
    3 days per week

    US Bank

    Saint Paul, MN
    4 days ago
  • $105.4k - $124k

     ...what you excel at—all from Day One.Job DescriptionThe Systems Engineer creates and integrates solutions and services. The Systems Engineer...  ...directs or recommends enhancements for system performance, reliability, scalability, and capacity. Leads the configuration of... 
    Full time
    Local area
    Remote work
    3 days per week

    US Bank

    Saint Paul, MN
    4 days ago
  •  ...cash incentive awards.Salary Range$91,100.00 - $150,300.00Target Openings1What Is the Opportunity?Travelers is seeking a Software Engineer I to join our organization as we grow and transform our Technology landscape. Individual will complete intermediate end to end engineering... 
    Work experience placement
    Immediate start

    The Travelers Indemnity Company

    Saint Paul, MN
    2 days ago
  • $92.5k - $209.5k

     ...Job Description We’re hiring a Platform Software Engineer to help build the execution layer behind OCI’s AI and GPU growth motion....  ...Object Storage, Data Flow) to move training and inference data reliably across customer and internal pipelines. • Build and extend GPU... 
    Temporary work
    Flexible hours

    Oracle

    Saint Paul, MN
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!