Site Reliability Engineer
Scorpion Therapeutics
Job Responsibilities: Own the production-support model for the clinical application portfolio (on-call rotation, escalation procedures, incident-response playbooks). Establish/audit production-access and segregation-of-duties controls; maintain SOX ITGC and applicable FDA audit-ready evidence (access grants, role changes, privileged-action logs). Define and manage clinical SLOs, error budgets, and availability targets; track MTTD/MTTR and on-call burden; intervene before error budgets are spent. Lead incident response for high-severity events as incident commander/senior technical responder; coordinate teams, run post-incident reviews, drive remediation to closure. Develop runbooks and approved operational automation; enable safe, documented standard interventions. Drive automation-first operations (reduce toil; favor self-service tooling over ticket requests). Coordinate deployments, verify health, and own rollback decisions. Produce audit-ready change/deployment evidence via CI/CD pipelines. Hire/develop the App-SRE team; set expectations and career growth frameworks. Own application-layer observability (dashboards, alerts, SLO monitors) and represent App-SRE in leadership/compliance forums. Run the function AI-first (AI-assisted runbooks, incident analysis, operational tooling). Required Qualifications: BS in CS/SE/IS (or equivalent) + 10+ years SRE/DevOps/platform/production ops. 4+ years people-management/team-lead experience. Hands-on incident response leadership for Tier 1/business-critical systems. Experience defining/implementing SLOs/error budgets and on-call workflows. Experience building/maturing production support/on-call/SRE (runbooks, rotation design). Experience operating under regulated/financial-controls frameworks (e.g., SOX ITGCs, HIPAA, FDA/CAP/CLIA) with audit-capable documentation. Track record applying AI-assisted practice to operations/engineering. Preferred Qualifications: Clinical diagnostics/lab/molecular pathology/digital health domain experience. Direct SOX ITGC audit or CAP/CLIA inspection support with evidence gathering. Cloud-native observability knowledge (telemetry/tracing/APM). Deployment/release/rollback experience in continuous delivery. Player-coach ability; automation-toil reduction track record; experience presenting strategy/risk to executives. #J-18808-Ljbffr
- ...Software Engineer II - SRE RunOps Engineer7-Eleven is an iconic family of brands with over 86,000 locations, surpassing every retailer... ....The SRE RunOps Engineer 2 is responsible for ensuring the reliability, availability, and performance of the 7NOW delivery platform and...SuggestedWork experience placement
$138.4k - $173k
...infrastructure as well as help improve the reliability, quality of services and overall... ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability... ...about our locations by visiting our site.Compensation & BenefitsThe base salary that...SuggestedFull timeFlexible hours- ...generative AI and cloud-native platforms to advanced release engineering practices, our teams are redefining how financial technology... ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages...SuggestedH1bWork at officeRemote workVisa sponsorshipFlexible hours2 days per week
- Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources Willingness to work on-site at stated location in the job openingDepartment...SuggestedContract workFor contractorsWork experience placement
- ...cloud-native platforms to advanced release engineering practices, our teams are redefining how... ...alert behavior preferred Exposure to reliability engineering concepts such as SLOs/SLIs and... ...office#LI-KC1#GMFjobsAbout The Role:The Site Reliability Engineer under the general...SuggestedWork experience placementH1bWork at officeRemote workVisa sponsorshipFlexible hoursShift work2 days per week
- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...Full time
- ...Administrator / SRE in Dallas to own production Java environments, middleware, and cloud automation. You will optimize performance, drive reliability, and mentor teammates while aligning with enterprise security and AI-enabled integrations. You will work across Java apps, IBM...
- ...Experience should include full product experience (APM, Logs, setting up monitoring, alerts, dashboards). Overall looking for a good Reliability Engineer that will support our environments by setting up alerting, monitoring strict SLA's and engaging to determine issues. Must...
$120.6k - $150.9k
About the RoleWe are looking for a highly motivated, high-potential Staff Site Reliability Engineer (SRE) to join our team as a technical leader and drive transformative impact across WEX’s platform reliability and operational excellence.This is a particularly exciting...Full timeFlexible hours$167.7k - $245.2k
...Cisco Meraki, we are responsible for building and growing the cloud that supports these customers and their networks. As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze...Full timeTemporary workLocal areaFlexible hours- Site Reliability Engineer - Vice PresidentSite Reliability Engineering (SRE) is an engineering discipline that combines software and systems engineering to build and run scalable, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for improving...
- ...and continuously improving the platforms that power TI's digital integration, automation and DevOps capabilities. As an IT Site Reliability Engineer within the Enterprise Platforms team, you will serve as the primary technical platform owner for TI's Apigee Edge private...Local area
$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...requiredBachelor’s Degree or additional equivalent work or military experience preferredLicenses and CertificationsSAFe 5.0 Release Train Engineer (RTE certification) within 90 Days required SAFe 5.0 SAFe Program Consultant (SPC certification) preferredWhat We Offer: Generous...Work at officeFlexible hours2 days per week
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...a company that values diversity, integrity, and growth. Role Overview PDI Technologies is looking for a Manager, Site Reliability Engineering to lead the SRE organization supporting Paylo, PDI’s payments, loyalty, and fuel-pricing product suite. This role owns the...
- Required Technical Skills • Linux bash, Python, pip, conda, Node/npm, rpm, GNU tools (g++,make,configure)Desirable Technical Skills • Public cloud platform experience• Bash script development• Python development, package management, package building and testing• Source ...
- ...Role : Site Reliability Engineer II (SRE II) Data & Intelligence Location: Dallas, TX / Overland Park, KS / Atlanta, GA / Bellevue, WA (Onsite) Job Summary The Site Reliability Engineer II (SRE II) is responsible for ensuring the reliability, scalability...
$125k - $140k
...character, perspective, and passion for achieving great things in the world are equally as important to us. The role The Site Reliability Engineer is a fundamental piece of the Site Reliability Engineering team. Site Reliability Engineering is accountable for the...Full timeLocal areaRemote workFlexible hours- ...asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing... ...- Manage, administer, and maintain all internet and intranet sites - Research/analyze data processing functions, methods and procedures...
- ...Job Description Job Description Loan Platform Configuration Engineer (.NET) About the Role This role sits at the center of a production loan servicing environment and focuses on configuring and administering a third‐party loan platform , not just building applications...Contract work
- ...country! We are seeking a highly experienced Senior Systems Engineer with expert-level mastery of the Microsoft enterprise technology... ...troubleshoot, and optimize system performance, availability, reliability, and security across on-premises and cloud environments...
$132.23k - $176.31k
...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...Full timeTemporary workRemote work$107.48k - $143.31k
...employees. What You'll Be Doing Lead and apply regional reliability engineering strategies to improve equipment performance, uptime, and... ...management skills with the ability to support multiple sites remotely. Willingness to travel to the plant and corporate...Temporary workRemote workFlexible hours- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...Work experience placement
- ...Senior Software Systems Engineer Immediate need for a talented Senior Software Systems Engineer with experience in Telecom Industry. This is an 11+ months contract opportunity with long-term potential and is located in Irving, TX. Please review the job description...Contract workWork experience placementImmediate start
- ...fulfill travel worldwide.SRE Software Systems Engineer IV - Data Intelligence and AI... ...Systems Engineer, you will drive platform reliability, auto-scaling and cloud cost efficiency... ...running smoothly. This role requires strong Site Reliability Engineering discipline, problem...Full timeWorldwideFlexible hoursWeekend work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Irving, TX
- site safety Irving, TX
- junior website developer Irving, TX
- construction site safety Irving, TX
- IT site lead Irving, TX
- website content developer Irving, TX
- historic site Irving, TX
- site services specialist Irving, TX
- official site Irving, TX
- site leader Irving, TX




