Site Reliability Engineer - Incident Management/Resiliency (Hybrid)
Enova International
We are interested in every qualified candidate who is eligible to work in the United States. However, we are not able to sponsor visas or take over sponsorship at this time. About the Role Resilience Engineering is a subset of the Site Reliability Engineering team that strives to foster a culture of continuous improvement through incident analysis, process evolution, and problem‑solving. We work closely with teams across Tech, Product, and Operations through our Production Incident process to uncover system weaknesses, learn from failures, and make our technology more reliable. In this role, you’ll play a key role in enhancing the resiliency of our systems. Your work will focus on our incident response, reporting and analysis processes, enabling the organization to better prepare for and respond to complex system failures. You’ll drive efforts to optimize how we manage unexpected outages, from leading real‑time incident response to facilitating post‑incident reviews. What you’ll be doing Lead production incidents as part of our PI PIC (or Incident Commander) rotation after completing training, ensuring clear communication and resolution. Capture and maintain detailed documentation of incidents, contributing factors, and learnings in formal incident reports. Deliver documentation that is clear, comprehensive, and accessible to different types of audiences in a timely manner within the established SLAs. Facilitate and document blameless post‑incident reviews that promote learning and continuous improvement. Collect and analyze incident data to identify systemic issues, risks, and trends. Work on improvements to how we collect, analyze, and learn from system failures. Requirements 2+ years experience in a technology or analyst role (e.g., Software Engineering, Systems, Operations, SRE, or Product). A strong interest in how complex distributed systems operate—and how to make them more reliable. Analytical and problem‑solving skills with a systems‑thinking mindset. Strong communication skills, both verbal and written, with the ability to tailor messaging to technical and non‑technical audiences. Comfort with ambiguity, and the ability to turn vague problems into actionable insights. Demonstrated maturity, sound judgment, and organizational awareness. Ability to coordinate the resolution of major incidents and post‑incident reviews following Enova’s Incident Management Process Ability to seamlessly shift between high‑urgency incident response and structured project work, with strong organizational skills and the capacity to manage projects independently. Nice to have Experience leading resolution of major system outages or production incidents. Experience driving large‑scale technical or process changes. Compensation This position includes various levels within our career ladder. The actual annual salary will be determined based on qualifications, skills, experience, and level assessed during the hiring process and may fall outside of the ranges shown. Budgeted annual salary ranges: Additional compensation for this role may include a bonus. All full‑time employees are eligible to participate in Company benefits, described in more detail here. Our hybrid roles require in‑office work Tuesday through Thursday, with remote flexibility on Mondays and Fridays. This schedule fosters collaboration, team connection, and strategic planning, enhancing communication and effectiveness to drive results. Health, dental, and vision insurance including mental health benefits 401(k) matching plus a roth option (U.S. Based employees only) PTO & paid holidays off Sabbatical program (for eligible roles) Summer hours (for eligible roles) Paid parental leave DEI groups (B.L.A.C.K. @ Enova, HOLA @ Enova, Women @ Enova, Pride @ Enova, South Asians @ Enova, APEX @ Enova, and Parents @ Enova) Employee recognition and rewards program Charitable matching and a paid volunteer day…Plus so much more! About Enova Enova International is a leading financial technology company that provides online financial services through our AI and machine learning‑powered Colossus™ platform. We serve non‑prime consumers and businesses alike, while offering world‑class technology and services to traditional banks—in order to create accessible credit for millions. Being a values‑driven organization is at the core of Enova’s success. We live our values by listening to our customers, challenging assumptions, thinking big, setting high expectations, and hiring and developing the best. Through our values and our commitment to making Enova an awesome place to work, we maintain an environment of inclusion and culture where our employees can thrive. You can learn more about Enova’s values and culture here. It is our policy to provide equal employment opportunity for all persons and not discriminate in employment decisions by placing the most qualified person in each job, without regard to any other classification protected by federal, state, or local law. California Applicants: Click here to review our California Privacy Policy for Job Applicants. Seniority level Mid‑Senior level Employment type Full‑time Job function Engineering and Information Technology Industries Technology, Information and Internet Referrals increase your chances of interviewing at Enova International by 2x #J-18808-Ljbffr Enova International
$108.08k - $172.5k
...with development and platform engineering teams to migrate and... ...production systems, facilitate incident management and conduct post-incident reviews... ...application operational resiliency, fault tolerance and... ...Telecommuting permitted on a hybrid schedule as determined by the...SuggestedFull timeRemote workWorldwide- ...in ensuring system reliability at one of the world’... ...financial institutions.As a Site Reliability Engineer II at JPMorgan Chase... ...to speed up incident triage, troubleshooting... ...to maintain resiliency, security, and auditability... ...and asset management. We offer a competitive...Suggested
$158.5k - $172k
...OpportunityAs a Senior Engineer on the Runtime... ...team is responsible for managing our centralized Enterprise... ...position driving continuous reliability, deep system... ...on-call rotation as an incident responder, leading postmortem... ...(e.g., EKS, GKE).Our hybrid model requires 3 days...SuggestedFull timeTemporary workWork at officeFlexible hours3 days per week- ...mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase... ...environment to accelerate incident triage, troubleshooting, and... ...supporting SRE practices for Data management/migration platforms and... ...hosted on public/private/hybrid cloud environments, including...Suggested
$132.1k - $220.1k
Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week... ...Kafka event bus.Incident Command & Systemic Resilience: Lead the response for... ...design and ArgoCD for managing immutable infrastructure...SuggestedFull timeWork at officeLocal areaWorldwide2 days per week$118.3k - $219.8k
Are you excited to lead Site Reliability Engineering teams that keep mission-critical, 24/7 services... ...proactive monitoring, effective incident management, and continuous optimisation. Through... ...You’ll strengthen reliability and resilience across a group of products, reduce...Full timeLocal area$130k - $165k
...Title: Senior Software Engineer Company:... ...Technology Team : Site Reliability Engineering... ...and innovative claims management technology, transforming... ...frequency and severity of incidents Build, maintain, and... ...technology stack remains resilient, scalable, and high-performing...Full timeTemporary workLocal areaRemote workVisa sponsorshipWork visaFlexible hours- ...-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within... ...environment to accelerate incident triage, troubleshooting, and... ...SRE practices for Data management/migration platforms and products... ...hosted on public/private/hybrid cloud environments, including...
$130k - $180k
...accomplishment.Being a Senior Site Reliability Engineer at iManage Means…... ...deliver scalable, resilient platforms that... ...observability, change management, and service... ...critical events. Leading incident management and post-... ...your personal data.#LI-Hybrid#LI-LM1 SummaryLocation:...Work at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$130k - $150k
...-premise systems and hybrid infrastructure, meaning... ...Position OverviewThe Site Reliability Engineer (SRE) helps ensure... ...observability, and strengthen incident response. The SRE... .../SLOs), implement resilient architectures, and... ...VMware Site Recovery Manager (SRM), SAN technologies...Work at officeWork from home3 days per week$100.7k - $167.8k
Job SummaryThe Site Reliability Engineer III is a pivotal architect of stability for CME Clearing... ...operations, you ensure our risk management services remain resilient and high-performing for customers... ...rotations and a leader in post-incident forensic reviews.Collaborate for...Full timeWorldwide$140k - $170k
We are looking for a Senior Site Reliability Engineer to work as part of a lean,... ...deployments. Participate in incident response, perform root... ...reusable modules/patterns and managing changes through review and... ..., tuition reimbursement, a hybrid working environment for most...Full timeWork experience placementFlexible hours$194k - $267k
...an experienced Staff Site Reliability Engineer to join Okta's... ...facing systems.Lead incident response efforts and... ...scalability, performance, and resilience.Ensure all... ...networking, and traffic management.Strategic experience... ...FAR) 2.101#LI-SM1#LI-Hybrid P17611_3494641The annual...Local areaWorldwideFlexible hours$125.04k - $187.56k
...and more. Overview The Site Reliability Engineer (SRE) III is... ...automation, observability, incident response, and infrastructure... ...that improve system resilience. Applicants must be... ...basis. Our flexible/hybrid work schedule... ...indicators (SLIs). Build and manage microservices-based...Full timeWork at officeRemote workFlexible hours$127.6k - $191.4k
Staff Reliability Engineer - IE07KEWe’re determined to make a difference... ....This role will have a Hybrid work schedule, with the... ...in real-time.Pipeline Resiliency & Automation: Collaborate... ...operational efficiency.Incident and Problem Management (Data Focus): Lead the response...Full timeTemporary workWork at office3 days per week- ...operations. This is a team for engineers who want real technical... ...architect highly available, resilient APIs serving millions of requests... ...internal stakeholders to define reliable data contracts and service... ...Participate in production support and incident response Continuously...Full timeSeasonal work
$135k - $145k
...currently seeking a Senior Service Reliability Engineer to embed with Fitch... ...and the Executive Program Management Office (EPMO). Driven by our... ...Teams integrations—and reduce incidents via telemetry-driven... ...ceremonies.Why Choose Fitch:Hybrid Work Environment: 2 to 3 days...Temporary workWork at officeImmediate start2 days per week3 days per week- ...off Platform Engineering & AI Operations... ...Lead Full Time | Hybrid | Chicago Metro... ...reducing utility bills, managing demand charges,... ..., supporting resilience, and helping... ...engineering delivery, reliability, security, and... ...security review agents, incident summarizers,...Full timeRemote workWork from homeFlexible hours
$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range... ...observability and alerting systems.The Fleet Management team provides the core runtime... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...Work at officeLocal areaRemote workWorldwideFlexible hours- ...oprecruiting.comJob Title: Site Reliability Engineer (Infrastructure & Systems)Location... ...automate network workflows.Incident Escalation: Serve as a Tier... ...(such as Kubernetes) in hybrid or public cloud settings.... .../ GCP).Familiarity with managed container orchestrators such...Local area
- ...areas, including asset management, private equity, M&A,... ...OverviewThe Cloud Solutions Engineer is responsible for... ...secure, scalable, and resilient solutions across the... ...Microsoft Entra ID (including hybrid identity with Entra... ...using Azure Site Recovery, Azure Backup...Work at office
$240k - $375k
...innovative wealth management, asset servicing,... ...for Platform Engineering leads the strategy... ...scalability, security, and reliability by delivering... ...stability and resilience.Deliver developer... ..., including incident management, reliability... ..., DevOps, and Site Reliability Engineering...H1bWorldwideFlexible hours$100k - $140k
DepartmentBSD CTD - Platform Engineering - GDCAbout the DepartmentThe... ..., & AI/ML infrastructure management across the open-source software... ..., and addressing production incidents. For monitoring, staff... ...incidents. CI/CD pipelines are for hybrid cloud architecture on-...Full timeWork experience placement- ...StatesAs a Senior IT Systems Engineer, you'll own and operate the company... ..., from helpdesk and device management to cloud architecture and... ...Qualifications5+ years in a hybrid IT/infrastructure role at a small... ...domain, please disregard them and report the incident to us....
$127k - $249k
...can sit in our NYC HQ on a hybrid basis, or it can be fully remote... ...for an experienced Senior Engineer for our SRE, Atlas team to... ...OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong... ...of a reliable and resilient multi-cloud platform that hosts...Local areaRemote workWorldwideFlexible hours$112.5k - $187.5k
...will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy,... ...capacity and security work, or anchoring incident response with calm and maturity, your... ...the people around you. This is a hybrid position and involves regular performance...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$194k - $267k
...tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support... ...workloads, providing high resilience and operational efficiency.... ...with minimal downtime.Incident Management & Troubleshooting...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$175k - $225k
...responsible for ensuring the reliability, stability, and... ...— partnering with Engineering, Infrastructure, Cybersecurity... ...our systems are resilient, observable, and... ...automation-first thinking, manage third-party vendor... ...center during incidents and outages. This role...Full timeTemporary workWork at office$160k - $210k
...Join our Platform Engineering team, where you'... ...engineers across reliability initiativesAnalyze... ...deployments, and incident responses with... ...provisioning, scaling, and management across all... ...in DevOps, Site Reliability Engineering... ..., we follow a hybrid work schedule: In...Work at officeWorldwideMonday to FridayFlexible hours- ...and cities to manage energy and water... ...our fast-growing Resiliency Solutions... ...OverviewThe DevOps Engineer plays a critical... ...security, and reliability.Collaborate with... ...in DevOps, Site Reliability Engineering... ...monitoring and incident management... ...purchase program, hybrid work schedule,...Full timeWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer - Incident Management/Resiliency (Hybrid). Be the first to apply!
- site reliability engineer Chicago, IL
- site reliability engineer sre Chicago, IL
- site reliability engineer remote Chicago, IL
- website coordinator Chicago, IL
- on-site clinical research associate (traveling/remote) Chicago, IL
- site safety Chicago, IL
- junior website developer Chicago, IL
- construction site safety Chicago, IL
- IT site lead Chicago, IL
- website content developer Chicago, IL



