Site Reliability Engineer - Incident Management/Resiliency (Hybrid)
Enova International
We are interested in every qualified candidate who is eligible to work in the United States. However, we are not able to sponsor visas or take over sponsorship at this time. About the Role Resilience Engineering is a subset of the Site Reliability Engineering team that strives to foster a culture of continuous improvement through incident analysis, process evolution, and problem‑solving. We work closely with teams across Tech, Product, and Operations through our Production Incident process to uncover system weaknesses, learn from failures, and make our technology more reliable. In this role, you’ll play a key role in enhancing the resiliency of our systems. Your work will focus on our incident response, reporting and analysis processes, enabling the organization to better prepare for and respond to complex system failures. You’ll drive efforts to optimize how we manage unexpected outages, from leading real‑time incident response to facilitating post‑incident reviews. What you’ll be doing Lead production incidents as part of our PI PIC (or Incident Commander) rotation after completing training, ensuring clear communication and resolution. Capture and maintain detailed documentation of incidents, contributing factors, and learnings in formal incident reports. Deliver documentation that is clear, comprehensive, and accessible to different types of audiences in a timely manner within the established SLAs. Facilitate and document blameless post‑incident reviews that promote learning and continuous improvement. Collect and analyze incident data to identify systemic issues, risks, and trends. Work on improvements to how we collect, analyze, and learn from system failures. Requirements 2+ years experience in a technology or analyst role (e.g., Software Engineering, Systems, Operations, SRE, or Product). A strong interest in how complex distributed systems operate—and how to make them more reliable. Analytical and problem‑solving skills with a systems‑thinking mindset. Strong communication skills, both verbal and written, with the ability to tailor messaging to technical and non‑technical audiences. Comfort with ambiguity, and the ability to turn vague problems into actionable insights. Demonstrated maturity, sound judgment, and organizational awareness. Ability to coordinate the resolution of major incidents and post‑incident reviews following Enova’s Incident Management Process Ability to seamlessly shift between high‑urgency incident response and structured project work, with strong organizational skills and the capacity to manage projects independently. Nice to have Experience leading resolution of major system outages or production incidents. Experience driving large‑scale technical or process changes. Compensation This position includes various levels within our career ladder. The actual annual salary will be determined based on qualifications, skills, experience, and level assessed during the hiring process and may fall outside of the ranges shown. Budgeted annual salary ranges: Additional compensation for this role may include a bonus. All full‑time employees are eligible to participate in Company benefits, described in more detail here. Our hybrid roles require in‑office work Tuesday through Thursday, with remote flexibility on Mondays and Fridays. This schedule fosters collaboration, team connection, and strategic planning, enhancing communication and effectiveness to drive results. Health, dental, and vision insurance including mental health benefits 401(k) matching plus a roth option (U.S. Based employees only) PTO & paid holidays off Sabbatical program (for eligible roles) Summer hours (for eligible roles) Paid parental leave DEI groups (B.L.A.C.K. @ Enova, HOLA @ Enova, Women @ Enova, Pride @ Enova, South Asians @ Enova, APEX @ Enova, and Parents @ Enova) Employee recognition and rewards program Charitable matching and a paid volunteer day…Plus so much more! About Enova Enova International is a leading financial technology company that provides online financial services through our AI and machine learning‑powered Colossus™ platform. We serve non‑prime consumers and businesses alike, while offering world‑class technology and services to traditional banks—in order to create accessible credit for millions. Being a values‑driven organization is at the core of Enova’s success. We live our values by listening to our customers, challenging assumptions, thinking big, setting high expectations, and hiring and developing the best. Through our values and our commitment to making Enova an awesome place to work, we maintain an environment of inclusion and culture where our employees can thrive. You can learn more about Enova’s values and culture here. It is our policy to provide equal employment opportunity for all persons and not discriminate in employment decisions by placing the most qualified person in each job, without regard to any other classification protected by federal, state, or local law. California Applicants: Click here to review our California Privacy Policy for Job Applicants. Seniority level Mid‑Senior level Employment type Full‑time Job function Engineering and Information Technology Industries Technology, Information and Internet Referrals increase your chances of interviewing at Enova International by 2x #J-18808-Ljbffr Enova International
$108.08k - $172.5k
...with development and platform engineering teams to migrate and... ...production systems, facilitate incident management and conduct post-incident reviews... ...application operational resiliency, fault tolerance and... ...Telecommuting permitted on a hybrid schedule as determined by the...SuggestedRemote workWorldwide- ...Contract Lifecycle Management, a Fortune Great... .... This is a hybrid role. Office... ...platform and champion engineering excellence... ...driving architectural resilience, and solving... ...direction for the Site Reliability Engineering team... ...to respond to incidents that impact Ironclad...SuggestedFull timeContract workWork at office
- ...in ensuring system reliability at one of the world'... ...institutions. As a Site Reliability Engineer II at JPMorgan Chase... ...environment to speed up incident triage,... ...controls to maintain resiliency, security, and auditability... ...processing and asset management. We offer a competitive...SuggestedLocal area
$132.1k - $220.1k
Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week... ...Kafka event bus.Incident Command & Systemic Resilience: Lead the response for... ...design and ArgoCD for managing immutable infrastructure...SuggestedFull timeWork at officeLocal areaWorldwide2 days per week- ...the future of digital wealth management by building tech-forward solutions... ...oriented Application Support Engineer to join our production... ...work. Duties/Responsibilities Incident and Escalation Response Serve... ...Environment This role operates in a hybrid environment, in office 3 days...SuggestedWork experience placementWork at officeWork from homeFlexible hours3 days per week
$130k - $180k
.... Being a Senior Site Reliability Engineer at iManage Means…... ...and deliver scalable, resilient platforms that support... ..., change management, and service scalability... ...events. Leading incident management and post-... ...personal data. #LI-Hybrid #LI-LM1 Powered by...Work at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$158.5k - $172k
...OpportunityAs a Senior Engineer on the Runtime... ...team is responsible for managing our centralized Enterprise... ...position driving continuous reliability, deep system... ...on-call rotation as an incident responder, leading postmortem... ...(e.g., EKS, GKE).Our hybrid model requires 3 days...Full timeTemporary workWork at officeFlexible hours3 days per week$140k - $175k
Senior Software Engineer (Hybrid - US) Boston, Massachusetts... ...as needed. Lead incident response, root-cause... ...application performance, reliability, scalability, capacity... ...and strengthen system resilience across teams. Minimum... ...Assurance, Project Management, DevOps, business stakeholders...Work at officeRemote work$175k - $225k
...leading alternative credit manager and a trusted... ...committed to building resilient, scalable, and modern... ...seeking a Vice President, Reliability Engineering & Technology... ...error budgets, and post-incident learning. Technology... ...New York office. #LI-hybrid A reasonable estimate...Temporary workWork at office- ...mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase... ...environment to accelerate incident triage, troubleshooting, and... ...supporting SRE practices for Data management/migration platforms and... ...hosted on public/private/hybrid cloud environments, including...
- ...operations. This is a team for engineers who want real technical... ...architect highly available, resilient APIs serving millions of requests... ...internal stakeholders to define reliable data contracts and service... ...Participate in production support and incident response Continuously...Full timeSeasonal work
- ...time off Platform Engineering & AI Operations Lead Full Time | Hybrid | Chicago Metro... ...reducing utility bills, managing demand charges,... ..., supporting resilience, and helping... ...engineering delivery, reliability, security, and... ...security review agents, incident summarizers,...Full timeRemote workWork from homeFlexible hours
$147k - $210k
...the backbone of Enova's engineering organization —... ...that improve platform reliability, developer productivity... ...prioritized work. Manage timeline feasibility and... ..., particularly during incident resolution and complex... ...Benefits & Perks: ~ Our hybrid roles require in-...Full timeSummer workWork at officeLocal areaRemote workMonday to Friday$135k - $145k
...currently seeking a Senior Service Reliability Engineer to embed with Fitch... ...and the Executive Program Management Office (EPMO). Driven by our... ...Teams integrations—and reduce incidents via telemetry-driven... ...ceremonies.Why Choose Fitch:Hybrid Work Environment: 2 to 3 days...Temporary workWork at officeImmediate start2 days per week3 days per week$130k - $170k
...Senior Site Reliability Engineer About Us Founded in 2014, we offer the industry’... ...leadership in suitability and risk management with industry‑leading education... ...of observability tools, incident response processes, and resilience strategies, shifting the organization...Full timeFlexible hoursShift work$150k - $200k
...Technology We are seeking a Site Reliability Engineer to join our team and... ...automated, scalable, and resilient infrastructure in support... ...line support; responding to incidents quickly and appropriately... ...and documentation/knowledge management skills This role must...Full timeWork at office$250k - $350k
...quantitative researchers, engineers, traders, and... ..., uptime, and resilience of the trading... ...end users during incidents Monitor systems in... ...throughput, and reliability Qualifications Minimum... ...support, site reliability, or infrastructure... ...-on experience managing Kubernetes in a...Full time$152k - $205k
...a systems-minded engineer who is happiest when... ...you want to own reliability for a platform that... ...looking for a Senior Site Reliability... ...across observability, incident response, capacity... ...Build secure, resilient, and cost-efficient... ...shipping boring. Manage infrastructure through...Local areaRemote workWork from homeVisa sponsorship$136.2k - $214.01k
...prevent data loss, and build resilience across their people and AI... ...The Role As a Senior Site Reliability Engineer at Proofpoint you will... ...hours should any alerts or incidents arise. • Collaborate with... ...• Experience automating management and operational tasks using...Full timeFlexible hours- ...HudsonEmail: : Job Title: Site Reliability Engineer (Infrastructure & Systems)Location... ...automate network workflows.Incident Escalation: Serve as a Tier... ...(such as Kubernetes) in hybrid or public cloud settings.... .../ GCP).Familiarity with managed container orchestrators such...Local area
$212.5k - $275k
...of the Cboe Platform Engineering group focuses around building... ...production, improve resiliency of the service and... ...evolution. The Senior Manager of Platform... ...running workloads across hybrid infrastructure focused... ...ensuring scalability, reliability, and security best practices...- ...systems that must perform reliably under real-time market... ...highly collaborative, engineering-driven, and focused on... ...time Investigate incidents across distributed... ...processes and improve system resilience Partner closely... ...of experience in site reliability, systems engineering...
$190.8k - $267.1k
...the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure... ...also take ownership of risk management, ensuring the reliability... ...to enhance system resilience. Your expertise will drive... ...issues. Practice sustainable incident response, and drive structural...Work experience placementHome officeFlexible hours$160k - $210k
...Join our Platform Engineering team, where you'... ...engineers across reliability initiativesAnalyze... ...deployments, and incident responses with... ...provisioning, scaling, and management across all... ...in DevOps, Site Reliability Engineering... ..., we follow a hybrid work schedule: In...Work at officeWorldwideMonday to FridayFlexible hours$100k - $140k
...Department BSD CTD - Platform Engineering - GDC About the... ...automation, & AI/ML infrastructure management across the open-source... ..., and addressing production incidents. For monitoring, staff wrangle... ...incidents. CI/CD pipelines are for hybrid cloud architecture on-...Work experience placement$112.5k - $187.5k
...will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy,... ...consequential work on the platform. This is a hybrid position and involves regular... ...models calm, effective, and blameless incident response. Serves as a significant technical...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$160k - $200k
...ability to move, manage, and optimize liquidity... ...This is an engineering-first role with a... ...observability and reliability engineering work:... ...customers. The incident management program... ...~7+ years in Site Reliability Engineering... ...to build team resilience. Knowledge of...Full timeWork at officeLocal area$194k - $267k
...tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support... ...workloads, providing high resilience and operational efficiency.... ...with minimal downtime.Incident Management & Troubleshooting...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$60 - $68 per hour
...hiring for a Linux DevOps Engineer Position type... ...: Chicago, IL 60661 (Hybrid 3 days in office)... ...Tower . Develop, manage, and optimize Ansible... ...tasks and improve system reliability through scripting and... ...health, investigate incidents, and drive root cause...Hourly payContract workTemporary workFor contractorsWork experience placementWork at officeImmediate startWorldwideFlexible hours- ...Description As a Software Engineer – Software Reliability in our product team,... ...develop scalable, resilient systems using Python... ...Apply Site Reliability Engineering... ...Participate in post-incident reviews and drive root... ...processing and asset management. We offer a competitive...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer - Incident Management/Resiliency (Hybrid). Be the first to apply!
- site reliability engineer sre Chicago, IL
- site reliability engineer Chicago, IL
- site reliability engineer remote Chicago, IL
- IT site lead Chicago, IL
- site safety Chicago, IL
- website content developer Chicago, IL
- site leader Chicago, IL
- on-site clinical research associate (traveling/remote) Chicago, IL
- junior website developer Chicago, IL
- historic site Chicago, IL



