Average salary: $135,000 /yearly
More statsGet new jobs by email
- ...rated for Part-Time Encompass Health: Where Nursing Meets Heart, Home, and Healing Are you seeking a nursing career deeply rooted in purpose, close to your heart and home? Encompass Health offers a transformative journey where your expertise as a Registered Nurse...SuggestedFull timePart timeRelocation packageFlexible hoursNight shift
- ...Iowa to inspect product, process and material across all product lines. You will partner with production to contain issues, determine root cause, and drive corrective actions while following IATF 16949 procedures. You may be asked to work flexible shifts to ensure...SuggestedFlexible hoursDay shift
$18.5 - $25.35 per hour
...the Express and Bonobos brands worldwide. About Express Express is a multichannel apparel brand dedicated to a design philosophy rooted in modern, confident and effortless style whether dressing for work, everyday or special occasions. Since its launch in 1980, the brand...SuggestedHourly payFull timePart timeWorldwideNight shift- ...development of all employees. Maintain documentation for workplace inspections, risk assessments, and task training. Assist management with Root Cause Analysis and incident investigations. Work with the Site/Production Manager on best practices and provide coverage when...SuggestedFor contractorsShift workNight shiftWeekend work
$21 - $25 per hour
...all proceeds back to the park beautification efforts. A permanent stand was eventually built…and the rest is Shack history! With our roots in fine dining and giving back to the community, we are committed to high quality food served with a high level of hospitality. Our...SuggestedHourly payWeekly payPermanent employmentTemporary workLocal areaFlexible hoursShift workNight shift- ...with leading retailers. With a team of over 150 employees and more than 365 live events annually, we continue to expand while staying rooted in community. At the center of it all is Chuck the GOAT, a symbol of self-belief, positivity, and showing up as the greatest...SuggestedFull timeCasual workShift workWeekend workDay shift
$80k - $140k
...Lead incident and problem management - Troubleshoot production issues across all layers, participate in on-call rotation, and drive root cause analysis and corrective actions. Drive continuous improvement - Identify opportunities to simplify, automate, and modernize...SuggestedFlexible hoursShift work$60 - $70 per hour
...deployment patterns, operational procedures, and troubleshooting guidance. Participate in production support, incident response, and root-cause analysis as appropriate. Independently own technical work and drive complex problems through resolution with limited...SuggestedHourly payContract workRemote work- ...improvement, service resilience, and uptime management. Incident Management & RCA - Major Incident Management (MIM), outage tracking, root cause analysis, problem management, and MTTR reduction. Failure Analysis & Risk Assessment - FMEA, risk quantification,...SuggestedFor contractors
$70 per hour
...applications using Dynatrace. Configure dashboards, alerts, and observability solutions. Investigate production issues and perform root cause analysis. Support application, API, integration, and infrastructure monitoring. Respond to incidents and participate in an on-...SuggestedContract workImmediate startRemote work$100k - $120k
...Origami's time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying preventative measures to minimize future disruptions. They also...SuggestedFull timeTemporary workWork experience placementRemote workFlexible hours- ...Incident Management & Support - Implement and enhance system monitoring, alerting, and observability. - Lead incident response, root cause analysis, and postmortem reviews. - Drive continuous improvement and oversee break/fix operations. Security & Compliance...Suggested
- ...development lifecycle. Support containerized workloads and cloud-native applications. Troubleshoot production issues, perform root cause analysis, and implement long-term reliability improvements. Optimize deployment strategies, release automation, and...Suggested3 days per week
- ...Implement and enhance monitoring, alerting, and observability frameworks for proactive issue detection. Lead incident response, root cause analysis (RCA), and postmortem reviews. Drive continuous improvement by identifying systemic issues and implementing...Suggested
- ...applications, infrastructure, and cloud services interact in production; demonstrated ability to troubleshoot production issues, perform root cause analysis, and drive long-term reliability improvements. ~-Proficiency with Python, Bash, or similar scripting languages;...Suggested
$114.3k - $235.32k
...Building and supporting CI/CD automation and deployment workflows using GitHub Actions Participating in incident response, root cause analysis, and post-incident improvement initiatives Reducing operational toil through scripting, tooling, and process automation...Work at officeLocal areaRemote workRelocationRelocation package$20 per hour
...perform performance tuning and implement long-term stability improvements. Respond to and resolve production incidents; perform root cause analysis and drive corrective actions through blameless postmortems. Rotate through the team's on-call schedule to keep...Permanent employmentFull timeImmediate startRelocation packageFlexible hoursWeekend work- ...support for production applications and platforms. Monitor system health, troubleshoot incidents, restore service, and perform root cause analysis. Participate in on-call support and production change activities. Platform Support (Kubernetes & Kafka)...Remote work
- ...~8 Required Experience defining and managing SLIs, SLOs, and error budgets ~8 Required Familiarity with incident management, root cause analysis (RCA), and postmortems ~8 Required Experience integrating security and compliance into operational workflows...Contract workLocal areaRemote work
- ...resiliency, performance, and operational excellence. Define and drive adoption of SLOs, SLIs, Error Budgets, Incident Management, and Root Cause Analysis. Drive automation initiatives that reduce operational overhead and improve service reliability. Serve as...Contract work
- ...Kubernetes clusters (pods, resources, scaling behavior) Linux OS tuning (CPU, memory, I/O, ulimits, networking) Identify root causes and propose clear, actionable engineering solutions. Distributed Systems Design Design, review, and influence high-performance...Long term contractContract work
- ...and peak business events. Implement and improve monitoring, observability, alerting, and service health dashboards. Lead Root Cause Analysis (RCA) and problem management activities. Identify reliability risks and implement proactive solutions to improve...Long term contract3 days per week
$65.12 per hour
...protocols. Implement and enhance monitoring, alerting, and observability for proactive detection. Lead incident response, root cause analysis, and postmortems with corrective actions. Oversee break/fix operations to ensure timely resolution and low business...Contract work3 days per week- ...and operational tools through secure conversational interfaces. Design intelligent remediation workflows for incident detection, root cause analysis, log analysis, and operational troubleshooting. Develop secure Infrastructure-as-Code automation using Terraform...Contract work
- ...tasks, incident reduction, and proactive reliability improvements. Monitor platform health, troubleshoot complex issues, and lead root cause analysis efforts to minimize downtime and improve system resiliency. Collaborate closely with engineering, platform, and...Remote work
$70 - $80 per hour
...they impact users. Respond to production incidents, participate in on-call rotations, and lead post-incident reviews to drive root cause analysis and reliability improvements. Collaborate with software engineering and security teams to ensure new services and...Hourly payContract work- ...and repetitive processes using scripting and Infrastructure as Code (IaC). Lead incident response activities, troubleshooting, root cause analysis (RCA), and post-incident reviews. Collaborate with development, infrastructure, and platform teams to improve system...
- ...including metrics, logging, tracing, alerting, and production diagnostics. Experience with production troubleshooting, incident response, root-cause analysis, and operational readiness. Experience with Terraform or another Infrastructure as Code technology. Experience...Long term contractLocal areaImmediate startRelocation
- ...development productivity, troubleshooting, documentation, and automation. Participate in production support, incident response, root cause analysis, and continuous improvement initiatives. Required Qualifications ~8+ years of experience in Site...Contract workLocal areaRemote work
- ...security to translate business needs into technical designs that align with organizational goals and societal impact. Perform root cause analysis for incidents using observability data and logs from Splunk dynatrace and other tools to reduce recurrence and improve...Permanent employmentContract workWork experience placementRemote work