Reliability Engineer
$98.18k - $115.5kU.S. Bank
At U.S. Bank, we’re on a journey to do our best. Helping the customers and businesses we serve to make better and smarter financial decisions, enabling the communities we support to grow and succeed in the right ways, all more confidently and more often—that’s what we call the courage to thrive. We believe it takes all of us to bring our shared ambition to life, and each person is unique in their potential. A career with U.S. Bank gives you a wide, ever-growing range of opportunities to discover what makes you thrive. Try new things, learn new skills and discover what you excel at—all from Day One.As a wholly owned subsidiary of U.S. Bank, Elavon is committed to building the platforms and ecosystems that help over 1.5 million customers around the world to achieve their financial goals—no matter what they need. From transaction processing to customer service, to driving innovation and launching new products, we’re building a range of tailored payment solutions powered by the latest technology. As part of our team, you can explore what motivates and energizes your career goals: partnering with our customers, our communities, and each other.Job DescriptionResponsibilitiesThe Reliability Observability Engineer 3 is responsible for enabling reliable, measurable, and supportable application operations across a broad portfolio of production applications. This role helps ensure application teams have actionable visibility into service health, customer experience, dependencies, performance, and failure conditions through effective use of metrics, logs, traces, dashboards, alerts, synthetic monitoring, and service-level indicators.As a senior-level Reliability Engineer specializing in observability, this role partners closely with product owners, application engineering teams, SRE teams, and business stakeholders to translate customer journeys and business outcomes into measurable reliability objectives. The role establishes and maintains best-practice processes for documenting, governing, reviewing, and improving user journeys, SLIs, SLOs, synthetic monitoring, dashboards, alerts, telemetry standards, and related observability assets. The engineer provides senior technical guidance, identifies observability gaps through incident and performance analysis, drives continuous improvement, and helps ensure teams have the data, processes, and operating discipline needed to detect issues earlier, reduce customer impact, and improve overall service reliability.Lead the definition, documentation, implementation, and continuous improvement of Observability across Critical Customer Journeys, ensuring alignment between Observability Strategy, business outcomes, and reliability objectives.Design, implement, and govern Service Level Indicators (SLIs), Service Level Objectives (SLOs), Error Budgets, and reliability metrics for enterprise applications and services.Establish and maintain Observability Governance Frameworks for Dashboards, Alerts, Synthetic Monitoring, Telemetry Standards, and lifecycle management of observability assets.Translate business and technical requirements into scalable Observability Architectures, including Instrumentation Standards, Monitoring Strategies, Tagging Frameworks, and Alerting Models.Partner with Product Owners, Application Engineering, Site Reliability Engineering (SRE), and Operations Teams to ensure applications are production-ready and fully instrumented for reliability measurement.Develop and maintain executive and operational Service Health Dashboards that provide insights into Availability, Latency, Customer Impact, Dependency Performance, and SLO Compliance.Analyze Telemetry Data, Incident Trends, Problem Records, and Alert Performance to identify observability gaps, reduce alert fatigue, and improve detection accuracy.Provide technical leadership and mentorship on Distributed Tracing, Logging, Metrics Collection, Synthetic Testing, Application Performance Monitoring, Real User Monitoring,Monitoring Design Patterns, and Alert Governance Best Practices.Lead the definition, documentation, and ongoing refinement of critical user journeys in partnership with product owners, engineering teams, SRE, operations, and business stakeholders to ensure observability practices are aligned to customer experience, business outcomes, and operational risk.Define, document, and govern appropriate service-level indicators and service-level objectives for applications and key capabilities, including availability, latency, error rate, throughput, dependency health, and other measurements that reflect meaningful customer and business impact.Establish and maintain a best-practice process for identifying, approving, implementing, reviewing, and retiring user journeys, SLIs, SLOs, dashboards, monitors, synthetic tests, alerts, and related observability artifacts.Translate product and engineering requirements into actionable observability designs that specify telemetry needs, measurement methods, tagging standards, dashboard requirements, alerting thresholds, ownership, evidence expectations, and operational runbook linkages.Partner with product and engineering teams during design, build, release, and production-readiness activities to ensure applications are instrumented to validate critical customer journeys, measure reliability outcomes, and support effective incident detection and triage.Develop, maintain, and continuously improve dashboards and reporting that communicate service health, SLO performance, error-budget posture, customer impact, dependency performance, alert effectiveness, and trends to technical teams and leadership stakeholders.Analyze telemetry, incidents, problem records, alert history, customer-impacting events, and SLO performance trends to identify observability gaps, reduce alert noise, improve detection accuracy, and recommend reliability improvements.Provide senior-level guidance, coaching, and standards interpretation to engineering, SRE, and operations teams on observability design patterns, SLI/SLO selection, customer journey monitoring, synthetic monitoring, logging, tracing, metrics, and alert governance.Maintain an authoritative inventory of observability assets, including user journeys, SLIs, SLOs, dashboards, monitors, alerts, synthetic tests, ownership assignments, review cadence, and evidence of ongoing compliance with approved observability standards.Basic Qualifications- Bachelor's degree, or equivalent work experience- Five to seven years of relevant work experience in business and risk analysis, IT Service Management, production support, product/project management, or application developmentPreferred Skills/ExperienceExpertise in Observability Engineering, Site Reliability Engineering (SRE), or Reliability Engineering.Strong knowledge of SLIs, SLOs, Error Budgets, and Customer Journey Monitoring.Demonstrated ability to understand stakeholder needs and guide the development of reliability requirements for large, complex multi-system products.Hands-on experience with APM, RUM, synthetics, monitoring, logging, tracing, and telemetry frameworks.Proficiency with Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry.Experience building, standardizing, and tuning operational dashboards and actionable alertsthat communicate service health, customer impact, dependency health, performance trends, failure conditions, severity, ownership, routing, and runbook linkage.Strong understanding of distributed systems, microservices, cloud platforms, and Kubernetes.Ability to leverage incident analysis, RCA, and performance data to drive reliability improvements.Excellent stakeholder management, communication, and technical leadership skills.Location expectationsThis role requires working from a U.S. Bank location three (3) or more days per week.If there’s anything we can do to accommodate a disability during any portion of the application or hiring process, please refer to ourdisability accommodations for applicants.Benefits:Our approach to benefits and total rewards considers our team members’ whole selves and what may be needed to thrive in and outside work. That's why our benefits are designed to help you and your family boost your health, protect your financial security and give you peace of mind. Our benefits include the following:Healthcare (medical, dental, vision)Basic term and optional term life insuranceShort-term and long-term disabilityPregnancy disability and parental leave401(k) and employer-funded retirement planPaid vacation (from two to five weeks depending on salary grade and tenure)Up to 11 paid holiday opportunitiesAdoption assistanceSick and Safe Leave accruals of one hour for every 30 worked, up to 80 hours per calendar year unless otherwise provided by lawReview our full benefits available by employment status here. U.S. Bank is an equal opportunity employer. We consider all qualified applicants without regard to race, religion, color, sex, national origin, age, sexual orientation, gender identity, disability or veteran status, and other factors protected under applicable law.E-VerifyU.S. Bank participates in the U.S. Department of Homeland Security E-Verify program in all facilities located in the United States and certain U.S. territories. The E-Verify program is an Internet-based employment eligibility verification system operated by the U.S. Citizenship and Immigration Services. Learn more about theE-Verify program.The salary range reflects figures based on the primary location, which is listed first. The actual range for the role may differ based on the location of the role. In addition to salary, U.S. Bank offers a comprehensive benefits package, including incentive and recognition programs, equity stock purchase 401(k) contribution and pension (all benefits are subject to eligibility requirements). Pay Range: $98,175.00 - $115,500.00U.S. Bank will consider qualified applicants with arrest or conviction records for employment. U.S. Bank conducts background checks consistent with applicable local laws, including the Los Angeles County Fair Chance Ordinance and the California Fair Chance Act as well as the San Francisco Fair Chance Ordinance. U.S. Bank is subject to, and conducts background checks consistent with the requirements of Section 19 of the Federal Deposit Insurance Act (FDIA). In addition, certain positions may also be subject to the requirements of FINRA, NMLS registration, Reg Z, Reg G, OFAC, the NFA, the FCPA, the Bank Secrecy Act, the SAFE Act, and/or federal guidelines applicable to an agreement, such as those related to ethics, safety, or operational procedures.Applicants must be able to comply with U.S. Bank policies and procedures including the Code of Ethics and Business Conduct and related workplace conduct and safety policies.Posting may be closed earlier due to high volume of applicants.SummaryLocation: Brookfield, WI; Atlanta, GA; Hopkins, MN; Cupertino, CA; Charlotte, NC; New York, NY; Chicago, IL; Gresham, OR; Englewood, CO; Cincinnati, OH; Irving, TX; Earth City, MOType: Full time
$225.5k - $338.5k
...passionate and growing team?-?we’d?love to have you apply!? About the Role: Join our dynamic Quality & Reliability Organization as a Principal Reliability Engineer, a pivotal role responsible for defining, driving, and ensuring the highest standards of product reliability...SuggestedLocal area$332k
...fueled by great technology—and amazing people.NVIDIA's hardware reliability is foundational to some of the world's most demanding... ...Director of Board and System Level Reliability, you will set the engineering standard for how NVIDIA's products perform and thrive across...SuggestedFull timeWork at office- ...meeting the growing demand for environmentally conscious production methods. Job Summary: The primary purpose of the Reliability Engineer (focused on L10/L11 execution) is to own the physical, environmental, and mechanical stress-testing validation of integrated...SuggestedWork at officeLocal area
$100k - $136.5k
...AreApplied Materials is the global leader in materials science and engineering solutions that are at the foundation of virtually every new... ...damage or injuryAssist in preparing & presenting safety and reliability reports;Assist with root cause analysis of field failures and...SuggestedFull time$150k - $230k
...excellence has earned us several prestigious awards, such as Best Engineering Team, Best Company for Diversity, Compensation, and Work-Life... ...you'll work with We are seeking an Optical Transceiver Reliability and Qualification Engineer to lead the reliability testing...SuggestedFull timeContract work$50 per hour
...next-generation space systems. As part of our central specialty engineering team, your technical insight won't just analyze risks—it will... ...in the product lifecycle, your next mission starts here As a Reliability Engineer, you will:• Drive Mission Success: Support...Full timeTemporary workWork experience placementCasual workFlexible hours- ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available database and data platforms. The role combines database engineering, site reliability engineering, Linux systems administration...
$40 - $55 per hour
...traceability and purity of biological substances and products. Job Description Eurofins E&E is seeking a Senior Reliability Test Engineer – Environmental Simulation Lab to perform, document, and oversee product testing in accordance with national and...Hourly payFull timeContract workWork at officeMonday to Friday- Get AI-powered advice on this job and more exclusive features. We are looking for Reliability Test Engineer based at Cupertino, CA (Onsite) | Fulltime position. If interested, please share your resume at ****@*****.*** or you can reach me at (***) ***-**** . Role...Full timeAfternoon shift
$134.9k - $185k
...an inventive research and development company that designs and engineers high-profile electronic devices. Amazon devices started with... ...produced devices like Fire tablets, Fire TV, and Amazon Echo. Amazon reliability team aims to develop reliable and robust products that delight...Local areaFlexible hours- ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation...
- ...in Cupertino, California, invites an experienced CDN Solutions Engineer to join the Content Delivery Network Solutions team. You will... ...and collaborate with engineering groups across Apple to ensure reliable delivery at scale. The ideal candidate has 4+ years in CDNs and...
- ...ServiceNow in Santa Clara, CA, seeks a Staff Software Engineer – SRE & AIOps to drive infrastructure automation, resilience, and toil... ...remediation for global engineering teams. Embedded within the Site Reliability & Database Engineering organization, you will architect SRE...
$276.1k - $311.4k
...Vehicle Software SRE team from the ground up — defining its charter, hiring its founding engineers, establishing the operating model, and creating the technical strategy that makes reliability a first-class property of the software running on our vehicles. You'll work in a...Permanent employmentFull timeWork at officeWork from home- ...function to support one of the world’s fastest-growing AI inference services, powered by the Wafer-Scale Engine (WSE). This team will help deliver world-class, ultra-reliable inference infrastructure for leading model builders such as OpenAI and other frontier labs.As a...Shift work
- ...Overview Title: Site Reliability Engineer SRE – ML platform Location: Austin, TX or Sunnyvale, CA Employment type: Full-time • Seniority: Mid-Senior level • ONLY W2 Responsibilities Continuous Deployment using GitHub Actions, Flux, Kustomize Design and implement cloud...Full time
- ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable...
$170k - $200k
...We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance...Full timeWorldwide$248k - $396.75k
Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production systems with exceptional efficiency, resilience, and availability. It combines software and systems engineering practices with...Full time$133.5k - $183.5k
Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced... ...our winning team that is focused on Quality Engineering and Reliability Engineering. We are working with engineers and scientists across...Full timeWorldwide$148k - $235.75k
...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer...Full time- ...design by customizing MES tool per business needs Education Requirements, Ideal Experience: Associate’s degree in Industrial Engineering or IT related field Minimum of 0-3 years’ relevant experience Experience in C#, Delphi desired Knowledge of the...Work at office
- ...of Huobi globe spanning infrastructure. • Work with engineering teams to make sure new features and changes are deployed quickly... .... • Constantly improve our system performance and reliability through better tools, process and monitoring system. •...Worldwide
$101k - $161k
...excellence has earned us several prestigious awards, such as Best Engineering Team, Best Company for Diversity, Compensation, and Work-Life... ...do.Job DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s CloudVision-as-a-...$131.1k - $222.76k
...Job Summary:We are now looking for a Principal Reliability Engineer. The position is an individual contributor role, working on NPD Reliability Engineering & qualifications, and MP change qualifications as needed.Job Description & Responsibilities:For NPD projects, the...Local area$167.3k - $284.4k
...invest 15% of sales back into R&D. Our expert teams of physicists, engineers, data scientists and problem-solvers work together with the... ...process. Job Description/Preferred Qualifications Sr. Reliability Engineer - SEM Systems Join a team solving complex reliability...Minimum wageWork experience placementFlexible hours$95.91k - $144.6k
...Senior Reliability EngineerThe annual base pay range for this position is $95,913 - $144,599. Our salary ranges are determined by role... ...predict reliability in the field. Actively manage reliability engineering projects in support of product and process development...Work experience placement$91.6k - $234k
...Reliability EngineerWe are looking for a Reliability Engineer with a strong background in high voltage systems to join our reliability team. In this role, you will play a key role in designing reliability into our ground-breaking vehicle products. As a successful candidate...Hourly payFull timeFlexible hours$190k - $222.5k
...Senior Reliability EngineerWe're ALSO, an electric mobility company originally conceived as a part of Rivian. We're a passionate team of... ...0-50x more efficient.ALSO is looking for a Senior Reliability Engineer to play a key role in developing and leading the reliability of...Work at officeLocal areaRemote work1 day per week- ...Join to apply for the Robotics Reliability Engineer role at Matic Robots . Company Overview Each year, 2.5 trillion hours are spent on household chores. At Matic, we’re on a mission to recapture that lost time, and we’re doing it by revolutionizing home robotics. Our...Full timeWork experience placementWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Reliability Engineer. Be the first to apply!


