Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Reliability Engineer 3 (Observability Specialist)

$98.18k - $115.5k

U.S. Bank

At U.S. Bank, we’re on a journey to do our best. Helping the customers and businesses we serve to make better and smarter financial decisions, enabling the communities we support to grow and succeed in the right ways, all more confidently and more often—that’s what we call the courage to thrive. We believe it takes all of us to bring our shared ambition to life, and each person is unique in their potential. A career with U.S. Bank gives you a wide, ever-growing range of opportunities to discover what makes you thrive. Try new things, learn new skills and discover what you excel at—all from Day One. As a wholly owned subsidiary of U.S. Bank, Elavon is committed to building the platforms and ecosystems that help over 1.5 million customers around the world to achieve their financial goals—no matter what they need. From transaction processing to customer service, to driving innovation and launching new products, we’re building a range of tailored payment solutions powered by the latest technology. As part of our team, you can explore what motivates and energizes your career goals: partnering with our customers, our communities, and each other. Job DescriptionResponsibilitiesThe Reliability Observability Engineer 3 is responsible for enabling reliable, measurable, and supportable application operations across a broad portfolio of production applications. This role helps ensure application teams have actionable visibility into service health, customer experience, dependencies, performance, and failure conditions through effective use of metrics, logs, traces, dashboards, alerts, synthetic monitoring, and service-level indicators.As a senior-level Reliability Engineer specializing in observability, this role partners closely with product owners, application engineering teams, SRE teams, and business stakeholders to translate customer journeys and business outcomes into measurable reliability objectives. The role establishes and maintains best-practice processes for documenting, governing, reviewing, and improving user journeys, SLIs, SLOs, synthetic monitoring, dashboards, alerts, telemetry standards, and related observability assets. The engineer provides senior technical guidance, identifies observability gaps through incident and performance analysis, drives continuous improvement, and helps ensure teams have the data, processes, and operating discipline needed to detect issues earlier, reduce customer impact, and improve overall service reliability.Lead the definition, documentation, implementation, and continuous improvement of Observability across Critical Customer Journeys, ensuring alignment between Observability Strategy, business outcomes, and reliability objectives.Design, implement, and govern Service Level Indicators (SLIs), Service Level Objectives (SLOs), Error Budgets, and reliability metrics for enterprise applications and services.Establish and maintain Observability Governance Frameworks for Dashboards, Alerts, Synthetic Monitoring, Telemetry Standards, and lifecycle management of observability assets.Translate business and technical requirements into scalable Observability Architectures, including Instrumentation Standards, Monitoring Strategies, Tagging Frameworks, and Alerting Models.Partner with Product Owners, Application Engineering, Site Reliability Engineering (SRE), and Operations Teams to ensure applications are production-ready and fully instrumented for reliability measurement.Develop and maintain executive and operational Service Health Dashboards that provide insights into Availability, Latency, Customer Impact, Dependency Performance, and SLO Compliance.Analyze Telemetry Data, Incident Trends, Problem Records, and Alert Performance to identify observability gaps, reduce alert fatigue, and improve detection accuracy.Provide technical leadership and mentorship on Distributed Tracing, Logging, Metrics Collection, Synthetic Testing, Application Performance Monitoring, Real User Monitoring,Monitoring Design Patterns, and Alert Governance Best Practices.Lead the definition, documentation, and ongoing refinement of critical user journeys in partnership with product owners, engineering teams, SRE, operations, and business stakeholders to ensure observability practices are aligned to customer experience, business outcomes, and operational risk.Define, document, and govern appropriate service-level indicators and service-level objectives for applications and key capabilities, including availability, latency, error rate, throughput, dependency health, and other measurements that reflect meaningful customer and business impact.Establish and maintain a best-practice process for identifying, approving, implementing, reviewing, and retiring user journeys, SLIs, SLOs, dashboards, monitors, synthetic tests, alerts, and related observability artifacts.Translate product and engineering requirements into actionable observability designs that specify telemetry needs, measurement methods, tagging standards, dashboard requirements, alerting thresholds, ownership, evidence expectations, and operational runbook linkages.Partner with product and engineering teams during design, build, release, and production-readiness activities to ensure applications are instrumented to validate critical customer journeys, measure reliability outcomes, and support effective incident detection and triage.Develop, maintain, and continuously improve dashboards and reporting that communicate service health, SLO performance, error-budget posture, customer impact, dependency performance, alert effectiveness, and trends to technical teams and leadership stakeholders.Analyze telemetry, incidents, problem records, alert history, customer-impacting events, and SLO performance trends to identify observability gaps, reduce alert noise, improve detection accuracy, and recommend reliability improvements.Provide senior-level guidance, coaching, and standards interpretation to engineering, SRE, and operations teams on observability design patterns, SLI/SLO selection, customer journey monitoring, synthetic monitoring, logging, tracing, metrics, and alert governance.Maintain an authoritative inventory of observability assets, including user journeys, SLIs, SLOs, dashboards, monitors, alerts, synthetic tests, ownership assignments, review cadence, and evidence of ongoing compliance with approved observability standards.Basic Qualifications- Bachelor's degree, or equivalent work experience- Five to seven years of relevant work experience in business and risk analysis, IT Service Management, production support, product/project management, or application developmentPreferred Skills/ExperienceExpertise in Observability Engineering, Site Reliability Engineering (SRE), or Reliability Engineering.Strong knowledge of SLIs, SLOs, Error Budgets, and Customer Journey Monitoring.Demonstrated ability to understand stakeholder needs and guide the development of reliability requirements for large, complex multi-system products.Hands-on experience with APM, RUM, synthetics, monitoring, logging, tracing, and telemetry frameworks.Proficiency with Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, or OpenTelemetry.Experience building, standardizing, and tuning operational dashboards and actionable alerts that communicate service health, customer impact, dependency health, performance trends, failure conditions, severity, ownership, routing, and runbook linkage.Strong understanding of distributed systems, microservices, cloud platforms, and Kubernetes.Ability to leverage incident analysis, RCA, and performance data to drive reliability improvements.Excellent stakeholder management, communication, and technical leadership skills.Location expectationsThis role requires working from a U.S. Bank location three (3) or more days per week.If there’s anything we can do to accommodate a disability during any portion of the application or hiring process, please refer to our disability accommodations for applicants.Benefits:Our approach to benefits and total rewards considers our team members’ whole selves and what may be needed to thrive in and outside work. That's why our benefits are designed to help you and your family boost your health, protect your financial security and give you peace of mind. Our benefits include the following:Healthcare (medical, dental, vision)Basic term and optional term life insuranceShort-term and long-term disabilityPregnancy disability and parental leave401(k) and employer-funded retirement planPaid vacation (from two to five weeks depending on salary grade and tenure)Up to 11 paid holiday opportunitiesAdoption assistanceSick and Safe Leave accruals of one hour for every 30 worked, up to 80 hours per calendar year unless otherwise provided by lawReview our full benefits available by employment status here. U.S. Bank is an equal opportunity employer. We consider all qualified applicants without regard to race, religion, color, sex, national origin, age, sexual orientation, gender identity, disability or veteran status, and other factors protected under applicable law.E-VerifyU.S. Bank participates in the U.S. Department of Homeland Security E-Verify program in all facilities located in the United States and certain U.S. territories. The E-Verify program is an Internet-based employment eligibility verification system operated by the U.S. Citizenship and Immigration Services. Learn more about the E-Verify program.The salary range reflects figures based on the primary location, which is listed first. The actual range for the role may differ based on the location of the role. In addition to salary, U.S. Bank offers a comprehensive benefits package, including incentive and recognition programs, equity stock purchase 401(k) contribution and pension (all benefits are subject to eligibility requirements). Pay Range: $98,175.00 - $115,500.00U.S. Bank will consider qualified applicants with arrest or conviction records for employment. U.S. Bank conducts background checks consistent with applicable local laws, including the Los Angeles County Fair Chance Ordinance and the California Fair Chance Act as well as the San Francisco Fair Chance Ordinance. U.S. Bank is subject to, and conducts background checks consistent with the requirements of Section 19 of the Federal Deposit Insurance Act (FDIA). In addition, certain positions may also be subject to the requirements of FINRA, NMLS registration, Reg Z, Reg G, OFAC, the NFA, the FCPA, the Bank Secrecy Act, the SAFE Act, and/or federal guidelines applicable to an agreement, such as those related to ethics, safety, or operational procedures.Applicants must be able to comply with U.S. Bank policies and procedures including the Code of Ethics and Business Conduct and related workplace conduct and safety policies.Posting may be closed earlier due to high volume of applicants.SummaryLocation: Brookfield, WI; Atlanta, GA; Hopkins, MN; Cupertino, CA; Charlotte, NC; New York, NY; Chicago, IL; Gresham, OR; Englewood, CO; Cincinnati, OH; Irving, TX; Earth City, MOType: Full time

Vacancy posted 16 hours ago
Similar jobs that could be interesting for youBased on the Reliability Engineer 3 (Observability Specialist) in Cupertino, CA vacancy
  • $114k - $171k

     ...enabling solutions for global security. Our Engineering and Sciences (E&S) organization pushes...  ...naval surface ships and submarines. The Reliability and System Safety Engineering Department...  .../Systems Safety Engineer Level 3 /4 we are seeking will be an individual... 
    Suggested
    Full time
    Work experience placement
    Relocation package
    Shift work

    Northrop Grumman

    Sunnyvale, CA
    2 days ago
  • $148k - $235.75k

     ....Join our team of innovative engineers who are building an AI Data Center...  ..., high-volume telemetry into reliable, job-centric insights and...  ...of reliability for an observability/AIOps platform: SLOs/SLIs, on...  ...USD - 235,750 USD for Level 3, and 176,000 USD - 276,000 USD... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $120k - $180k

     ...distributed systems, processing almost 3 trillion events per day and...  ...the Role:At CrowdStrike, our engineering organization depends on...  ...ownership to operate reliably, scale safely, harden for security...  ...multi-cloud environmentsBuild observability - Implement metrics,... 
    Suggested
    Full time
    Work experience placement
    Work at office
    Local area

    CrowdStrike

    Sunnyvale, CA
    4 days ago
  • $169k - $338k

     ...Summary...As a Distinguished AI/ML Engineer within Walmart Global Tech's Site Reliability Engineering organization, you...  ...-store systems.Build intelligent observability and monitoring systems using ML-driven...  ...for all engineering teams, 3) Build tools/automate to prevent... 
    Suggested
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    2 days ago
  • $83.4k - $125.2k

     ...provider of mission-enabling solutions for global security. Our Engineering and Sciences (E&S) organization pushes the boundaries of innovation...  ...Northrop Grumman is seeking Mechanical Design Engineers Level 2/3 to join our team based out of Sunnyvale, CA. The Marine Systems... 
    Suggested
    Full time
    Work experience placement
    Remote work
    Relocation package
    Shift work

    Northrop Grumman

    Sunnyvale, CA
    3 days ago
  • $91.8k - $137.6k

     ...part of history, they're making history.We are looking for you to join our Launcher Electronics Group as an Electrical Design Engineer Level 2/ 3 for the development and ongoing support of new and existing electrical designs used in surface ship and submarine launch... 
    Full time
    Work experience placement
    Work at office
    Relocation package
    Shift work

    Northrop Grumman

    Sunnyvale, CA
    21 hours ago
  • $150k - $195k

     ...growing, and we are looking for engineers with passion for automation....  ..., workload management, observability, and storage services. We build...  ...improve the scalability and reliability of internal processes. Participate...  .... Minimum Qualifications 3 years of Devops/SRE experience... 
    Full time
    Worldwide

    Fortinet, Inc.

    Sunnyvale, CA
    21 hours ago
  • $136k - $218.5k

    We are seeking Systems Quality and Reliability Engineer to join our LPU team!NVIDIA has continuously reinvented itself over two decades. Our invention...  ...The base salary range is 136,000 USD - 218,500 USD for Level 3, and 168,000 USD - 264,500 USD for Level 4.You will also be... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ...Reliability Engineer Location: Cupertino, CA Job Summary: The candidate will work with an innovative Reliability Engineering team in the hardware...  ...and specifications Education: Engineering degree and 3+ years of experience in a related field or industry required.... 

    Yantran LLC

    Cupertino, CA
    4 days ago
  •  ...Reliability EngineerThe primary purpose of the Reliability Engineer (focused on L10/L11 execution) is to own the physical, environmental, and mechanical stress-testing...  ...survival gatekeeper — subjecting fully assembled, 3,000+ lb liquid-cooled AI clusters to extreme accelerated... 
    Work at office
    Local area

    Foxconn Technology Group

    Santa Clara, CA
    3 days ago
  •  ...Reliability Engineering Support – Eco Softgoods RELLocation: Cupertino, CA Onsite/Remote: OnsiteSkills Qualifications: BS or MS in Mechanical, Electrical...  ..., Automation, Materials Science, Electrical, etc.). 3+ years in a reliability, test, or quality engineering role Experience... 
    Remote work

    Yantran LLC

    Cupertino, CA
    21 hours ago
  • $91.8k - $137.6k

     ...Systems sector is seeking an Aeronautical/Mechanical Engineer2/3 to join a highly talented, motivated, and collaborative Performance...  ...for Level 2: · Bachelor’s degree in Aeronautical/Aerospace Engineering, Mechanical Engineering, Structural Engineering, Physics, or a closely... 
    Full time
    Relocation package
    Shift work

    Northrop Grumman

    Sunnyvale, CA
    3 days ago
  • $134.9k - $185k

     ...and development company that designs and engineers high-profile electronic devices. Amazon devices...  ..., Fire TV, and Amazon Echo. Amazon reliability team aims to develop reliable and robust...  ...in mechanical engineering or equivalent- 3+ years of relevant experience.- Experience... 
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    2 days ago
  • $136k - $218.5k

     ...Embedded markets. As a Silicon Speed Features Engineer, you will co-design system-level speed...  ..., hardware, firmware/software, process/reliability, and operations teams to co-design system...  ...range is 136,000 USD - 218,500 USD for Level 3, and 168,000 USD - 264,500 USD for Level... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $114.1k - $268.18k

     ...is currently seeking a Lead Specialist, Cloud Security to join our Managed...  ...between cybersecurity, engineering, infrastructure, application,...  ...available for H-1B, L-1, TN, O-1, E-3, H-1B1, F-1, J-1, OPT, CPT...  ...a calendar of holidays to be observed during the year and provides... 
    H1b
    Local area

    KPMG

    Santa Clara, CA
    2 days ago
  • $27 - $63 per hour

     ...Esthetician to join our team as a Treatment Specialist (Receptionist). This role is expected to...  ...20, Santa Clara, CA 95054 Schedule: 2-3 shifts a week, minimum 4 weekend days a...  ...Participate in hands-on training and observation of treatment protocols for development purposes... 
    Hourly pay
    Part time
    Shift work
    Weekend work

    OrangeTwist

    Santa Clara, CA
    1 day ago
  •  ...Construction Safety Specialist We are hiring a Construction Safety...  ...term position for a duration of 3.5 months....  ...contractors. Provide daily observations of on-site safety practices....  ...safety, occupational health, engineering, or related degree is preferred... 
    Full time
    For contractors
    Fixed term contract
    For subcontractor
    Casual work
    Work at office
    Local area
    Shift work

    Erm LLC

    Sunnyvale, CA
    1 day ago
  •  ...on this job and more exclusive features. We are looking for Reliability Test Engineer based at Cupertino, CA (Onsite) | Fulltime position. If interested...  ...– Cupertino, CA (Onsite with work swing shift.(12-9PM or 3-11PM) ) Job Description: 3-4 years of hands-on experience... 
    Full time
    Afternoon shift

    RHP Soft Inc

    Cupertino, CA
    3 days ago
  • $114k - $171k

     ...provider of mission-enabling solutions for global security. Our Engineering and Sciences (E&S) organization pushes the boundaries of innovation...  ...and submarines. We are seeking a LabVIEW Systems Engineer Level 3 with experience in LabVIEW, data acquisition, and signal... 
    Full time
    Relocation package
    Shift work

    Northrop Grumman

    Sunnyvale, CA
    16 hours ago
  • $150k - $250k

     ...firm specializing in aerospace solutions is seeking a Board Reliability Simulation Engineer to lead reliability analysis and simulation for avionics...  ...Bachelor’s or Master’s degree in relevant fields and at least 3 years of experience in reliability engineering.... 

    E-Space

    Saratoga, CA
    4 days ago
  • 3 days ago Be among the first 25 applicants If interested or know someone who...  ...****@*****.***. Title: Reliability Engineering Support Position Type: Full Time Permanent...  ...supplier technologies. The Engineering Specialist (ES) will provide technical expertise... 
    Permanent employment
    Full time

    Confidencial

    Cupertino, CA
    3 days ago
  • $83.4k - $125.2k

     ...provider of mission-enabling solutions for global security. Our Engineering and Sciences (E&S) organization pushes the boundaries of innovation...  ...(three shifts available) Mechanical Production Engineer Level 2/3 to participate in the development and ongoing support of new... 
    Full time
    Relocation package
    Monday to Thursday
    Shift work
    Night shift
    Day shift

    Northrop Grumman

    Sunnyvale, CA
    2 days ago
  •  ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available database...  ..., Terraform, service discovery, and secrets management. ~3+ years of Linux systems engineering experience, including... 

    Neshent Technologies

    Santa Clara, CA
    4 days ago
  •  ...Child Life Specialist II, Certified Per Diem Primary Location Santa Clara, California Facility...  .../or staff on services, as required; and observing and identifying changes in patients...  ...Minimum Qualifications: Minimum three (3) years of experience in a CLS I or equivalent... 
    Daily paid
    Work experience placement
    Work at office
    Shift work

    Kaiser Permanente

    Santa Clara, CA
    1 day ago
  • $230k - $250k

     ...foundation for autonomous networking, giving engineers and AI agents the ability to know the...  ...done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep...  ...how we think about availability, observability, incident response, and operational excellence... 
    Night shift

    Forward Networks

    Santa Clara, CA
    21 hours ago
  • $170k - $200k

    We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining...  ...(Linux servers, network devices, databases etc,.) Improve observability with logging, monitoring, alerting, and tracing tools (e.g... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    2 days ago
  •  ..., and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure...  ...patterns with precision and cost-efficiency.Advanced Observability: Define our monitoring and alerting philosophy using New Relic... 
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    21 hours ago
  • $128k - $216k

     ...another millions of times a day - quickly, reliably, and securely. Any time you swipe your...  ...a successful Senior Site Reliability Engineer do at Fiserv?As a Senior Site...  ...manage complex RDBMS/Document storage.Observability: Implement advanced monitoring and tracing... 
    Full time
    Worldwide

    Fiserv

    Sunnyvale, CA
    16 hours ago
  • $160k - $240k

     ...another millions of times a day - quickly, reliably, and securely. Any time you swipe your...  ...does a successful Site Reliability Engineer do at Fiserv?You will join our global team...  ...and alerting systems to ensure strong observability across services.Participate in on-call... 
    Full time

    Fiserv

    Sunnyvale, CA
    4 days ago
  • $168k - $270.25k

     ...artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing,...  ...experience, MS with 5+ years of experience or Ph.D. with 3 years of experience.8+ years of experience crafting technology... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Reliability Engineer 3 (Observability Specialist). Be the first to apply!