Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer III

Chase

Site Reliability Engineer III

There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.

As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team, you will solve complex and broad business problems with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions. You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform.

Job Responsibilities

  • Guides and assists others in the areas of building appropriate level designs and gaining consensus from peers where appropriate, supporting adoption of site reliability engineering best practices within your team, including modern technologies such as Databricks, Snowflake, AWS, and Kubernetes
  • Collaborates with other software engineers and teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines, while supporting application development and production environments
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements, including developing and supporting AI/ML solutions for incident resolution
  • Implements infrastructure, configuration, and network as code for the applications and platforms in your remit
  • Collaborates with technical experts, key stakeholders, and team members to resolve complex problems, coordinate incident management coverage, and proactively address issues using service level indicators and objectives before they impact customers
  • Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to SLO outcomes, and supporting intelligent automation for troubleshooting
  • Familiar with availability, reliability, scalability, and solutions in their applications and works with partners to improve these outcomes iteratively, including participation in root cause analysis and implementing production changes
  • Proactively recognizes road blocks and identifies improvements to solve business problems, including exploring new technologies where appropriate
  • Required qualifications, capabilities, and skills
  • Formal training or certification on site reliability engineering concepts and 3+ years applied experience (NAMR/APAC – India/ LATAM/ Hong Kong)
  • Proficient in site reliability culture and principles and familiarity with how to implement site reliability within an application or platform, including strong understanding of SLI/SLO/SLA and error budgets
  • Proficient in at least one programming language such as Python, Java/Spring Boot, and.Net, including Python or PySpark for AI/ML modeling and automation to reduce operational toil by building tools for repeated tasks
  • Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data sensitivity, including developing AI/ML solutions for troubleshooting and incident resolution
  • Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements, while ensuring compliance with risk controls and company-wide standards
  • Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., Cloud, AI, Android, etc.), including hands-on experience in system design, resiliency, testing, operational stability, and disaster recovery
  • Experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, and others
  • Experience with continuous integration and continuous delivery tooling, along with experience running production incident calls and managing incident resolution in collaboration with cross-functional teams
  • Familiarity with container and container orchestration and troubleshooting common networking technologies and issues, along with ability to work collaboratively in teams and build meaningful relationships to achieve common goals

Preferred qualifications, capabilities, and skills

  • 4+ years in an SRE or production support role with AWS Cloud, Databricks, Snowflake or similar Technologies.
  • AWS and Databricks certifications.

This position is subject to Section 19 of the Federal Deposit Insurance Act. As such, an employment offer for this position is contingent on JPMorganChase's review of criminal conviction history, including pretrial diversions or program entries

Vacancy posted 18 hours ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer III in Palo Alto, CA vacancy
  •  ...Job Title: Software Engineer III Location: Mountain View , CA, US, 2 days a week hybrid SDG Tuesday/Wed Top Skill Software Engineer Contractor, Python Qualifications: - A minimum of 5 years of professional experience specifically in... 
    Suggested
    For contractors
    Self employment
    Worldwide
    2 days per week

    Apex Informatics

    Mountain View, CA
    4 days ago
  •  ...everyone. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that...  ...and conducting employment, background and reference checks; (iii) establishing an employment relationship or entering into an... 
    Suggested
    Full time
    Contract work
    Local area

    Rivian and Volkswagen Group Technologies

    Palo Alto, CA
    1 day ago
  • $103.75k - $174.75k

     ...AI Engineer III New York, NY, United States Palo Alto, CA, United States (Hybrid) Job Description At American Express, our...  ...while meeting the high standards for security, explainability, reliability, and compliance required in financial services. We partner... 
    Suggested
    Full time
    Internship
    Shift work

    American Express

    Palo Alto, CA
    19 hours ago
  • $103.75k - $174.75k

     ...AI Engineer III - Agentic AI New York, NY, United States Phoenix, AZ, United States...  ...standards for security, explainability, reliability, and compliance required in financial services...  ...surrogacy ~ Free access to global on-site wellness centers staffed with nurses and... 
    Suggested
    Full time
    Internship
    Work at office
    Local area
    Remote work
    Visa sponsorship
    Flexible hours
    Shift work
    3 days per week

    American Express

    Palo Alto, CA
    4 days ago
  •  ...professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of...  ...professionals and position yourself among the top echelon in site reliability.  As a Senior Lead Site Reliability Engineer at JPMorgan Chase... 
    Suggested

    J.P. Morgan

    Palo Alto, CA
    5 days ago
  •  ...The Role We're looking for a Senior Site Reliability Engineer to own the reliability, scalability, and operational excellence of the production systems that power Nectar's platform. We run high-volume data ingestion pipelines and real-time AI agents on top of a fast... 
    Remote work

    Nectar Social

    Palo Alto, CA
    4 days ago
  • $189k - $232k

     ...Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck.... 
    Full time
    Work at office
    Local area
    2 days per week

    Earnin

    Mountain View, CA
    2 days ago
  • $137.77k - $194.59k

     ...distributed team of roughly 80 scientists and engineers building and operating Rubin's petascale...  ...Your role: \n You will own the reliability and robustness of Rubin Observatory's...  ...nature of this position, SLAC is open to on-site, hybrid, and remote work options. \n \... 
    Remote work
    Flexible hours
    Night shift

    Stanford University

    Menlo Park, CA
    19 hours ago
  • $170k - $250k

     ...Site Reliability Engineer (SRE) Location: San Francisco, CA / Palo Alto, CA Company Stage of Funding: Growth-Stage AI Infrastructure Company ($80M Raised) Office Type: Onsite (4 Days Per Week) Salary: $170,000–$250,000 + Competitive Equity We're representing a rapidly... 
    Work at office
    Visa sponsorship
    Flexible hours

    Recruiting from Scratch

    Palo Alto, CA
    1 day ago
  • $150k - $175k

     ...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed... 
    Remote work

    ASAPP

    Mountain View, CA
    1 day ago
  •  ...Senior Site Reliability Engineer Latitude AI is building the future of Ford's autonomy roadmap to make travel safer, less stressful, and more enjoyable for everyone. Bringing this vision to scale, our fully in-house developed hands-free ADAS platform will debut on Ford... 
    Work at office
    Immediate start

    Latitude AI

    Palo Alto, CA
    35 minutes ago
  • $165k - $190k

     ...Site Reliability Engineer Palo Alto, California, USA Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast... 
    Work from home
    Flexible hours

    Obsidian Security

    Palo Alto, CA
    2 days ago
  • $170k - $230k

     ...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community,... 
    Work at office
    Local area
    1 day per week

    Mithril

    Palo Alto, CA
    2 days ago
  • $100k - $200k

    OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about... 
    Full time

    OPPO

    Palo Alto, CA
    2 days ago
  • $151.6k - $245.3k

     ...Site Reliability Engineer Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture... 

    Palo Alto Networks

    Palo Alto, CA
    19 hours ago
  •  ...Lead Site Reliability Engineer Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within... 

    Chase

    Palo Alto, CA
    1 day ago
  • $200k - $260k

     ...Lead Site Reliability Engineer Glean is the Work AI platform that helps everyone work smarter with AI. What began as the industry's most advanced enterprise search has evolved into a full-scale Work AI ecosystem, powering intelligent Search, an AI Assistant, and scalable... 
    Work at office
    Home office
    Flexible hours

    Softbank Investment Advisers

    Mountain View, CA
    1 day ago
  • $217.57k - $260k

     ...job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA;...  ...which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure... 
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    Shift work

    ID.me

    Mountain View, CA
    19 hours ago
  • $252k - $308k

     ...Staff Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck... 
    Full time
    Work at office
    2 days per week

    Earnin

    Mountain View, CA
    4 days ago
  • $117k - $234k

     ...Position Summary... What you'll do... “Immigration sponsorship is not available for this role" The Software Engineer III will lead the development, and delivery of scalable software solutions aligned with business objectives. This role involves analyzing requirements... 
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    2 days ago
  •  ...Position: Software Engineer III Location: Cupertino, California Duration: Contract Job ID: 171424 Job Overview: We are seeking a highly skilled and experienced Software Engineer III to join our team in Cupertino, California. The ideal candidate will bring a strong technical... 
    Full time
    Contract work

    Pinnacle Group

    Cupertino, CA
    5 days ago
  • $70 - $74 per hour

     ...Akkodis is seeking a Software Engineer III for a Contract with a client in Cupertino, CA. The ideal candidate must have strong experience...  ...Kafka, Spark, or Flink. Optimize system performance and reliability while ensuring user privacy and data integrity. Contribute... 
    Hourly pay
    Contract work
    Temporary work
    Local area

    Akkodis

    Cupertino, CA
    2 days ago
  • $95 - $119 per hour

     ...About the Role We are seeking an experienced Backend Software Engineer to design and build scalable systems that ensure compliance of...  ...for large-scale data processing Ensure high performance, reliability, and scalability of backend services Contribute to... 
    Hourly pay
    Contract work

    US Tech Solutions

    Mountain View, CA
    2 days ago
  •  ...Software Developer III This candidate will be part of a team that is responsible for enabling web services. These web services...  ...like Jenkins Bachelor degree in Computer Science, Software Engineering or related field with 7+ years of professional software development... 

    InterSources

    Sunnyvale, CA
    1 hour ago
  • $172.53k - $201.38k

     ...explicitly states otherwise, all roles are on-site five days per week at one of our offices...  ...Overview ID.me is seeking a Software Engineer III to join the Trust Service team, where we...  ...that captures what was verified and how reliably it was verified. Every access decision... 
    Full time
    Contract work
    Temporary work
    Work at office
    Remote work
    Flexible hours

    ID.me

    Mountain View, CA
    11 days ago
  • $186k - $232.5k

     ...future that’s more connected, more intelligent, more sustainable for everyone. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that underpins the build pipelines that allow our companies to... 
    Full time
    Local area

    Rivian and Volkswagen Group Technologies

    Palo Alto, CA
    1 day ago
  •  ...Site Reliability Engineering Tech Lead Palo Alto, California, United States DataHub is an AI & Data Context Platform adopted by over 3,000 enterprises, including Apple, CVS Health, Netflix, and Visa. Innovated jointly with a thriving open-source community of 13,... 
    Remote work
    Home office
    Flexible hours

    Acryl Data, Inc.

    Palo Alto, CA
    2 days ago
  • $140k - $200k

     ...planet. We are a team of mission-driven engineers with experience across aerospace, robotics...  ...the systems that enable a robust, highly reliable data link between the remote pilot and aircraft...  ..., (ii) U.S. lawful permanent resident, (iii) refugee under 8 U.S.C. * 1157, or (iv)... 
    Permanent employment
    Remote work

    Reliable Robotics Corporation

    Mountain View, CA
    3 days ago
  •  ...Senior Site Reliability Engineer Location: Remote Duration: 12 month contract to start IV Process: 1-3 Round IV process International Tech Top Skills: Java Python NodeJS -DevOps Engineer should work here too Main Responsibilities:... 
    Contract work
    Local area
    Remote work

    My3Tech Inc

    Sunnyvale, CA
    3 days ago
  •  ...Site Reliability Engineer Forward is transforming how the world's most complex networks are managed and secured. Founded in 2013 by four Stanford Ph.D.s, we built the industry's first network digital twin — a mathematically precise model of the production network that... 
    Night shift

    Forward Networks Inc

    Santa Clara, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer III. Be the first to apply!