Site Reliability Engineer III
Chase
Site Reliability Engineer III
There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.
As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team, you will solve complex and broad business problems with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions. You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform.
Job Responsibilities
- Guides and assists others in the areas of building appropriate level designs and gaining consensus from peers where appropriate, supporting adoption of site reliability engineering best practices within your team, including modern technologies such as Databricks, Snowflake, AWS, and Kubernetes
- Collaborates with other software engineers and teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines, while supporting application development and production environments
- Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements, including developing and supporting AI/ML solutions for incident resolution
- Implements infrastructure, configuration, and network as code for the applications and platforms in your remit
- Collaborates with technical experts, key stakeholders, and team members to resolve complex problems, coordinate incident management coverage, and proactively address issues using service level indicators and objectives before they impact customers
- Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to SLO outcomes, and supporting intelligent automation for troubleshooting
- Familiar with availability, reliability, scalability, and solutions in their applications and works with partners to improve these outcomes iteratively, including participation in root cause analysis and implementing production changes
- Proactively recognizes road blocks and identifies improvements to solve business problems, including exploring new technologies where appropriate
- Required qualifications, capabilities, and skills
- Formal training or certification on site reliability engineering concepts and 3+ years applied experience (NAMR/APAC – India/ LATAM/ Hong Kong)
- Proficient in site reliability culture and principles and familiarity with how to implement site reliability within an application or platform, including strong understanding of SLI/SLO/SLA and error budgets
- Proficient in at least one programming language such as Python, Java/Spring Boot, and.Net, including Python or PySpark for AI/ML modeling and automation to reduce operational toil by building tools for repeated tasks
- Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data sensitivity, including developing AI/ML solutions for troubleshooting and incident resolution
- Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements, while ensuring compliance with risk controls and company-wide standards
- Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., Cloud, AI, Android, etc.), including hands-on experience in system design, resiliency, testing, operational stability, and disaster recovery
- Experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, and others
- Experience with continuous integration and continuous delivery tooling, along with experience running production incident calls and managing incident resolution in collaboration with cross-functional teams
- Familiarity with container and container orchestration and troubleshooting common networking technologies and issues, along with ability to work collaboratively in teams and build meaningful relationships to achieve common goals
Preferred qualifications, capabilities, and skills
- 4+ years in an SRE or production support role with AWS Cloud, Databricks, Snowflake or similar Technologies.
- AWS and Databricks certifications.
This position is subject to Section 19 of the Federal Deposit Insurance Act. As such, an employment offer for this position is contingent on JPMorganChase's review of criminal conviction history, including pretrial diversions or program entries
- ...Job Title: Software Engineer III Location: Mountain View , CA, US, 2 days a week hybrid SDG Tuesday/Wed Top Skill Software Engineer Contractor, Python Qualifications: - A minimum of 5 years of professional experience specifically in...SuggestedFor contractorsSelf employmentWorldwide2 days per week
- ...everyone. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that... ...and conducting employment, background and reference checks; (iii) establishing an employment relationship or entering into an...SuggestedFull timeContract workLocal area
$103.75k - $174.75k
...AI Engineer III New York, NY, United States Palo Alto, CA, United States (Hybrid) Job Description At American Express, our... ...while meeting the high standards for security, explainability, reliability, and compliance required in financial services. We partner...SuggestedFull timeInternshipShift work$103.75k - $174.75k
...AI Engineer III - Agentic AI New York, NY, United States Phoenix, AZ, United States... ...standards for security, explainability, reliability, and compliance required in financial services... ...surrogacy ~ Free access to global on-site wellness centers staffed with nurses and...SuggestedFull timeInternshipWork at officeLocal areaRemote workVisa sponsorshipFlexible hoursShift work3 days per week- ...professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of... ...professionals and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase...Suggested
- ...The Role We're looking for a Senior Site Reliability Engineer to own the reliability, scalability, and operational excellence of the production systems that power Nectar's platform. We run high-volume data ingestion pipelines and real-time AI agents on top of a fast...Remote work
$189k - $232k
...Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck....Full timeWork at officeLocal area2 days per week$137.77k - $194.59k
...distributed team of roughly 80 scientists and engineers building and operating Rubin's petascale... ...Your role: \n You will own the reliability and robustness of Rubin Observatory's... ...nature of this position, SLAC is open to on-site, hybrid, and remote work options. \n \...Remote workFlexible hoursNight shift$170k - $250k
...Site Reliability Engineer (SRE) Location: San Francisco, CA / Palo Alto, CA Company Stage of Funding: Growth-Stage AI Infrastructure Company ($80M Raised) Office Type: Onsite (4 Days Per Week) Salary: $170,000–$250,000 + Competitive Equity We're representing a rapidly...Work at officeVisa sponsorshipFlexible hours$150k - $175k
...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed...Remote work- ...Senior Site Reliability Engineer Latitude AI is building the future of Ford's autonomy roadmap to make travel safer, less stressful, and more enjoyable for everyone. Bringing this vision to scale, our fully in-house developed hands-free ADAS platform will debut on Ford...Work at officeImmediate start
$165k - $190k
...Site Reliability Engineer Palo Alto, California, USA Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast...Work from homeFlexible hours$170k - $230k
...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community,...Work at officeLocal area1 day per week$100k - $200k
OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about...Full time$151.6k - $245.3k
...Site Reliability Engineer Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture...- ...Lead Site Reliability Engineer Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within...
$200k - $260k
...Lead Site Reliability Engineer Glean is the Work AI platform that helps everyone work smarter with AI. What began as the industry's most advanced enterprise search has evolved into a full-scale Work AI ecosystem, powering intelligent Search, an AI Assistant, and scalable...Work at officeHome officeFlexible hours$217.57k - $260k
...job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA;... ...which can be found here. Role Overview The Staff Site Reliability Engineer, Infrastructure role is building a high-scale infrastructure...Full timeTemporary workWork at officeRemote workFlexible hoursShift work$252k - $308k
...Staff Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck...Full timeWork at office2 days per week$117k - $234k
...Position Summary... What you'll do... “Immigration sponsorship is not available for this role" The Software Engineer III will lead the development, and delivery of scalable software solutions aligned with business objectives. This role involves analyzing requirements...Full timeTemporary workPart time- ...Position: Software Engineer III Location: Cupertino, California Duration: Contract Job ID: 171424 Job Overview: We are seeking a highly skilled and experienced Software Engineer III to join our team in Cupertino, California. The ideal candidate will bring a strong technical...Full timeContract work
$70 - $74 per hour
...Akkodis is seeking a Software Engineer III for a Contract with a client in Cupertino, CA. The ideal candidate must have strong experience... ...Kafka, Spark, or Flink. Optimize system performance and reliability while ensuring user privacy and data integrity. Contribute...Hourly payContract workTemporary workLocal area$95 - $119 per hour
...About the Role We are seeking an experienced Backend Software Engineer to design and build scalable systems that ensure compliance of... ...for large-scale data processing Ensure high performance, reliability, and scalability of backend services Contribute to...Hourly payContract work- ...Software Developer III This candidate will be part of a team that is responsible for enabling web services. These web services... ...like Jenkins Bachelor degree in Computer Science, Software Engineering or related field with 7+ years of professional software development...
$172.53k - $201.38k
...explicitly states otherwise, all roles are on-site five days per week at one of our offices... ...Overview ID.me is seeking a Software Engineer III to join the Trust Service team, where we... ...that captures what was verified and how reliably it was verified. Every access decision...Full timeContract workTemporary workWork at officeRemote workFlexible hours$186k - $232.5k
...future that’s more connected, more intelligent, more sustainable for everyone. Role Summary We are seeking an experienced Site Reliability Engineer to help design, build, and operate the infrastructure that underpins the build pipelines that allow our companies to...Full timeLocal area- ...Site Reliability Engineering Tech Lead Palo Alto, California, United States DataHub is an AI & Data Context Platform adopted by over 3,000 enterprises, including Apple, CVS Health, Netflix, and Visa. Innovated jointly with a thriving open-source community of 13,...Remote workHome officeFlexible hours
$140k - $200k
...planet. We are a team of mission-driven engineers with experience across aerospace, robotics... ...the systems that enable a robust, highly reliable data link between the remote pilot and aircraft... ..., (ii) U.S. lawful permanent resident, (iii) refugee under 8 U.S.C. * 1157, or (iv)...Permanent employmentRemote work- ...Senior Site Reliability Engineer Location: Remote Duration: 12 month contract to start IV Process: 1-3 Round IV process International Tech Top Skills: Java Python NodeJS -DevOps Engineer should work here too Main Responsibilities:...Contract workLocal areaRemote work
- ...Site Reliability Engineer Forward is transforming how the world's most complex networks are managed and secured. Founded in 2013 by four Stanford Ph.D.s, we built the industry's first network digital twin — a mathematically precise model of the production network that...Night shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer III. Be the first to apply!
- site reliability engineer Palo Alto, CA
- site reliability engineer sre Palo Alto, CA
- IT site lead Palo Alto, CA
- junior website developer Palo Alto, CA
- site safety Palo Alto, CA
- site services specialist Palo Alto, CA
- site recruiter Palo Alto, CA
- website content developer Palo Alto, CA
- site leader Palo Alto, CA
- on-site clinical research associate (traveling/remote) Palo Alto, CA


