Site Reliability Engineer III- Production Management
JPMorgan Chase
There’s nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment Bank, Production management team, you will solve complex and broad business problems with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions. You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform.Job ResponsibilitiesGuides and assists others in the areas of building appropriate level designs and gaining consensus from peers where appropriate, supporting adoption of site reliability engineering best practices within your teamCollaborates with other software engineers and teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelinesUses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.Continuously monitor operational signals (alerts, dashboards, service health indicators, and desk/user feedback) and drive rapid engagement, ownership, and coordination when issues emerge.Enhance production observability and reliability by improving instrumentation, dashboards, and alert quality; use trend and incident data to reduce noise, prevent recurrence, and measurably lower MTTR and incident frequency.Own and improve the support operating model: maintain/run runbooks, ensure clean handoffs and escalation paths, manage shift coverage, and drive disciplined post-incident follow-through (actions, owners, due dates).Partner with engineering and adjacent service owners to drive root-cause fixes and stronger change/release standards, identify cross-service dependencies early, and build lightweight automation/self-service tooling to reduce manual steps and speed repeatable triage.Lead L1/L2 production support using SRE practices: quickly triage incidents, investigate likely causes, apply safe mitigations/workarounds, and provide timely, clear stakeholder updates through to resolution.Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to SLO outcomes.Required qualifications, capabilities, and skillsFormal training or certification on site reliability engineering concepts and 3+ years applied experience Proficient in site reliability culture and principles and familiarity with how to implement site reliability within an application or platformProficient in at least one programming language such as Python, Java/Spring Boot, and .NetWorking knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data sensitivityAbility to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirementsProficient knowledge of software applications and technical processes within a given technical discipline (e.g., Cloud, AI, Android, etc.)Practical cloud experience in AWS supporting production services (visibility/troubleshooting, deployments, and operational hygiene)Ability to write code to automate support work (e.g., Python, Java, or similar) and reduce operational toilExperience operating in time-sensitive environments (market hours or similar), with strong incident discipline and stakeholder communicationStrong debugging fundamentals across distributed systems (logs/metrics, latency analysis, dependency failures, data issues)Preferred qualifications, capabilities, and skillsPrior Markets experience, especially pricing / risk / market data support (Fixed Income a plus)Experience with Datadog (metrics/logs/traces), alert tuning, dashboards, and basic SLO conceptsFamiliarity with messaging/streaming patterns (Kafka/MQ) and data quality checks in execution and pricing pipelinesExperience applying AI-assisted tooling to reduce support toil JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world’s most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management. We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process. We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.JPMorgan Chase & Co. is an Equal Opportunity Employer, including Disability/VeteransJ.P. Morgan’s Commercial & Investment Bank is a global leader across banking, markets, securities services and payments. Corporations, governments and institutions throughout the world entrust us with their business in more than 100 countries. The Commercial & Investment Bank provides strategic advice, raises capital, manages risk and extends liquidity in markets around the world. Full timePosting Date: 2026-07-14
- ...recognized firm, driven by pride in ownership.As a Senior Manager of Site Reliability Engineering at JPMorgan Chase within the Corporate Investment Bank... ...sensitive data.Provide North America leadership for production management teams supporting trading desks across...SuggestedBank staffShift work
- ...direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Commercial & Investment Bank, Production Management team, you hold a leadership role in your team, demonstrate strong...Suggested
$120k - $165k
...the future of our communities.This is a Lead Software Production Management & Reliability Engineering position at Director level which is part of the job family... ...OverviewThe Wealth Management Production Management Site Reliability Engineer position is a highly visible/...SuggestedTemporary workWork at office- ...impact. As a Software Engineer III at JPMorgan Chase within... ...focused on specific products and projects, providing... ...our systems are reliable and easy to operate.Keep... ...processing and asset management. We offer a competitive... ...health care coverage, on-site health and wellness...Suggested
- ...take your software engineering career to the next... ...Software Engineer III at JPMorgan Chase... ...leading technology products in a secure,... ...code quality and reliability.Leverages enterprise... ...Git source control management.Experience with modularization... ...care coverage, on-site health and...Suggested
- ...to take your software engineering career to the next level... ...a Software Engineer III at JPMorganChase within... ...-leading technology products in a secure, stable, and... ...processing and asset management. We offer a competitive... ...care coverage, on-site health and wellness centers...
- ...to take your software engineering career to the next level... ...a Software Engineer III at JPMorgan Chase within... ...-leading technology products that are secure, stable... ...Typescript and "state management" Hands on... ...health care coverage, on-site health and wellness centers...Full time
$138.1k - $198.2k
...that simply works. The SRE Engineering Enablement Team supports... ...tools, code review, artifact management, CI, education, and documentation... ...at Cisco. Your Impact As a Site Reliability Engineer, you will be at... ...on and build great products Lead the design and platform...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$158.5k - $172k
...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will... ...ecosystems. Our team is responsible for managing our centralized Enterprise Logging... ...high-impact position driving continuous reliability, deep system optimization, and automation...Full timeWork at office3 days per week$136k - $253k
...of the RoleAdvanced Content Engineering (ACE) is seeking a Staff Software... ...pipeline and search index management systems — from the Kafka-... ...exception, and constant delivery to production is the expectation. This is... ...APIs, while maintaining reliable, observable, and performant...Full timeTemporary workWork at officeLocal areaFlexible hours2 days per week3 days per week$150k - $250k
...We DoAt Goldman Sachs, our Engineers don't just make things - we... ...Banking & Markets business, the Site Reliability Engineering (SRE) team... ...codebases, and raise the bar for production-quality automation across... ...at every layer.Establish and manage SLIs, SLOs, and error budgets...Full timeTemporary workPart time$195k - $275k
...investment banking, securities, investment management and wealth management services. The... ...of the technical solutions behind the products and services used by the Morgan Stanley... ...Management, and the Chief Operating Office.The Reliability Operations (RO) within WMT is...Temporary workWork at officeWorldwideNight shift$150k - $160k
Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech... ...for Vue 3 applications and WordPress.Manage GCP infrastructure (Cloud Run, GKE) to... ...unrelenting focus on the quality of the products and the people. The philosophy has...Work at officeLocal area$150k - $190k
Senior Site Reliability Engineer, VPAt Morgan Stanley, we advise, originate, trade, manage and distribute capital for governments, institutions and individuals, and always... ...financial IT community. The position in the WM Product Technology team is focused on delivering...Temporary workWorldwideFlexible hoursWeekend work- ..., and deliver top-notch technology products.As a Senior Lead Software Engineer - Precious Metals / Front Office /... ...global priorities. Acts as the direct manager for the NYC based software... ...comprehensive health care coverage, on-site health and wellness centers, a retirement...For contractorsWork at office
$160k - $240k
Senior Software Engineer - Service Mesh Security and Configuration Location... ...of virtually all Bloomberg products and services, responsible for reliably handling billions of dollars of transactions... ...of this, the BAS Enterprise team manages critical infrastructure that...Temporary workFor contractorsWork experience placementFlexible hours$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you... ...opportunity to support mission‑critical production systems while collaborating with... ...error budget practices to proactively manage reliability. Identify capacity constraints...- ...Karsun Solutions in the DMV area is seeking a Site Reliability Manager to ensure reliability, scalability, and performance of our systems. You will lead a team focusing on Application Reliability, DevSecOps, and Platform Lifecycle Management. The ideal candidate has 1...
- ...customers who rely on us for production AI workloads, including... ...medalists, and experienced engineering and product leaders with decades... ...company, we seek to improve our reliability dramatically while scaling... ...with auto scaling, fleet management, and capacity planning at scale...
$141.8k - $195k
...the choice, control, and flexibility to manage and analyze telemetry for both humans... ...Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock... ...create, deploy, test, and ship Cribl products. Not often do you get to be part of something...Temporary workRemote work- ...Job DescriptionSenior DevOps Engineer OverviewLead the design and... ...infrastructure provisioning and management using advanced tools like... ...and Kubernetes.Ensure reliability, scalability, and observability... ...AI systems and workloads in production environments.Implement security...
- ...the full AI lifecycle, from development through production. Trusted by Woven by Toyota, AXA, UiPath,... ...raised $60M in Series C funding from Wellington Management, CRV, Next47 and Y Combinator. The role As a Solutions Engineer at Encord, you will be the core technical expert...Work at office
$190k - $259k
...Viam is the engineering platform for robotics and automation, empowering... ...and scale—from prototype to production. Our platform makes robotics... ..., observable, and manageable at scale. About the Team... ...fleets of machines to be managed reliably at scale. As a Senior or...Full timeWork at office3 days per week$150k - $225k
...platform. Our flagship product, the Intelligent... ...simplify complex wealth management and accounting, foster... ...seeking a Senior Software Engineer to report to our Lead... ...office workflows into reliable software Contribute to... ...Frequent company off-sites and team-building events...Work at office- ...CPG brands from deduction to production plan. We unify cash... ..., disputes, trade promotion management, forecasting, demand planning... ...momentum. As a Senior Software Engineer (Platform/Core) , you will help... ...operational workflows into reliable, scalable software. Location...Local areaRelocation packageNight shift
$230k - $350k
...operating layer for wealth management. Their platform combines generative... ...processes, enhance advisor productivity, and deliver better client... .... As they scale their engineering team, they are looking for... ...systems for performance, reliability, latency, and cost efficiencyContribute...Permanent employmentWork at office3 days per week$180k - $230k
...critical role in ensuring that both production and development environments... ...smoothly, securely, and reliably. This role leverages advanced... ...curious MLOps/DevOps Engineers with deep expertise in machine... ...supporting infrastructure.Build and manage cloud‑native infrastructure...Full timeWork at officeRemote work$120k - $142k
...Overview & Responsibilities: At The New York Times, our Site Reliability Engineering (SRE) team is central to how we design, test, and operate... ...customer experiences. We're looking for a Technical Product Manager to lead the strategy for reliability programs and platforms...Local areaFlexible hours$207k - $301k
...team of SREs. Drive technical execution, task management, and project delivery.Drive service reliability and performance, including end-to-end... ...critical systems in close partnership with Product Management and Engineering teams.Lead the team in building automation to...$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied... ...integrity of high-visibility products and platforms and the environments they... ...privilege/RBAC, deploy approvals, secrets management) in partnership with security and risk...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer III- Production Management. Be the first to apply!
- site reliability engineer New York, NY
- site reliability engineer sre New York, NY
- site reliability engineer remote New York, NY
- after school site coordinator New York, NY
- site services specialist New York, NY
- construction site safety New York, NY
- site merchandiser New York, NY
- site leader New York, NY
- official site New York, NY
- website content developer New York, NY


