Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer - Production Management

$61k - $101k
Full-time

J.P. Morgan

Salary: $61,000 - 101,000 per year Requirements:

  • Formal training or certification in site reliability engineering, plus 3+ years of applied experience
  • Strong grasp of site reliability culture and principles, with the ability to apply SRE practices within an application or platform
  • Proficiency in at least one programming language such as Python, Java/Spring Boot, or .NET
  • Working knowledge of enterprise-authorized AI capabilities in the workplace to support SRE workflows, with disciplined validation habits and awareness of data sensitivity
  • Ability to verify AI-assisted operational recommendations before making changes, escalate when uncertain, and follow data sensitivity requirements
  • Solid understanding of software applications and technical processes within a technical discipline such as cloud, AI, or Android
  • Practical AWS experience supporting production services, including visibility, troubleshooting, deployments, and operational hygiene
  • Ability to write code to automate support tasks and reduce operational toil
  • Experience working in time-sensitive environments such as market hours, with strong incident discipline and stakeholder communication
  • Strong debugging fundamentals across distributed systems, including logs and metrics analysis, latency investigation, dependency failures, and data issues
  • Preferred: prior markets experience, especially pricing, risk, or market data support; Fixed Income experience is a plus
  • Preferred: experience with Datadog metrics, logs, traces, alert tuning, dashboards, and basic SLO concepts
  • Preferred: familiarity with messaging and streaming patterns such as Kafka or MQ, plus data quality checks in execution and pricing pipelines
  • Preferred: experience applying AI-assisted tooling to reduce support toil
Responsibilities:
  • Guide and support peers in building appropriate design approaches and reaching alignment, while promoting site reliability engineering best practices across the team
  • Work with software engineers and partner teams to design, develop, test, and implement deployment and reliability solutions using automated CI/CD pipelines
  • Use enterprise-authorized AI capabilities to speed up incident triage, troubleshooting, and post-incident analysis while validating outputs and handling data according to security requirements
  • Continuously monitor operational signals such as alerts, dashboards, service health indicators, and user feedback, and drive rapid response, ownership, and coordination when issues arise
  • Improve production observability and reliability by enhancing instrumentation, dashboards, and alert quality; use trend and incident data to reduce noise, prevent repeat incidents, and lower MTTR and incident frequency
  • Own and refine the support operating model by maintaining runbooks, ensuring clear handoffs and escalation paths, managing shift coverage, and driving disciplined post-incident follow-up
  • Partner with engineering and other service owners to implement root-cause fixes, strengthen change and release standards, identify cross-service dependencies early, and build lightweight automation or self-service tools to reduce manual work and speed up repeatable triage
  • Lead L1 and L2 production support using SRE practices: triage incidents quickly, investigate likely causes, apply safe mitigations or workarounds, and provide timely stakeholder updates through resolution
  • Apply enterprise-authorized AI capabilities to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements linked to SLO outcomes
Technologies:
  • AI
  • AWS
  • Android
  • CI/CD
  • Cloud
  • Datadog
  • Support
  • Java
  • Kafka
  • Python
  • Security
  • Spring
  • Spring Boot
  • ASP.NET

More:

hackajob is partnering with J.P. Morgan to connect exceptional professionals with this Site Reliability Engineer III opportunity in the Commercial & Investment Bank, Production management team. We are modernizing complex, mission-critical systems through code, cloud infrastructure, and strong operational discipline, with a focus on availability, reliability, scalability, and continuous improvement. JPMorganChase is one of the worlds oldest financial institutions, serving consumers, businesses, and major corporate, institutional, and government clients across more than 100 countries. We offer a competitive total rewards package, with base salary based on role, experience, skills, and location, and eligible roles may include commission or discretionary incentive compensation. Our benefits can include comprehensive health care, wellness centers, retirement savings, backup childcare, tuition reimbursement, mental health support, and financial coaching. We value diversity and inclusion, provide reasonable accommodations where needed, and operate as an equal opportunity employer, including Disability/Veterans. The role sits within J.P. Morgans Commercial & Investment Bank, a global leader in banking, markets, securities services, and payments.

last updated 37 week of 2026

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer - Production Management in New York, NY vacancy
  •  ...innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment Bank, Production management team, you will solve complex and broad business problems with simple and... 
    Suggested
    Shift work

    JP Morgan Chase

    New York, NY
    5 days ago
  •  ...direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Commercial & Investment Bank, Production Management team, you hold a leadership role in your team, demonstrate strong... 
    Suggested

    JP Morgan Chase

    New York, NY
    2 days ago
  •  ...TekWissen is a global workforce management provider headquartered in Ann Arbor, Michigan...  ...clients world-wide. Job Title: SRE Engineer Job Type: Temporary Assignment...  ...Months Job Description: ~ Production Support background with Financial... 
    Suggested
    Temporary work

    TekWissen LLC

    New York, NY
    3 days ago
  • $153k - $210k

     ...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid...  ...opportunity to support mission-critical production systems while collaborating with talented...  ...budget practices to proactively manage reliability. Identify capacity constraints... 
    Suggested
    Full time

    Ridgeline

    New York, NY
    1 day ago
  • $130k - $250k

     ...We DoAt Goldman Sachs, our Engineers don't just make things - we...  ...journey here.Securities Frontline Site Reliability Engineers (SREs) play a...  ...environment. The Frontline SRE Production Engineering team owns the...  ...eliminate operational toil, and manage robust UAT and Production... 
    Suggested
    Full time
    Temporary work
    Part time
    Immediate start

    Goldman Sachs

    New York, NY
    3 days ago
  •  ...who specialize in creativity, management, and more. Our operating principles...  ...! About the Team Clay's product is increasingly powered by AI...  ...-adjacent teammates, and other engineers to make sure agents aren't just capable, but reliable, steerable, and worth trusting... 
    Full time

    Clay Labs

    New York, NY
    1 day ago
  • $10k

     ...entrepreneurship and innovation worldwide. Our Product Bubble is the only fully visual AI...  ...-grade hosting, security, database management, and automatic scaling that grows with...  ...the team: We are expanding our AI Engineering Team to build next-generation AI-... 
    Full time
    Work at office
    Worldwide
    Flexible hours

    Bubble

    New York, NY
    1 day ago
  • $158.5k - $172k

     ...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will...  ...ecosystems. Our team is responsible for managing our centralized Enterprise Logging...  ...high-impact position driving continuous reliability, deep system optimization, and automation... 
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    New York, NY
    4 days ago
  •  ...Platform Cloud Nerve Center. This leader will establish the production environment, engineering patterns, and operational capabilities needed to...  ...engineering outcomes.Cloud security and edge: Lead traffic management, web application protection, private connectivity,... 
    Worldwide
    Flexible hours

    The Bank of New York Mellon

    New York, NY
    1 day ago
  • $136k - $253k

     ...of the RoleAdvanced Content Engineering (ACE) is seeking a Staff Software...  ...pipeline and search index management systems — from the Kafka-...  ...exception, and constant delivery to production is the expectation. This is...  ...APIs, while maintaining reliable, observable, and performant... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    2 days per week
    3 days per week

    Thomson Reuters

    New York, NY
    6 days ago
  • $150k - $220k

     ...way. The Role: As an engineering organization, we pride...  ...creative activity. Engineering managers enable engineers to do their...  ..., and purpose. The Manager, Site Reliability Engineering will lead Forge’...  ..., Security, Compliance, and Product teams to improve reliability... 
    Local area

    Forge Global

    New York, NY
    3 days ago
  • $150k - $250k

     ...We DoAt Goldman Sachs, our Engineers don't just make things - we...  ...Banking & Markets business, the Site Reliability Engineering (SRE) team...  ...codebases, and raise the bar for production-quality automation across...  ...at every layer.Establish and manage SLIs, SLOs, and error budgets... 
    Full time
    Temporary work
    Part time

    Goldman Sachs

    New York, NY
    3 days ago
  • $190k - $259k

     ...companies to configure, control, and manage physical devices—from a...  ...in New York City. The Engineering Challenge Building software...  ...sits at the intersection of product engineering and platform infrastructure...  ...of machines to be managed reliably at scale. As a Senior or... 
    Full time
    Work at office
    3 days per week

    Viam, Inc.

    New York, NY
    1 day ago
  •  ...Join us as a Python engineer to collaborate with Market Risk...  ...you to transform MVPs into reliable production solutions and support workflows...  ...will support JPMorgan's Risk Managers in understanding firmwide exposures...  ...health care coverage, on-site health and wellness centers,... 
    Full time

    JPMorgan Chase & Co.

    New York, NY
    1 day ago
  •  ...Senior Platform Engineer Nomic is the domain-specific AI platform...  ...make our agents run well in production: orchestrating rollouts...  ...keeping inference fast and reliable, and maintaining industry-leading...  ...disaster recovery, and cost management. Security posture. Access... 
    Remote work
    Flexible hours

    Nomic

    New York, NY
    4 days ago
  • $150k - $160k

    Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech...  ...for Vue 3 applications and WordPress.Manage GCP infrastructure (Cloud Run, GKE) to...  ...unrelenting focus on the quality of the products and the people. The philosophy has... 
    Work at office
    Local area

    Haymarket Media Group

    New York, NY
    4 days ago
  •  ...Senior Site Reliability Engineer (SRE) Plenful is hiring a Senior Site Reliability Engineer (SRE) to keep our production systems reliable, performant, and scalable as we grow. This role is centered on operating real systems at scale — not just building infrastructure... 
    Full time
    Work at office
    Remote work
    Flexible hours
    2 days per week

    Plenful

    New York, NY
    13 hours ago
  • $195k - $275k

     ...investment banking, securities, investment management and wealth management services. The...  ...of the technical solutions behind the products and services used by the Morgan Stanley...  ...Management, and the Chief Operating Office.The Reliability Operations (RO) within WMT is... 
    Temporary work
    Work at office
    Worldwide
    Night shift

    Morgan Stanley

    New York, NY
    2 days ago
  • $150k - $190k

    Senior Site Reliability Engineer, VPAt Morgan Stanley, we advise, originate, trade, manage and distribute capital for governments, institutions and individuals, and always...  ...financial IT community. The position in the WM Product Technology team is focused on delivering... 
    Temporary work
    Worldwide
    Flexible hours
    Weekend work

    Morgan Stanley

    New York, NY
    4 days ago
  •  ...environments — not prototypes, not demos. We need engineer's who can own the entire agent stack: a production frontend, a robust backend, a properly secured API...  ...or Node.js backend that orchestrates agent logic, manages state, and exposes clean APIs; and containerized,... 
    Full time
    Temporary work

    Trilagen

    New York, NY
    1 day ago
  • $150k - $225k

     ...platform. Our flagship product, the Intelligent...  ...simplify complex wealth management and accounting, foster...  ...seeking a Senior Software Engineer to report to our Lead...  ...office workflows into reliable software Contribute to...  ...Frequent company off-sites and team-building events... 
    Work at office

    Asseta Inc

    New York, NY
    3 days ago
  •  ...The Opportunity Join our engineering team as a Full Stack...  ...handle sensitive medical data reliably, scale gracefully under load...  ...medical imaging data at scale Production-grade CI/CD pipelines with...  ...clinical workflows and data management What You Bring ~3+ years... 
    Worldwide

    NUC S.A.I.

    New York, NY
    4 days ago
  •  ...Senior Software Engineer (Platform/Core)Confido is the AI infrastructure...  ...brands from deduction to production plan. We unify cash...  ..., disputes, trade promotion management, forecasting, demand planning...  ...operational workflows into reliable, scalable software.Location:... 
    Local area
    Relocation package

    Confido

    New York, NY
    13 hours ago
  • $10k

     ...innovation worldwide. Our Product Bubble is the only fully...  ...hosting, security, database management, and automatic scaling that...  ...team executes on our scaling, reliability, and enablement efforts to...  ...is reshaping the traditional engineering stack for the entire world.... 
    Full time
    Currently hiring
    Work at office
    Remote work
    Worldwide
    Relocation
    Relocation package

    Bubble

    New York, NY
    1 day ago
  •  ...Software EngineerAs a Software Engineer at Wealth.com, you will help...  ...tax planning, and AI-driven products used by tens of thousands of...  ...with enterprise wealth management platformsTake features from...  ...decisions around scalability, reliability, security, and performanceParticipate... 
    Temporary work
    Work at office
    Flexible hours
    Shift work

    Wealth LLC

    New York, NY
    5 days ago
  •  ...pairing end-to-end workforce management technology with a rapidly...  ...Why this role exists The Engineering Team enables Shiftsmart to anticipate...  ...ideation and design through production and deployment, for example,...  ...build highly-performant and reliable technical features, tools,... 
    Hourly pay
    Work experience placement
    Work at office
    Flexible hours
    Shift work
    Night shift
    Weekend work

    Shiftsmart Inc

    New York, NY
    13 hours ago
  • $137.77k - $178.25k

     ...Engineering Role at BeaconYou'll join the engineering team behind Beacon...  ...banks, hedge funds, asset managers, and insurance companies....  ...intersection of financial data and production software — ingesting market...  ...transform investment data reliably at scaleIntegrations with... 

    Clearwater Analytics

    New York, NY
    5 days ago
  •  ...new portfolio of agentic software products in the industries we have spent decades...  ...do this. These teams bring product management, design, and engineering together with the autonomy to move...  ...for the architecture, security, reliability, and quality of the product. You are... 
    Permanent employment
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    New York, NY
    2 days ago
  • $180k - $230k

     ...critical role in ensuring that both production and development environments...  ...smoothly, securely, and reliably. This role leverages advanced...  ...curious MLOps/DevOps Engineers with deep expertise in machine...  ...supporting infrastructure.Build and manage cloud‑native infrastructure... 
    Full time
    Work at office
    Remote work

    iCapital Network

    New York, NY
    4 days ago
  •  ...full AI lifecycle, from development through production.   Trusted by Woven by Toyota, AXA,...  ...raised $60M in Series C funding from Wellington Management, CRV, Next47 and Y Combinator.   The role As a Solutions Engineer at Encord, you will be the core technical... 
    Flexible hours

    Encord

    New York, NY
    21 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer - Production Management. Be the first to apply!