Site Reliability Engineer
J.P. Morgan
hackajob is collaborating with J.P. Morgan to connect them with exceptional professionals for this role. JOB DESCRIPTION There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology, Infrastructure Platforms team, you will solve complex and broad business problems with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions. You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform. Job responsibilities Owns end-to-end reliability for the ServiceNow platform, including incident response, SLO/observability, automation, and platform performance engineering Troubleshoots incidents/problems; leads blameless postâmortems and drives nonârecurrence actions Defines/defends SLIs/SLOs and error budgets; improves telemetry, monitoring, alerting, and performance testing review Monitors platform and product performance metrics; partners with engineering on remediation Identifies and delivers stability/resiliency/performance improvements; supports platform/vendor upgrades Automates repetitive operational work to reduce toil Performs design reviews for platform/tenant implementations; partners with architecture on platform standards Holistically analyzes ServiceNow instances to identify/remediate resource contention across all layers of the stack; explores platform Java and JavaScript to understand ServiceNow application behavior Improves throughput via JVM memory allocation and garbage collection tuning; optimizes relational database performance via query refactoring and/or database tuning Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements. Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to SLO outcomes. Required qualifications, capabilities, and skills Formal training or certification on site reliability engineering concepts and 3 years applied experience Proficient in site reliability culture and principles and familiarity with how to implement site reliability within an application or platform Proficient in at least one programming language such as Python, Java/Spring Boot, and .Net Strong enterprise platform support experience and practical SRE/observability fundamentals relevant to ServiceNow 10 years in technical support, software development, performance load testing, or professional services Expert-level understanding of web application stack components Scripting: one or more of JavaScript, Python, Perl, Unix shell, Windows PowerShell Experience working with/debugging object-oriented code like Java Strong RDBMS experience (Oracle/MySQL/SQL Server/PostgreSQL) and performance tuning; JVM performance tuning (memory/GC) and troubleshooting in complex systems Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data sensitivity Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements Preferred qualifications, capabilities, and skills ServiceNow CSA (or higher) Prior ServiceNow architecture exposure (training can be provided) Cloud/SaaS experience Memory management and dump analysis (Java heap dump analysis preferred) ITSM/ITIL/CMDB fundamentals Linux/Unix or Windows Server administration Browser/network troubleshooting tools (e.g., Firebug, Chrome DevTools, Fiddler) ABOUT US JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world's most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management. We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process. We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation. JPMorgan Chase & Co. is an Equal Opportunity Employer, including Disability/Veterans ABOUT THE TEAM Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we're setting our businesses, clients, customers and employees up for success.aa415a4b-8b21-40fc-a65c-70d2b25ca29a
$200.7k - $250.9k
...washed away in a flood in 1942, the Royal Engineers rebuilt it. Then it washed away again in... ...opportunities for improvements in reliability/observability/performance/preparedness and... ...candidate for the role: Has past Site Reliability Engineering or DevOps experience...Suggested- ...Lambda Inc. in San Francisco is seeking a Storage Engineer to own the reliability, performance, and capacity of our production storage fleet across multiple data centers, using a software-defined data plane. You will build monitoring, dashboards, and alerting for storage...Suggested
$194k - $237k
...employer, at the date of hire. This position is ineligible for employment Visa sponsorship. Role Summary The Principal Site Reliability Engineer applies software engineering and systems engineering practices to improve the reliability, resilience, scalability, and...SuggestedHourly payWork at officeImmediate startVisa sponsorshipWork visaFlexible hours$200k - $240k
...Senior Site Reliability Engineer (SRE)Location: San Francisco, CAWork Model: OnsiteIndustry: Renewable EnergyComp: $200,000 - $240,000 We’re partnering with a fast-growing energy technology company looking for a Senior Site Reliability Engineer to take ownership of...Suggested- ...Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering... ...You will define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role includes...Suggested
- ...troubleshooting, providing rubric-based written feedback. This role requires hands-on Kubernetes expertise in EKS/GKE/AKS or self-managed clusters, with strong scripting in Go, Python, or TypeScript, and ability to document findings clearly for engineering #J-18808-Ljbffr
$350k
...with leading AI companies and infrastructure providers to build reliable, high-performance platforms supporting next-generation AI workloads. This opportunity is for a Staff Site Reliability Engineer to lead the reliability of large-scale GPU infrastructure, covering...$200k - $240k
...systems across all product teams. You will collaborate closely with engineering leadership, product managers, and cross-functional teams to... ...and Helm ~ Understand the importance of performant and reliable systems ~ Education - Ideally looking for a B.A. / B.S. degree...Work at officeImmediate start3 days per week- ...Site Reliability Engineer (SRE) FLUIX is building the AI operating system that plans, designs, and optimizes AI infrastructure. We are based in Silicon Valley. We specialize in providing AI-driven solutions for data centers and power providers, leveraging cutting-edge...Work at officeWeekend work
- ...Senior Engineering Role at Salesforce Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here... ...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with...WorldwideWeekend work
$200k - $300k
...Site Reliability Engineer Title of Role: Site Reliability Engineer Location: San Francisco, onsite Company Stage of Funding: Venture Round — Healthcare, AI Office Type: Onsite Salary: $200K–$300K Company Description We're representing a dynamic company...Work at office$81.1k - $187k
...Site Reliability Engineer 3 We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving...Temporary workImmediate startFlexible hoursShift work- ...had design and make. Operate is the phase that tells you what actually happened, and it is ours. We’re looking for a Site Reliability Engineer to help advance MaintainX’s reliability, observability, and developer autonomy as we scale our platform. In this role,...Remote work
$260k - $300k
...software agents. We're the makers of Devin, the first AI software engineer. Our team is extremely talent-dense. Among our founding... ...faster than anyone expects. You will own both the production reliability of our user-facing products and the platform engineering that...- ...access to life-saving treatment. What We Look for in a Great Engineer You have the intensity and technical mastery to own... ...support high-velocity feature release while maintaining the highest reliability. DevX Support: Support Developer Experience (DevX) work to...Work at office
- ...The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure platform. You'll be building and operating the core systems that power agentic AI at scale. Your mission: keep...
- ...human would. We're a small team of former Google and Stripe engineers, including the founding team of Google Wallet, dedicated to... ...The Role We're looking for a skilled and passionate Site Reliability Engineer to join our team. As a SRE, you'll be responsible...Remote work1 day per week
- ...globe. Join us on this journey to redefine resource management-and change lives along the way. The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You...Temporary workWorldwide
$98.58k - $138.02k
...Site Reliability Engineer II Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back-office operations for restaurants. Restaurant365's culture is focused on empowering...Work at office- ...Site Reliability Engineer Specter's mission is to help automate the physical world. Today, we build video sensors with state-of-the-art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical...Remote work
$130k - $200k
...Site Reliability Engineer Nscale is the GPU cloud built for AI. We run high-performance, cost-efficient infrastructure for AI-native startups and global enterprises, from bare metal up through the platform services teams actually build on. Our culture runs on ownership...Shift work- ...Arena Intelligence Engineer Arena Intelligence is looking for an engineer to build the core infrastructure that sits beneath our online... ...foundational infrastructure for our users that scales, is reliable, and makes the complexities of operating this infrastructure at...Permanent employmentShift work
$150k
...Site Reliability Engineer San Francisco, CA About The Role We are seeking an experienced Site Reliability Engineer (SRE) with a strong focus on DevSecOps to join our growing engineering team. In this role, you will oversee and maintain the reliability, security...- ...JOB DESCRIPTION Project Outline: We are looking for a Site Reliability Engineer with experience in incident response. In this role, you will help Shipt understand where we can improve stability and reliability. There will be a focus on the intersection of systems...
- ...Site Reliability Engineer We are looking for a dynamic engineer to join our rapidly growing SRE team. As an SRE, you will report to our VP of Technical Operations and be responsible for operating an extremely high performance and scalable, low latency platform built...Relocation package
$170k - $220k
...Senior Site Reliability Engineer Supio is a trusted AI platform purpose-built for law firms, reshaping how data drives impactful outcomes. Our innovative approach blends technology with deep legal expertise, making us a leader in our field. We go beyond surface-level...Work at officeRemote workFlexible hours- ...enterprise that runs the real economy. Learn more about our vision in our manifesto. About the Role We're looking for a Site Reliability Engineer to take the lead on scaling our operational resilience as we grow. You'll own the stability, observability, and debugging...WorldwideShift work
- ...guarantees and certifications. We're hiring staff-level SREs to help run and evolve that infrastructure, working alongside the senior engineers already on the team. You'll contribute to architecture decisions for how we deploy, observe, and secure the platform, and help...Remote workFlexible hours
$165k - $227k
...opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity We are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly...Local areaWorldwideFlexible hours- ...design of information and operational support systems. Required Skills/Qualifications: BS/MS degree in Computer Science, Engineering, or a related subject. Equivalent experience accepted. Proven working experience in installing, configuring, and troubleshooting...Full timeWork experience placementRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre San Francisco, CA
- site reliability engineer San Francisco, CA
- site reliability engineer remote San Francisco, CA
- site recruiter San Francisco, CA
- site services specialist San Francisco, CA
- junior website developer San Francisco, CA
- official site San Francisco, CA
- on site coordinator San Francisco, CA
- site leader San Francisco, CA
- historic site San Francisco, CA


