Senior Site Reliability Engineer
Bank of America ATM
At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. We do this by driving Responsible Growth and delivering for our clients, teammates, communities and shareholders every day.
Being a Great Place to Work is core to how we drive Responsible Growth. This includes our commitment to being an inclusive workplace, attracting and developing exceptional talent, supporting our teammates’ physical, emotional, and financial wellness, recognizing and rewarding performance, and how we make an impact in the communities we serve. Bank of America is committed to an in-office culture with specific requirements for office-based attendance and which allows for an appropriate level of flexibility for our teammates and businesses based on role-specific considerations. At Bank of America, you can build a successful career with opportunities to learn, grow, and make an impact. Join us!This job is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing observability designs through instrumentation and dashboards, identifying root causes of complex/impactful issues, partnering with cross functional teams to deliver sustainable design patterns, and driving early adoption of non-functional production support requirements. Job expectations include automating services to improve reliability and efficiency and influencing a culture of innovation and continuous improvement.
Position Summary:
The Senior Site Reliability Engineer acts as an advanced senior individual contributor responsible for designing, implementing, and maturing reliability engineering capabilities. The role focuses on complex technical problem solving, reliability architecture, automation strategy, observability maturity, platform resiliency, and operational excellence. The ideal candidate has experience driving the evolution of traditional infrastructure operations toward a modern reliability engineering model through automation, observability, AIOps, and engineering-first operational practices that improve service reliability and reduce operational toil.
Responsibilities:
- Designs solutions to visualize key production support metrics enabling Operational Readiness and Site Reliability Engineer teams to identify scenarios requiring intervention
- Develops software solutions and/or improved processes to address work identified as ‘toil’ by collaborating with key partners to identify, track and remediate processes to free time allocated to reliability
- Partners with Development and Infrastructure teams to create error budget policies prioritizing reliability stories that fall below Service Level Objective (SLO) thresholds and suggests code optimizations, additional instrumentation and/or logging structures to gain service reliability visibility
- Identifies and plans for capacity bottlenecks, vulnerabilities and opportunities for reliability improvement, such as low level error rates and 'noise', and reduces manual support effort and/or improves system reliability
- Assesses monitoring for new changes with development partners and works with monitoring tools team to monitor dashboards and enhance application and system monitoring designs
- Engages as a subject matter expert in incident triage efforts, failure scenario modelling and works with the Problem Manager to diagnose root causes for complex/high impact incident/problem management investigations
- Collaborates with Development and Infrastructure teams to understand technical solutions and develop Service Level Indicators and SLOs to measure/improve the reliability of the services they support
- Champions modern SRE, observability, and AIOps practices by driving automation, reliability engineering, and operational excellence across platform services.
- Leads intelligent operations initiatives, including automated remediation, predictive monitoring, event correlation, and data-driven decision-making.
- Serves as a senior technical authority for platform reliability, resiliency, observability, automation, and production readiness, while leading key initiatives such as secondary-region readiness, network observability, GenAI platform health monitoring, and enterprise dashboard automation.
- Defines and matures SLIs, SLOs, alerting standards, and service health reporting across on-premises and cloud platforms.
- Develops reusable IaC modules, automation frameworks, and CI/CD patterns to improve consistency, compliance, and operational quality. Identifies reliability risks, drives remediation and automation strategies, leads major incident investigations, and partners with security and governance teams to embed IAM, policy-as-code, vulnerability remediation, and audit readiness into cloud operations.
Required Qualifications:
- 5+ years of experience in platform, systems, or infrastructure engineering, with a strong focus on automation and integration.
- Demonstrated experience transforming traditional infrastructure operations, systems administration, or platform support functions into modern engineering-led operating models utilizing Site Reliability Engineering (SRE), observability, automation, and AIOps practices.
- ·Proven experience implementing and operationalizing SRE principles, including Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, reliability metrics, incident management modernization, and continuous service improvement programs.
- Hands-on experience designing and implementing enterprise observability solutions leveraging metrics, logs, traces, telemetry, synthetic monitoring, and service health analytics to improve operational visibility and platform reliability.
- Experience reducing operational toil through automation, self-healing capabilities, event-driven remediation, Infrastructure as Code (IaC), and intelligent operational workflows.
- Demonstrated ability to identify operational inefficiencies and drive engineering solutions that improve availability, resiliency, scalability, and operational effectiveness.
- Experience partnering with infrastructure, application, security, and platform teams to establish reliability engineering standards, observability frameworks, production readiness requirements, and operational best practices.
Desired Qualifications:
- Experience driving organizational adoption of SRE culture and practices, including operational maturity assessments, reliability reviews, automation programs, post-incident learning, and engineering enablement.
- Experience supporting or implementing AIOps platforms and capabilities, including event correlation, anomaly detection, predictive analytics, automated remediation, noise reduction, and operational intelligence.
- Experience with modern observability platforms.
- Experience with CI/CD pipelines and infrastructure-as-code (e.g., Terraform, Ansible)
- Familiarity with containerization and orchestration platforms
- Experience in financial services or highly regulated environments
- Ability to communicate complex technical concepts to non-technical stakeholders
- Prior experience mentoring or leading engineering operations teams
Skills:
- Architecture
- Collaboration
- Innovative Thinking
- Result Orientation
- Solution Design
- Adaptability
- Analytical Thinking
- Influence
- Stakeholder Management
- Technical Strategy Development
- Other
- Resource Management
- Financial and Succession Planning
Shift:
1st shift (United States of America)Hours Per Week:
40- ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern...SeniorLocal area
$152.6k - $191.5k
...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include... ...continuous improvement. Position Summary: The Senior Site Reliability Engineer acts as an advanced senior...SeniorFull timeWork at officeShift workDay shift- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Chief Data & Analytics...SeniorWork at office
- LiveRamp, a data collaboration platform leader, is seeking a Senior Site Reliability Engineer in San Francisco with 5+ years of SRE/DevOps experience. The role focuses on deployments, 24/7 support across regions, and establishing SRE best practices. Strong skills in Terraform...Senior
$174k - $252k
Senior Software Engineer, Site Reliability Engineering corporate_fare Google place Seattle, WA, USA ; Kirkland, WA, USA Mid Experience driving progress, solving problems, and mentoring more junior team members; deeper expertise and applied knowledge within relevant area...SeniorTemporary work- ...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of... ...MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment...Full timeWork at officeLocal areaRemote work
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex...Work at office
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational Services team, you hold a leadership role in your team, demonstrate...
- ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new opportunities...
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer focussed on Operations Excellence, you will have the opportunity to shape how we respond to, learn from, and...
$132.23k - $176.31k
...shape the future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem....SeniorFull timeTemporary workRemote work- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...
- ...QualificationsBA degree or higherComputer Science or related disciplines.7+ years working with technical teams to define software business requirementsSummaryFunction: Information TechnologyExperience level: Mid-Senior LevelIndustry: Information Technology And ServicesSenior
$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will... ...Professional development From entry-level employees to senior leaders, we believe there’s always room to learn. We offer...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week- ...Job Description Job Description BCforward is currently seeking a highly motivated SRE Software Engineer. Job Title: SRE Software Engineer Location: Jersey City, NJ Duration: Temp - 12 months Job Description We are seeking a Software Engineer-Other...Temporary work
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial Investment Banking team of Fraud Prevention , you will solve complex and broad...
- ...Citigroup Inc. in Jersey City, NJ, seeks a Senior Principal Platform & AI Architect with 10+ years of hands-on cloud, API, and AI experience to lead design, development, and security of next-generation cloud platforms and enterprise APIs. You will work hands-on in...Senior
- Job ID: 25677433Reference Number: 25-00616Title: Senior Python DeveloperLocation: jersey city, NJJob Type: Full Time/ContractPosted... ...developers with 6+ years' experience Experience with PySpark Should also have a strong SQL background and have data engineering experienceSeniorFull time
- Role SummaryWe are seeking a Senior Java MuleSoft Engineer P4 with strong expertise in enterprise integration API development and cloudbased integration platforms The ideal candidate will have handson experience in MuleSoft Anypoint Platform Java Spring Boot APIled architecture...Senior
- ...focused on improving the security, release reliability, maintainability, and operational... ...critical banking applications.Forward Deployed Engineers will conduct targeted, hands-on... ...rapid release cadence.The role combines senior software engineering, DevOps, SRE, test...Senior
- ...collaborative company where innovation, creativity, ownership, and impact are part of everyday work. ABOUT THE ROLE:As a Senior Software Engineer at Global-e, you will design and deliver the core services behind our global logistics platform. You will drive innovation...SeniorWork at officeWorldwide
$130k - $150k
Senior Software Engineer - AI/Agentic SystemsPearson Learning Studio (PLS) - REIA Squad | Hoboken, NJPearson Learning Studio is seeking a Senior... ...prompts to improve response quality, reasoning, and reliability.Build scalable AI services and APIs using Python.Platform...SeniorFull time$108k - $216k
...PermanentCompany: WalmartBusiness Segment: Home OfficeRole summary: The Senior Software Engineer will lead the delivery of scoped features and models... ..., uphold engineering excellence, and support system reliability while mentoring peers and contributing to a high-...SeniorFull timeTemporary workPart time- We're looking for a talented, senior engineering professional ready to take their career to new heights... ..., storage, security controls, and reliability enabling product teams to deploy... ...comprehensive health care coverage, on-site health and wellness centers, a retirement...Senior
- ...leader in catastrophe-exposed property insurance, is seeking a Senior Software Engineer to join our Policy Onboarding and Home Services team. As a... ...after they purchase a policy, ensuring a seamless, reliable, and modern experience. You'll contribute to architectural...SeniorFor contractorsLive inRemote work
- ...to understand business requirements and translate them into technical requirements ~ Demonstrate a solid understanding of core engineering principles ~ Familiarity with modern software engineering practices and continuous integration and delivery ~ Comfortable working...SeniorRemote jobFull timeHome office
$140k - $200k
...asked to participate in an on-site interview as part of the... ...every collector, old or new. Our engineering mission is to democratize technology... ....We’re looking for a Senior Platform Engineer to join our... ...observability), improving the reliability, scalability, and cost-efficiency...SeniorFull timeRemote workWorldwideFlexible hoursShift work- ...matter and your skills drive change.As a Senior Lead Software Engineer, Cloud Platform at JPMorgan Chase in... ...development of secure, scalable, and reliable cloud infrastructure and platform... ...comprehensive health care coverage, on-site health and wellness centers, a...SeniorWork at officeShift work
- ...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of... ...team will provide more details.Job Summary:MUFG is seeking a Senior Software Engineer with payment check processing experience. In this role you...SeniorFull timeWork at officeLocal areaRemote work
$134.2k - $258.3k
...processProvides advanced technical expertise to maximize efficiency, reliability and value from current solutions, infrastructure and emerging... ...strong working relationships with peers across Development & Engineering and Architecture teams, collaborating to develop and engineer...SeniorSummer holidayLocal areaFlexible hoursShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Jersey City, NJ
- site reliability engineer sre Jersey City, NJ
- senior network engineer remote Jersey City, NJ
- senior app developer Jersey City, NJ
- senior manager legal Jersey City, NJ
- sr project manager Jersey City, NJ
- senior commercial counsel Jersey City, NJ
- senior manager tax Jersey City, NJ
- senior construction accountant Jersey City, NJ
- senior associate Jersey City, NJ




