Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

Bank of America ATM

At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. We do this by driving Responsible Growth and delivering for our clients, teammates, communities and shareholders every day.

Being a Great Place to Work is core to how we drive Responsible Growth. This includes our commitment to being an inclusive workplace, attracting and developing exceptional talent, supporting our teammates’ physical, emotional, and financial wellness, recognizing and rewarding performance, and how we make an impact in the communities we serve.

Bank of America is committed to an in-office culture with specific requirements for office-based attendance and which allows for an appropriate level of flexibility for our teammates and businesses based on role-specific considerations.

At Bank of America, you can build a successful career with opportunities to learn, grow, and make an impact. Join us!

This job is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing observability designs through instrumentation and dashboards, identifying root causes of complex/impactful issues, partnering with cross functional teams to deliver sustainable design patterns, and driving early adoption of non-functional production support requirements. Job expectations include automating services to improve reliability and efficiency and influencing a culture of innovation and continuous improvement.

Position Summary:

The Senior GCP Site Reliability Engineer acts as an advanced senior individual contributor responsible for designing, implementing, and maturing reliability engineering capabilities. The role focuses on complex technical problem solving, reliability architecture, automation strategy, observability maturity, platform resiliency, and operational excellence. The ideal candidate can operate across both strategy and execution: defining patterns, building automation, mentoring engineers, resolving hard platform issues, and driving long-term reliability improvements.

Key Responsibilities:

  • Serve as a senior technical authority for GCP platform reliability, resiliency, observability, automation, and production readiness
  • Design and implement advanced reliability patterns for Azure landing zones, private networking, DNS, firewalls, regional resiliency, service health checks, and workload onboarding
  • Lead complex platform reliability initiatives such as secondary-region readiness, egress/ingress observability, private DNS resolver monitoring, GenAI platform health checks, and enterprise dashboard automation
  • Define and mature SLIs, SLOs, reliability indicators, alerting standards, and service health reporting for Azure platform services
  • Develop reusable Terraform modules, automation frameworks, and CI/CD patterns that improve consistency, compliance, and operational quality
  • Drive observability improvements using Log Analytics, Dynatrace, Resource Graph, dashboards, and enterprise monitoring tools
  • Identify systemic reliability risks and translate them into engineering roadmaps, remediation plans, automation opportunities, and operational controls
  • Lead deep technical investigations for major incidents, recurring problems, platform defects, or service degradation events
  • Partner with security and governance teams to integrate IAM, policy-as-code, vulnerability remediation, control validation, and audit readiness into Azure platform operations
  • Provide technical design input for new Azure services and workloads to ensure operational readiness before production adoption
  • Mentor SRE engineers and raise the technical bar for automation, troubleshooting, documentation, resiliency design, and production support
  • Create executive-ready technical summaries, reliability narratives, and recommendations for leadership review
  • Establish reusable standards for runbooks, dashboards, health checks, problem records, post-incident reviews, and production-readiness gates
  • Designs solutions to visualize key production support metrics enabling Operational Readiness and Site Reliability Engineer teams to identify scenarios requiring intervention
  • Develops software solutions and/or improved processes to address work identified as ‘toil’ by collaborating with key partners to identify, track and remediate processes to free time allocated to reliability
  • Partners with Development and Infrastructure teams to create error budget policies prioritizing reliability stories that fall below Service Level Objective (SLO) thresholds and suggests code optimizations, additional instrumentation and/or logging structures to gain service reliability visibility
  • Identifies and plans for capacity bottlenecks, vulnerabilities and opportunities for reliability improvement, such as low level error rates and 'noise', and reduces manual support effort and/or improves system reliability
  • Assesses monitoring for new changes with development partners and works with monitoring tools team to monitor dashboards and enhance application and system monitoring designs
  • Engages as a subject matter expert in incident triage efforts, failure scenario modelling and works with the Problem Manager to diagnose root causes for complex/high impact incident/problem management investigations
  • Collaborates with Development and Infrastructure teams to understand technical solutions and develop Service Level Indicators and SLOs to measure/improve the reliability of the services they support
  • see position summary required/desired qualifications

Required Qualifications:

  • 4+ years of experience in cloud infrastructure engineering, platform engineering, or cloud operations, with exposure to Google Cloud Platform (GCP)
  • Strong hands-on experience with Infrastructure as Code (IaC), including practical use of Terraform or Terraform Enterprise for infrastructure provisioning
  • Solid understanding of software engineering fundamentals, including version control, code quality, and basic testing practices for infrastructure code
  • Experience developing and maintaining Terraform modules and infrastructure configurations to support automated cloud environments
  • Familiarity with CI/CD pipelines for infrastructure deployment, including automated build, test, and release processes
  • Working knowledge of DevSecOps practices, including integrating security and compliance checks into automated workflows
  • Good understanding of GCP services and cloud architecture fundamentals, including networking (VPCs, IAM, load balancing)
  • Exposure to policy-as-code, governance, and compliance requirements in enterprise environments
  • Experience supporting automation and standardization efforts to improve consistency and efficiency in cloud deployments
  • Understanding of monitoring, logging, and observability tools to support system reliability and performance
  • Hands-on experience with incident response, troubleshooting, and root cause analysis in cloud or distributed systems

Desired Qualifications:

  • Ability to collaborate effectively with engineering, architecture, and security teams to support reliable and secure platform operations
  • Strong problem-solving and analytical skills, with the ability to diagnose and resolve infrastructure issues
  • Effective communication skills, with the ability to work within cross-functional teams and document technical solutions clearly
  • Interest in learning and applying emerging technologies and automation techniques (including AI/ML where applicable) to improve platform reliability

Skills:

  • Architecture
  • Collaboration
  • Innovative Thinking
  • Result Orientation
  • Solution Design
  • Adaptability
  • Analytical Thinking
  • Influence
  • Stakeholder Management
  • Technical Strategy Development

Shift:

1st shift (United States of America)

Hours Per Week: 

40

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Jersey City, NJ vacancy
  •  ...are looking for people just like you. Join our team and help us develop game-changing, high-quality solutions.As a Senior Lead Site Reliability Engineer at JPMorganChase within the Core Engineering Solutions team of Consumer and Community Banking, you are an integral part... 
    Senior

    JP Morgan Chase

    Jersey City, NJ
    1 day ago
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern... 
    Senior
    Local area

    2T Consulting

    Jersey City, NJ
    5 days ago
  • $120k - $175k

     ...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's... 
    Senior
    Full time
    Remote work
    Work visa
    Flexible hours

    GrabJobs

    Jersey City, NJ
    3 days ago
  • Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the enterprise technology... 
    Senior

    JP Morgan Chase

    Jersey City, NJ
    1 day ago
  • $113.3k - $205.52k

     ...is important to maintain our strong culture, achieve our goals, and thrive as #OneJamf. What you'll do at Jamf: As a Senior Site Reliability Engineer, you'll help us balance development velocity with the reliability our customers depend on. You'll partner with engineering... 
    Senior
    Work at office
    Remote work
    Worldwide
    Flexible hours

    GrabJobs

    Jersey City, NJ
    1 day ago
  • $140k - $150k

    WORK OPTION: Remote_________________The NBA is hiring a Senior Site Reliability Engineer (SRE) - Messaging & Collaboration to ensure the availability, performance, and reliability of enterprise messaging and collaboration platforms, including Microsoft Exchange Online... 
    Senior
    Full time
    Temporary work
    Local area
    Remote work
    Weekend work

    National Basketball Association

    Secaucus, NJ
    8 hours ago
  •  ...PVH (Tommy Hilfiger/Calvin Klein) seeks a Senior Software Engineer to own the reliability and performance of our Kubernetes-based data platform across multi-region deployments. You will design scalable infrastructure, optimize deployment pipelines, and strengthen security... 
    Senior

    PVH (Tommy Hilfiger/Calvin Klein)

    Livingston, NJ
    4 days ago
  • $175k - $185k

     ...together. Come join our team as we develop new ways to improve the lives of working Americans. About the role: This Senior Database Reliability Engineer is responsible for supporting all production database systems so that they run smoothly. We run MySQL 8.4 on Google... 
    Senior
    Daily paid
    Remote work
    Home office
    Flexible hours

    GrabJobs

    Newark, NJ
    1 day ago
  • $141.8k

     ...their best work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl... 
    Senior
    Temporary work
    Remote work

    GrabJobs

    Newark, NJ
    1 day ago
  • $210k - $220k

     ...secure and private by design, it’s popular with security, IT, engineering, finance, and other security-focused teams. At Tines, we'...  ..., and we’re looking for others to join us on our journey. Senior Site Reliability Engineer - Government Cloud You'll join the team... 
    Senior
    Work at office
    Remote work

    GrabJobs

    Newark, NJ
    4 days ago
  • $160k - $240k

     ...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users.... 
    Senior
    Permanent employment
    Full time
    Remote work
    Work from home
    Relocation
    Flexible hours

    GrabJobs

    Newark, NJ
    1 day ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex... 
    Work at office

    JP Morgan Chase

    Jersey City, NJ
    1 day ago
  •  ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team, you hold a leadership role in your team, demonstrate... 
    Local area

    JP Morgan Chase

    Jersey City, NJ
    2 days ago
  •  ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology and Enterprise Technology Team, you draw upon your advanced knowledge to... 
    Local area

    JP Morgan Chase

    Jersey City, NJ
    4 days ago
  •  ...QualificationsBA degree or higherComputer Science or related disciplines.7+ years working with technical teams to define software business requirementsSummaryFunction: Information TechnologyExperience level: Mid-Senior LevelIndustry: Information Technology And Services
    Senior

    Sonsoft

    Jersey City, NJ
    4 days ago
  •  ...We are looking for a  Senior   Software Engineer. Our product is built from day one as a modern API First microservice architecture (Scala, ReactJS, TypeScript) with true continuous delivery and a deep investment in automated testing. This role reports to the Director... 
    Senior
    Work at office

    Kastel Staffing

    Hoboken, NJ
    18 days ago
  •  ...exceptional professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of...  ...and position yourself among the top echelon in site reliability. As an Associate Site Reliability Engineer at JPMorgan Chase... 
    Worldwide

    J.P. Morgan

    Jersey City, NJ
    1 day ago
  •  ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise... 

    Longfinch Technologies

    Jersey City, NJ
    21 hours ago
  • $130k - $180k

     ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and...  ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to... 
    Temporary work
    Work at office
    Immediate start
    Remote work
    Flexible hours

    GrabJobs

    Jersey City, NJ
    8 hours ago
  •  ...itD is seeking a Site Reliability Engineer to develop and enhance automation solutions that improve the reliability, scalability, and operational efficiency of large-scale cloud infrastructure. The ideal candidate will bring hands-on experience in site reliability engineering... 
    Work experience placement
    Remote work

    GrabJobs

    Jersey City, NJ
    1 day ago
  • $189.59k - $228k

    Pershing LLC seeks Senior Vice President, Release Train Engineer in Jersey City, NJ, to perform Agile Release Train (ART) or product. Facilitate development events and processes and assist the teams in delivering value. Communicate with stakeholders, escalate impediments... 
    Senior
    Temporary work
    Work at office
    Remote work
    Worldwide
    Flexible hours

    The Bank of New York Mellon

    Jersey City, NJ
    4 days ago
  • $119k - $170k

     ...the greater good, come make your next move with Zscaler. Our Engineering team built the world’s largest cloud security platform from...  ...cloud-first strategy. We’re looking for an experienced Staff Site Reliability Engineer (Federal) to join our Government Cloud team.... 
    Full time
    Work at office
    Local area
    Worldwide
    Night shift

    GrabJobs

    Jersey City, NJ
    3 days ago
  •  ...Site Reliability Engineer III There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer... 

    Chase

    Jersey City, NJ
    3 days ago
  •  ...are looking for people just like you. Join our team and help us develop game-changing, high-quality solutions. As a Lead Site Reliability Engineer at JPMorganChase within the Corporate sector, Enterprise Technology team, you are an integral part of a team that develops... 
    Work at office

    J.P. Morgan

    Jersey City, NJ
    1 day ago
  • $113.1k - $232.3k

    Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will...  ...Professional development From entry-level employees to senior leaders, we believe there’s always room to learn. We offer... 
    Work at office
    Local area
    Visa sponsorship
    Flexible hours
    3 days per week

    Deloitte

    Jersey City, NJ
    3 days ago
  •  ...leader in catastrophe-exposed property insurance, is seeking a Senior Software Engineer to join our Policy Onboarding and Home Services team. As a...  ...after they purchase a policy, ensuring a seamless, reliable, and modern experience. You'll contribute to architectural... 
    Senior
    For contractors
    Live in
    Remote work

    SageSure

    Jersey City, NJ
    13 days ago
  • Job ID: 25677433Reference Number: 25-00616Title: Senior Python DeveloperLocation: jersey city, NJJob Type: Full Time/ContractPosted...  ...developers with 6+ years' experience Experience with PySpark Should also have a strong SQL background and have data engineering experience
    Senior
    Full time

    HAN Staffing

    Jersey City, NJ
    8 hours ago
  • $176.72k - $265.08k

     ...CitiCiti’s Equities Technology organization is seeking a hands‑on Senior Software Engineer to join the Equities Timeseries group, a strategic,...  ...and productionize analytics, translating research ideas into reliable, production‑grade time‑series systems.Drive high standards... 
    Senior
    Full time

    Citigroup

    Jersey City, NJ
    3 days ago
  • $157k - $235k

     ...Talkdesk is looking for a Senior Solutions Engineer specializing in Banking, Financial Services, and Insurance (FSI). You will engage throughout the sales lifecycle, offering industry-centered expertise, conducting research, and designing impactful presentations. The role... 
    Senior

    Talkdesk

    Jersey City, NJ
    4 days ago
  •  ...collaborative company where innovation, creativity, ownership, and impact are part of everyday work. ABOUT THE ROLE:As a Senior Software Engineer at Global-e, you will design and deliver the core services behind our global logistics platform. You will drive innovation... 
    Senior
    Work at office
    Worldwide

    Global-e

    Hoboken, NJ
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!