Senior Site Reliability Engineer
Bank of America
Senior Site Reliability Engineer
Jersey City, New Jersey;Chandler, Arizona; Plano, Texas
To proceed with your application, you must be at least 18 years of age.
Acknowledge (
Bank of America employees are required to meet all posting eligibility requirements prior to applying for any new position.
Acknowledge (
Refer a friend
To proceed with your application, you must be at least 18 years of age.
Acknowledge (
Bank of America employees are required to meet all posting eligibility requirements prior to applying for any new position.
Acknowledge (
Job Description:
At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. We do this by driving Responsible Growth and delivering for our clients, teammates, communities and shareholders every day.
Being a Great Place to Work is core to how we drive Responsible Growth. This includes our commitment to being an inclusive workplace, attracting and developing exceptional talent, supporting our teammates' physical, emotional, and financial wellness, recognizing and rewarding performance, and how we make an impact in the communities we serve.
Bank of America is committed to an in-office culture with specific requirements for office-based attendance and which allows for an appropriate level of flexibility for our teammates and businesses based on role-specific considerations.
At Bank of America, you can build a successful career with opportunities to learn, grow, and make an impact. Join us!
Job Description:
This job is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing observability designs through instrumentation and dashboards, identifying root causes of complex/impactful issues, partnering with cross functional teams to deliver sustainable design patterns, and driving early adoption of non-functional production support requirements. Job expectations include automating services to improve reliability and efficiency and influencing a culture of innovation and continuous improvement.
Position Summary:
The Senior Azure Site Reliability Engineer acts as an advanced senior individual contributor responsible for designing, implementing, and maturing reliability engineering capabilities across the enterprise Azure platform. The role focuses on complex technical problem solving, reliability architecture, automation strategy, observability maturity, platform resiliency, and operational excellence. The ideal candidate can operate across both strategy and execution: defining patterns, building automation, mentoring engineers, resolving hard platform issues, and driving long-term reliability improvements.
Environment and Platform Scope
Enterprise-scale, governed Azure platform supporting infrastructure, platform services, and application workload enablement
Primary tools and practices: Microsoft Azure, Terraform / Terraform Enterprise, Azure Monitor, Azure Log Analytics, Dynatrace, CI/CD pipelines, Git, Python, PowerShell, Bash, ServiceNow/Jira or equivalent workflow tools
Core domain areas: Azure landing zones, private networking, DNS, firewalls, private endpoints, observability, incident response, problem management, canary health checks, production readiness, and cloud governance
Role supports Azure platform reliability, operational readiness, observability, automation, and enterprise controls.
Responsibilities:
Designs solutions to visualize key production support metrics enabling Operational Readiness and Site Reliability Engineer teams to identify scenarios requiring intervention
Develops software solutions and/or improved processes to address work identified as 'toil' by collaborating with key partners to identify, track and remediate processes to free time allocated to reliability
Partners with Development and Infrastructure teams to create error budget policies prioritizing reliability stories that fall below Service Level Objective (SLO) thresholds and suggests code optimizations, additional instrumentation and/or logging structures to gain service reliability visibility
Identifies and plans for capacity bottlenecks, vulnerabilities and opportunities for reliability improvement, such as low level error rates and 'noise', and reduces manual support effort and/or improves system reliability
Assesses monitoring for new changes with development partners and works with monitoring tools team to monitor dashboards and enhance application and system monitoring designs
Engages as a subject matter expert in incident triage efforts, failure scenario modelling and works with the Problem Manager to diagnose root causes for complex/high impact incident/problem management investigations
Collaborates with Development and Infrastructure teams to understand technical solutions and develop Service Level Indicators and SLOs to measure/improve the reliability of the services they support
Lead complex platform reliability initiatives such as secondary-region readiness, egress/ingress observability, private DNS resolver monitoring, GenAI platform health checks, and enterprise dashboard automation.
Define and mature SLIs, SLOs, reliability indicators, alerting standards, and service health reporting for Azure platform services.
Develop reusable Terraform modules, automation frameworks, and CI/CD patterns that improve consistency, compliance, and operational quality.
Drive observability improvements using Azure Monitor, Log Analytics, Dynatrace, Resource Graph, dashboards, and enterprise monitoring tools.
Identify systemic reliability risks and translate them into engineering roadmaps, remediation plans, automation opportunities, and operational controls.
Partner with security and governance teams to integrate IAM, policy-as-code, vulnerability remediation, control validation, and audit readiness into Azure platform operations.
Provide technical design input for new Azure services and workloads to ensure operational readiness before production adoption.
Mentor SRE engineers and raise the technical bar for automation, troubleshooting, documentation, resiliency design, and production support.
Create executive-ready technical summaries, reliability narratives, and recommendations for leadership review.
Required Qualifications:
Advanced experience in Azure platform engineering, SRE, cloud infrastructure, or enterprise cloud operations.
Deep knowledge of Microsoft Azure architecture, including networking, identity, compute, PaaS, monitoring, security, governance, and resiliency patterns.
Strong experience designing and developing Terraform modules and infrastructure-as-code automation in enterprise environments.
Strong understanding of Azure landing zones, hub-and-spoke networking, ExpressRoute or enterprise connectivity, private endpoints, DNS, routing, firewalls, and workload isolation.
Advanced observability experience with Azure Monitor, Log Analytics, Dynatrace, dashboards, alerting, metrics, and platform telemetry.
Experience with SRE operating models, SLIs, SLOs, incident response, problem management, toil reduction, and production-readiness reviews.
Strong scripting or programming experience with Python, PowerShell, Bash, Java, or similar languages.
Experience operating highly available Azure IaaS and PaaS services in enterprise-scale environments.
Ability to influence architecture and engineering decisions across multiple technical teams.
Strong communication skills with the ability to translate complex engineering topics into actionable recommendations.
Desired Qualifications:
Microsoft Azure certification strongly preferred.
Experience in financial services, regulated technology, or large enterprise infrastructure organizations.
Experience with AKS, ACR, Kubernetes, container networking, CI/CD pipelines, and container observability.
Experience with Azure AI Foundry, OpenAI/GenAI platform operations, model-serving observability, or AI platform readiness.
Experience with FinOps, governance dashboards, resource hygiene, cost visibility, security posture, and compliance reporting.
Experience creating technical roadmaps, reliability scorecards, production-readiness frameworks, or executive-level reliability reporting.
Skills:
Architecture
Collaboration
Innovative Thinking
Result Orientation
Solution Design
Adaptability
Analytical Thinking
Influence
Stakeholder Management
Technical Strategy Development
Other
Terraform
Python
Shift:
1st shift (United States of America)
Hours Per Week:
40
Bank of America and its affiliates consider for employment and hire qualified candidates without regard to race, religious creed, religion, color, sex, sexual orientation, genetic information, gender, gender identity, gender expression, age, national origin, ancestry, citizenship, protected veteran or disability status or any factor prohibited by law, and as such affirms in policy and practice to support and promote the concept of equal employment opportunity, in accordance with all applicable federal, state, provincial and municipal laws. The company also prohibits discrimination on other bases such as medical condition, marital status or any other factor that is irrelevant to the performance of our teammates.
View your "Know your Rights ( " poster.
View the LA County Fair Chance Ordinance ( .
Bank of America aims to create a workplace free from the dangers and resulting consequences of illegal and illicit drug use and alcohol abuse. Our Drug-Free Workplace and Alcohol Policy ("Policy") establishes requirements to prevent the presence or use of illegal or illicit drugs or unauthorized alcohol on Bank of America premises and to provide a safe work environment.
Bank of America is committed to an in-office culture with specific requirements for office-based attendance and which allows for an appropriate level of flexibility for our teammates and businesses based on role-specific considerations. Should you be offered a role with Bank of America, your hiring manager will provide you with information on the in-office expectations associated with your role. These expectations are subject to change at any time and at the sole discretion of the Company. To the extent you have a disability or sincerely held religious belief for which you believe you need a reasonable accommodation from this requirement, you must seek an accommodation through the Bank's required accommodation request process before your first day of work.
This communication provides information about certain Bank of America benefits. Receipt of this document does not automatically entitle you to benefits offered by Bank of America. Every effort has been made to ensure the accuracy of this communication. However, if there are discrepancies between this communication and the official plan documents, the plan documents will always govern. Bank of America retains the discretion to interpret the terms or language used in any of its communications according to the provisions contained in the plan documents. Bank of America also reserves the right to amend or terminate any benefit plan in its sole discretion at any time for any reason.
Investment products offered through MLPF&S and insurance and annuity products offered through MLLA:
Are Not FDIC Insured Are Not Bank Guaranteed May Lose Value
Are Not Deposits Are Not Insured by Any Federal Government Agency Are not a condition to Any Banking Service or Activity
Merrill Lynch, Pierce, Fenner & Smith Incorporated (also referred to as "MLPF&S" or "Merrill") makes available certain investment products sponsored, managed, distributed or provided by companies that are affiliates of Bank of America Corporation ("BofA Corp."). MLPF&S is a registered broker-dealer, registered investment adviser, Member SIPC and a wholly owned subsidiary of BofA Corp. Insurance and annuity products are offered through Merrill Lynch Life Agency Inc., a licensed insurance agency and wholly owned subsidiary of Bank of America Corporation.
Trust, fiduciary and investment management services are provided by Bank of America, N.A., Member FDIC and wholly owned subsidiary of Bank of America Corporation ("BofA Corp.").
Bank of America Private Bank is a division of Bank of America, N.A.
Banking products are provided by Bank of America, N.A. and affiliated banks, Members FDIC and wholly owned subsidiaries of Bank of America Corporation.
© 2026 Bank of America Corporation. All rights reserved.
$117k - $209.33k
...Full time 26WD99276 Job Requisition ID # 26WD99276 Position Overview Want to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products...SeniorFull timeFor contractorsRemote work- ...the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines... ...the team for you. Your Impact You will be the most senior technical individual contributor on the team — setting the...Senior
- ...TechMContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8... ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will...SeniorRemote work
- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Corporate and Investment...Senior
- Job Title Design and maintain highly available, scalable, and fault-tolerant systems Implement and manage monitoring, logging, and observability tools Grafana, Prometheus, etc. Participate in incident management, RCA, and postmortems Automate operational tasks...Suggested
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and broad business problems with simple and straightforward...
$96.8k - $145.2k
...want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a Site Reliability Engineer (Onsite Hybrid) to join our team in Plano, Texas (US-TX), United States (US). Job Responsibilities Include: Own...Temporary workWork at officeRemote workFlexible hours- ...Job Description Insight Global is seeking a Site Reliability Engineer for a top enterprise client. This role focuses on supporting and administering... ...stability across AWS and hybrid environments. This is a senior-level opportunity with leadership visibility, especially...
- ...tasks using scripting and tools Python Bash etc Collaborate with development infrastructure and support teams to improve system reliability Drive adoption of SRE practices like SLIs SLOs and error budgets Ensure performance optimisation capacity planning and...Permanent employmentTemporary workWork experience placement
$63.68 - $71.68 per hour
Genesis10 is currently seeking a Site Reliability Engineer (SRE) Lead for a hybrid position (3 days onsite per week) with a Global Financial Institution located in Plano, TX or Charlotte, NC. This is a 12+ month contract opportunity. This role will lead reliability engineering...Hourly payPermanent employmentContract work3 days per week$96.8k - $145.2k
...Site Reliability Engineer (Onsite Hybrid) NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a...Temporary workFlexible hours- ...Site Reliability Engineer Hybrid Onsite Worker is required to work onsite 2-3 days per week in Phoenix, AZ OR Plano, TX Main Responsibilities ~ Experience in leading observability initiatives as lead engineer. ~ Development and implementation of build release...Work experience placement2 days per week3 days per week
$159.75k - $225.5k
...moments that reflect who they uniquely are.We are seeking a Senior Principal Software Engineer to lead the design and development of next-generation... ...commitment to Diversity, Equity, and Inclusion on our Career Site.The compensation package for this role is based on...SeniorRemote work- ...Description Join the Enterprise Justice team at Tyler Technologies as an Senior Software Engineer! As a full stack engineer, you will work with a wide range of technologies such as .NET, Angular, HTML5, SQL, and Cloud/AWS building scalable solutions that empower...SeniorFull time
- ...to 40% to client locations throughout the U.S. Special Skills: Position requires a Bachelor's Degree in Computer Science, Engineering or Related Field and five (5) years of experience. The aforementioned five (5) years of experience must be progressively...SeniorFull time
$69.99 per hour
...Job Description Job Description Job Title: Site Reliability Engineer (SRE) Location: Plano, TX / Pennington, NJ / Charlotte, NC Duration: Contract - 12 months Pay Range: $69.99/hr (W2) Job ID: 409158 About BCforward BCforward is a leading global...Contract work- ...Job Description Job Description BCforward is currently seeking a highly motivated Site Reliability Engineer Job Title: Site Reliability Engineer Location: Columbus, OH/Plano, TX or Tampa, FL Duration: Contract to hire 4 months Job Description We are...Contract work
$61k - $101k
...Salary: $61,000 - 101,000 per year Requirements: Formal training or certification in site reliability engineering concepts, plus 3+ years of hands-on experience Experience supporting SRE practices for data management or migration platforms and products Familiarity...Full time$86.8k - $165.2k
...than 100 years of experience and renowned engineering expertise to meet the needs of today’s... ...solutions.Raytheon has an opportunity for a Senior Software Engineer to join our Mission... ...of whether the role is designated as on-site, hybrid or remote.The salary range for this...SeniorTemporary workWork experience placementWork at officeLocal areaRemote workRelocationFlexible hours- Create the future of e-health together with us by becoming a Senior Software DeveloperAt CompuGroup Medical we have the mission of building ground-breaking solutions for digital healthcare. Our vision is revolutionizing how healthcare professionals produce, access, and...SeniorFull timeRemote workFlexible hours
$61k - $101k
...Salary: $61,000 - 101,000 per year Requirements: We expect formal training or certification in site reliability engineering, along with 5+ years of applied experience We need deep expertise in reliability, scalability, performance, security, enterprise architecture...Full time$140k - $200k
...experiencing exponential growth. Overview We're looking for a Senior Software Engineer to join our Core Experiences Team. This team builds and... ...strategically, and is passionate about designing clear, reliable APIs and simple systems that directly enhance the user...SeniorFull timeRemote work- Site Reliability Engineer (SRE)Mandatory skills: Azure DevOps (ADO), GitHub & GitHub Actions, JFrog ArtifactoryWill require end2end testing, performance testing, failover.Overall understanding of the Infra and Application to conduct load testing, failure recovery, KPIs,...Senior
- Hands-on practical experience in system design, application development, testing, and operational stabilityProficient in coding in one or more languages. Experience in developing, debugging, and maintaining code in a large corporate environment with one or more modern programming...Senior
- ...Senior Java Spring Boot DeveloperLocation: Richardson, TX 75082 (Onsite) Duration: 06 MonthsMust Have Skills: Springboot Java MicroservicesNice... ...practices. Optimize application performance, scalability, and reliability. Minimum years of experience 10 yearsTop 3 responsibilities:...Senior
- ...Job Description Job Description Senior Infrastructure Engineer Azure Platform Engineering Location: Plano, TX Work Schedule: Hybrid,... ...successful candidate will also contribute to platform standards, reliability, security, observability, and technical strategy. Key...SeniorContract workRemote work3 days per week
- ...Senior Platform Engineer At RTX, the world's largest aerospace and defense company, 185,000 great minds are united by purpose and inspired... ...C++, C#, Go). Experience troubleshooting performance, reliability, and availability issues across Linux/Windows based platforms...SeniorRelocation
- ...Senior Platform Engineer Our client, a Health Insurance company, is looking for a Senior Platform... ...capabilities that improve software delivery, reliability, scalability, security, and developer... ..., CI/CD & Engineering Productivity Site Reliability Engineering (SRE) &...SeniorWork experience placement
- ...Senior DeveloperVisa status: U.S. Citizens and those authorized to work in the U.S. are encouraged to apply. Tax Terms: W2, 1099 Corp-Corp or 3rd Parties: YesPosition Title: Senior DeveloperCandidate should be Citizen (Security clearance) Mandatory Skills:SQL Server Integration...Senior
- Senior Software EngineerThe Senior Software Engineer will help build and enhance a data fabric to support mission-critical GEOINT allocation and decision analytics... ...with a team across multiple geographic sites. This is an onsite role located in Richardson, Texas...SeniorRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Plano, TX
- site reliability engineer sre Plano, TX
- senior associate attorney Plano, TX
- senior developer Plano, TX
- senior aws cloud engineer Plano, TX
- remote senior salesforce administrator Plano, TX
- senior marketing operations manager Plano, TX
- senior manager tax Plano, TX
- senior property accountant Plano, TX
- senior tax Plano, TX




