Site Reliability Engineer Lead
Bank of America ATM
Job Description:At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. We do this by driving Responsible Growth and delivering for our clients, teammates, communities and shareholders every day.Being a Great Place to Work is core to how we drive Responsible Growth. This includes our commitment to being an inclusive workplace, attracting and developing exceptional talent, supporting our teammates’ physical, emotional, and financial wellness, recognizing and rewarding performance, and how we make an impact in the communities we serve.Bank of America is committed to an in-office culture with specific requirements for office-based attendance and which allows for an appropriate level of flexibility for our teammates and businesses based on role-specific considerations.At Bank of America, you can build a successful career with opportunities to learn, grow, and make an impact. Join us!Job Description:This job is responsible for building and leading a team to deliver technology products and services that meet business outcomes. Key responsibilities include developing a technology strategy, ensuring technology solutions comply with applicable standards, promoting design, engineering, and organizational practices, and advocating and advancing modern, Agile solution delivery practices. Job expectations may include coaching, mentoring, providing feedback and hands on career development, identifying emerging talent, fostering leadership skills, and managing stakeholders.Overview:Seeking a seasoned Site Reliability Engineering (SRE) Leader to drive the reliability, scalability, and performance of critical Infrastructure Automation platforms. This role will lead the design and implementation of SRE practices across a federated technology ecosystem, ensuring operational excellence through automation, observability, and resilient architecture.The ideal candidate will bring deep expertise in distributed systems, cloud-native infrastructure, SaaS application support and DevOps/SRE principles, along with strong leadership and collaboration skills to influence cross-functional engineering and Production management teams and drive continuous improvement in service reliability.Responsibilities:SRE Strategy & Governance:Define and implement SRE frameworks, including SLIs/SLOs/SLAs, error budgets, and incident response protocols.Establish governance models for reliability engineering across distributed teams.Champion a culture of observability, proactive monitoring, and continuous feedback loops.Reactive & Proactive Problem Management:Lead root cause analysis (RCA) and post-incident reviews to identify systemic issues and prevent recurrence.Implement proactive problem detection using telemetry, anomaly detection, and trend analysis.Collaborate with engineering and operations teams to eliminate toil and reduce incident frequency and impact.Capacity & Performance Management:Develop and maintain capacity models to ensure systems scale efficiently with business demand.Monitor performance trends and lead optimization efforts across infrastructure and applications.Partner with finance and engineering teams to align capacity planning with cost and growth objectives.Platform Reliability & Automation:Drive automation of operational tasks including deployments, scaling, and recovery.Integrate reliability tooling with CI/CD pipelines, ITSM platforms (e.g., ServiceNow), and observability systems.Incident Management & Operational Excellence:Oversee major incident response, escalation, and communication processes.Develop and maintain runbooks, playbooks, and escalation protocols.Drive continuous improvement through blameless retrospectives and operational reviews.Technical Leadership:Serve as a senior technical advisor and thought leader in SRE and platform engineering.Mentor and guide SRE teams and partner with engineering leaders across the enterprise.Provide input on staffing, tooling strategy, and budget planning for reliability initiatives.Managerial Responsibilities:This position may also have responsibilities for managing associates. At Bank of America, all managers at this level demonstrate the following responsibilities, in addition to those specific to the role, listed above.Opportunity & Inclusion Champion: Models an inclusive environment for employees and clients, aligned to company Great Place to Work goals.Manager of Process & Data: Demonstrates deep process knowledge, operational excellence and innovation through a focus on simplicity, data based decision making and continuous improvement.Enterprise Advocate & Communicator: Communicates enterprise decisions, purpose, and results, and connects to team strategy, priorities and contributions.Risk Manager: Ensures proper risk discipline, controls and culture are in place to identify, escalate and debate issues.People Manager & Coach: Provides inspection, coaching and feedback to motivate, differentiate and improve performance.Financial Steward: Actively manages expenses and budgets in alignment with objectives, making sound financial decisions.Enterprise Talent Leader: Assesses talent and builds bench strength for roles across the organization.Driver of Business Outcomes: Delivers results by effectively prioritizing, inspecting and appropriately delegating team work.Required Qualifications:10+ years of experience in systems engineering, DevOps, or SRE roles in large-scale environments.Deep understanding of Linux/Unix & Windows systems, networking, and distributed computing.Proven experience with observability stacks (e.g., Dynatrace, Grafana, Splunk, OpenTelemetry).Expertise in infrastructure-as-code and automation tools (e.g., Terraform, Ansible, Python).Strong knowledge of cloud platforms and container orchestration (Kubernetes).Demonstrated success in leading incident response and driving systemic improvements.Experience with capacity planning, performance tuning, and cost optimization.Excellent communication and stakeholder management skills, including executive engagement.Desired Qualifications:Experience with ITIL/ITSM processes and integration with platforms like ServiceNow.Familiarity with security and compliance in regulated industries (e.g., financial services).Background in performance engineering and infrastructure analytics.Experience developing dashboards and metrics for operational health and reliability.Skills:InfluenceRisk ManagementSolution DesignStakeholder ManagementTechnical Strategy DevelopmentAnalytical ThinkingApplication DevelopmentCollaborationResult OrientationSolution Delivery ProcessAgile PracticesArchitectureAutomationData ManagementDevOps PracticesShift:1st shift (United States of America)Hours Per Week: 40SummaryLocation: Plano; Chandler; CharlotteType: Full time
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and broad business problems with simple and straightforward solutions...Suggested
$152.6k - $191.5k
...for partnering with leaders across engineering and technology to define objective reliability goals for services. Key... ....Position Summary:The Senior GCP Site Reliability Engineer acts as an advanced... ...benefits eligible. We provide industry-leading benefits, access to paid time off...SuggestedFull timeWork at officeDay shift$117k - $209.33k
...Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable,... ...and implementing operational automation at scaleExperience leading or participating in Gamedays, disaster recovery exercises,...SuggestedFull timeFor contractorsRemote work- OpenArc - Empowering Your Career. As a leading IT staffing firm, we are dedicated to connecting talented professionals with your... ...ideal opportunities. We are currently seeking a qualified Site Reliability Engineer (AI & Agentic Systems) to join our client’s organization...Suggested
$96.8k - $145.2k
...thinking organization, apply now.We are currently seeking a Site Reliability Engineer (Onsite Hybrid) to join our team in Plano, Texas (US-TX),... ...through responsible innovation. We are one of the world's leading AI and digital infrastructure providers, with unmatched capabilities...SuggestedTemporary workWork at officeRemote workFlexible hours- ...: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8 to... ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will have...Remote work
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platforms, Web Hosting team , you hold a leadership role in your...Work experience placement
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office, Identity & Access Management team, you are the non-functional...Work at office
- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Corporate and Investment...
$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...Full timeTemporary workWork experience placementRemote work$48 per hour
...Site Reliability Engineer Trident Consulting is seeking a Site Reliability Engineer for one of our clients in Richardson, Texas or Scottsdale, Arizona. Job Title: Site Reliability Engineer Location: Richardson, Texas or Scottsdale, Arizona (Onsite) Length of Assignment...Contract work$141k - $208k
...over 250 percent year over year, ClickHouse leads the market in real-time analytics, data... ...committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building...Local areaRemote workHome officeFlexible hours- ...Site Reliability Engineer Hybrid Onsite Worker is required to work onsite 2-3 days per week in Phoenix, AZ OR Plano, TX Main Responsibilities ~ Experience in leading observability initiatives as lead engineer. ~ Development and implementation of build release...Work experience placement2 days per week3 days per week
- ...Site Reliability Engineer II Our client, a leading organization in the financial services industry, is seeking a Site Reliability Engineer II to join their team. As a Site Reliability Engineer II, you will be part of the Infrastructure Support Department supporting...Weekly pay
- ...Site Reliability Engineer (SRE) Our client, a IT Services and Consulting company, is looking for a Site Reliability Engineer (SRE) for their Plano, TX/Hybrid location. Requirements: Years of experience required: 11+ Mandatory skills: Azure DevOps (ADO), GitHub...
- ...broader AIOps strategy for TFS team members. You will lead end-to-end product vision across production and sub-productions, translating reliability needs into a measurable strategy. You will partner with SRE engineering, platform teams, and stakeholders to deliver...
- ...Site Reliability Engineer III Location: Plano, Texas (Hybrid) Duration: 18 months Role Overview We are seeking a Site Reliability Engineer to support the production and operations of critical applications. This role focuses on establishing and improving monitoring...Work experience placement
$69.62k - $99.45k
...progress and enhancing lives by providing reliable, high-speed connectivity solutions... ...for you! Job Summary As a Site Reliability Engineer II, you will be a primary driver in the... ...and Observability. You will lead the transition away from "toil" (manual...Permanent employmentLocal area$96.8k - $145.2k
...Site Reliability Engineer (Onsite Hybrid) NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking...Temporary workFlexible hours- ...outage situations • Single voice for our line of business and for cross line of business incidents • Influence senior technology leads across organizations to ensure timely resolution of incidents Required skills • Experience in troubleshooting, resolving, and...Remote work
- ...Senior Lead Site Reliability Engineer Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan...
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...Full timeWork at officeLocal areaRemote workWork from home- Company DescriptionSonsoft , Inc. is a USA based corporation duly organized under the laws of the Commonwealth of Georgia. Sonsoft Inc. is growing at a steady pace specializing in the fields of Software Development, => Software Consultancy and Information Technology Enabled...Permanent employmentFull timeH1b
- Company DescriptionSonsoft , Inc. is a USA based corporation duly organized under the laws of the Commonwealth of Georgia. Sonsoft Inc. is growing at a steady pace specializing in the fields of Software Development, => Software Consultancy and Information Technology Enabled...Permanent employmentFull timeH1b
- Company DescriptionSonsoft , Inc. is a USA based corporation duly organized under the laws of the Commonwealth of Georgia. Sonsoft Inc. is growing at a steady pace specializing in the fields of Software Development, => Software Consultancy and Information Technology Enabled...Permanent employmentFull timeH1bLocal area
- Sprouts Farmers Market, Inc. is looking for an Assistant Meat Manager in Murphy, Texas. This role is vital to enhancing customer experiences by managing the meat and seafood department effectively. The ideal candidate will assist in supervising a team, ensuring excellent...
- ...specifications, troubleshoots and testing.Actively involved with requirement understanding and analysis.Work closely with functional leads/PMs to understand the partner integration requirements.Additional responsibilities as deemed necessary for the role.Qualifications1+...Permanent employmentFull timeH1b
$66.73 per hour
...Description We are seeking a Site Reliability Engineer (SRE) to help establish and scale our client's Google Cloud Platform (GCP) SRE practice... ...About TEKsystems and TEKsystems Global Services We’re a leading provider of business and technology services. We accelerate...Contract workTemporary work$160k - $240k
...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users. We...Permanent employmentFull timeRemote workWork from homeRelocationFlexible hours- ...Valued at US$8 billion and backed by world-leading investors including T. Rowe Price, Visa,... ...division and is responsible for the reliability, performance, security, and automation of... ...is to make databases invisible: product engineers should be able to provision, scale, and...Worldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer Lead. Be the first to apply!

