Lead Site Reliability Engineer
SunTrust Investment Services, Inc.
Site Reliability Engineering LeadThe Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability, and incident management while collaborating across multiple business and technology teams. Responsibilities include leading major incident responses, driving problem management, and implementing automation to reduce service downtime. The role involves standardizing observability practices, mentoring SRE team members, and contributing to enterprise-wide reliability frameworks. Candidates require 7+ years of experience, expertise in distributed systems, Kubernetes, automation scripting, and strong leadership in incident management.Essential Duties And Responsibilities Following is a summary of the essential functions for this job. Other duties may be performed, both major and minor, which are not mentioned below. Specific activities may change from time to time.Implements software architecture and engineering approaches for complex initiatives within the job area, contributing to technical plans and working to achieve operational targets with major impact on results.Adopts and refines advanced software engineering standards, practices, and governance mechanisms for the job area, influencing how multiple teams improve quality, reliability, and delivery.Collaborates with senior engineers, product partners, and architecture teammates to shape technology approaches for the domain, providing deep technical insight and proposing solution patterns that inform local roadmaps and priorities.Leads the end-to-end technical design and implementation of scalable, secure, and highly available software solutions for the job area, producing patterns and examples that other technical professionals can follow.Independently troubleshoots and resolves complex technical issues in the area of responsibility, designing innovative architectures and performance, reliability, and scalability improvements that advance business objectives.Provides ongoing technical guidance, coaching, and training to other engineers, delegating and reviewing work from lower-level technical professionals and raising the technical bar through design reviews and knowledge sharing.Evaluates emerging technologies and techniques relevant to the job area, building prototypes and solution concepts that contribute measurable input into new features, products, or capabilities.Contributes to the development of long-term technical goals and plans for the area of responsibility through well-reasoned recommendations, design proposals, and implementation experience.Leads large or complex initiatives within the job area, coordinating and delegating technical work that may span outside the immediate team, and ensuring cohesive, high-quality outcomes with limited supervision.Qualifications Required Qualifications The requirements listed below are representative of the knowledge, skill and/or ability required. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.Bachelor's degree in Computer Science, Software Engineering, or related field.Minimum of 7 years of professional experience in software development.Deep knowledge of multiple programming languages, software architecture, and design principles.Deep understanding of software development lifecycle, testing, deployment, and security practices.Preferred QualificationsAdvanced degree in Computer Science or related technical discipline.Professional certifications such as Certified Software Development Professional (CSDP) or equivalent.Deep expertise in cloud-native architectures, microservices, container orchestration, and DevOps.Strong familiarity with Agile frameworks, continuous integration/continuous deployment (CI/CD), and enterprise innovation management.7+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Infrastructure Operations.Deep hands-on experience with distributed systems, container orchestration (Kubernetes), and cloud-native operational tooling.Proficiency with automation and scripting languages (Python, Go, PowerShell, Ansible).Strong understanding of observability platforms (Splunk, Dynatrace) and event-driven monitoring.Proven leadership in major incident management and cross-team technical coordination.Strong grasp of networking, Linux/Unix internals, and modern infrastructure patterns.Excellent communication skills, including executive-level situational awareness during critical incidents.Demonstrated ability to influence technical roadmaps and drive adoption of reliability best practices.Preferred QualificationsFinancial services or regulated industry experience.Experience enabling large-scale SRE transformations or modernization initiatives.Familiarity with chaos engineering, resilience assessments, and service failure modeling.Exposure to hybrid-cloud and multi-cloud operational frameworks.Experience contributing to or leading Center for Enablement functions or Communities of Practice.Key ResponsibilitiesIncident & Problem Management LeadershipLead major and high-severity incident response efforts, focusing on diagnosing technical root causes therein, and driving multi-team technical resolution.Drive problem management to closure, ensuring systemic fixes replace recurring operational risks.Establish and maintain standardized incident playbooks, escalation paths, and communication frameworks.Reliability Engineering & AutomationArchitect and deliver automation solutions that eliminate toil, reduce MTTR, and increase service resilience.Implement intelligent alerting, anomaly detection, and event correlation leveraging AI and AIOps tools.Guide and enforce SLO/SLI adoption across product teams, ensuring metrics inform decision-making and prioritization.Observability & Operational ExcellenceEnhance telemetry coverage across logs, metrics, traces, and events using platforms such as Dynatrace and Splunk.Define and standardize enterprise observability practices, dashboards, and KPIs.Ensure operational readiness of applications and platforms through resiliency testing, chaos engineering, and failure-mode validation.Cross-Functional Leadership & InfluencePartner with Delivery, Architecture, Security, and Risk teams to embed reliability and resilience into design and execution.Act as a change agent to elevate operational maturity and drive transformative improvements across Wholesale.Lead workshops, maturity assessments, and enablement sessions through the SRE C4E and Communities of Practice.Standardization & DocumentationDevelop, maintain, and enforce runbooks, response playbooks, and automated recovery patterns.Contribute to enterprise SRE frameworks, templates, and maturity models.Promote consistent adoption of best practices across domains and lines of business.Mentorship & Technical DevelopmentCoach and mentor Associate, Professional, and Senior SREs to build technical depth and operational discipline.Provide thought leadership in SRE methodologies, cloud-native operational patterns, and automated reliability engineering.Candidate must be willing to work onsite Monday - Friday at either office in Charlotte NC, Raleigh NC, or Atlanta, GA.
- Role Profile:We are evolving our Site Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence... ...a continuous improvement approach to everything they do!Lead the establishment of SRE foundations for new projects...SuggestedFull timeShift work
- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...SuggestedWork at officeImmediate start
$118.6k - $195.68k
...SummaryThe Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat... ...Red Hat Product Engineering teamsDesign software tests and lead peer reviews to increase the quality of our codebaseHelp and...SuggestedPermanent employmentFull timeContract workWork experience placementWork at officeRemote workFlexible hours- ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and...SuggestedPermanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...SuggestedFull timeTemporary workWork experience placementRemote work- ...About Us Epic Kids is the leading digital reading platform built for kids 12 and under, trusted by millions of children, educators... ...and literacy. About the Role We're looking for a Senior Site Reliability Engineer to drive the stability, observability, and reliability of...Remote work
$110k - $270k
...and communities. The Role Join our dynamic team as a Senior Site Reliability Engineer on the Vault Platform team, where you'll ensure the... ...global customers (across North America, Europe, and Asia) Lead Incident Management: During an incident, effectively lead triage...Work at officeLocal areaRemote workWork from homeMonday to FridayFlexible hours$119k - $170k
...the greater good, come make your next move with Zscaler. Our Engineering team built the world’s largest cloud security platform from... ...cloud-first strategy. We’re looking for an experienced Staff Site Reliability Engineer (Federal) to join our Government Cloud team....Full timeWork at officeLocal areaWorldwideNight shift- ...SRE Support Engineer - Observability While this position is not currently open, we are interviewing strong candidates for upcoming opportunities... ...support across Slack and tickets, improving monitoring reliability, and reducing incident impact through better triage,...Remote work
- ...each enterprise customer, around the clock. The Senior Manager leads the 24x7 incident response, change management, and release delivery... ...-on technical depth with people leadership, drives AI-driven automation, and collaborates across engineering, product, #J-18808-Ljbffr...
$120k - $137.5k
...Site Reliability / DevOps EngineerLocation: Raleigh, North Carolina, USType: Full-timeDepartment... ...is seeking a motivated SRE/DevOps Engineer with strong observability experience to... ...for next stepsAbout eClerxeClerx is a leading provider of productized services, bringing...Work experience placement- ...SoftPro is the nation's leading provider of real estate closing and title insurance software. A division of Fidelity National... ...are we looking for? SoftPro is seeking a well-rounded Site Reliability Engineer (SRE) to join our Cloud Operations Team in our Raleigh,...Hourly payWork at officeRemote work
$174k - $252k
Senior Software Engineer, Site Reliability Engineering X Note: By applying to this position you will have an opportunity to share your preferred... ...large-scale distributed systems. 2 years of experience leading projects and providing technical leadership. Preferred qualifications...Full time$100k - $153k
...Job Title CaaS Private Site Reliability EngineerCorporate Title Assistant Vice PresidentLocation Cary, NCIn short – an essential part of... ...technology strategy while transforming it within a robust, hands-on engineering culture. Learning is a key element of our people strategy,...Work at officeWork from homeShift work- ...MetLife is seeking a highly skilled Lead Mainframe CICS/MQ Site Reliability Engineer to modernize and optimize mainframe platforms. You will bridge legacy reliability with future-ready solutions, architecting, designing, and directing modernization initiatives across...
$125k - $185k
...Position Overview J ob Title: CaaS Private Site Reliability Engineer Corporate Title: Vice President Location: Cary, NC Who we... ...Overview As a CaaS Private Site Reliability Engineer, you will lead reliability, resilience, and operational excellence for the...Full timeWork at officeWork from homeShift work$91.4k - $187k
...configuration design to meet workflow requirements and make recommendations to clients. You will mitigate solution risks and issues and lead client meetings and events. You will execute workflow and process improvement strategies that advance how healthcare is provided...Contract workTemporary workLocal areaFlexible hours$125k - $185k
...Position Overview Job Title Platform Engineering & Site Reliability Engineering VP Corporate Title Vice President Location Cary, NC... ...matching gift and volunteer programs What You’ll Do Lead the design, development, and evolution of cloud-native platform...Work at officeWork from homeShift work$150k - $165k
...collaborate, and innovate with some of the most talented people out there, WalkMe is the place for you!WalkMe is hiring a AI Automation Lead, Post-Sales Operations to serve as the Operations focal point for our Post-Sales CSG functions. The CSG Operations team is the right...Full timeWork experience placementImmediate startRemote workShift work$110k - $160.91k
...This Role MattersAs the demand for reliable, resilient, and sustainable... ...essential. This role is focused on leading and overseeing routing and siting efforts for electric transmission... ...experts across environmental, land use, engineering, and stakeholder engagement disciplines...Full timeFixed term contractCasual workWork at officeRemote workFlexible hoursShift work$116.2k - $229.1k
...communication skillsMeticulous attention to detail and quality of work productAbility to build and sustain professional relationshipsAbility to lead projects or workstreamsAbility to manage and prioritize multiple tasks in a fast-paced and dynamic environmentStrong interpersonal...Local area$161.93k - $243.5k
Job DescriptionSetidegrasib Launch Lead (NSCLC), US OncologyAbout Astellas Astellas is a global life sciences company committed to turning innovative science into VALUE for patients. We provide transformative therapies in disease areas that include oncology, ophthalmology...Work at officeRemote workWork from homeWorldwideFlexible hours- Job DescriptionHi,Hope you are doing great. We have an urgent positions with our direct client, kindly send me your updated resume if suitable for following job des. Title: Android LeadLocation: Raleigh, NCDuration: Contract Note: Do provide Visa copy as well as I-94 Copy...Contract workWork experience placementRemote work
- ...integrating the front end with RESTful API’s and event driven microservices is beneficialExperience with Cloud architecture and engineering is beneficial, ideally on Microsoft AzureGood communication skills - both written and verbalStrong analytical and problem-solving...Full time
$148k - $195k
Make your mark for patientsWe are looking for a Global Regulatory Affairs (GRA) Device Lead who is strategic, collaborative, and regulatory focused to join us in our Regulatory CMC & Devices team, based in our Raleigh, NC office.About the roleThe Global Regulatory Affairs...Permanent employmentWork at officeLocal area- ...but around the world. We believe building engineering is more than systems and structures, it’s... ...This isn’t just a job, it’s a chance to lead innovation, engineer impact, and build a... ...legacy of excellence.HDR is is looking for a Site Solutions Team Lead to assist with site...Local area
$92.5k - $124.25k
Ready to lead? Join ERM and help shape the future of sustainable energy. Bring your expertise to a team that values innovation, collaboration... ...Impact IsAs a Managing Consultant, Electric Transmission Line Siting Lead, you’ll provide a key role in ERM’s electric transmission...Full timeFixed term contractCasual workWork at officeRemote workFlexible hours- ...become part of a global team of over 50,000 planners, designers, engineers, scientists, digital innovators, program and construction... ...candidate is experienced delivering in a fast-paced environment while leading complex projects, leading diverse and geographically...Work at officeLocal areaWorldwideVisa sponsorshipFlexible hours
- ...information on Guerbet, go to and follow Guerbet on Linkedin, Twitter, Instagram and YoutubeWHAT WE ARE LOOKING FORThe QC Chemistry Lead operates in full compliance with applicable hygiene, safety, quality, and regulatory standards, and in alignment with the organization...Local area
- ...Microbiology and Sterility Assurance Manager, the QC Microbiology Lead is responsible for the day-to-day operations of the QC... ...the skills to interact with formulation, filling, validation, engineering, and packaging personnel in support of those functions. The qualified...Local areaImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead Site Reliability Engineer. Be the first to apply!
- lead engineer Raleigh, NC
- lead operating engineer Raleigh, NC
- site reliability engineer sre Raleigh, NC
- site reliability engineer Raleigh, NC
- on-site clinical research associate (traveling/remote) Raleigh, NC
- website coordinator Raleigh, NC
- junior website developer Raleigh, NC
- site leader Raleigh, NC
- historic site Raleigh, NC
- website content developer Raleigh, NC


