Site Reliability Engineering
Truist Inc
The position is described below. If you want to apply, click the Apply Now button at the top or bottom of this page. After you click Apply Now and complete your application, you'll be invited to create a profile, which will let you see your application status and any communications. If you already have a profile with us, you can log in to check status.Need Help?If you have a disability and need assistance with the application, you can request a reasonable accommodation. Send an email to Accessibility (accommodation requests only; other inquiries won't receive a response).Regular or Temporary:RegularLanguage Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability, and incident management while collaborating across multiple business and technology teams.Responsibilities include leading major incident responses, driving problem management, and implementing automation to reduce service downtime.The role involves standardizing observability practices, mentoring SRE team members, and contributing to enterprise-wide reliability frameworks.Candidates require 7+ years of experience, expertise in distributed systems, Kubernetes, automation scripting, and strong leadership in incident management.ESSENTIAL DUTIES AND RESPONSIBILITIESFollowing is a summary of the essential functions for this job. Other duties may be performed, both major and minor, which are not mentioned below. Specific activities may change from time to time.1. Implements software architecture and engineering approaches for complex initiatives within the job area, contributing to technical plans and working to achieve operational targets with major impact on results.2. Adopts and refines advanced software engineering standards, practices, and governance mechanisms for the job area, influencing how multiple teams improve quality, reliability, and delivery.3. Collaborates with senior engineers, product partners, and architecture teammates to shape technology approaches for the domain, providing deep technical insight and proposing solution patterns that inform local roadmaps and priorities.4. Leads the end-to-end technical design and implementation of scalable, secure, and highly available software solutions for the job area, producing patterns and examples that other technical professionals can follow.5. Independently troubleshoots and resolves complex technical issues in the area of responsibility, designing innovative architectures and performance, reliability, and scalability improvements that advance business objectives.6. Provides ongoing technical guidance, coaching, and training to other engineers, delegating and reviewing work from lower-level technical professionals and raising the technical bar through design reviews and knowledge sharing.7. Evaluates emerging technologies and techniques relevant to the job area, building prototypes and solution concepts that contribute measurable input into new features, products, or capabilities.8. Contributes to the development of long-term technical goals and plans for the area of responsibility through well-reasoned recommendations, design proposals, and implementation experience.9. Leads large or complex initiatives within the job area, coordinating and delegating technical work that may span outside the immediate team, and ensuring cohesive, high-quality outcomes with limited supervision.QualificationsRequired QualificationsThe requirements listed below are representative of the knowledge, skill and/or ability required. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.1. Bachelor’s degree in Computer Science, Software Engineering, or related field.2. Minimum of 7 years of professional experience in software development.3. Deep knowledge of multiple programming languages, software architecture, and design principles.4. Deep understanding of software development lifecycle, testing, deployment, and security practices.Preferred Qualifications1. Advanced degree in Computer Science or related technical discipline.2. Professional certifications such as Certified Software Development Professional (CSDP) or equivalent.3. Deep expertise in cloud-native architectures, microservices, container orchestration, and DevOps.4. Strong familiarity with Agile frameworks, continuous integration/continuous deployment (CI/CD), and enterprise innovation management.5. 7+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Infrastructure Operations. 6. Deep hands‑on experience with distributed systems, container orchestration (Kubernetes), and cloud-native operational tooling.7. Proficiency with automation and scripting languages (Python, Go, PowerShell, Ansible).8. Strong understanding of observability platforms (Splunk, Dynatrace) and event-driven monitoring.9. Proven leadership in major incident management and cross-team technical coordination.10. Strong grasp of networking, Linux/Unix internals, and modern infrastructure patterns.11. Excellent communication skills, including executive-level situational awareness during critical incidents.12. Demonstrated ability to influence technical roadmaps and drive adoption of reliability best practices.Preferred QualificationsFinancial services or regulated industry experience.Experience enabling large-scale SRE transformations or modernization initiatives. Familiarity with chaos engineering, resilience assessments, and service failure modeling.Exposure to hybrid-cloud and multi-cloud operational frameworks.Experience contributing to or leading Center for Enablement functions or Communities of Practice.Key ResponsibilitiesIncident & Problem Management LeadershipLead major and high-severity incident response efforts,focusing on diagnosing technical rootcausestherein, and drivingmulti-team technical resolution.Drive problem management to closure, ensuring systemic fixes replace recurring operational risks.Establish and maintain standardized incident playbooks, escalation paths, and communication frameworks.Reliability Engineering & AutomationArchitect and deliver automation solutions that eliminate toil, reduce MTTR, and increase service resilience.Implement intelligent alerting, anomaly detection, and event correlation leveragingAI andAIOps tools.Guide and enforce SLO/SLI adoption across product teams, ensuring metrics inform decision-making and prioritization.Observability & Operational ExcellenceEnhance telemetry coverage across logs, metrics, traces, and events using platforms such as Dynatrace and Splunk.Define and standardize enterprise observability practices, dashboards, and KPIs.Ensure operational readiness of applications and platforms through resiliency testing, chaos engineering, and failure-mode validation.Cross-Functional Leadership & InfluencePartner with Delivery, Architecture, Security, and Risk teams to embed reliability and resilience into design and execution.Act as a change agent to elevate operational maturity and drive transformative improvements acrossWholesale.Lead workshops, maturity assessments, and enablement sessions through the SRE C4E and Communities of Practice.Standardization & DocumentationDevelop, maintain, and enforce runbooks, response playbooks, and automated recovery patterns.Contribute to enterprise SRE frameworks, templates, and maturity models. Promote consistent adoption of best practices across domains and lines of business.Mentorship & Technical DevelopmentCoach and mentor Associate, Professional, and Senior SREs to build technical depth and operational discipline.Provide thought leadership in SRE methodologies, cloud-native operational patterns, and automated reliability engineering.For this opportunity, Truist will not sponsor an applicant for work visa status or employment authorization, nor will we offer any immigration-related support for this position (including, but not limited to H-1B, F-1 OPT, F-1 STEM OPT, F-1 CPT, J-1, TN-1 or TN-2, E-3, O-1, or future sponsorship for U.S. lawful permanent residence status.)Candidate must be willing to work onsite Monday - Friday at either office in Charlotte NC, Raleigh NC, or Atlanta, GA.General Description of Available Benefits for Eligible Employees of Truist Financial Corporation: All regular teammates (not temporary or contingent workers) working 20 hours or more per week are eligible for benefits, though eligibility for specific benefits may be determined by the division of Truist offering the position.Truist offers medical, dental, vision, life insurance, disability, accidental death and dismemberment, tax-preferred savings accounts, and a 401k plan to teammates. Teammates also receive no less than 10 days of vacation (prorated based on date of hire and by full-time or part-time status) during their first year of employment, along with 10 sick days (also prorated), and paid holidays. For more details on Truist’s generous benefit plans, please visit our Benefits site. Depending on the position and division, this job may also be eligible for Truist’s defined benefit pension plan, restricted stock units, and/or a deferred compensation plan. As you advance through the hiring process, you will also learn more about the specific benefits available for any non-temporary position for which you apply, based on full-time or part-time status, position, and division of work.Truist is an Equal Opportunity Employer that does not discriminate on the basis of race, gender, color, religion, citizenship or national origin, age, sexual orientation, gender identity, disability, veteran status, or other classification protected by law. Truist is a Drug Free Workplace.EEO is the LawE-VerifyIER Right to WorkJob SummaryJob number: R Profession: Technology
- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...SuggestedWork at officeImmediate start
$118.6k - $195.68k
Job SummaryThe Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at scale...SuggestedPermanent employmentFull timeContract workWork experience placementWork at officeRemote workFlexible hours- ...Site Reliability Engineer The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability...SuggestedWork at officeLocal areaImmediate startMonday to Friday
$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...SuggestedWork experience placementWork at office$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...SuggestedPermanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$168k - $200k
...is passionate about creating transformative change in healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the forefront of building and operating a resilient, observable, and...Remote work- Role Profile:We are evolving our Site Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a Senior SRE, you will be a senior hands‑on technical person help...Full timeShift work
$55k - $151.47k
...ApplicableSpecialismIFS - Internal Firm Services - OtherManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...Full timeH1b$105.79k - $141.05k
...delivers on-demand networking at scale. As Lead SRE, you'll own the reliability of that platform — partnering with operations teams and... ..., and automation, and you'll coordinate across architecture, engineering, and systems development organizations to measurably improve...Temporary workRemote work$84.9k - $209.5k
...About the Opportunity Help ensure healthcare professionals can reliably access the applications they depend on to deliver patient care. Oracle Health is seeking a Principal Site Reliability Engineer to strengthen the reliability, performance, security, and...Temporary workImmediate startFlexible hoursShift work- ...Site Reliability Engineer Based in Jacksonville, FL, Cary, NC, location for a Fulltime position. Should be having cloud engineering experience and acting as the SME on operation automation and monitoring, identifying TOIL within the teams existing systems and processes...Full time
- ...Site Reliability Engineer (SRE) - Security Infrastructure Position Summary We are seeking an SRE to support reliability, scalability, and operational excellence for a large-scale network security transformation initiative. This role will focus on monitoring, automation...
- ...technologies to enable scalable, secure, and reliable business operations. Applies strong... ...infrastructure.3. Manages infrastructure engineering projects and processes aligned with... ...benefit plans, please visit our Benefits site. Depending on the position and division,...Permanent employmentFull timePart timeWork experience placementH1bRemote workWork visaShift workWeekend workDay shift
$125k - $185k
...Position Overview J ob Title: CaaS Private Site Reliability Engineer Corporate Title: Vice President Location: Cary, NC Who we are: In short – an essential part of Deutsche Bank’s technology solution, developing applications for key business areas...Full timeWork at officeWork from homeShift work$130k - $180k
Position OverviewPower your future with Qualus as a Lead Relay Settings Engineer. In this role you will perform Protective Relay Design & Coordination: Design, specify, calculate settings, and coordinate protective relays and relay control schemes. Do you have 7+ years...Temporary workFlexible hours$136.09k - $168.11k
...looking to move fast and make a significant impact in an exciting space, you're in the right place!We are seeking a Lead Solution Engineer to join our North America GTM team, specifically supporting our Industry Verticals organization across SLED (State, Local & Education...Local areaFlexible hours$100k - $153k
...Position Overview J ob Title CaaS Private Site Reliability Engineer Corporate Title Assistant Vice President Location Cary, NC Who we are: In short – an essential part of Deutsche Bank’s technology solution, developing applications for key business...Full timeWork at officeWork from homeShift work- ...AI infrastructure, working with server, cloud, and platform engineering teams.Operationalize machine learning workflows and support AI... ...implement system enhancements to improve performance, scalability, reliability, and cost efficiency.Collaborate across divisions to support...Full timeWork at officeRemote work
- ..., Iowa, Oklahoma, California, Pennsylvania and Florida. Job Description CaptiveAire is looking for a strong Senior Software Engineer to join the CASLink team onsite at our corporate location in Raleigh. CASLink is an IoT Solution serving as CaptiveAire's proprietary...Full timeRelocation packageFlexible hours
- ...Description DomainTools is hiring a Senior Software Engineer to join our Engineering Team. This role is perfect for a creative problem-solver with deep expertise in large-scale data engineering and a passion for technical leadership and continuous improvement. You...Full timeImmediate startFlexible hours
$23 - $27 per hour
...education, licensure requirements and/or skill level. Hourly Rate: $23 - $27/hr. OVERVIEW To perform embedded software engineering design services and other related software development support services throughout all stages of the product development life...Hourly payContract workWork experience placementSummer workInternship$140k - $200k
...– Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and... ...→ testing → release → maintenance. Ensure quality, reliability, and consistency across releases. Identify, diagnose, and resolve...Full timeWork at office- ...and want to make a meaningful impact, we'd love to hear from you. Position Summary Vadum is seeking a talented Software Engineer to join our growing engineering team. In this role, you will research, design, develop, and test software solutions that address...Full timeFlexible hours
- ...VAST Data is looking for a Software Engineer -A new college Grad to join our growing team! This is a great opportunity to be part of one of the fastest-growing infrastructure companies in history, an organization that is in the center of the hurricane being created by...Full time
$184k - $287.5k
...inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.We are looking for a dedicated engineer for the Senior Systems Software Engineer role, focusing on GPU Performance at Scale. At NVIDIA, this role is uniquely positioned...Full timeRemote work$130k - $170k
About Us:BW Design Group is a fully integrated architecture, engineering, construction, system integration, and consulting firm committed to helping our clients realize their most critical goals from Strategy to Commercialization. As the only firm born from a manufacturing...Full timeFlexible hours- ...software applications Actively participate in resolving production issues and recommend preventive strategies to enhance system reliability Maintain detailed records of code, testing techniques, and support activities to enrich the knowledge base and assist other similar...Full timeTemporary workRelocation
- ...Jewelers Mutual is seeking a highly skilled Senior Software Engineer to lead the technical evolution of our agent-facing web applications... ...modernization efforts, and a passion for building intuitive, reliable B2C and B2B applications. About Jewelers Mutual Jewelers...Full time
- ...looking for an experienced software developer with strong skills in C, C++, or Python to join our V-Force team —a specialized engineering squad embedded within R&D. this is a hands-on development position focused on solving the most challenging and high-impact problems...Full time
- ...enrollment journey for all. We are seeking a Senior Software Engineer to lead the design, development, and optimization of our call... ...platform. You will play a critical role in building scalable, reliable, and intelligent telephony solutions that enable seamless customer...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineering. Be the first to apply!
- IT site lead Raleigh, NC
- site safety Raleigh, NC
- website content developer Raleigh, NC
- site leader Raleigh, NC
- on-site clinical research associate (traveling/remote) Raleigh, NC
- junior website developer Raleigh, NC
- historic site Raleigh, NC
- on site coordinator Raleigh, NC
- site recruiter Raleigh, NC
- construction site safety Raleigh, NC



