Site Reliability Engineer
SunTrust Investment Services, Inc.
Site Reliability Engineering LeadThe Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability, and incident management while collaborating across multiple business and technology teams. Responsibilities include leading major incident responses, driving problem management, and implementing automation to reduce service downtime. The role involves standardizing observability practices, mentoring SRE team members, and contributing to enterprise-wide reliability frameworks. Candidates require 7+ years of experience, expertise in distributed systems, Kubernetes, automation scripting, and strong leadership in incident management.Essential Duties And Responsibilities Following is a summary of the essential functions for this job. Other duties may be performed, both major and minor, which are not mentioned below. Specific activities may change from time to time.Implements software architecture and engineering approaches for complex initiatives within the job area, contributing to technical plans and working to achieve operational targets with major impact on results.Adopts and refines advanced software engineering standards, practices, and governance mechanisms for the job area, influencing how multiple teams improve quality, reliability, and delivery.Collaborates with senior engineers, product partners, and architecture teammates to shape technology approaches for the domain, providing deep technical insight and proposing solution patterns that inform local roadmaps and priorities.Leads the end-to-end technical design and implementation of scalable, secure, and highly available software solutions for the job area, producing patterns and examples that other technical professionals can follow.Independently troubleshoots and resolves complex technical issues in the area of responsibility, designing innovative architectures and performance, reliability, and scalability improvements that advance business objectives.Provides ongoing technical guidance, coaching, and training to other engineers, delegating and reviewing work from lower-level technical professionals and raising the technical bar through design reviews and knowledge sharing.Evaluates emerging technologies and techniques relevant to the job area, building prototypes and solution concepts that contribute measurable input into new features, products, or capabilities.Contributes to the development of long-term technical goals and plans for the area of responsibility through well-reasoned recommendations, design proposals, and implementation experience.Leads large or complex initiatives within the job area, coordinating and delegating technical work that may span outside the immediate team, and ensuring cohesive, high-quality outcomes with limited supervision.Qualifications Required Qualifications The requirements listed below are representative of the knowledge, skill and/or ability required. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.Bachelor's degree in Computer Science, Software Engineering, or related field.Minimum of 7 years of professional experience in software development.Deep knowledge of multiple programming languages, software architecture, and design principles.Deep understanding of software development lifecycle, testing, deployment, and security practices.Preferred QualificationsAdvanced degree in Computer Science or related technical discipline.Professional certifications such as Certified Software Development Professional (CSDP) or equivalent.Deep expertise in cloud-native architectures, microservices, container orchestration, and DevOps.Strong familiarity with Agile frameworks, continuous integration/continuous deployment (CI/CD), and enterprise innovation management.7+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Infrastructure Operations.Deep hands-on experience with distributed systems, container orchestration (Kubernetes), and cloud-native operational tooling.Proficiency with automation and scripting languages (Python, Go, PowerShell, Ansible).Strong understanding of observability platforms (Splunk, Dynatrace) and event-driven monitoring.Proven leadership in major incident management and cross-team technical coordination.Strong grasp of networking, Linux/Unix internals, and modern infrastructure patterns.Excellent communication skills, including executive-level situational awareness during critical incidents.Demonstrated ability to influence technical roadmaps and drive adoption of reliability best practices.Preferred QualificationsFinancial services or regulated industry experience.Experience enabling large-scale SRE transformations or modernization initiatives.Familiarity with chaos engineering, resilience assessments, and service failure modeling.Exposure to hybrid-cloud and multi-cloud operational frameworks.Experience contributing to or leading Center for Enablement functions or Communities of Practice.Key ResponsibilitiesIncident & Problem Management LeadershipLead major and high-severity incident response efforts, focusing on diagnosing technical root causes therein, and driving multi-team technical resolution.Drive problem management to closure, ensuring systemic fixes replace recurring operational risks.Establish and maintain standardized incident playbooks, escalation paths, and communication frameworks.Reliability Engineering & AutomationArchitect and deliver automation solutions that eliminate toil, reduce MTTR, and increase service resilience.Implement intelligent alerting, anomaly detection, and event correlation leveraging AI and AIOps tools.Guide and enforce SLO/SLI adoption across product teams, ensuring metrics inform decision-making and prioritization.Observability & Operational ExcellenceEnhance telemetry coverage across logs, metrics, traces, and events using platforms such as Dynatrace and Splunk.Define and standardize enterprise observability practices, dashboards, and KPIs.Ensure operational readiness of applications and platforms through resiliency testing, chaos engineering, and failure-mode validation.Cross-Functional Leadership & InfluencePartner with Delivery, Architecture, Security, and Risk teams to embed reliability and resilience into design and execution.Act as a change agent to elevate operational maturity and drive transformative improvements across Wholesale.Lead workshops, maturity assessments, and enablement sessions through the SRE C4E and Communities of Practice.Standardization & DocumentationDevelop, maintain, and enforce runbooks, response playbooks, and automated recovery patterns.Contribute to enterprise SRE frameworks, templates, and maturity models.Promote consistent adoption of best practices across domains and lines of business.Mentorship & Technical DevelopmentCoach and mentor Associate, Professional, and Senior SREs to build technical depth and operational discipline.Provide thought leadership in SRE methodologies, cloud-native operational patterns, and automated reliability engineering.Candidate must be willing to work onsite Monday - Friday at either office in Charlotte NC, Raleigh NC, or Atlanta, GA.
- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...SuggestedWork at officeImmediate start
$118.6k - $195.68k
Job SummaryThe Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at scale...SuggestedPermanent employmentFull timeContract workWork experience placementWork at officeRemote workFlexible hours- Role Profile:We are evolving our Site Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a Senior SRE, you will be a senior hands‑on technical person help...SuggestedFull timeShift work
$120k - $137.5k
...Site Reliability / DevOps EngineerLocation: Raleigh, North Carolina, USType: Full-timeDepartment: BFSIJob SummaryeClerx is seeking a motivated SRE/DevOps Engineer with strong observability experience to join our growing Platform Engineering team. This team is responsible...SuggestedWork experience placement$110k - $270k
...committed to making a positive impact on its customers, employees, and communities. The Role Join our dynamic team as a Senior Site Reliability Engineer on the Vault Platform team, where you'll ensure the scalability and reliability of our enterprise applications. You'll...SuggestedWork at officeLocal areaRemote workWork from homeMonday to FridayFlexible hours$119k - $170k
...the greater good, come make your next move with Zscaler. Our Engineering team built the world’s largest cloud security platform from... ...cloud-first strategy. We’re looking for an experienced Staff Site Reliability Engineer (Federal) to join our Government Cloud team....Full timeWork at officeLocal areaWorldwideNight shift$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...Full timeTemporary workWork experience placementRemote work- ...partnering with two technical leads to ensure health, currency, and security of thousands of VMs. We seek a leader who balances hands-on technical depth with people leadership, drives AI-driven automation, and collaborates across engineering, product, #J-18808-Ljbffr...
- ...SRE Support Engineer - Observability While this position is not currently open, we are interviewing strong candidates for upcoming opportunities... ...support across Slack and tickets, improving monitoring reliability, and reducing incident impact through better triage,...Remote work
- ...SoftPro has won this prestigious award 14 times since 2012! What are we looking for? SoftPro is seeking a well-rounded Site Reliability Engineer (SRE) to join our Cloud Operations Team in our Raleigh, NC office or as a remote employee. This team supports our...Hourly payWork at officeRemote work
- ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
$160k - $240k
...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users. We...Permanent employmentFull timeRemote workWork from homeRelocationFlexible hours$174k - $252k
Senior Software Engineer, Site Reliability Engineering X Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Raleigh, NC, USA; Durham, NC, USA . Bachelor’s degree in Computer Science, Engineering...Full time$151k - $297k
...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB's cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and...Local areaRemote workWorldwideFlexible hours$100k - $153k
...Job Title CaaS Private Site Reliability EngineerCorporate Title Assistant Vice PresidentLocation Cary, NCIn short – an essential part of... ...technology strategy while transforming it within a robust, hands-on engineering culture. Learning is a key element of our people strategy,...Work at officeWork from homeShift work- ...MetLife is seeking a highly skilled Lead Mainframe CICS/MQ Site Reliability Engineer to modernize and optimize mainframe platforms. You will bridge legacy reliability with future-ready solutions, architecting, designing, and directing modernization initiatives across MetLife...
$125k - $185k
...Position Overview J ob Title: CaaS Private Site Reliability Engineer Corporate Title: Vice President Location: Cary, NC Who we are: In short – an essential part of Deutsche Bank’s technology solution, developing applications for key business areas...Full timeWork at officeWork from homeShift work$125k - $185k
...Position Overview Job Title Platform Engineering & Site Reliability Engineering VP Corporate Title Vice President Location Cary, NC Who we are In short – an essential part of Deutsche Bank’s technology solution, developing applications for key business...Work at officeWork from homeShift work$100k - $153k
...Position Overview J ob Title CaaS Private Site Reliability Engineer Corporate Title Assistant Vice President Location Cary, NC Who we are: In short – an essential part of Deutsche Bank’s technology solution, developing applications for key business...Full timeWork at officeWork from homeShift work$127k - $249k
...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...workloads. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Full timeLocal areaRemote workWorldwideFlexible hours- METRIX IT SOLUTIONS INC is seeking an L2 SRE focused on Integration and E‑Commerce platforms to ensure system availability and performance across APIs and data flows. You will diagnose incidents, perform RCA, and collaborate with cross‑functional teams to implement robust...
- ...Reliability EngineerThe Enviva team is driven by our shared vision for a renewable energy future... ...& CBM SupportSupport regional sites with continuous improvement of CBM programs... ...visits.Collaborate with other Regional Engineers, Corporate Reliability, and site teams to...For contractorsWork at officeLocal areaVisa sponsorship
- ...Reliability Test Engineer (Manufacturing)Location: Greensboro / Raleigh / Charlotte, NCDuration: FulltimeJob Description:Skills Desired:4+ years of experience in Test Eng and Lab.Electro-mechanical systems, environmental testing, and design validationElectrical functional...
- ...AI infrastructure, working with server, cloud, and platform engineering teams.Operationalize machine learning workflows and support AI... ...implement system enhancements to improve performance, scalability, reliability, and cost efficiency.Collaborate across divisions to support...Full timeWork at officeRemote work
$96k - $132k
...saves lives Join our dynamic team as a Senior Software Systems Engineer in the R&D/Software organization, where you will play a pivotal... ..., please speak with your recruiter or visit our Benefits site: Benefits | Baxter Equal Employment Opportunity Baxter is...Temporary workLocal areaWork visaFlexible hours$125k - $191.7k
...Job Description Hybrid: This role is categorized as hybrid/Remote Role: As a Senior Software Systems Engineer on the Software Validation team within the AV organization, you will play a critical role in leading the strategy and execution of validation efforts...Local areaRemote workWork from homeFlexible hours- ...patterns, and modern DevOps methodologiesMinimum RequirementsBachelor’s Degree in Arts/Sciences (BA/BS) Computer Science, Software Engineering, or related discipline.3-5 years Experience in a Systems Analyst, Business Analyst, Developer or similar role And3-5 years Strong...Full time
- ...enrollment journey for all. We are seeking a Senior Software Engineer to lead the design, development, and optimization of our call... ...platform. You will play a critical role in building scalable, reliable, and intelligent telephony solutions that enable seamless customer...Full timeRemote work
$127k - $171k
What You'll DoWe’re looking for a Senior Software Engineer to join our Retirement & Income Solutions (RIS) business. In this role, you’ll serve as an engineering leader providing architectural oversight for the modernization of our process orchestration and workflow platform...Hourly payPermanent employmentTemporary workWork experience placementH1bWork at office- ...Preferred.Role Responsibilities: (what they will be doing)The CMDB Platform Engineer is responsible for building and maintaining platform-native capabilities that make service and configuration data reliable, enforceable, and operationally useful. This role translates Service-...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Raleigh, NC
- site reliability engineer Raleigh, NC
- on-site clinical research associate (traveling/remote) Raleigh, NC
- website coordinator Raleigh, NC
- junior website developer Raleigh, NC
- site leader Raleigh, NC
- historic site Raleigh, NC
- website content developer Raleigh, NC
- construction site safety Raleigh, NC
- official site Raleigh, NC



