Site Reliability Engineer
SunTrust Investment Services, Inc.
Site Reliability Engineer
The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability, and incident management while collaborating across multiple business and technology teams. Responsibilities include leading major incident responses, driving problem management, and implementing automation to reduce service downtime. The role involves standardizing observability practices, mentoring SRE team members, and contributing to enterprise-wide reliability frameworks. Candidates require 7+ years of experience, expertise in distributed systems, Kubernetes, automation scripting, and strong leadership in incident management.
Essential Duties And Responsibilities Following is a summary of the essential functions for this job. Other duties may be performed, both major and minor, which are not mentioned below. Specific activities may change from time to time.
- Implements software architecture and engineering approaches for complex initiatives within the job area, contributing to technical plans and working to achieve operational targets with major impact on results.
- Adopts and refines advanced software engineering standards, practices, and governance mechanisms for the job area, influencing how multiple teams improve quality, reliability, and delivery.
- Collaborates with senior engineers, product partners, and architecture teammates to shape technology approaches for the domain, providing deep technical insight and proposing solution patterns that inform local roadmaps and priorities.
- Leads the end-to-end technical design and implementation of scalable, secure, and highly available software solutions for the job area, producing patterns and examples that other technical professionals can follow.
- Independently troubleshoots and resolves complex technical issues in the area of responsibility, designing innovative architectures and performance, reliability, and scalability improvements that advance business objectives.
- Provides ongoing technical guidance, coaching, and training to other engineers, delegating and reviewing work from lower-level technical professionals and raising the technical bar through design reviews and knowledge sharing.
- Evaluates emerging technologies and techniques relevant to the job area, building prototypes and solution concepts that contribute measurable input into new features, products, or capabilities.
- Contributes to the development of long-term technical goals and plans for the area of responsibility through well-reasoned recommendations, design proposals, and implementation experience.
- Leads large or complex initiatives within the job area, coordinating and delegating technical work that may span outside the immediate team, and ensuring cohesive, high-quality outcomes with limited supervision.
Qualifications Required Qualifications The requirements listed below are representative of the knowledge, skill and/or ability required. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.
- Bachelor's degree in Computer Science, Software Engineering, or related field.
- Minimum of 7 years of professional experience in software development.
- Deep knowledge of multiple programming languages, software architecture, and design principles.
- Deep understanding of software development lifecycle, testing, deployment, and security practices.
Preferred Qualifications
- Advanced degree in Computer Science or related technical discipline.
- Professional certifications such as Certified Software Development Professional (CSDP) or equivalent.
- Deep expertise in cloud-native architectures, microservices, container orchestration, and DevOps.
- Strong familiarity with Agile frameworks, continuous integration/continuous deployment (CI/CD), and enterprise innovation management.
- 7+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Infrastructure Operations.
- Deep hands-on experience with distributed systems, container orchestration (Kubernetes), and cloud-native operational tooling.
- Proficiency with automation and scripting languages (Python, Go, PowerShell, Ansible).
- Strong understanding of observability platforms (Splunk, Dynatrace) and event-driven monitoring.
- Proven leadership in major incident management and cross-team technical coordination.
- Strong grasp of networking, Linux/Unix internals, and modern infrastructure patterns.
- Excellent communication skills, including executive-level situational awareness during critical incidents.
- Demonstrated ability to influence technical roadmaps and drive adoption of reliability best practices.
Preferred Qualifications
- Financial services or regulated industry experience.
- Experience enabling large-scale SRE transformations or modernization initiatives.
- Familiarity with chaos engineering, resilience assessments, and service failure modeling.
- Exposure to hybrid-cloud and multi-cloud operational frameworks.
- Experience contributing to or leading Center for Enablement functions or Communities of Practice.
Key Responsibilities
- Incident & Problem Management Leadership
- Lead major and high-severity incident response efforts, focusing on diagnosing technical root causes therein, and driving multi-team technical resolution.
- Drive problem management to closure, ensuring systemic fixes replace recurring operational risks.
- Establish and maintain standardized incident playbooks, escalation paths, and communication frameworks.
- Reliability Engineering & Automation
- Architect and deliver automation solutions that eliminate toil, reduce MTTR, and increase service resilience.
- Implement intelligent alerting, anomaly detection, and event correlation leveraging AI and AIOps tools.
- Guide and enforce SLO/SLI adoption across product teams, ensuring metrics inform decision-making and prioritization.
- Observability & Operational Excellence
- Enhance telemetry coverage across logs, metrics, traces, and events using platforms such as Dynatrace and Splunk.
- Define and standardize enterprise observability practices, dashboards, and KPIs.
- Ensure operational readiness of applications and platforms through resiliency testing, chaos engineering, and failure-mode validation.
- Cross-Functional Leadership & Influence
- Partner with Delivery, Architecture, Security, and Risk teams to embed reliability and resilience into design and execution.
- Act as a change agent to elevate operational maturity and drive transformative improvements across Wholesale.
- Lead workshops, maturity assessments, and enablement sessions through the SRE C4E and Communities of Practice.
- Standardization & Documentation
- Develop, maintain, and enforce runbooks, response playbooks, and automated recovery patterns.
- Contribute to enterprise SRE frameworks, templates, and maturity models.
- Promote consistent adoption of best practices across domains and lines of business.
- Mentorship & Technical Development
- Coach and mentor Associate, Professional, and Senior SREs to build technical depth and operational discipline.
- Provide thought leadership in SRE methodologies, cloud-native operational patterns, and automated reliability engineering.
Candidate must be willing to work onsite Monday - Friday at either office in Charlotte NC, Raleigh NC, or Atlanta, GA.
- ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises...SuggestedPermanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...SuggestedWork at officeImmediate start
$118.6k - $195.68k
Job SummaryThe Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at scale...SuggestedPermanent employmentFull timeContract workWork experience placementWork at officeRemote workFlexible hours- Role Profile:We are evolving our Site Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a Senior SRE, you will be a senior hands‑on technical person help...SuggestedFull timeShift work
- ...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...crucial workloads. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...SuggestedFull timeRemote workWorldwide
- ...Site Reliability Engineer Based in Jacksonville, FL, Cary, NC, location for a Fulltime position. Should be having cloud engineering experience and acting as the SME on operation automation and monitoring, identifying TOIL within the teams existing systems and processes...Full time
- ...Site Reliability Engineer (SRE) – Security Infrastructure We are seeking an SRE to support reliability, scalability, and operational excellence for a large-scale network security transformation initiative. This role will focus on monitoring, automation, operational...
- ...technologies to enable scalable, secure, and reliable business operations. Applies strong... ...infrastructure.3. Manages infrastructure engineering projects and processes aligned with... ...benefit plans, please visit our Benefits site. Depending on the position and division,...Permanent employmentFull timePart timeWork experience placementH1bRemote workWork visaShift workWeekend workDay shift
$125k - $185k
...Position Overview J ob Title: CaaS Private Site Reliability Engineer Corporate Title: Vice President Location: Cary, NC Who we are: In short – an essential part of Deutsche Bank’s technology solution, developing applications for key business areas...Full timeWork at officeWork from homeShift work$130k - $180k
Position OverviewPower your future with Qualus as a Lead Relay Settings Engineer. In this role you will perform Protective Relay Design & Coordination: Design, specify, calculate settings, and coordinate protective relays and relay control schemes. Do you have 7+ years...Temporary workFlexible hours$100k - $153k
...Position Overview J ob Title CaaS Private Site Reliability Engineer Corporate Title Assistant Vice President Location Cary, NC Who we are: In short – an essential part of Deutsche Bank’s technology solution, developing applications for key business...Full timeWork at officeWork from homeShift work- ...AI infrastructure, working with server, cloud, and platform engineering teams.Operationalize machine learning workflows and support AI... ...implement system enhancements to improve performance, scalability, reliability, and cost efficiency.Collaborate across divisions to support...Full timeWork at officeRemote work
$23 - $27 per hour
...education, licensure requirements and/or skill level. Hourly Rate: $23 - $27/hr. OVERVIEW To perform embedded software engineering design services and other related software development support services throughout all stages of the product development life...Hourly payContract workWork experience placementSummer workInternship- ...a multi-disciplinary collaborative team of hardware, software engineers and security specialists to design and implement a secure cloud... ...strive to maximize team & department performanceWilling to work on-site, daily, at our Raleigh, NC locationPreferred Experience &...Full timeFlexible hours
$130k - $170k
About Us:BW Design Group is a fully integrated architecture, engineering, construction, system integration, and consulting firm committed to helping our clients realize their most critical goals from Strategy to Commercialization. As the only firm born from a manufacturing...Full timeFlexible hours- ...enrollment journey for all. We are seeking a Senior Software Engineer to lead the design, development, and optimization of our call... ...platform. You will play a critical role in building scalable, reliable, and intelligent telephony solutions that enable seamless customer...Full timeRemote work
- ...essential functions.1. Bachelor's degree in business, finance, engineering, information systems, computer science, analytics, or a related... ...on Truist’s generous benefit plans, please visit our Benefits site. Depending on the position and division, this job may also be eligible...Full timePart timeWork at officeShift workDay shift
$168.98k - $218.68k
...real-time decisioning use casesEngineering & InfrastructureLead engineering of scalable, secure platform infrastructure on AWS (S3, Lake... ...change managementDrive operational excellence across reliability, cost management (FinOps for data and AI workloads), observability...Full timeContract workFor contractorsLocal area$77k - $202k
...ApplicableSpecialismData, Analytics & AIManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a GenAI Python Systems Engineer - Senior Associate, you will play a pivotal role in transforming raw data into actionable insights, enabling informed decision-...Full timeH1b$63k - $140k
...ApplicableSpecialismData, Analytics & AIManagement LevelAssociateJob Description & SummaryThe OpportunityAs a GenAI Python Systems Engineer - Experienced Associate, you will leverage advanced technologies and techniques to design and develop robust data solutions for...Full timeH1b$60 - $80 per hour
...Systems is hiring a Mainframe Systems Programmer for their CICS environment. Pay Range: $70-75/hr. 100% Remote Our Mainframe Engineering team is looking for an experienced, senior level Systems Programmer to support the CICS environment. This role will entail the...Full timeRemote work- ...including server, storage, data center, mobile, RF, networking, industrial, business equipment, and automotive. Position : Reliability Engineer Location : Raleigh, NC RESPONSIBILITIES : Test, Validation & Qualification Develop and execute...Flexible hours
$97k - $143k
...Division is currently seeking a Lead Test Engineering to join our team in Raleigh, NC.... ...productivity, but employees should expect to be on-site as necessary to meet team and project... ...-based products to ensure power reliability in the most demanding applications. Our...Full timeFor contractorsWork experience placementH1bWork at officeLocal areaRemote workWork from homeVisa sponsorshipRelocation package$147k - $210k
...Experience developing accessible technologies.Proficiency in code and system health, diagnosis and resolution, and software test engineering.Google's software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact...$95k - $138k
...DoWe’re looking for an Experienced Software Engineer to join our Retirement & Income Solutions... ...engineering fundamentals and a focus on reliable, well-defined system interactions.You’ll... ...an expectation for this role to work on site at least 3 days a week.Job LevelWe’ll consider...Hourly payPermanent employmentTemporary workWork experience placementH1bWork at office3 days per week- Job Summary:We are currently looking for Software Engineering interns to join us in these locations:Boston, MALowell, MARaleigh, NCYou will work closely with a senior mentor to gain technical knowledge and experience in your field, and cooperate with a broader international...Contract workInternshipWork at officeRemote workFlexible hours
- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking Software Engineers who have a passion for technology and education to help us add new features to our extremely successful educational products and build new, innovative...Work at office
- ...intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge,...Internship
- ...Reference26-20362Salary$75.48 / hourRemote50% Remote Job Description Project Overview We are seeking a highly specialized Software Engineer for a long-term contract engagement focusing on the modernization of the CLIENT Certificate System (RHCS). The primary objective...Long term contractFor contractorsRemote work
- ...CarolinaOnsiteFull Time$120k - $140kSenior Software Engineer — Edge PlatformA client of ours is an... ...high-rate sensor and controller data and reliably forwarding it to the cloud over... ...suppliers, attend conferences and meetings, and visit customer sites (hatcheries).LI-JD17Work at officeRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!



