Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas
Goldman Sachs
Site Reliability Engineer - Vice PresidentSite Reliability Engineering (SRE) is an engineering discipline that combines software and systems engineering to build and run scalable, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for improving the availability and reliability of the firm’s most critical platform services and ensures they meet the requirements of our internal and external users. It is also responsible for firmwide policies and standards focused on firm’s digital resilience. We are looking for engineers who are motivated to collaborate with our businesses to build and run sustainable production systems, which can evolve and adapt to changes in our fast-paced, global business environment.The SRE team develops and maintains platforms and tools which help other Engineering teams in Goldman Sachs to build and operate reliable and resilient systems. These systems span on-premises datacenters and multiple public cloud environments. The platforms we offer include central logging, monitoring, agents and alerting and we provide tools to drive adoption and improvements to capacity planning, operational readiness assessments, production incident postmortems, SLIs / SLOs, and deployment automation including canary releases.The products and services we provide to our internal customers are used by thousands of engineers every day. We believe that reliability is the most important feature of any system, and we are devoted to giving our engineers the platforms and tools they need to build and operate reliable products.Role OverviewAs a Site Reliability Engineer (SRE) at Goldman Sachs, you will be a pivotal leader in ensuring the availability, reliability, and scalability of the firm's most critical platform applications and services. You will combine deep software and systems engineering expertise to architect, build, and run large-scale, massively distributed, fault-tolerant systems. This role involves providing technical leadership, mentoring senior engineers, and collaborating closely with internal teams and executive stakeholders to build and operate sustainable production systems that can adapt to our dynamic global business environment. You will drive a culture of continuous improvement, championing the adoption of advanced SRE principles and best practices across the organization. ResponsibilitiesStrategic Reliability & Performance: Drive the strategic direction for availability, scalability, and performance of mission-critical applications and platform services, ensuring alignment with firm-wide objectives.Architectural Leadership: Lead the design, build, and implementation of highly available, resilient, and scalable infrastructure and application architectures.Advanced Automation & Tooling: Architect and develop sophisticated platforms, tools, and automation solutions to eliminate toil, optimize operational workflows, and enhance deployment processes across the enterprise.Complex Incident Management & Post-Mortem Analysis: Lead critical incident response, conduct in-depth root cause analysis for systemic issues, and implement long-term preventative measures to significantly enhance system stability and resilience.System Design & Capacity Planning: Partner with development teams to embed reliability into application design from inception, provide expert system design consulting, and lead comprehensive capacity planning initiatives for future growth.Observability & Insights: Define and implement advanced monitoring, high volume logging with multi-user query capabilities, and tracing strategies to provide deep, actionable insights into application performance, infrastructure health, and user experience.Technical Vision & Mentorship: Provide technical vision, lead complex technical projects, conduct rigorous code reviews, enforce SDLC best practices, and actively mentor and develop senior and staff-level engineers.Technology Evaluation & Adoption: Stay at the forefront of industry trends and advancements, evaluating and integrating cutting-edge tools and frameworks to significantly improve operational efficiency and reliability.On-Call Leadership: Participate in and lead on-call rotations, providing expert guidance and hands-on support for critical system incidents.QualificationsExperience: Minimum of 6+ years of hands-on experience in Site Reliability Engineering, with a proven track record in architecting, designing, building, and maintaining highly available, scalable, and fault-tolerant systems at an enterprise level.Technical Proficiency:Exceptional programming skills in one or more major languages such as Java, Python, Go with a focus on building robust, scalable software.Extensive hands-on experience with cloud platforms (e.g., AWS, GCP) and deep expertise in containerization and orchestration technologies (e.g., Docker, Kubernetes).Mastery of Infrastructure as Code (IaC) tools (e.g., Terraform, CloudFormation) and configuration management tools (e.g., Puppet, Chef, Ansible).Advanced proficiency in Prompt Engineering and Retrieval-Augmented Generation (RAG) architectures to automate complex SRE workflows, such as the generation of Infrastructure as Code (IaC), dynamic runbooks, and incident response summaries.Profound understanding of Linux internals, networking, distributed systems, and advanced system performance tuning.Expertise in designing and implementing comprehensive monitoring, alerting, logging and tracing solutions (e.g., Prometheus, Grafana, ELK stack, Datadog, PagerDuty).Deep experience with CI/CD tools and practices (e.g., Jenkins, GitLab, Maven).Strong foundation in databases and distributed systems.Exceptional problem-solving abilities and analytical skills, with a track record of resolving complex technical challenges.Preferred Experience:Experience with Distributed Databases like Elastic SearchExperience with working on GCP Big QueryExperience with messaging Systems Like KafkaEducation: Advanced degree (Bachelor’s or Mas ter's or PhD) in Computer Science or a related technical field involving coding and/or systems engineering, or equivalent practical experience.Soft Skills: Superior communication, collaboration, and interpersonal skills, with the ability to influence technical direction, lead cross-functional initiatives, and effectively engage with global teams and executive leadership. Proven ability to work independently, manage multiple complex stakeholders, and drive significant organizational change.Posting Date: 2026-04-13
- ...EngineeringWe are Compliance Engineering, a global team of more than 5... ...build and operate a suite of platforms and applications that prevent... ...Compliance application portfolio.SRE at Goldman Sachs combines... ...changes that improve capacity and reliability.Practicing sustainable...Suggested
- We are Compliance Engineering, a global team of more than 300 engineers... ...build and operate a suite of platforms and applications that prevent... ...application portfolio.SRE at Goldman Sachs combines software... ...changes that improve capacity and reliability.Practicing sustainable...Suggested
$207k - $301k
...with multiple stakeholders across the Site Reliability Engineering and Developer organizations.Serve as... ...systemsSite Reliability Engineering (SRE) combines software and systems engineering... ...the next generation of Google platforms, we make Google's product portfolio possible...Suggested- ...Technology group delivers secure, reliable technology solutions that... ...Application Support Engineer, you will help power DTCC'... ...Trade Processing (ITP) platforms that support cross-border... ...and settlement.Leveraging Site Reliability Engineering (SRE) principles, you will support...SuggestedRemote workFlexible hoursAfternoon shift
- Role: Vice President - AI EngineerDivision: Risk Engineering - Market RiskLocation: Dallas, Americas About Goldman SachsAt Goldman Sachs, we... ...of the risk metrics, our platform is continuously growing and... ...performance, scalability, and reliability in distributed and cloud‑...Suggested
- Job Duties: Vice President, Software Engineering with Goldman Sachs Bank USA in Dallas, Texas. Perform detailed technical design and development of data intensive capabilities... .... Develop, test, rollout, and support data platform feature and contribute to the vision, roadmap,...
$197.3k - $313.7k
...Job Category Software Engineering Job Details About... ...Job Title: Director, Site Reliability Engineering Location:... ...NY; San Francisco, CA; Dallas, TX About the Role We... ...you will transform our SRE function—moving our engineering... ...Engineering, Platform, Architecture, Security...Full timeImmediate start- ...Compliance, Surveillance & Models-Software Engineering, Vice President, Dallas Job Description HOW YOU WILL... ...‑grade standards of quality, reliability, and maintainability Hands‑on... ...services architecture Cloud‑based data platforms such as Snowflake Apache Spark...
- ...LanternLantern is the specialty care platform connecting people with the... ...an experienced Senior Site Reliability Engineer to champion the reliability,... ...will define and implement SRE practices, drive incident management... ...- at least 3 days/wk in our Dallas, TX officesOn-Call: This...
- ...the ability to think outside the box? We are seeking a Vice President based in Dallas to report to the COO of Wealth Solutions and partner with... ...cross-divisional businesses such as HCM, CWS, Operations, Engineering, Compliance and Legal to develop and execute project...Work experience placement
- ...candidate will work from Goldman Sachs’ Dallas office on multiple regional and global... ...ResponsibilitiesWe are seeking a Privacy Vice President with excellent analytical,... ...in partnership with Legal, Compliance, Engineering, Risk and Business teamsMonitoring and...Work experience placementWork at officeLocal area
- ...about Goldman Sachs' employee referral program including the monetary incentives offered. Job Duties: Vice President, Risk Governance with Goldman Sachs & Co. LLC in Dallas, Texas. Responsible for insurance procurement, administration, and compliance for the firm’s global...Contract work
- ...identify compliance, conduct, and reputational risks, and refine firm controls as appropriate. CTG’s global team (with locations in Dallas, New York, Salt Lake City, London, Warsaw, Tokyo, Singapore, and Hong Kong) is comprised of individuals with varying backgrounds...Work at office
$104.9k - $174.7k
...Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve... ...operating monitoring and uptime platforms such as Grafana, Pingdom, and UptrendsStrong... ...- USA - Nationwide; Boca Raton, FL; Dallas, TX; Florida; Georgia; Home based-...Full timeWork at officeLocal areaRemote workWork from home- ...part of Transaction Banking within the Platform Solutions Division. We are responsible for... ...with Product, Digital, Sales, and Engineering to build out the next generation capabilities... ...Singapore, Bengaluru, London, New York, and Dallas. All our offices work closely together...Work at officeLocal area
- Goldman Sachs & Co. LLC. Corporate Insurance & Advisory Vice President - Corporate Insurance Risk Manager (Dallas) Corporate Insurance & Advisory manages the Firm’s global commercial insurance needs and advises its investing businesses on insurance-related risk. The team...Contract workLocal area
- ...Audit, Data Analytics, Technology Audit, Vice President, Dallas The Goldman Sachs Group, Inc. is a... ...effective controls by assessing the reliability of financial reports, monitoring the... ...cyber-security and technology risk, and engineering.RESPONSIBILITIESDevelop and maintain...
$180k - $200k
...experience and market knowledge.The open position is for a Vice President based in the New York or Dallas office who will be dedicated to executing client... ...across the entire financial advisory services platform;Develop creative content, thought leadership and collaboration...Full timeWork at officeWorldwide- ...foundations that support the firm’s AI and analytics capabilities. This role sits within the engineering effort to develop a modern Lakehouse and AI data platform that enables reliable, well-governed and high-performing data use across the firm.At Goldman Sachs, engineering...
- ...Contract Pay Rate: $40/Hr. W2 Experience: 3-5 Years Overview We are seeking a remote Junior SRE/DevOps Engineer role. The ideal candidate has foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes, and is enthusiastic about growing in a DevOps‑driven...Long term contractContract workInternshipRemote work
$40 per hour
...A technology solutions provider is seeking a remote Junior SRE/DevOps Engineer. The ideal candidate should have foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes. Responsibilities include gaining experience in a DevOps-driven environment. Applicants...Long term contractInternshipRemote work- ...Senior Java Applications Administrator / SRE in Dallas to own production Java environments,... .... You will optimize performance, drive reliability, and mentor teammates while aligning with... ...to improve operational efficiency and platform strategy. #J-18808-Ljbffr...
- ...seeking a highly experienced and strategic Vice President of Purchasing. This role will serve on... ...of Purchasing in both Austin and Dallas divisions to achieve and exceed cost and... ...during design development. Lead value engineering efforts and continuously identify opportunities...For contractors
- We are seeking an AI Engineer with 5+ years of experience to join the Liquidity Risk technology... ...Agents' performance, scalability, and reliability in distributed and cloud‑based... ...those agents.Experience with AWS Bed Rock platform especially using AWS Agent core for deploying...
- What we doAt Goldman Sachs, our Engineers don’t just make things - we make things possible. We change the world by connecting people and... ...: Mosaic, an end-to-end open architecture technology platform and a digital product that provides trading, settlement, analytics...
- ...and application inventory systems. Our platforms are used firm-wide by all business units... ...used by thousands of users across our engineering organization. The successful candidate... ...application inventory platformsEnsure the reliability, scalability and performance of...
- What We DoAt Goldman Sachs, our Engineers don’t just make things - we make things possible. Change the world by connecting people and... ...part of WM Engineering at Goldman Sachs, the WM Cloud Enablement Platform team is responsible for enabling the use of public cloud...
$200k - $250k
...Ascensus is the leading independent technology and service platform powering savings plans across America, providing products and expertise... ...Newton, MA; New York City, NY; Boston, MA; Washington, DC; Philadelphia, PA; Charlotte, NC; Chicago, IL; Dallas, TXType: Full time...Full timePart timeWork at officeRemote work$200k - $215k
...Vice President of PreconstructionDallas-Fort Worth, TX | $200,000 to $... ...through Friday, based in the Dallas-Fort Worth metropolitan area... ...Construction Management, Civil Engineering, or a related fieldPreferred... ...with estimating software platforms commonly used in industrial...Part timeFor subcontractorWork at officeRelocation packageMonday to Friday$302.2k - $324.9k
...recruitment process and interview prep.Pariveda is seeking a Vice President for our Dallas office. In this role, you will establish and deepen... ...objectives.Bachelor’s Degree in MIS, Computer Science, Math, Engineering, or comparable experience.Legally authorized to work for...Part timeWork at officeLocal areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas. Be the first to apply!
- platform engineer Dallas, TX
- platform engineering manager Dallas, TX
- client platform engineer Dallas, TX
- senior platform engineer Dallas, TX
- platform developer Dallas, TX
- site reliability engineer Dallas, TX
- site reliability engineer sre Dallas, TX
- vice president education Dallas, TX
- vice president manufacturing Dallas, TX
- vice president Dallas, TX






