Vice President - Site Reliability Engineering (SRE) - The Core Engineering
Goldman Sachs
Vice President - Site Reliability Engineering (SRE) – The Core EngineeringWHAT WE DOSite Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this VP role, you will help engineer highly reliable, observable, and resilient platforms that support critical business services at scale. You will collaborate with multiple engineering teams to continually improve our production system architecture, facilitate fast delivery of new services, and reduce downtime.This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization.Key ResponsibilitiesPartner with engineering leadership to establish service level objectives (SLOs), service level indicators (SLIs), and error budgets.Collaborate with product developers to architect highly available, fault-tolerant, and self-healing systems. Conduct architectural reviews and introduce patterns like circuit breakers, graceful degradation, and rate limiting.Reduce operational toil by building automation, tooling, and self-service capabilities that remove repetitive manual work.Improve production readiness through load testing, performance tuning, capacity forecasting, and reliability reviews.Lead the response to complex, multi-system production incidents. Facilitate blameless post-mortems to identify root causes and drive long-term preventative actions.Promote sustainable operations by helping design healthy on-call models, clear escalation paths, and balanced pager responsibilities.WHAT WE ARE LOOKING FORCore Technical SkillsStrong proficiency in at least one major programming language (e.g., Java, Python, or Node.js) with a focus on writing clean, maintainable code for tooling and automation.Hands-on experience with Infrastructure as Code (IaC) frameworks such as Terraform, Ansible, or CloudFormation.Deep understanding of containerization and orchestration technologies, specifically Docker and Kubernetes (K8s), including service meshes and ingress controllers.Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud-native architectures.Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch)Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software design.Knowledge of networking protocols and load balancing strategies in a distributed systems environment.Core Competencies & Soft SkillsAbility to analyze complex, distributed systems holistically and understand how individual components interact under load.Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority.Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders.Highly motivated, pro-active and capable of multi-tasking under pressure in a fast-paced environment without compromising quality.Commitment to fostering a blameless culture where failures are treated as opportunities to learn and improve systems.Interest in financial markets and technology.Preferred QualificationsBachelor’s degree in Computer Science, System Engineering, or a related technical field that involves programming.7 to 10 years of experience ABOUT GOLDMAN SACHSThe Goldman Sachs Group, Inc. is a leading global investment banking, securities and investment management firm that provides a wide range of financial services to a substantial and diversified client base that includes corporations, financial institutions, governments and individuals. Founded in 1869, the firm is headquartered in New York and maintains offices in all major financial centers around the world.Posting Date: 2026-09-17
- ...Software Reliability Engineer Good software has to run where customers need it. For many of Retool's largest customers, that means running... ...clarity they would expect from any critical system. Retool's Core Infrastructure team owns the systems that make this possible:...Suggested
- ...Senior Site Reliability Engineer (SRE) Our client is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies....SuggestedLocal area
- ...Sr. Site Reliability Engineer (SRE) New York City, NY - LOCALS ONLY Hybrid, 3 days 6-Month Contract 10-15 years Our client is seeking a Senior Site Reliability Engineer (SRE) with 10–15 years of experience to support front-office trading systems in a production...SuggestedContract workLocal area
$80k - $95k
...join our dynamic team supporting the company’s users, applications, and web-based product offerings. In this role, the Site Reliability Engineer (SRE) will play a key role in maintaining resources at peak efficiency to guarantee staff are able to perform their...SuggestedRemote workVisa sponsorshipWork visa$120k - $180k
...in space and defense. You will be the first dedicated Site Reliability Engineer and own critical infrastructure end to end. This is a greenfield... ...operate cloud and on-premises infrastructure as the sole SRE. Architect migrations from AWS into on-premises and air-gapped...SuggestedPermanent employmentFull timeRelocation package- ...Versana is seeking a motivated SRE/DevOps Engineer with strong observability... ...indicators. • Improve system reliability and resiliency. • Conduct... ...5+ years of experience as a Site Reliability Engineer or... ...Proven track record leveraging core observability concepts, end-...Work experience placementLocal area
$150k - $160k
Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech Site Reliability Engineer (SRE) to join the... ...balancing programmatic monetization via Prebid.js against Core Web Vitals. In this hands-on role you will ensure our web...Work at officeLocal area$140k - $215k
...trillions of events per day. As a Principal SRE, you will operate at the intersection of our Core Platform and Embedded Reliability charters: building the foundational... ...on, while embedding directly with product engineering teams and their leadership to drive reliability...Full timeWork experience placementWork at officeLocal area2 days per week3 days per week$150k - $300k
...What We Do At Goldman Sachs, our Engineers don't just make things - we make things possible... ...Banking & Markets business, the Site Reliability Engineering (SRE) team ensures the availability, resilience, and performance of core business services that underpin a global...Full timeTemporary workPart time- ...infrastructure that has to be both reliable and low-latency to influence... ...in production # AI is a core part of how you work - you’ve... ...scoping, etc. (gist of harness engineering) # You have experience building... ...haves ~4+ years in SRE, DevOps, or infrastructure/platform...Temporary workImmediate start
$189k - $283.6k
...Role As a member of the SRE team, you will proactively and reactively improve the reliability of Block's platform and... ...desire to perform and grow as an engineer ~5+ years of software development... ..., based solely on the core competencies required of the role...Full timeLocal areaRemote workRelocation packageFlexible hoursShift work- ...Exp: 8-12 Years Client: Amex Job Description: SRE Engineer (This is not a Devops role, strictly need an SRE Engineer, who... ...manager as well) This is an SRE role supporting the B2B and Core Services. Skills: SRE. REST Web Services. Kafka....
- ...Senior Site Reliability Engineer (SRE) Plenful is hiring a Senior Site Reliability Engineer (SRE) to keep our production systems reliable, performant... ...Define and implement SLIs, SLOs, and error budgets across core services. Own production system health: uptime, latency...Full timeWork at officeRemote workFlexible hours2 days per week
$195k - $275k
...& Markets Technology, Client Reporting, Core Processing, Private and International Wealth... ...and the Chief Operating Office. The Reliability Operations (RO) within WMT is responsible... .... Demonstrated leadership in driving SRE practices across a technology stack...Full timeTemporary workWork at officeWorldwideNight shift- ...provider headquartered in Ann Arbor, Michigan that offers strategic talent solutions to our clients world-wide. Job Title: SRE Engineer Job Type: Temporary Assignment Work Location: New York, NY, 10001 Work Type: Hybrid Duration: 12+ Months...Temporary work
$500 per month
...diverse group of experienced engineers, traders, and brokerage professionals... .... If you align with our core values—Stay Curious, Have... ...to apply. Your Role: As a Site Reliability Engineer at Alpaca, you'll help... ...while still being a well-rounded SRE the rest of the week....Home office- ...development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands,... ...workflows for accuracy and reliability. Work with AWS, Azure, GCP, Kubernetes... ...DevOps Cloud Infrastructure Site Reliability Engineering (SRE) Platform...Remote jobFor contractors
$140k - $155k
...Salary: $140,000 - 155,000 per year Requirements: Over 8 years of experience in Site Reliability Engineering, DevOps, or Production Engineering Demonstrated leadership capabilities as a technical lead or team supervisor, with a focus on overseeing and guiding engineers...Full time$254k - $407k
...Vice President, Software Engineering Mastercard is seeking a Vice President of Engineering... ...will bring together core engineering products, developer... ...experience, platform engineering, SRE, DevOps, cloud engineering,... ...reimbursement or on-site fitness facilities; eligibility...Full timePart timeFlexible hours$207k - $300k
...by pushing for changes that improve reliability and velocity.Practice sustainable... ...Master's degree in Computer Science or Engineering.Experience mentoring engineers and... ...across cross-functional teams.Site Reliability Engineering (SRE) combines software and systems engineering...- ...Job title : Platform Engineer / SRE-DevOps Engineer Amazon Redshift Location: Remote Duration: 3+Months Key Responsibilities Design, deploy, and maintain Amazon Redshift environments and supporting AWS infrastructure. Automate infrastructure...Remote work
$150k - $190k
...Senior Site Reliability Engineer, VP At Morgan Stanley, we advise, originate, trade, manage and distribute capital for governments, institutions and individuals, and always do so with a standard of excellence. We are a leading global financial services firm that conducts...Full timeTemporary workWorldwideFlexible hoursWeekend work$129k - $152k
...are a growing team of world-class engineering, operations, medical affairs, marketing... ...some roles requiring you to be on-site in a location. Cleerly has... ...highly skilled, experienced Site Reliability Engineer (SRE) to join the core technical team of our growing next...Full timeRemote work$180.5k - $236.91k
...Oscar. We're hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering team.... ...family. About the role: Our Core Technology teams build and maintain... ...technical domains such as DevOps, site reliability, and cloud best practices Lead the...Full timeWork at officeFlexible hours$150k - $160k
...the Senior Cloud Architect (Network & SRE) , you will serve as a senior technical... ...between advanced cloud networking and site reliability engineering. You will be responsible for designing... ...optimization and reliability as a core component of our infrastructure strategy...Permanent employmentFull timeRemote work$115k - $160k
...a transformative journey as a Senior Site Reliability Engineer - AVP - Credit Trade Floor. At Barclays... .... You will join the IB Credit SRE team, applying reliability engineering... ...are known when they occur. Assistant Vice President Expectations To advise and influence...Hourly payWork at office- ...paths and high-volume event processing in a fast-moving startup environment. We are looking for a Platform/Infrastructure engineer with 4+ years in SRE or DevOps, strong GCP experience, and hands-on work with serverless systems, IAM, and observability. #J-18808-Ljbffr
$191k - $226k
...healthier, faster. About the role We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of the... ...operating production cloud infrastructure at scale in an SRE, DevOps, or platform engineering role ~ Deep expertise with...Remote workWork visaFlexible hours$182.8k - $247.3k
...learners around the world. About the role... As a Senior Site Reliability Engineer, you will work closely with both product and platform... ...distributed systems and drive operational excellence Support core infrastructure (i.e understand, diagnose, and debug these...Work experience placement$100k - $250k
...Role Roadmap As a member of Kalshi's engineering team, you'll help build the next-... ...You'll Do Improve observability, reliability, and service availability by defining and... ...operational burden Collaborate with core infrastructure engineers to performance-...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Vice President - Site Reliability Engineering (SRE) - The Core Engineering. Be the first to apply!
- vice president program management New York, NY
- senior vice president human resources New York, NY
- vice president sales New York, NY
- vice president innovation New York, NY
- vice president logistics New York, NY
- vp product management New York, NY
- vice president global communications New York, NY
- vice president strategic initiatives New York, NY
- vice president information technology New York, NY
- vice president of revenue cycle New York, NY



