Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Vice President - Site Reliability Engineering (SRE) - The Core Engineering

Goldman Sachs

Vice President - Site Reliability Engineering (SRE) – The Core EngineeringWHAT WE DOSite Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this VP role, you will help engineer highly reliable, observable, and resilient platforms that support critical business services at scale. You will collaborate with multiple engineering teams to continually improve our production system architecture, facilitate fast delivery of new services, and reduce downtime.This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization.Key ResponsibilitiesPartner with engineering leadership to establish service level objectives (SLOs), service level indicators (SLIs), and error budgets.Collaborate with product developers to architect highly available, fault-tolerant, and self-healing systems. Conduct architectural reviews and introduce patterns like circuit breakers, graceful degradation, and rate limiting.Reduce operational toil by building automation, tooling, and self-service capabilities that remove repetitive manual work.Improve production readiness through load testing, performance tuning, capacity forecasting, and reliability reviews.Lead the response to complex, multi-system production incidents. Facilitate blameless post-mortems to identify root causes and drive long-term preventative actions.Promote sustainable operations by helping design healthy on-call models, clear escalation paths, and balanced pager responsibilities.WHAT WE ARE LOOKING FORCore Technical SkillsStrong proficiency in at least one major programming language (e.g., Java, Python, or Node.js) with a focus on writing clean, maintainable code for tooling and automation.Hands-on experience with Infrastructure as Code (IaC) frameworks such as Terraform, Ansible, or CloudFormation.Deep understanding of containerization and orchestration technologies, specifically Docker and Kubernetes (K8s), including service meshes and ingress controllers.Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud-native architectures.Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch)Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software design.Knowledge of networking protocols and load balancing strategies in a distributed systems environment.Core Competencies & Soft SkillsAbility to analyze complex, distributed systems holistically and understand how individual components interact under load.Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority.Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders.Highly motivated, pro-active and capable of multi-tasking under pressure in a fast-paced environment without compromising quality.Commitment to fostering a blameless culture where failures are treated as opportunities to learn and improve systems.Interest in financial markets and technology.Preferred QualificationsBachelor’s degree in Computer Science, System Engineering, or a related technical field that involves programming.7 to 10 years of experience ABOUT GOLDMAN SACHSThe Goldman Sachs Group, Inc. is a leading global investment banking, securities and investment management firm that provides a wide range of financial services to a substantial and diversified client base that includes corporations, financial institutions, governments and individuals. Founded in 1869, the firm is headquartered in New York and maintains offices in all major financial centers around the world.Posting Date: 2026-09-17

Vacancy posted 11 hours ago
Similar jobs that could be interesting for youBased on the Vice President - Site Reliability Engineering (SRE) - The Core Engineering in New York, NY vacancy
  •  ...Software Reliability Engineer Good software has to run where customers need it. For many of Retool's largest customers, that means running...  ...clarity they would expect from any critical system. Retool's Core Infrastructure team owns the systems that make this possible:... 
    Suggested

    re-tool®

    New York, NY
    3 days ago
  •  ...Senior Site Reliability Engineer (SRE) Our client is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies.... 
    Suggested
    Local area

    E-Solutions

    New York, NY
    19 hours ago
  •  ...Sr. Site Reliability Engineer (SRE) New York City, NY - LOCALS ONLY Hybrid, 3 days 6-Month Contract 10-15 years Our client is seeking a Senior Site Reliability Engineer (SRE) with 10–15 years of experience to support front-office trading systems in a production... 
    Suggested
    Contract work
    Local area

    RIT Solutions

    New York, NY
    3 days ago
  • $80k - $95k

     ...join our dynamic team supporting the company’s users, applications, and web-based product offerings. In this role, the Site Reliability Engineer (SRE) will play a key role in maintaining resources at peak efficiency to guarantee staff are able to perform their... 
    Suggested
    Remote work
    Visa sponsorship
    Work visa

    RANE Network

    New York, NY
    a month ago
  • $120k - $180k

     ...in space and defense. You will be the first dedicated Site Reliability Engineer and own critical infrastructure end to end. This is a greenfield...  ...operate cloud and on-premises infrastructure as the sole SRE. Architect migrations from AWS into on-premises and air-gapped... 
    Suggested
    Permanent employment
    Full time
    Relocation package

    Raydar

    New York, NY
    23 days ago
  •  ...Versana is seeking a motivated SRE/DevOps Engineer with strong observability...  ...indicators. • Improve system reliability and resiliency. • Conduct...  ...5+ years of experience as a Site Reliability Engineer or...  ...Proven track record leveraging core observability concepts, end-... 
    Work experience placement
    Local area

    Versana

    New York, NY
    21 days ago
  • $150k - $160k

    Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech Site Reliability Engineer (SRE) to join the...  ...balancing programmatic monetization via Prebid.js against Core Web Vitals. In this hands-on role you will ensure our web... 
    Work at office
    Local area

    Haymarket Media Group

    New York, NY
    19 hours ago
  • $140k - $215k

     ...trillions of events per day. As a Principal SRE, you will operate at the intersection of our Core Platform and Embedded Reliability charters: building the foundational...  ...on, while embedding directly with product engineering teams and their leadership to drive reliability... 
    Full time
    Work experience placement
    Work at office
    Local area
    2 days per week
    3 days per week

    CrowdStrike

    New York, NY
    19 hours ago
  • $150k - $300k

     ...What We Do At Goldman Sachs, our Engineers don't just make things - we make things possible...  ...Banking & Markets business, the Site Reliability Engineering (SRE) team ensures the availability, resilience, and performance of core business services that underpin a global... 
    Full time
    Temporary work
    Part time

    Socket

    New York, NY
    1 day ago
  •  ...infrastructure that has to be both reliable and low-latency to influence...  ...in production # AI is a core part of how you work - you’ve...  ...scoping, etc. (gist of harness engineering) # You have experience building...  ...haves ~4+ years in SRE, DevOps, or infrastructure/platform... 
    Temporary work
    Immediate start

    BAM VENTURES LLC

    New York, NY
    1 day ago
  • $189k - $283.6k

     ...Role As a member of the SRE team, you will proactively and reactively improve the reliability of Block's platform and...  ...desire to perform and grow as an engineer ~5+ years of software development...  ..., based solely on the core competencies required of the role... 
    Full time
    Local area
    Remote work
    Relocation package
    Flexible hours
    Shift work

    Block USA

    New York, NY
    4 days ago
  •  ...Exp: 8-12 Years Client: Amex Job Description: SRE Engineer (This is not a Devops role, strictly need an SRE Engineer, who...  ...manager as well) This is an SRE role supporting the B2B and Core Services. Skills: SRE. REST Web Services. Kafka.... 

    IVidTek, Inc.

    New York, NY
    1 day ago
  •  ...Senior Site Reliability Engineer (SRE) Plenful is hiring a Senior Site Reliability Engineer (SRE) to keep our production systems reliable, performant...  ...Define and implement SLIs, SLOs, and error budgets across core services. Own production system health: uptime, latency... 
    Full time
    Work at office
    Remote work
    Flexible hours
    2 days per week

    Plenful

    New York, NY
    2 days ago
  • $195k - $275k

     ...& Markets Technology, Client Reporting, Core Processing, Private and International Wealth...  ...and the Chief Operating Office. The Reliability Operations (RO) within WMT is responsible...  .... Demonstrated leadership in driving SRE practices across a technology stack... 
    Full time
    Temporary work
    Work at office
    Worldwide
    Night shift

    Morgan Stanley

    New York, NY
    a month ago
  •  ...provider headquartered in Ann Arbor, Michigan that offers strategic talent solutions to our clients world-wide. Job Title: SRE Engineer Job Type: Temporary Assignment Work Location: New York, NY, 10001 Work Type: Hybrid Duration: 12+ Months... 
    Temporary work

    TekWissen LLC

    New York, NY
    2 days ago
  • $500 per month

     ...diverse group of experienced engineers, traders, and brokerage professionals...  .... If you align with our core values—Stay Curious, Have...  ...to apply. Your Role: As a Site Reliability Engineer at Alpaca, you'll help...  ...while still being a well-rounded SRE the rest of the week.... 
    Home office

    Alpaca

    New York, NY
    15 days ago
  •  ...development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands,...  ...workflows for accuracy and reliability. Work with AWS, Azure, GCP, Kubernetes...  ...DevOps Cloud Infrastructure Site Reliability Engineering (SRE) Platform... 
    Remote job
    For contractors

    YO AI Labs

    New York, NY
    23 days ago
  • $140k - $155k

     ...Salary: $140,000 - 155,000 per year Requirements: Over 8 years of experience in Site Reliability Engineering, DevOps, or Production Engineering Demonstrated leadership capabilities as a technical lead or team supervisor, with a focus on overseeing and guiding engineers... 
    Full time

    EPAM Systems

    New York, NY
    4 days ago
  • $254k - $407k

     ...Vice President, Software Engineering Mastercard is seeking a Vice President of Engineering...  ...will bring together core engineering products, developer...  ...experience, platform engineering, SRE, DevOps, cloud engineering,...  ...reimbursement or on-site fitness facilities; eligibility... 
    Full time
    Part time
    Flexible hours

    Dynamic Yield

    New York, NY
    1 day ago
  • $207k - $300k

     ...by pushing for changes that improve reliability and velocity.Practice sustainable...  ...Master's degree in Computer Science or Engineering.Experience mentoring engineers and...  ...across cross-functional teams.Site Reliability Engineering (SRE) combines software and systems engineering... 

    Google

    New York, NY
    11 hours ago
  •  ...Job title : Platform Engineer / SRE-DevOps Engineer Amazon Redshift Location: Remote Duration: 3+Months Key Responsibilities Design, deploy, and maintain Amazon Redshift environments and supporting AWS infrastructure. Automate infrastructure... 
    Remote work

    Conch Technologies Inc

    New York, NY
    1 day ago
  • $150k - $190k

     ...Senior Site Reliability Engineer, VP At Morgan Stanley, we advise, originate, trade, manage and distribute capital for governments, institutions and individuals, and always do so with a standard of excellence. We are a leading global financial services firm that conducts... 
    Full time
    Temporary work
    Worldwide
    Flexible hours
    Weekend work

    Morgan Stanley

    New York, NY
    a month ago
  • $129k - $152k

     ...are a growing team of world-class engineering, operations, medical affairs, marketing...  ...some roles requiring you to be on-site in a location.  Cleerly has...  ...highly skilled, experienced Site Reliability Engineer (SRE) to join the core technical team of our growing next... 
    Full time
    Remote work

    Cleerly

    New York, NY
    19 hours ago
  • $180.5k - $236.91k

     ...Oscar. We're hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering team....  ...family. About the role: Our Core Technology teams build and maintain...  ...technical domains such as DevOps, site reliability, and cloud best practices Lead the... 
    Full time
    Work at office
    Flexible hours

    Oscar Health

    New York, NY
    14 days ago
  • $150k - $160k

     ...the Senior Cloud Architect (Network & SRE) , you will serve as a senior technical...  ...between advanced cloud networking and site reliability engineering. You will be responsible for designing...  ...optimization and reliability as a core component of our infrastructure strategy... 
    Permanent employment
    Full time
    Remote work

    Liveperson

    New York, NY
    8 days ago
  • $115k - $160k

     ...a transformative journey as a Senior Site Reliability Engineer - AVP - Credit Trade Floor. At Barclays...  .... You will join the IB Credit SRE team, applying reliability engineering...  ...are known when they occur. Assistant Vice President Expectations To advise and influence... 
    Hourly pay
    Work at office

    Barclays

    New York, NY
    2 days ago
  •  ...paths and high-volume event processing in a fast-moving startup environment. We are looking for a Platform/Infrastructure engineer with 4+ years in SRE or DevOps, strong GCP experience, and hands-on work with serverless systems, IAM, and observability. #J-18808-Ljbffr

    BAM VENTURES LLC

    New York, NY
    1 day ago
  • $191k - $226k

     ...healthier, faster. About the role We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of the...  ...operating production cloud infrastructure at scale in an SRE, DevOps, or platform engineering role ~ Deep expertise with... 
    Remote work
    Work visa
    Flexible hours

    Garner Health

    New York, NY
    4 days ago
  • $182.8k - $247.3k

     ...learners around the world. About the role... As a Senior Site Reliability Engineer, you will work closely with both product and platform...  ...distributed systems and drive operational excellence Support core infrastructure (i.e understand, diagnose, and debug these... 
    Work experience placement

    United States Digital Space LLC

    New York, NY
    1 day ago
  • $100k - $250k

     ...Role Roadmap As a member of Kalshi's engineering team, you'll help build the next-...  ...You'll Do Improve observability, reliability, and service availability by defining and...  ...operational burden Collaborate with core infrastructure engineers to performance-... 
    Local area

    Kalshi

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Vice President - Site Reliability Engineering (SRE) - The Core Engineering. Be the first to apply!