Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Goldman Sachs Group, Inc.

Vice President - Site Reliability Engineering (SRE) – The Core Engineering

New York, NY, United States

Job Description
Vice President - Site Reliability Engineering (SRE) – The Core Engineering
WHAT WE DO

Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this VP role, you will help engineer highly reliable, observable, and resilient platforms that support critical business services at scale. You will collaborate with multiple engineering teams to continually improve our production system architecture, facilitate fast delivery of new services, and reduce downtime.

This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization.

Key Responsibilities
  • Partner with engineering leadership to establish service level objectives (SLOs), service level indicators (SLIs), and error budgets.
  • Collaborate with product developers to architect highly available, fault‑tolerant, and self‑healing systems. Conduct architectural reviews and introduce patterns like circuit breakers, graceful degradation, and rate limiting.

Reduce operational toil by building automation, tooling, and self‑service capabilities that remove repetitive manual work.

Improve production readiness through load testing, performance tuning, capacity forecasting, and reliability reviews.

Lead the response to complex, multi‑system production incidents. Facilitate blameless post‑mortems to identify root causes and drive long‑term preventative actions.

Promote sustainable operations by helping design healthy on‑call models, clear escalation paths, and balanced pager responsibilities.

WHAT WE ARE LOOKING FOR
Core Technical Skills
  • Strong proficiency in at least one major programming language (e.g., Java, Python, or Node.js) with a focus on writing clean, maintainable code for tooling and automation.

Hands‑on experience with Infrastructure as Code (IaC) frameworks such as Terraform, Ansible, or CloudFormation.

Deep understanding of containerization and orchestration technologies, specifically Docker and Kubernetes (K8s), including service meshes and ingress controllers.

Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud‑native architectures.

  • Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch)

Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software design.

Knowledge of networking protocols and load balancing strategies in a distributed systems environment.

Core Competencies & Soft Skills

Ability to analyze complex, distributed systems holistically and understand how individual components interact under load.

Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority.

  • Ability to translate complex technical issues into clear, actionable insights for both technical and non‑technical stakeholders.

Highly motivated, pro‑active and capable of multi‑tasking under pressure in a fast‑paced environment without compromising quality.

Commitment to fostering a blameless culture where failures are treated as opportunities to learn and improve systems.

Interest in financial markets and technology.

Preferred Qualifications

Bachelor’s degree in Computer Science, System Engineering, or a related technical field that involves programming.

7 to 10 years of experience

ABOUT GOLDMAN SACHS

The Goldman Sachs Group, Inc. is a leading global investment banking, securities and investment management firm that provides a wide range of financial services to a substantial and diversified client base that includes corporations, financial institutions, governments and individuals. Founded in 1869, the firm is headquartered in New York and maintains offices in all major financial centers around the world.

Job Info
  • Job Identification 182983
  • Job Category Vice President
  • Posting Date 09/17/2026, 12:09 AM
  • Locations New York, NY, United States
Healthcare & Medical Services

We believe who you are makes you better at what you do. We're committed to fostering and advancing diversity and inclusion in our own workplace and beyond by ensuring every individual within our firm has a number of opportunities to grow professionally and personally

We offer competitive vacation policies based on employee level and office location. We promote time off from work to recharge by providing generous vacation entitlements and a minimum of three weeks expected vacation usage each year.

Financial Wellness & Retirement

We assist employees in saving and planning for retirement, offer financial support for higher education, and provide a number of benefits to help employees prepare for the unexpected. We offer live financial education and content on a variety of topics to address the spectrum of employees’ priorities.

Health

We offer a medical advocacy service for employees and family members facing critical health situations, and counseling and referral services through the Employee Assistance Program (EAP). We provide Global Medical, Security and Travel Assistance and a Workplace Ergonomics Program. We also offer state‑of‑the‑art on‑site health centers in certain offices.

Fitness

To encourage employees to live a healthy and active lifestyle, some of our offices feature on‑site fitness centers. For eligible employees we typically reimburse fees paid for a fitness club membership or activity (up to a pre‑approved amount).

We offer on‑site child care centers that provide full‑time and emergency back‑up care, as well as mother and baby rooms and homework rooms. In every office, we provide advice and counseling services, expectant parent resources and transitional programs for parents returning from parental leave. Adoption, surrogacy, egg donation and egg retrieval stipends are also available.

Benefits at Goldman Sachs

Read more about the full suite of class‑leading benefits our firm has to offer..

Learn More

#J-18808-Ljbffr
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in New York, NY vacancy
  • $153k - $210k

     ...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating... 
    Suggested
    Full time

    Ridgeline

    New York, NY
    1 day ago
  •  ...to meet you. Our Enterprise Information Technology (EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In this role, you will move beyond traditional infrastructure maintenance... 
    Suggested
    Permanent employment
    Full time
    H1b
    Local area
    Remote work
    Shift work

    Jack Henry & Associates

    New York, NY
    1 day ago
  •  ...Morgan with security as a key differentiator.We pride ourselves on our "hands off culture" by having very few meetings and giving engineers and creators a broad scope of responsibility and autonomy. You'll be working in-person 5 days/week in our lower-manhattan NYC office... 
    Suggested
    Full time
    Work at office

    Hadrius

    New York, NY
    1 day ago
  •  ...production paths and high-volume event processing in a fast-moving startup environment. We are looking for a Platform/Infrastructure engineer with 4+ years in SRE or DevOps, strong GCP experience, and hands-on work with serverless systems, IAM, and observability. #J-188... 
    Suggested

    BAM VENTURES LLC

    New York, NY
    4 days ago
  • $191k - $226k

     ...incentives to steer members to the care that helps them get healthier, faster. About the role We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of the cloud infrastructure powering Garner's products and AI/ML... 
    Suggested
    Remote work
    Work visa
    Flexible hours

    Garner Health

    New York, NY
    2 days ago
  • $195k - $275k

     ...Management, Capital Markets Application & Data Services, Deployment Planning & Release Management, and the Chief Operating Office.The Reliability Operations (RO) within WMT is responsible for providing swift, courteous, and knowledgeable customer service to end users of the... 
    Temporary work
    Work at office
    Worldwide
    Night shift

    Morgan Stanley

    New York, NY
    4 days ago
  • $150k - $300k

     ...What We Do At Goldman Sachs, our Engineers don't just make things - we make things possible. Change the world by connecting people...  ...Within the firm's Global Banking & Markets business, the Site Reliability Engineering (SRE) team ensures the availability, resilience,... 
    Full time
    Temporary work
    Part time

    Socket

    New York, NY
    4 days ago
  • $180.5k - $236.91k

     ...Senior Software Engineer, Cloud Infrastructure / SRENew York, New York, United StatesHi, we're Oscar. We're hiring a Senior Software...  ...your team's business and technical domains such as DevOps, site reliability, and cloud best practicesLead the planning, execution and release... 
    Full time
    Work at office
    Flexible hours

    Oscar Health

    New York, NY
    4 days ago
  • $160k - $230k

     ...Cloud AWS Support Reliability Engineer (SRE)When you work at the New York Fed, you have the opportunity to make an impact in our communities and across the nation. Our mission-driven, curious, and dedicated colleagues apply their diverse perspectives and unique talents... 
    Full time

    Federal Reserve System

    New York, NY
    4 days ago
  • $115k - $160k

     ...with Barclays to connect them with exceptional professionals for this role. Embark on a transformative journey as a Senior Site Reliability Engineer - AVP - Credit Trade Floor. At Barclays, our vision is clear –to redefine the future of banking and help craft innovative... 
    Hourly pay
    Work at office

    Barclays

    New York, NY
    7 hours ago
  •  ...real-time decisioning, and infrastructure that has to be both reliable and low-latency to influence behavior in the moment. We’ve...  ...rate limiting, and permissions scoping, etc. (gist of harness engineering) # You have experience building or orchestrating AI/agent workflows... 
    Temporary work
    Immediate start

    BAM VENTURES LLC

    New York, NY
    4 days ago
  • $182.8k - $247.3k

     ...to develop education for our half a billion (and growing!) learners around the world. About the role... As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure the company’s sophisticated distributed... 
    Work experience placement

    United States Digital Space LLC

    New York, NY
    4 days ago
  •  ...Site Reliability Engineer Our Client, a multinational telecommunications technology company is seeking a Site Reliability Engineer (SRE I) to join our Video Platform Engineering Team. As a Level 1 SRE, you will work closely with senior engineers to respond to incidents... 
    Temporary work

    Elite Technical

    New York, NY
    17 days ago
  • $120k - $150k

     ...allows each person to achieve personal success and add value to our teams and communities.We are currently looking for a Site Reliability Engineer to join our Platform Engineering team in New York, NY.About the RoleJoin our Platform Engineering team as a Site Reliability... 
    Full time

    Piper Sandler Companies

    New York, NY
    7 days ago
  •  ...Job Description Job Description Location: New York, NY, USA Exp: 8-12 Years Client: Amex Job Description: SRE Engineer (This is not a Devops role, strictly need an SRE Engineer, who has great analytical skills and is a good incident manager as well)... 

    IVID TEK INC

    New York, NY
    26 days ago
  • $120k - $180k

     ...people, and works with high-profile manufacturers including leaders in space and defense. You will be the first dedicated Site Reliability Engineer and own critical infrastructure end to end. This is a greenfield opportunity to architect the path from AWS to on-premises... 
    Permanent employment
    Full time
    Relocation package

    Raydar

    New York, NY
    26 days ago
  • $167.7k - $245.2k

     ...platform. This team is responsible for architecting, delivering, and maintaining our FedRAMP offering. Your ImpactAs a FedRAMP Site Reliability Engineer(SRE), you will lead the operations and architecture of our mission-critical Federal region, ensuring peak availability,... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    New York, NY
    a month ago
  • $500 per month

     ...accounts. Our global team is a diverse group of experienced engineers, traders, and brokerage professionals who are working to...  ...significant impact, we encourage you to apply. Your Role: As a Site Reliability Engineer at Alpaca, you'll help keep our brokerage platform... 
    Home office

    Alpaca

    New York, NY
    18 days ago
  • $80k - $95k

     ...join our dynamic team supporting the company’s users, applications, and web-based product offerings. In this role, the Site Reliability Engineer (SRE) will play a key role in maintaining resources at peak efficiency to guarantee staff are able to perform their functions... 
    Remote work
    Visa sponsorship
    Work visa

    RANE Network

    New York, NY
    a month ago
  • $141k - $216.6k

     ...—it means helping shape the future of emergency response and building a safer, more connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability, and operational excellence of our Unified Call (UC) platform—the mission-critical... 
    Work experience placement
    Work at office

    Axon

    New York, NY
    a month ago
  • The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also... 
    Full time
    Work experience placement
    Remote work

    Shutterstock

    New York, NY
    4 days ago
  • $150k - $220k

     ...teams, and innovators in this way. The Role: As an engineering organization, we pride ourselves on engineering as a creative...  ...can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team responsible for keeping... 
    Local area

    Forge Global

    New York, NY
    2 days ago
  • $130k - $153k

     ...our customers, and in our growing commitment to land stewardship and recreational access. WHAT YOU WILL DO onX is seeking a Site Reliability Engineer to build and maintain the infrastructure that enables our developers to ship reliably at scale. You'll manage onX's... 
    Full time
    Part time
    Work at office
    Remote work
    Flexible hours

    ONX, Inc.

    New York, NY
    5 days ago
  • $200k - $250k

    Hudson River Trading (HRT) is seeking a Senior Site Reliability Engineer to join our growing Enterprise SRE team. This team is responsible for developing and maintaining productivity service infrastructure for the entire firm, both on-prem and in the cloud. They ensure... 
    Work at office
    Local area
    Immediate start

    Hudson River Trading

    New York, NY
    7 days ago
  •  ...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives critical to our multi-region disaster recovery roadmap. You will architect and implement robust observability pipelines, ensure deep... 

    EPAM Systems

    New York, NY
    10 days ago
  • $140k - $155k

     ...Salary: $140,000 - 155,000 per year Requirements: Over 8 years of experience in Site Reliability Engineering, DevOps, or Production Engineering Demonstrated leadership capabilities as a technical lead or team supervisor, with a focus on overseeing and guiding engineers... 
    Full time

    EPAM Systems

    New York, NY
    7 days ago
  • $131k - $164k

     ...Role OverviewYou’re a seasoned Site Reliability Engineer who loves owning complex infrastructure, making things run faster, safer, and with less manual effort. In this Staff‑level role, you’ll design and operate VMware‑based private cloud platforms that power mission‑critical... 
    Work at office
    Local area
    Visa sponsorship
    Flexible hours

    Diligent

    New York, NY
    a month ago
  • $220k - $260k

     ...is headquartered in New York and brings deep expertise in finance and AI. About the Role We’re looking for a Staff Site Reliability Engineer to lead the evolution of Tabs’ platform as we scale. In this role, you’ll operate as a senior individual contributor, partnering... 
    Full time
    Contract work
    Work at office

    Tabs

    New York, NY
    26 days ago
  •  ...Description A major financial services company in NYC is growing its team rapidly, and they are looking for a Senior DevOps Engineer / Site Reliability Engineer who can join. If you’re passionate about high-availability, reliability, automation, we’d be excited to talk... 

    The Greene Group

    New York, NY
    24 days ago
  •  ...About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial...  ...Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to... 
    Full time

    Anthropic

    New York, NY
    11 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!