Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Manager, Site Reliability Engineer

$150k - $220k

Forge Global

At Forge, we know our team is our greatest asset. As technology innovators in the private market, our vision is to deliver a richer future for everyone. We live that vision through our values of being bold, accountable, and humble. We experience the value that our vision brings to the world every day, helping the teams behind the greatest innovations of our generation, from space travel to artificial intelligence, and more. With liquidity solutions, exclusive data and insights, a custody offering, and a vibrant marketplace, Forge’s goal is to build the best-in-class technology infrastructure to power a global private market that is transparent, accessible, and seamless for companies, their employees, and investors. Through Forge, employees can sell their private shares, employers can reward shareholders with pre-IPO liquidity and individual and institutional investors can participate in private unicorn growth. Forge's differentiated global marketplace addresses rising demand among individual and institutional investors for exposure to private company stocks and is building a growing network effect. Our ability to offer these powerful financial solutions has generated incredible interest from investors, demand from customers, and a need to grow our team to meet the needs of more companies, teams, and innovators in this way. The Role: As an engineering organization, we pride ourselves on engineering as a creative activity. Engineering managers enable engineers to do their best work by maintaining a culture and environment where engineers can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team responsible for keeping Forge systems highly available for customers, while partnering closely with Platform, Engineering, Security, Compliance, and Product teams to improve reliability, observability, incident response, and operational maturity. This is an opportunity for a hands-on technical leader who can coach engineers, improve production operations, and help Forge build and run secure, scalable, and highly reliable products.Responsibilities: Manage Forge’s Site Reliability Engineering team responsible for keeping Forge systems highly available for customers.Drive strong incident management practices in partnership with engineering teams, including response, mitigation, follow-up, and post-incident learning.Build, improve, and manage observability infrastructure in partnership with Platform Engineering, including monitoring, alerting, dashboards, and operational metrics.Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support faster response and mitigation.Champion reliability best practices across engineering, including service ownership, operational readiness, disaster recovery, and production support standards.Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.Collaborate with engineering teams to troubleshoot production issues, identify recurring problems, and improve system reliability.Hire, coach, mentor, and manage performance for SRE team members while supporting career development and team health.• Partner with Security, Compliance, and Risk partners to ensure reliability and infrastructure practices meet the needs of a regulated business.Qualifications: 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.Bachelor’s degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.Strong technical judgment, communication skills, and ability to influence across engineering and non-engineering stakeholders.Preferred Qualifications: Experience in FinTech, financial services, or another regulated industry.Experience with AWS and/or Azure cloud platforms.Familiarity with Kubernetes, container platforms, infrastructure-as-code, Terraform, Ansible, or similar automation tooling.Experience with observability platforms such as Datadog, CloudWatch, or similar tools.Experience improving developer experience through paved-road platforms, standardization, and self-service infrastructure capabilities.Experience supporting growth-stage companies where speed, scale, reliability, and operational discipline must be balanced.For residents of San Francisco/Bay Area, CA or New York, NY the annual salary range for this role is $150,000-$220,000 + annual bonus. Final offers may vary from the amount listed based on geography, candidate experience and expertise, annual bonus, and other factors.Upon offer, we conduct background checks that include employment and education verification, state, and county criminal history searches as well as fingerprint and drug test. Forge is proud to be an equal opportunity employer committed to supporting a diverse and inclusive workplace. Our employment decisions are made without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), gender, gender identity, gender expression, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, marital status, sexual orientation, veteran status, or any other characteristic protected by federal, state, or local laws.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Manager, Site Reliability Engineer in San Francisco, CA vacancy
  • $210.38k - $243.21k

    Manager, Site Reliability Engineer (Hybrid in South San Francisco)About the RoleWe are seeking an experienced and hands-on Site Reliability Engineering (SRE) Manager to lead our Site Operations and infrastructure initiatives. This role is responsible for ensuring the reliability... 
    Suggested

    Twist Bioscience

    San Francisco, CA
    2 days ago
  •  ...builds the platforms and tooling that help engineering teams develop, deploy, and operate...  ...default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll...  ...habits and tooling.Architect and manage the SLO and error-budget framework, empowering... 
    Suggested
    Permanent employment
    Work experience placement
    Work at office
    Local area

    Plaid Financial

    San Francisco, CA
    3 days ago
  •  ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology,...  ...exposure (training can be provided)Cloud/SaaS experienceMemory management and dump analysis (Java heap dump analysis preferred)ITSM/... 
    Suggested

    JP Morgan Chase

    San Francisco, CA
    1 day ago
  • $152.5k - $205k

     ...a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate...  ...experience, including authoring reusable modules, managing state and environments, and delivering infrastructure changes... 
    Suggested
    Flexible hours

    Circle

    San Francisco, CA
    3 days ago
  •  ...principles to see it in full.About the teamThe Engineering team at Airwallex is a diverse group of...  ..., working together to build scalable, reliable, and secure products that empower...  ...Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work closely... 
    Suggested
    Temporary work
    Local area

    Airwallex

    San Francisco, CA
    4 days ago
  • $152.5k - $205k

     ...everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common...  ..., observability, access controls, auditability, and cost management. You will troubleshoot production issues, document operational... 
    Flexible hours

    Circle

    San Francisco, CA
    4 days ago
  • $190.8k - $267.1k

     ...while helping Reddit grow its business. The reliability of our Ads systems directly impacts...  ...Reliability team partners closely with Ads Engineering to improve reliability, scalability,...  ...advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate,... 
    For contractors
    Work experience placement

    Reddit

    San Francisco, CA
    4 days ago
  • $114.3k - $235.32k

     ...advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-...  ...and HelmSupporting infrastructure provisioning and change management through Terraform/TerragruntBuilding and supporting CI/CD... 
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    5 days ago
  • $113.4k - $162k

     ...conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd,...  ...GitHub, Terraform, Ansible, or similar tools to build and manage cloud infrastructure efficiently.Incident Management Expert... 
    Temporary work

    TextNow

    San Francisco, CA
    2 days ago
  • $127k - $249k

    The TeamPlatform Engineering sits within SRE and builds the core infrastructure...  ...Engineering, the Fabric team manages the global network substrate...  ...role in engineering the reliable, globally connected, multi-...  ...seeking a talented Senior Site Reliability Engineer (SRE) with... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    5 days ago
  • $117k - $209.33k

     ...Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable,...  ...such as SLOs/SLIs, production readiness, incident management, observability, resilience testing, and toil reduction. Success... 
    Full time
    For contractors

    Autodesk

    San Francisco, CA
    5 days ago
  • $148.5k - $223.9k

     ...future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with...  ...eliminate toil and improve operational efficiency.Incident Management: Lead the coordinated response to incidents as an... 
    Full time
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    3 days ago
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and...  ...cross-functionallyNice-to-HaveExperience with cloud and managed services (e.g. AWS)Experience supporting data-intensive platforms... 

    Alembic

    San Francisco, CA
    4 days ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical...  ...and alerting systems.The Fleet Management team provides the core runtime environment...  ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    1 day ago
  • $106k - $130k

     ...ineligible for employment Visa sponsorship.Role Summary The Senior Site Reliability Engineer applies software engineering and systems engineering...  ...as Code, automation, testing, incident response, capacity management, resilience, and operational readiness. Identify recurring... 
    Hourly pay
    Full time
    Immediate start
    Visa sponsorship
    Work visa
    Flexible hours

    Early Warning

    San Francisco, CA
    5 days ago
  • $195k - $257.5k

     ...is a stakeholder.What you’ll be responsible for:As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and...  ...on:Operate and scale production blockchain infrastructure, managing full nodes across networks such as Arc, Ethereum, Solana,... 
    Flexible hours

    Circle

    San Francisco, CA
    2 days ago
  • $204k - $306k

     ...We're all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity,...  ...week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    Okta

    San Francisco, CA
    4 days ago
  • $55k - $151.47k

     ...LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in...  ...data integrity and accessibility- Leading incident management and resolution efforts to maintain operational continuityWhat... 
    Full time
    H1b

    PwC

    San Francisco, CA
    3 days ago
  • $194k - $267k

     ...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk...  ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    3 days ago
  • $220k - $235k

     ...this space is still largely uncharted. Engineers here are building agentic solutions to...  ...the forefront of AI choose Ironclad to manage their contracts.We’re consistently recognized...  ...and strategic direction for the Site Reliability Engineering team and our broader Cloud... 
    Full time
    Contract work
    Work at office

    Ironclad

    San Francisco, CA
    3 days ago
  • $194k - $267k

     ..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    5 days ago
  • $153k - $191.3k

     ...manufacturing, data processing, and software engineering, our office is a truly inspiring mix of...  ...environments, to guarantee the reliability, scalability, and availability of our services...  ..., particularly resource optimization, management, and cluster tuning in a constrained... 
    Full time
    Temporary work
    For contractors
    Work at office
    Local area
    Remote work
    Home office
    3 days per week

    Planet Labs PBC

    San Francisco, CA
    1 day ago
  •  ...and the U.S. Special Forces. The Role We're hiring a Site Reliability Engineer to own the operational health of our connected sensor...  ...Systems Builder — Close the Loop Build and maintain fleet management systems: OTA update pipelines, device health tracking, remote... 
    Remote work

    Specter Services LLC

    San Francisco, CA
    2 days ago
  • $189k - $283.6k

     ...Afterpay is transforming the way customers manage their spending over time. TIDAL is a...  ...proactively and reactively improve the reliability of Block's platform and critical infrastructure...  ...strong desire to perform and grow as an engineer ~5+ years of software development... 
    Full time
    Relocation package
    Flexible hours
    Shift work

    Block Inc

    San Francisco, CA
    2 days ago
  •  ...culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of...  ...~6+ years of experience in Site Reliability Engineering, managing infrastructure and services at scale ~ History of end-to... 
    Immediate start
    Remote work
    Worldwide

    OutSystems

    San Francisco, CA
    2 days ago
  •  ...for As an SRE at Wordbricks, you will keep our systems fast, reliable, and boring. You'll own the infrastructure and operations...  ...systems Build and maintain CI/CD, observability, and alerting Manage cloud infrastructure across Cloudflare, AWS, and Vercel Lead... 
    Remote work
    Flexible hours

    Wordbricks, Inc.

    San Francisco, CA
    2 days ago
  •  ...daily users while enabling our engineering teams to ship fast. You'll...  ...automation and tooling that improves reliability and partnering with...  ...reliability best practices Manage and optimize our infrastructure...  ...you'll bring ~5+ years in Site Reliability Engineering, DevOps... 
    Work at office
    Work from home

    Gamma

    San Francisco, CA
    1 day ago
  • $210k - $240k

     ...Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual...  ..., deployment automation, rollback mechanisms, and config management Implement and maintain monitoring, alerting, and incident... 
    Full time

    Alembic Technologies

    San Francisco, CA
    5 days ago
  • $175k - $250k

     ...Senior Cloud Infrastructure Engineer Location: San Francisco, CA....  ...Remote unavailable. Modality: On-Site only. Must live within...  ...scalability, performance, and reliability across environments. What You...  ...powers AI workloads at scale Manage and automate GPU compute clusters... 
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    San Francisco, CA
    5 days ago
  • $98.58k - $138.02k

     ...office locations: Austin, TX; Irvine, CA; or Akron, OH. Site Reliability Engineer II will be responsible for supporting, enhancing, and...  ...Terraform, Ansible, or CloudFormation. Work within change management protocols to provide maximum uptime for production systems... 
    Work at office

    Restaurant365

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Manager, Site Reliability Engineer. Be the first to apply!