Manager, Site Reliability Engineer
$150k - $220kForge Global
At Forge, we know our team is our greatest asset. As technology innovators in the private market, our vision is to deliver a richer future for everyone. We live that vision through our values of being bold, accountable, and humble. We experience the value that our vision brings to the world every day, helping the teams behind the greatest innovations of our generation, from space travel to artificial intelligence, and more. With liquidity solutions, exclusive data and insights, a custody offering, and a vibrant marketplace, Forge’s goal is to build the best-in-class technology infrastructure to power a global private market that is transparent, accessible, and seamless for companies, their employees, and investors. Through Forge, employees can sell their private shares, employers can reward shareholders with pre-IPO liquidity and individual and institutional investors can participate in private unicorn growth. Forge's differentiated global marketplace addresses rising demand among individual and institutional investors for exposure to private company stocks and is building a growing network effect. Our ability to offer these powerful financial solutions has generated incredible interest from investors, demand from customers, and a need to grow our team to meet the needs of more companies, teams, and innovators in this way. The Role: As an engineering organization, we pride ourselves on engineering as a creative activity. Engineering managers enable engineers to do their best work by maintaining a culture and environment where engineers can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team responsible for keeping Forge systems highly available for customers, while partnering closely with Platform, Engineering, Security, Compliance, and Product teams to improve reliability, observability, incident response, and operational maturity. This is an opportunity for a hands-on technical leader who can coach engineers, improve production operations, and help Forge build and run secure, scalable, and highly reliable products.Responsibilities: Manage Forge’s Site Reliability Engineering team responsible for keeping Forge systems highly available for customers.Drive strong incident management practices in partnership with engineering teams, including response, mitigation, follow-up, and post-incident learning.Build, improve, and manage observability infrastructure in partnership with Platform Engineering, including monitoring, alerting, dashboards, and operational metrics.Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support faster response and mitigation.Champion reliability best practices across engineering, including service ownership, operational readiness, disaster recovery, and production support standards.Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.Collaborate with engineering teams to troubleshoot production issues, identify recurring problems, and improve system reliability.Hire, coach, mentor, and manage performance for SRE team members while supporting career development and team health.• Partner with Security, Compliance, and Risk partners to ensure reliability and infrastructure practices meet the needs of a regulated business.Qualifications: 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.Bachelor’s degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.Strong technical judgment, communication skills, and ability to influence across engineering and non-engineering stakeholders.Preferred Qualifications: Experience in FinTech, financial services, or another regulated industry.Experience with AWS and/or Azure cloud platforms.Familiarity with Kubernetes, container platforms, infrastructure-as-code, Terraform, Ansible, or similar automation tooling.Experience with observability platforms such as Datadog, CloudWatch, or similar tools.Experience improving developer experience through paved-road platforms, standardization, and self-service infrastructure capabilities.Experience supporting growth-stage companies where speed, scale, reliability, and operational discipline must be balanced.For residents of San Francisco/Bay Area, CA or New York, NY the annual salary range for this role is $150,000-$220,000 + annual bonus. Final offers may vary from the amount listed based on geography, candidate experience and expertise, annual bonus, and other factors.Upon offer, we conduct background checks that include employment and education verification, state, and county criminal history searches as well as fingerprint and drug test. Forge is proud to be an equal opportunity employer committed to supporting a diverse and inclusive workplace. Our employment decisions are made without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), gender, gender identity, gender expression, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, marital status, sexual orientation, veteran status, or any other characteristic protected by federal, state, or local laws.
$210.38k - $243.21k
Manager, Site Reliability Engineer (Hybrid in South San Francisco)About the RoleWe are seeking an experienced and hands-on Site Reliability Engineering (SRE) Manager to lead our Site Operations and infrastructure initiatives. This role is responsible for ensuring the reliability...Suggested- ...builds the platforms and tooling that help engineering teams develop, deploy, and operate... ...default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll... ...habits and tooling.Architect and manage the SLO and error-budget framework, empowering...SuggestedPermanent employmentWork experience placementWork at officeLocal area
- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology,... ...exposure (training can be provided)Cloud/SaaS experienceMemory management and dump analysis (Java heap dump analysis preferred)ITSM/...Suggested
$152.5k - $205k
...a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate... ...experience, including authoring reusable modules, managing state and environments, and delivering infrastructure changes...SuggestedFlexible hours- ...principles to see it in full.About the teamThe Engineering team at Airwallex is a diverse group of... ..., working together to build scalable, reliable, and secure products that empower... ...Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work closely...SuggestedTemporary workLocal area
$152.5k - $205k
...everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common... ..., observability, access controls, auditability, and cost management. You will troubleshoot production issues, document operational...Flexible hours$190.8k - $267.1k
...while helping Reddit grow its business. The reliability of our Ads systems directly impacts... ...Reliability team partners closely with Ads Engineering to improve reliability, scalability,... ...advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate,...For contractorsWork experience placement$114.3k - $235.32k
...advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-... ...and HelmSupporting infrastructure provisioning and change management through Terraform/TerragruntBuilding and supporting CI/CD...Work at officeLocal areaRelocationRelocation package$113.4k - $162k
...conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd,... ...GitHub, Terraform, Ansible, or similar tools to build and manage cloud infrastructure efficiently.Incident Management Expert...Temporary work$127k - $249k
The TeamPlatform Engineering sits within SRE and builds the core infrastructure... ...Engineering, the Fabric team manages the global network substrate... ...role in engineering the reliable, globally connected, multi-... ...seeking a talented Senior Site Reliability Engineer (SRE) with...Local areaRemote workWorldwideFlexible hours$117k - $209.33k
...Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable,... ...such as SLOs/SLIs, production readiness, incident management, observability, resilience testing, and toil reduction. Success...Full timeFor contractors$148.5k - $223.9k
...future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with... ...eliminate toil and improve operational efficiency.Incident Management: Lead the coordinated response to incidents as an...Full timeWorldwideWeekend work- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and... ...cross-functionallyNice-to-HaveExperience with cloud and managed services (e.g. AWS)Experience supporting data-intensive platforms...
$127k - $249k
Platform Engineering is the department within SRE that is responsible for a range of critical... ...and alerting systems.The Fleet Management team provides the core runtime environment... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...Work at officeLocal areaRemote workWorldwideFlexible hours$106k - $130k
...ineligible for employment Visa sponsorship.Role Summary The Senior Site Reliability Engineer applies software engineering and systems engineering... ...as Code, automation, testing, incident response, capacity management, resilience, and operational readiness. Identify recurring...Hourly payFull timeImmediate startVisa sponsorshipWork visaFlexible hours$195k - $257.5k
...is a stakeholder.What you’ll be responsible for:As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and... ...on:Operate and scale production blockchain infrastructure, managing full nodes across networks such as Arc, Ethereum, Solana,...Flexible hours$204k - $306k
...We're all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity,... ...week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions...Permanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week$55k - $151.47k
...LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in... ...data integrity and accessibility- Leading incident management and resolution efforts to maintain operational continuityWhat...Full timeH1b$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$220k - $235k
...this space is still largely uncharted. Engineers here are building agentic solutions to... ...the forefront of AI choose Ironclad to manage their contracts.We’re consistently recognized... ...and strategic direction for the Site Reliability Engineering team and our broader Cloud...Full timeContract workWork at office$194k - $267k
..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$153k - $191.3k
...manufacturing, data processing, and software engineering, our office is a truly inspiring mix of... ...environments, to guarantee the reliability, scalability, and availability of our services... ..., particularly resource optimization, management, and cluster tuning in a constrained...Full timeTemporary workFor contractorsWork at officeLocal areaRemote workHome office3 days per week- ...and the U.S. Special Forces. The Role We're hiring a Site Reliability Engineer to own the operational health of our connected sensor... ...Systems Builder — Close the Loop Build and maintain fleet management systems: OTA update pipelines, device health tracking, remote...Remote work
$189k - $283.6k
...Afterpay is transforming the way customers manage their spending over time. TIDAL is a... ...proactively and reactively improve the reliability of Block's platform and critical infrastructure... ...strong desire to perform and grow as an engineer ~5+ years of software development...Full timeRelocation packageFlexible hoursShift work- ...culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of... ...~6+ years of experience in Site Reliability Engineering, managing infrastructure and services at scale ~ History of end-to...Immediate startRemote workWorldwide
- ...for As an SRE at Wordbricks, you will keep our systems fast, reliable, and boring. You'll own the infrastructure and operations... ...systems Build and maintain CI/CD, observability, and alerting Manage cloud infrastructure across Cloudflare, AWS, and Vercel Lead...Remote workFlexible hours
- ...daily users while enabling our engineering teams to ship fast. You'll... ...automation and tooling that improves reliability and partnering with... ...reliability best practices Manage and optimize our infrastructure... ...you'll bring ~5+ years in Site Reliability Engineering, DevOps...Work at officeWork from home
$210k - $240k
...Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual... ..., deployment automation, rollback mechanisms, and config management Implement and maintain monitoring, alerting, and incident...Full time$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco, CA.... ...Remote unavailable. Modality: On-Site only. Must live within... ...scalability, performance, and reliability across environments. What You... ...powers AI workloads at scale Manage and automate GPU compute clusters...Full timeRemote workRelocationRelocation package$98.58k - $138.02k
...office locations: Austin, TX; Irvine, CA; or Akron, OH. Site Reliability Engineer II will be responsible for supporting, enhancing, and... ...Terraform, Ansible, or CloudFormation. Work within change management protocols to provide maximum uptime for production systems...Work at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineer. Be the first to apply!
- remediation manager San Francisco, CA
- treatment manager San Francisco, CA
- dining manager San Francisco, CA
- aesthetic manager San Francisco, CA
- macys manager San Francisco, CA
- conflicts manager San Francisco, CA
- at&t manager San Francisco, CA
- phlebotomy manager San Francisco, CA
- personal lines manager San Francisco, CA
- workday manager San Francisco, CA


