Manager, Site Reliability Engineer
$150k - $220kForge Global
At Forge, we know our team is our greatest asset. As technology innovators in the private market, our vision is to deliver a richer future for everyone. We live that vision through our values of being bold, accountable, and humble. We experience the value that our vision brings to the world every day, helping the teams behind the greatest innovations of our generation, from space travel to artificial intelligence, and more. With liquidity solutions, exclusive data and insights, a custody offering, and a vibrant marketplace, Forge’s goal is to build the best-in-class technology infrastructure to power a global private market that is transparent, accessible, and seamless for companies, their employees, and investors. Through Forge, employees can sell their private shares, employers can reward shareholders with pre-IPO liquidity and individual and institutional investors can participate in private unicorn growth. Forge's differentiated global marketplace addresses rising demand among individual and institutional investors for exposure to private company stocks and is building a growing network effect. Our ability to offer these powerful financial solutions has generated incredible interest from investors, demand from customers, and a need to grow our team to meet the needs of more companies, teams, and innovators in this way. The Role: As an engineering organization, we pride ourselves on engineering as a creative activity. Engineering managers enable engineers to do their best work by maintaining a culture and environment where engineers can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team responsible for keeping Forge systems highly available for customers, while partnering closely with Platform, Engineering, Security, Compliance, and Product teams to improve reliability, observability, incident response, and operational maturity. This is an opportunity for a hands-on technical leader who can coach engineers, improve production operations, and help Forge build and run secure, scalable, and highly reliable products.Responsibilities: Manage Forge’s Site Reliability Engineering team responsible for keeping Forge systems highly available for customers.Drive strong incident management practices in partnership with engineering teams, including response, mitigation, follow-up, and post-incident learning.Build, improve, and manage observability infrastructure in partnership with Platform Engineering, including monitoring, alerting, dashboards, and operational metrics.Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support faster response and mitigation.Champion reliability best practices across engineering, including service ownership, operational readiness, disaster recovery, and production support standards.Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.Collaborate with engineering teams to troubleshoot production issues, identify recurring problems, and improve system reliability.Hire, coach, mentor, and manage performance for SRE team members while supporting career development and team health.• Partner with Security, Compliance, and Risk partners to ensure reliability and infrastructure practices meet the needs of a regulated business.Qualifications: 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.Bachelor’s degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.Strong technical judgment, communication skills, and ability to influence across engineering and non-engineering stakeholders.Preferred Qualifications: Experience in FinTech, financial services, or another regulated industry.Experience with AWS and/or Azure cloud platforms.Familiarity with Kubernetes, container platforms, infrastructure-as-code, Terraform, Ansible, or similar automation tooling.Experience with observability platforms such as Datadog, CloudWatch, or similar tools.Experience improving developer experience through paved-road platforms, standardization, and self-service infrastructure capabilities.Experience supporting growth-stage companies where speed, scale, reliability, and operational discipline must be balanced.For residents of San Francisco/Bay Area, CA or New York, NY the annual salary range for this role is $150,000-$220,000 + annual bonus. Final offers may vary from the amount listed based on geography, candidate experience and expertise, annual bonus, and other factors.Upon offer, we conduct background checks that include employment and education verification, state, and county criminal history searches as well as fingerprint and drug test. Forge is proud to be an equal opportunity employer committed to supporting a diverse and inclusive workplace. Our employment decisions are made without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), gender, gender identity, gender expression, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, marital status, sexual orientation, veteran status, or any other characteristic protected by federal, state, or local laws.
$150k - $220k
...in this way. The Role: As an engineering organization, we pride ourselves on engineering... ...as a creative activity. Engineering managers enable engineers to do their best work... ..., mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team...SuggestedFull timeLocal area$151k - $297k
..., you will partner with SRE leaders and engineers to scale the platform that underpins all... ...program execution, strengthen production reliability practices, and coordinate cross-... ...criteria with SRE engineers and leaders. Manage dependencies across platform teams, keep...SuggestedLocal areaRemote workWorldwideFlexible hours$158.5k - $172k
...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will... ...ecosystems. Our team is responsible for managing our centralized Enterprise Logging... ...high-impact position driving continuous reliability, deep system optimization, and automation...SuggestedFull timeTemporary workWork at officeFlexible hours3 days per week$167.7k - $245.2k
...performance, efficiency, change management, monitoring, emergency... ...effective.We’re looking for talented engineers with a software or operations... ...teams to ensure the reliability, performance and security of... ...Please see the Cisco careers site to discover more benefits and...SuggestedFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week$141k - $216.6k
...—it means helping shape the future of emergency response and building a safer, more connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability, and operational excellence of our Unified Call (UC) platform—the mission-critical...SuggestedWork experience placementWork at office- ...innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment Bank, Production management team, you will solve complex and broad business problems with simple and...Shift work
$120k - $200k
...DaveContact Email: ****@*****.*** Reliability Engineer(SRE) ResponsibilitiesGlobal... ...Infrastructure Platform Deployment & Operations Manage the deployment, operation, and continuous... ...testing, automated recovery)SkillsBilingual Mandarin Site Reliability Engineer(SRE)Overseas$139k - $257.55k
...Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning,... ...themOwn patch, vulnerability, and golden-image lifecycle management across the fleet — triage, remediate, automateContribute to...Full timeTemporary workLocal areaRemote workWorldwide$200k - $250k
Hudson River Trading (HRT) is seeking a Senior Site Reliability Engineer to join our growing Enterprise SRE team. This team is responsible for developing... ...Technology team.In this role you’ll build, enhance, and manage a variety of internal systems - from directory services to...Work at officeLocal areaImmediate start$156k - $262k
...building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to... ...LLMs and the real world. The Role: Senior Site Reliability Engineer ~ Managing Kubernetes clusters across multiple environments and regions...Full timeTemporary workWork at officeImmediate startRemote work$110k - $120k
...expertise, scale, and technology.Job DescriptionJob Title: Site Reliability Engineer (SRE) / L3 Support EngineerGetting to know us:As a leading... ....Experience with infrastructure as code and configuration management.Strong scripting or programming skills (e.g. Python, Bash,...Ongoing contractFull timeCasual workRemote workFlexible hours- ...Cohere is a team of researchers, engineers, designers, and more, who are... ...-performance, scalable and reliable machine learning systems? Do... ...? We are looking for a Site Reliability Engineer to join... ...service systems that automate managing, deploying and operating services...Full timeWork experience placementWork at officeLocal areaRemote workHome office
$174k - $267k
...it” and who can rapidly self-educate on new concepts and tools. Position Overview: The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses...Permanent employmentFull timeWork at officeLocal areaWorldwideFlexible hours- ...direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Commercial & Investment Bank, Production Management team, you hold a leadership role in your team, demonstrate strong...
$131k - $164k
Position OverviewWe are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and automation... ...Directory for authentication, policy, and service account management across hybrid environments. Collaborate with network and...Work at officeLocal areaVisa sponsorshipFlexible hours$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$150k - $250k
What We DoAt Goldman Sachs, our Engineers don't just make things - we make things possible... ...Global Banking & Markets business, the Site Reliability Engineering (SRE) team ensures the... ...built in at every layer.Establish and manage SLIs, SLOs, and error budgets; drive blameless...Full timeTemporary workPart time$182k - $250.8k
...at Okta is the backbone of our platform's reliability and operational excellence. We are a forward-thinking group of engineers and leaders who believe that great infrastructure... ...for millions of users worldwide. As a Manager, Site Reliability Engineer, you'll lead this...Permanent employmentLocal areaRemote workWorldwideFlexible hoursWeekend workWeekday work$194k - $267k
...let's talk.The TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission... ...balancing, ingress, TLS, service networking, and traffic management.Strategic experience designing comprehensive observability...Local areaWorldwideFlexible hours$141.8k - $195k
...we give customers the choice, control, and flexibility to manage and analyze telemetry for both humans and agents, so they can... ...You'll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all...Temporary workRemote work- ...Site Reliability Engineer (SRE) Job Title Site Reliability Engineer (SRE) Job Summary We are seeking a skilled Site Reliability Engineer (SRE)... ...(RCA), and implement preventive measures. Configure and manage observability solutions including logging, monitoring, tracing...Flexible hours
- ...aligned to service health. Drive incident management: on-call readiness, triage, incident... .../rollback automation). Establish reliability standards: SLOs/SLIs, error budgets, production... .... Performance and reliability engineering: capacity planning, load/performance...
- ...Senior Site Reliability Engineer (SRE) Our client is seeking a Senior Site Reliability Engineer (SRE) with 10–15 years of experience to support... ...on troubleshooting complex trading infrastructure, managing observability, and collaborating across trading and technology...
- ...Chariot Engineering Hire Chariot's engineering hire will be responsible for taking the Chariot platform to the next level. You will... ...world while working as part of a small, fast-moving team. Manage and build out the backend server applications and components that...Work experience placementWork at officeWork from homeMonday to FridayMonday to Thursday
$150k - $175k
...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve... ...~ Knowledge of AWS services, containers and container management frameworks ~ Familiarity with message bus based systems...Remote work- ...DevOps Engineer Responsible for reliability and support of container platform on-prem and external clouds (Azure /AWS /Google) Monitor and troubleshoot... ...into systemic and latent reliability issues, incident management, problem management Identifying, analyzing, and...
- ...Site Reliability Engineer I, Abhishek, would like to share a job opportunity as Site Reliability Engineer in Jacksonville, FL, Cary, NC or New York, NY (Onsite) location for a Fulltime position. In case, if you are not comfortable with this location, please share your...Full timeWork visa
$195k - $275k
...providing a wide range of investment banking, securities, investment management and wealth management services. The Firm's employees serve... ...& Release Management, and the Chief Operating Office.The Reliability Operations (RO) within WMT is responsible for providing swift...Temporary workWork at officeWorldwideNight shift$150k - $190k
Senior Site Reliability Engineer, VPAt Morgan Stanley, we advise, originate, trade, manage and distribute capital for governments, institutions and individuals, and always do so with a standard of excellence. We are a leading global financial services firm that conducts...Temporary workWorldwideFlexible hoursWeekend work$130k - $250k
What We DoAt Goldman Sachs, our Engineers don't just make things - we make things possible... ...your journey here.Securities Frontline Site Reliability Engineers (SREs) play a critical role... ...tools to eliminate operational toil, and manage robust UAT and Production environments....Full timeTemporary workPart timeImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineer. Be the first to apply!

