Manager, Site Reliability Engineer
$150k - $220kForge Global
At Forge, we know our team is our greatest asset. As technology innovators in the private market, our vision is to deliver a richer future for everyone. We live that vision through our values of being bold, accountable, and humble. We experience the value that our vision brings to the world every day, helping the teams behind the greatest innovations of our generation, from space travel to artificial intelligence, and more.
With liquidity solutions, exclusive data and insights, a custody offering, and a vibrant marketplace, Forge's goal is to build the best-in-class technology infrastructure to power a global private market that is transparent, accessible, and seamless for companies, their employees, and investors. Through Forge, employees can sell their private shares, employers can reward shareholders with pre-IPO liquidity and individual and institutional investors can participate in private unicorn growth.
Forge's differentiated global marketplace addresses rising demand among individual and institutional investors for exposure to private company stocks and is building a growing network effect.
Our ability to offer these powerful financial solutions has generated incredible interest from investors, demand from customers, and a need to grow our team to meet the needs of more companies, teams, and innovators in this way.
The Role:
- Manage Forge's Site Reliability Engineering team responsible for keeping Forge systems highly available for customers.
- Drive strong incident management practices in partnership with engineering teams, including response, mitigation, follow-up, and post-incident learning.
- Build, improve, and manage observability infrastructure in partnership with Platform Engineering, including monitoring, alerting, dashboards, and operational metrics.
- Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support faster response and mitigation.
- Champion reliability best practices across engineering, including service ownership, operational readiness, disaster recovery, and production support standards.
- Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.
- Collaborate with engineering teams to troubleshoot production issues, identify recurring problems, and improve system reliability.
- Hire, coach, mentor, and manage performance for SRE team members while supporting career development and team health.
- 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.
- 10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.
- Bachelor's degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.
- Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.
- Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.
- Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.
- Strong technical judgment, communication skills, and ability to influence across engineering and non-engineering stakeholders.
- Experience in FinTech, financial services, or another regulated industry.
- Experience with AWS and/or Azure cloud platforms.
- Familiarity with Kubernetes, container platforms, infrastructure-as-code, Terraform, Ansible, or similar automation tooling.
- Experience with observability platforms such as Datadog, CloudWatch, or similar tools.
- Experience improving developer experience through paved-road platforms, standardization, and self-service infrastructure capabilities.
- Experience supporting growth-stage companies where speed, scale, reliability, and operational discipline must be balanced.
Forge is proud to be an equal opportunity employer committed to supporting a diverse and inclusive workplace. Our employment decisions are made without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), gender, gender identity, gender expression, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, marital status, sexual orientation, veteran status, or any other characteristic protected by federal, state, or local laws.
$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient... ...(SLOs), and error budget practices to proactively manage reliability. Identify capacity constraints and reliability...SuggestedFull time- ...EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In this... ...: • Everything as Code: Drive repository-led management across our public and private cloud environments to establish...SuggestedPermanent employmentFull timeH1bLocal areaRemote workShift work
- ...Job Title WHAT YOU'LL DO DAY-TO-DAY: The engineer will be responsible for developing and implementing new systems and services in the areas of infrastructure monitoring, configuration management, and automation. Additional responsibilities include upgrading and/or re...Suggested
- ...Triomics Backend Engineer Triomics is building the agentic AI layer for oncology EHRs. Cancer hospitals spend billions on highly trained... ...across both Triomics and customer environments (AWS, Azure) Manage Kubernetes clusters, containerized services, CI/CD, and release...SuggestedDay shift
- ...Site Reliability Engineer TXSE is building the next-generation exchange infrastructure to support transparent, efficient, and resilient capital markets. With SEC approval and $275MM in funding, we are currently hiring a Site Reliability Engineer to help with a greenfield...SuggestedCurrently hiring
- ...Sr. Site Reliability Engineer (SRE) New York City, NY - LOCALS ONLY Hybrid, 3 days 6-Month Contract 10-15 years Our client is seeking... ...on troubleshooting complex trading infrastructure, managing observability, and collaborating across trading and technology...Contract workLocal area
$111k - $218k
...The Site Reliability Engineering team designs and builds the global infrastructure on which we deploy our services, focusing on the above mentioned... ...Services, Google Compute, Microsoft Azure) Experience managing kubernetes clusters or some other container orchestration...Local areaWorldwideFlexible hours- ...USA Exp: 8-12 Years Client: Amex Job Description: SRE Engineer (This is not a Devops role, strictly need an SRE Engineer, who has great analytical skills and is a good incident manager as well) This is an SRE role supporting the B2B and Core Services....
$100k - $250k
...financial markets. Role Roadmap As a member of Kalshi's engineering team, you'll help build the next-generation financial... ..., and evolve. What You'll Do Improve observability, reliability, and service availability by defining and measuring key metrics...Local area- ...the bar. This is the place. The role As a Senior Site Reliability Engineer you'll join the founding SRE team at our new NYC engineering... ...and practical tradeoffs Strong observability, incident management, and on-call experience, as well as experience with cloud...Work at office
$115k - $125k
...Site Reliability Engineer New York City, NY Pico fuels the global capital markets community by providing exceptional market data services and customized managed infrastructure solutions. As financial industry experts at the center of markets and technology, we help...Work experience placementWork at officeWork from homeMonday to FridayFlexible hoursShift workWeekend workAfternoon shiftEarly shift$105k - $300k
...Site Reliability Engineer At Citadel, a leading investor in the world's financial markets, we aim to win together as one team to earn the... ...solutions for issues based on root cause analyses Own incident management and drive tactical and strategic solutions Lead by...- ...Lead Site Reliability Engineer Assume a critical role in defining the future of a globally recognized firm and have a direct and significant... ...Chase within the Commercial & Investment Bank, Production Management team, you hold a leadership role in your team, demonstrate...
$120k - $180k
...people, and works with high-profile manufacturers including leaders in space and defense. You will be the first dedicated Site Reliability Engineer and own critical infrastructure end to end. This is a greenfield opportunity to architect the path from AWS to on-premises...Permanent employmentFull timeRelocation package- ...Overview We are seeking an experienced Observability / Site Reliability Engineer (SRE) to design, scale, and maintain our enterprise monitoring... ...-native tools. Key Responsibilities GCP & Cloud Management: Architect, optimize, and maintain observability frameworks...Temporary work
- ...Hi Position: Site Reliability Engineer/ Developer Location:New York, NY, Phoenix, AZ ( Onsite ) Duration: 6-12 Months COntract... ...communication skills - able to explain concepts to product managers and business partners in ways that are relevant to them...Contract workWork experience placement
$80k - $95k
...applications, and web-based product offerings. In this role, the Site Reliability Engineer (SRE) will play a key role in maintaining resources at peak... ...** * Monitor systems for uptime, create and manage alerting * Triaging issues through log analysis and basic...Remote workVisa sponsorshipWork visa$115k - $160k
...looking for proven incident management capability, including rapid... .... We need strong systems engineering expertise, with deep knowledge... ...fixes that preserve reliability and performance. We require... ...innovative solutions. This Senior Site Reliability Engineer - AVP -...Full timeWork at office$61k - $101k
...year Requirements: Formal training or certification in site reliability engineering, plus 3+ years of applied experience Strong grasp of... ...maintaining runbooks, ensuring clear handoffs and escalation paths, managing shift coverage, and driving disciplined post-incident...Full timeShift work- ...Site Reliability Engineer Our Client, a multinational telecommunications technology company is seeking a Site Reliability Engineer (SRE I) to join our Video Platform Engineering Team. As a Level 1 SRE, you will work closely with senior engineers to respond to incidents...Temporary work
$500 per month
...broker-dealers, investment advisors, wealth managers, hedge funds, and crypto exchanges,... ...team is a diverse group of experienced engineers, traders, and brokerage professionals... ...encourage you to apply. Your Role: As a Site Reliability Engineer at Alpaca, you'll help keep...Home office- ...Overview: Dune Security is the world’s first User Adaptive Risk Management solution. Powered by AI, we quantify employee risk with... ..., more resilient organizations. The Role: As a Senior Site Reliability Engineer (SRE) at Dune Security, you will play a critical role in ensuring...Full timeWork at office
$151k - $297k
..., you will partner with SRE leaders and engineers to scale the platform that underpins all... ...program execution, strengthen production reliability practices, and coordinate cross-... ...criteria with SRE engineers and leaders. Manage dependencies across platform teams, keep...Local areaRemote workWorldwideFlexible hours- ...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives... ...future Active-Active) architectures Platform-as-Code (PaC): Manage and automate Datadog configurations (Monitors, Dashboards, Synthetics...
- ...team rapidly, and they are looking for a Senior DevOps Engineer / Site Reliability Engineer who can join. If you’re passionate about high... ...streamline and minimize errors. Deployment: Use configuration management software to automatically deploy updates and fixes into...
$220k - $260k
...is headquartered in New York and brings deep expertise in finance and AI. About the Role We’re looking for a Staff Site Reliability Engineer to lead the evolution of Tabs’ platform as we scale. In this role, you’ll operate as a senior individual contributor, partnering...Full timeContract workWork at office$195k - $275k
...providing a wide range of investment banking, securities, investment management and wealth management services. The Firm's employees serve... ...Release Management, and the Chief Operating Office. The Reliability Operations (RO) within WMT is responsible for providing swift...Full timeTemporary workWork at officeWorldwideNight shift$100k - $150k
...offering tremendous career growth potential. Job Title: Site Reliability Engineer Technical Lead Location: New Albany, NY (Onsite)... ...excellence of critical production systems. This role is less about managing project timelines and resources, and more about creating...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...together: Forward-deployed expertise in engineering, product, and research Mosaic, our in... ...Applied AI Engineers, Embedded Product Managers and Researchers motivated by diffusing... ...based. Expect a couple of days a month on site with a customer, more when a deployment...Full timeImmediate startShift work
- Sr Site Reliability Engineer (Linux, UNIX, Reliability Engineering, Python, C, C++, Java, DevOps) in New York City C, C++, DevOps Engineer, Java... .... • Participate in system design consulting, platform management and capacity planning. • Develop software and systems architectural...Permanent employmentFull timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineer. Be the first to apply!
- extraction manager New York, NY
- certification manager New York, NY
- senior manager tax New York, NY
- ranch manager New York, NY
- sterile manager New York, NY
- valuation manager New York, NY
- lean manager New York, NY
- employment manager New York, NY
- senior preconstruction manager New York, NY
- e-learning manager New York, NY



