Director, Site Reliability Engineering

$205k - $305k

Stellar

Director Of Site Reliability Engineering

Interested in working on cutting-edge blockchain technology and creating equitable access to the global financial system? Since 2014, the mission-driven team at the Stellar Development Foundation (SDF) has helped fuel the tremendous growth of the Stellar blockchain network, an open-source platform that operates at high-scale today. Developers and companies around the world build on it, and the SDF team is expanding to support the rapidly growing and changing Stellar ecosystem.

SDF is looking for a Director of Site Reliability Engineering to lead a small, high-leverage SRE team and help shape how engineering teams own, operate, and improve production services.

This is a senior engineering leadership role reporting to the CTO. You will set the vision, operating model, and culture for SRE while owning the core infrastructure services that help SDF engineering teams build, deploy, observe, and operate software with confidence.

Engineering teams at SDF own the services they build. SRE provides the frameworks, standards, shared infrastructure, tooling, observability practices, and enablement model that make strong service ownership possible across engineering.

You will be successful here if you bring strong technical judgment, pragmatic leadership, and the ability to influence through trust, clarity, and execution. SDF is a small, mission-driven foundation with a broad technical surface area, so this role requires leverage, ownership, and a bias toward solving the right problems over creating processes for its own sake.

In this role, you will:

Lead, coach, and develop a distributed SRE team, setting a clear vision, charter, operating model, priorities, and success measures.
Define and roll out a Service Ownership & Maturity Framework across engineering, with expectations that vary appropriately by service criticality.
Own and improve core engineering infrastructure services, including cloud foundations, Kubernetes and compute patterns, CI/CD, observability, secrets management, GitHub workflows, and infrastructure automation.
Help engineering teams become stronger owners and operators of their services through better standards, dashboards, runbooks, alerting, escalation paths, operational readiness, and deployment practices.
Make reliability, operational maturity, infrastructure health, and developer productivity more measurable through trusted metrics and practical operational intelligence.
Improve deployment automation, resilience, self-healing patterns, disaster recovery readiness, and service reliability based on actual impact and risk.
Mature incident response, escalation, postmortems, and on-call health across a geographically distributed team.
Build paved paths and self-service infrastructure that reduce toil, lower cognitive load, and help engineering teams move faster while strengthening ownership and reliability.
Partner closely with Security, Compliance, Legal, Finance, Procurement, and Corporate IT where infrastructure, access management, cloud operations, vendor review, or controls intersect with engineering.
Pragmatically evaluate AI-assisted and agentic workflows where they can improve infrastructure operations, service ownership, developer workflows, or toil reduction.

You have:

10+ years of experience in SRE, infrastructure engineering, platform engineering, cloud infrastructure, production operations, or closely related engineering roles.
5+ years of experience leading, managing, or formally developing infrastructure, SRE, platform, or reliability engineers.
Strong experience defining team charters, operating models, roadmaps, success measures, and engineering practices for infrastructure or reliability teams.
Deep technical judgment across cloud infrastructure, production operations, distributed systems, reliability tradeoffs, automation, and operational risk.
3+ years of experience with modern cloud infrastructure in AWS, GCP, or similar environments.
3+ years of experience with Kubernetes, container orchestration, infrastructure-as-code, declarative systems, CI/CD, and deployment safety.
Strong experience with observability, monitoring, alerting, logging, dashboards, SLOs/SLIs, incident response, postmortems, and on-call practices.
Experience helping product or application engineering teams improve service ownership, operational readiness, and production accountability.
A pragmatic approach to tooling: you understand when to build, buy, adapt, simplify, or retire systems based on the actual engineering problem.
The ability to operate effectively in a small or mid-sized engineering organization where influence comes from credibility, judgment, and outcomes rather than bureaucracy.
Clear executive communication skills and the ability to partner directly with a CTO and senior engineering leaders.

Bonus points if:

Experience leading SRE, infrastructure, or platform work in a lean, high-agency organization.
Experience supporting globally distributed teams or 24/7 operational coverage.
Experience improving developer productivity through paved paths, self-service infrastructure, automation, and reduced toil.
Experience with infrastructure security fundamentals, secrets management, access controls, cloud security practices, or compliance-related infrastructure controls.
Experience in financial services, regulated environments, blockchain, crypto, Web3, or other high-reliability technical ecosystems.
Experience evaluating vendors and infrastructure platforms with skepticism, technical rigor, and cost discipline.
Practical experience applying AI-assisted or agentic workflows to infrastructure, reliability, operations, observability, or developer productivity.

We offer competitive pay with a base salary range for this position of $205,000 - $305,000 depending on job-related knowledge, skills, experience, and location. In addition, we offer lumen-denominated grants along with the following perks and benefits:

Competitive health, dental & vision coverage with most plans covered at 100% for the employee + any dependents
Flexible time off + 15 company holidays including a company-wide holiday break
Generous paid parental leave for all parents, plus paid pregnancy disability leave for birthing parents
Gym reimbursement ($80 per month)
Life & ADD (up to $50K)
Short & Long term disability
401K with 4% match
Health & Dependent Care FSA Accounts
Commuter benefits with $250/month employer contribution
Health Savings Account (HSA) with monthly employer contribution
Family building benefits through Kindbody
Wellbeing benefits (One Medical, Rightway, Headspace)
L&D budget of $1,500/year
Daily lunch and snacks in office
Company retreats

About Stellar

Stellar is more than a blockchain. Powered by a decentralized, fast, scalable, and uniquely sustainable network made for financial products and services and a thriving and passionate ecosystem that includes a non-profit organization driven by a mission, Stellar is paving the path to unlock the world's economic potential through blockchain technology. Built with speed and low costs in mind, the Stellar network provides builders and financial institutions worldwide a platform to issue assets, and to send and convert currencies in real time creating real world utility. Founded in 2014, the Stellar Development Foundation (SDF) supports the continued development and growth of the Stellar network and also serves the ecosystem of NGOs, corporations, universities, small businesses, governments, and solo entrepreneurs building on the Stellar network through tooling, funding and strategic collaborations. Together, Stellar is where blockchain meets the real world.

About the Stellar Development Foundation

The Stellar Development Foundation (SDF) is a non-profit organization focused on working with and supporting change-makers to create equitable access to the global financial system through blockchain technology. SDF provides grants, investments, funding, and other awards to builders and organizations. SDF also develops resources and tooling on the Stellar network to help unlock real world utility. As a nonprofit foundation, SDF puts the health of the Stellar network and the Stellar ecosystem and its mission above all else.

We look forward to hearing from you!

By submitting your application, you are agreeing to our use and processing of your data in accordance with our Privacy Policy.

SDF is committed to diversity in its workforce and is proud to be an equal opportunity employer. SDF does not make hiring or employment decisions on the basis of race, color, religion, creed, gender, national origin, age, disability, veteran status, marital status, pregnancy, sex, gender expression or identity, sexual orientation, citizenship, or any other basis protected by applicable local, state or federal law.

Apply

Vacancy posted 3 days ago

Similar jobs that could be interesting for youBased on the Director, Site Reliability Engineering in New York, NY vacancy

Staff Technical Program Manager, Site Reliability Engineering
$126k - $248k
...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB’s cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and...
Suggested
Local area
Remote work
Worldwide
Flexible hours
MongoDB
New York, NY
3 days ago
Senior Software Engineer, Site Reliability Engineering
$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building resilient, highly available cloud platforms that enable engineering teams to move quickly and confidently? Do you enjoy automating complex operational...
Suggested
Ridge Line Services
New York, NY
3 days ago
Site Reliability Engineer
...paced, regulated financial environment. • Excellent communication and collaboration skills. Role Overview The System Engineer will be responsible for designing, implementing, and maintaining enterprise-level infrastructure solutions across Linux...
Suggested
Q1 Technologies
New York, NY
15 hours ago
Senior Site Reliability Engineer (SRE)
...human risk—the leading cause of cybersecurity breaches—and build safer, more resilient organizations. The Role: As a Senior Site Reliability Engineer (SRE) at Dune Security, you will play a critical role in ensuring our platform's stability, scalability, and security. You...
Suggested
Full time
Work at office
Dune Security
New York, NY
2 days ago
Site Reliability Engineer
$100k - $250k
...financial markets. Role Roadmap As a member of Kalshi's engineering team, you'll help build the next-generation financial... ..., and evolve. What You'll Do Improve observability, reliability, and service availability by defining and measuring key metrics...
Suggested
Local area
Kalshi
New York, NY
3 days ago
Site Reliability Engineer
...Applications Deployment Responsible for reliability and support of Container Platform on-... ...Perform blameless RCA, partner with engineering and operation teams across the... ...Additional Skills : Automation Process Engineer,Site Reliability Engineer,Full Stack DeveloperThis...
Kaav Inc.
New York, NY
2 days ago
Site Reliability Engineer II
$123k - $165k
...Site Reliability Engineer II Our engineering fleet is a horizontal set of teams providing engineering services across the organization. Our specific team provides reliability engineering and operational support to backend service development teams. Technology is...
Disney France
New York, NY
8 hours ago
Site Reliability Engineer (SRE)
...Site Reliability Engineer I, Abhishek, would like to share a job opportunity as Site Reliability Engineer in Jacksonville, FL, Cary, NC or New York, NY (Onsite) location for a Fulltime position. In case, if you are not comfortable with this location, please share your...
Full time
Work visa
Syntricate Technologies
New York, NY
3 days ago
Site Reliability Engineer
$125k - $150k
...Site Reliability Engineer Virtu is a leading financial firm that leverages cutting edge technology to deliver liquidity to the global markets and innovative, transparent trading solutions to our clients. As a market maker, Virtu provides deep liquidity that helps to...
Worldwide
Virtu Financial
New York, NY
15 hours ago
Senior Site Reliability Engineer, Fleet Management
$127k - $249k
THE TEAM Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational... ..., alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....
Work at office
Local area
Remote work
Worldwide
Flexible hours
MongoDB
New York, NY
4 days ago
Senior Site Reliability Engineer (SRE)
$100 per hour
...Where you'll create impact Improve reliability of our systems Build & maintain our... ...about new frameworks and solutions to engineering problems Fast-moving: you deploy daily... ...benefits ~401k benefits ~ On-site team culture - high collaboration, no bureaucracy...
Immediate start
Weekend work
DualEntry
New York, NY
15 hours ago
Senior Site Reliability Engineer (SRE)
...Senior Site Reliability Engineer (SRE) Our client is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies....
Local area
E-Solutions
New York, NY
4 days ago
SITE RELIABILITY ENGINEER
...DevOps Engineer DevOps teams in our Infrastructure Engineering group enable Company to continually disrupt the Insure tech space. Our teams build, maintain and deliver infrastructure that enables Company Life product teams to ship industry leading and innovative systems...
MRINetwork
New York, NY
3 days ago
Site Reliability Engineer
...Asia and the Middle East. We are creative, low-ego and team-spirited. The Role We are seeking highly experienced Site Reliability Engineers (SRE) to shape the reliability, scalability and performance of our platform and customer facing applications. You will work...
Relocation package
Mistral AI
New York, NY
1 day ago
Site Reliability Engineer
...Site Reliability Engineer Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling...
Flexible hours
Baseten
New York, NY
3 days ago
Senior Site Reliability Engineer
$189k - $283.6k
...the SRE team, you will proactively and reactively improve the reliability of Block's platform and critical infrastructure. You are metrics... ...~ A strong desire to perform and grow as an engineer ~5+ years of software development experience Technologies...
Full time
Local area
Remote work
Relocation package
Flexible hours
Shift work
Block USA
New York, NY
3 days ago
Site Reliability Engineer (SRE)
...Site Reliability Engineer (SRE) Job Title Site Reliability Engineer (SRE) Job Summary We are seeking a skilled Site Reliability Engineer (SRE) to build, automate, and maintain highly available, scalable, and reliable infrastructure and applications...
Flexible hours
Ova Technologies
New York, NY
15 hours ago
Site Reliability Engineer
$176.75k - $209.1k
...Site Reliability Engineer At Peloton, we view Platform as a Product. A phenomenal platform unlocks speed of development and learning. It allows us to scale easily, enabling our engineers to maximize attention on new features and capabilities. A key to crafting a phenomenal...
Temporary work
Peloton
New York, NY
15 hours ago
Senior Site Reliability Engineer
$150k - $175k
...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed...
Remote work
ASAPP
New York, NY
4 days ago
Site Reliability Engineer (SRE)
...self-healing, deployment/rollback automation). Establish reliability standards: SLOs/SLIs, error budgets, production readiness reviews... ..., and release risk controls. Performance and reliability engineering: capacity planning, load/performance analysis, resilience...
Bahwan CyberTek
New York, NY
1 day ago
Senior Site Reliability Engineer
...About the job Senior Site Reliability Engineer About the Company Stellar is a decentralized, public blockchain that gives developers the tools to create experiences that are more like cash than crypto. The network is faster, cheaper, and far more energy-efficient...
TechChain Talent
New York, NY
3 days ago
Senior Site Reliability Engineer
$225k - $325k
...What you'll do day-to-day Ensure the scalability, reliability, and observability of our systems to maintain and improve the firm's core infrastructure environment. Lead a range of engineering projects, from developing proprietary platforms for configuration...
Hourly pay
D. E. Shaw & Co.
New York, NY
1 day ago
Site Reliability Engineer
$152.5k - $219.2k
...global cloud platform. As a team of six engineers distributed across the US, Canada, and the... ...with a strong focus on automation, reliability, and operational excellence. We are one... ...Qualifications ~2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure...
Permanent employment
Full time
Temporary work
Local area
Worldwide
Flexible hours
CISCO, Inc.
New York, NY
3 days ago
Site Reliability Engineer
$115k - $125k
...Site Reliability Engineer New York City, NY Pico fuels the global capital markets community by providing exceptional market data services and customized managed infrastructure solutions. As financial industry experts at the center of markets and technology, we help...
Work experience placement
Work at office
Work from home
Monday to Friday
Flexible hours
Shift work
Weekend work
Afternoon shift
Early shift
Pico
New York, NY
3 days ago
Senior Site Reliability Engineer
$140k - $170k
...data security, navigate complex regulatory compliance and optimize business interactions. Role Description: As a Site Reliability Engineer, you will work with Agile engineering teams to provide production insight into running and operating software at-scale in...
Local area
Symphony
New York, NY
3 days ago
Staff Site Reliability Engineer - Kubernetes
$194k - $267k
...something more than once, automate it" and who can rapidly self-educate on new concepts and tools. Position Overview: The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and...
Permanent employment
Work at office
Local area
Worldwide
Flexible hours
Okta, Inc.
New York, NY
15 hours ago
Staff Site Reliability Engineer
...Staff Site Reliability Engineer Tabs is the leading AI-native revenue platform for modern finance and accounting teams. Tabs agents automate the entire contract-to-cash lifecycle, including billing, collections, revenue recognition, and reporting, to help teams eliminate...
Full time
Contract work
Work at office
TABS inc.
New York, NY
3 days ago
Staff Site Reliability Engineer
.... No one coasts. If you're driven by impact, pace, and raising the bar. This is the place. The Role As a Staff Site Reliability Engineer you'll play a lead role on the founding SRE team at our new NYC engineering hub. You'll own multi-team reliability and infrastructure...
Work at office
Legora
New York, NY
3 days ago
Senior/Staff Site Reliability Engineer
$175k - $230k
...Senior/Staff Site Reliability Engineer New York, New York, United States Sage is on a mission to improve care and quality of life for older adults, starting with those residing in senior living facilities. Falls are the leading cause of injury-related death among...
Apprenticeship
Work at office
Local area
Remote work
2 days per week
SAGE
New York, NY
3 days ago
Staff Site Reliability Engineer
$131k - $164k
...Position Overview We are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and automation frameworks, to join our global Infrastructure & Operations team. This role is a hands-on senior engineering position...
Work at office
Local area
Visa sponsorship
Flexible hours
Diligent
New York, NY
15 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Director, Site Reliability Engineering. Be the first to apply!