Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Director, Site Reliability Engineering

$205k - $305k

Stellar

Director Of Site Reliability Engineering

Interested in working on cutting-edge blockchain technology and creating equitable access to the global financial system? Since 2014, the mission-driven team at the Stellar Development Foundation (SDF) has helped fuel the tremendous growth of the Stellar blockchain network, an open-source platform that operates at high-scale today. Developers and companies around the world build on it, and the SDF team is expanding to support the rapidly growing and changing Stellar ecosystem.

SDF is looking for a Director of Site Reliability Engineering to lead a small, high-leverage SRE team and help shape how engineering teams own, operate, and improve production services.

This is a senior engineering leadership role reporting to the CTO. You will set the vision, operating model, and culture for SRE while owning the core infrastructure services that help SDF engineering teams build, deploy, observe, and operate software with confidence.

Engineering teams at SDF own the services they build. SRE provides the frameworks, standards, shared infrastructure, tooling, observability practices, and enablement model that make strong service ownership possible across engineering.

You will be successful here if you bring strong technical judgment, pragmatic leadership, and the ability to influence through trust, clarity, and execution. SDF is a small, mission-driven foundation with a broad technical surface area, so this role requires leverage, ownership, and a bias toward solving the right problems over creating processes for its own sake.

In this role, you will:

  • Lead, coach, and develop a distributed SRE team, setting a clear vision, charter, operating model, priorities, and success measures.
  • Define and roll out a Service Ownership & Maturity Framework across engineering, with expectations that vary appropriately by service criticality.
  • Own and improve core engineering infrastructure services, including cloud foundations, Kubernetes and compute patterns, CI/CD, observability, secrets management, GitHub workflows, and infrastructure automation.
  • Help engineering teams become stronger owners and operators of their services through better standards, dashboards, runbooks, alerting, escalation paths, operational readiness, and deployment practices.
  • Make reliability, operational maturity, infrastructure health, and developer productivity more measurable through trusted metrics and practical operational intelligence.
  • Improve deployment automation, resilience, self-healing patterns, disaster recovery readiness, and service reliability based on actual impact and risk.
  • Mature incident response, escalation, postmortems, and on-call health across a geographically distributed team.
  • Build paved paths and self-service infrastructure that reduce toil, lower cognitive load, and help engineering teams move faster while strengthening ownership and reliability.
  • Partner closely with Security, Compliance, Legal, Finance, Procurement, and Corporate IT where infrastructure, access management, cloud operations, vendor review, or controls intersect with engineering.
  • Pragmatically evaluate AI-assisted and agentic workflows where they can improve infrastructure operations, service ownership, developer workflows, or toil reduction.

You have:

  • 10+ years of experience in SRE, infrastructure engineering, platform engineering, cloud infrastructure, production operations, or closely related engineering roles.
  • 5+ years of experience leading, managing, or formally developing infrastructure, SRE, platform, or reliability engineers.
  • Strong experience defining team charters, operating models, roadmaps, success measures, and engineering practices for infrastructure or reliability teams.
  • Deep technical judgment across cloud infrastructure, production operations, distributed systems, reliability tradeoffs, automation, and operational risk.
  • 3+ years of experience with modern cloud infrastructure in AWS, GCP, or similar environments.
  • 3+ years of experience with Kubernetes, container orchestration, infrastructure-as-code, declarative systems, CI/CD, and deployment safety.
  • Strong experience with observability, monitoring, alerting, logging, dashboards, SLOs/SLIs, incident response, postmortems, and on-call practices.
  • Experience helping product or application engineering teams improve service ownership, operational readiness, and production accountability.
  • A pragmatic approach to tooling: you understand when to build, buy, adapt, simplify, or retire systems based on the actual engineering problem.
  • The ability to operate effectively in a small or mid-sized engineering organization where influence comes from credibility, judgment, and outcomes rather than bureaucracy.
  • Clear executive communication skills and the ability to partner directly with a CTO and senior engineering leaders.

Bonus points if:

  • Experience leading SRE, infrastructure, or platform work in a lean, high-agency organization.
  • Experience supporting globally distributed teams or 24/7 operational coverage.
  • Experience improving developer productivity through paved paths, self-service infrastructure, automation, and reduced toil.
  • Experience with infrastructure security fundamentals, secrets management, access controls, cloud security practices, or compliance-related infrastructure controls.
  • Experience in financial services, regulated environments, blockchain, crypto, Web3, or other high-reliability technical ecosystems.
  • Experience evaluating vendors and infrastructure platforms with skepticism, technical rigor, and cost discipline.
  • Practical experience applying AI-assisted or agentic workflows to infrastructure, reliability, operations, observability, or developer productivity.

We offer competitive pay with a base salary range for this position of $205,000 - $305,000 depending on job-related knowledge, skills, experience, and location. In addition, we offer lumen-denominated grants along with the following perks and benefits:

  • Competitive health, dental & vision coverage with most plans covered at 100% for the employee + any dependents
  • Flexible time off + 15 company holidays including a company-wide holiday break
  • Generous paid parental leave for all parents, plus paid pregnancy disability leave for birthing parents
  • Gym reimbursement ($80 per month)
  • Life & ADD (up to $50K)
  • Short & Long term disability
  • 401K with 4% match
  • Health & Dependent Care FSA Accounts
  • Commuter benefits with $250/month employer contribution
  • Health Savings Account (HSA) with monthly employer contribution
  • Family building benefits through Kindbody
  • Wellbeing benefits (One Medical, Rightway, Headspace)
  • L&D budget of $1,500/year
  • Daily lunch and snacks in office
  • Company retreats

About Stellar

Stellar is more than a blockchain. Powered by a decentralized, fast, scalable, and uniquely sustainable network made for financial products and services and a thriving and passionate ecosystem that includes a non-profit organization driven by a mission, Stellar is paving the path to unlock the world's economic potential through blockchain technology. Built with speed and low costs in mind, the Stellar network provides builders and financial institutions worldwide a platform to issue assets, and to send and convert currencies in real time creating real world utility. Founded in 2014, the Stellar Development Foundation (SDF) supports the continued development and growth of the Stellar network and also serves the ecosystem of NGOs, corporations, universities, small businesses, governments, and solo entrepreneurs building on the Stellar network through tooling, funding and strategic collaborations. Together, Stellar is where blockchain meets the real world.

About the Stellar Development Foundation

The Stellar Development Foundation (SDF) is a non-profit organization focused on working with and supporting change-makers to create equitable access to the global financial system through blockchain technology. SDF provides grants, investments, funding, and other awards to builders and organizations. SDF also develops resources and tooling on the Stellar network to help unlock real world utility. As a nonprofit foundation, SDF puts the health of the Stellar network and the Stellar ecosystem and its mission above all else.

We look forward to hearing from you!

Privacy Policy

By submitting your application, you are agreeing to our use and processing of your data in accordance with our Privacy Policy.

SDF is committed to diversity in its workforce and is proud to be an equal opportunity employer. SDF does not make hiring or employment decisions on the basis of race, color, religion, creed, gender, national origin, age, disability, veteran status, marital status, pregnancy, sex, gender expression or identity, sexual orientation, citizenship, or any other basis protected by applicable local, state or federal law.

Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the Director, Site Reliability Engineering in New York, NY vacancy
  • $126k - $248k

     ...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB's cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and... 
    Suggested
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    New York, NY
    5 days ago
  •  ...DevOps Engineer DevOps teams in our Infrastructure Engineering group enable Company to continually disrupt the Insure tech space. Our teams build, maintain and deliver infrastructure that enables Company Life product teams to ship industry leading and innovative systems... 
    Suggested

    MRINetwork

    New York, NY
    10 hours ago
  • $165k - $225k

     ...operationalize AI and run it as a true business performance engine delivering measurable value. For more, visit the Dataiku blog...  ...taking a look here. How you'll make an impact: As a Site Reliability Engineer (SRE) with advanced expertise in networking and security... 
    Suggested
    Work at office
    Flexible hours

    Dataiku

    New York, NY
    1 day ago
  • $100k - $250k

     ...financial markets. Role Roadmap As a member of Kalshi's engineering team, you'll help build the next-generation financial...  ..., and evolve. What You'll Do Improve observability, reliability, and service availability by defining and measuring key metrics... 
    Suggested
    Local area

    Kalshi

    New York, NY
    11 hours ago
  • $175k - $225k

     ...Site Reliability Engineer Chicago, IL or New York, NY Old Mission is a global proprietary trading firm that leverages state-of-the-art technology and research to identify and execute profitable trading strategies across multiple asset classes around the world. Our... 
    Suggested
    Full time
    Work at office
    Remote work
    Monday to Friday
    Flexible hours
    Rotating shift

    Old Mission Capital

    New York, NY
    3 days ago
  •  ...Applications Deployment Responsible for reliability and support of Container Platform on-...  ...Perform blameless RCA, partner with engineering and operation teams across the...  ...Additional Skills : Automation Process Engineer,Site Reliability Engineer,Full Stack DeveloperThis... 

    Kaav Inc.

    New York, NY
    4 days ago
  • $182.3k - $220k

     ...healthcare by putting patients first - and that mission depends on reliable, secure, and scalable systems. As a Senior SRE on the...  ...hardening infrastructure and building tools that empower our engineers to ship safely and confidently. You will work across teams... 
    Local area
    Flexible hours

    Modern Fertility

    New York, NY
    3 days ago
  • $89k - $178k

     ...Sr. Site Reliability Engineer I NYC Global HQ Hybrid (3 days per week in office) DV is the leader in digital performance solutions, helping our advertiser and agency partners verify the quality of their digital campaigns, optimize to improve performance and prove... 
    Work at office
    3 days per week

    DoubleVerify

    New York, NY
    2 days ago
  • $150k - $175k

     ...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed... 
    Remote work

    ASAPP

    New York, NY
    1 day ago
  •  ...About the job Senior Site Reliability Engineer About the Company Stellar is a decentralized, public blockchain that gives developers the tools to create experiences that are more like cash than crypto. The network is faster, cheaper, and far more energy-efficient... 

    TechChain Talent

    New York, NY
    1 day ago
  • $123k - $165k

     ...Site Reliability Engineer II Our engineering fleet is a horizontal set of teams providing engineering services across the organization. Our specific team provides reliability engineering and operational support to backend service development teams. Technology is... 

    Disney

    New York, NY
    2 days ago
  • $125k - $350k

     ...Site Reliability Engineer New York, Miami, Gurugram, London, Singapore, Sydney Job Description Opportunities may be available from time to time in any location in which the business is based for suitable candidates. If you are interested in a career with Citadel... 

    Citadel Securities

    New York, NY
    1 day ago
  •  ...Site Reliability Engineer I, Abhishek, would like to share a job opportunity as Site Reliability Engineer in Jacksonville, FL, Cary, NC or New York, NY (Onsite) location for a Fulltime position. In case, if you are not comfortable with this location, please share your... 
    Full time
    Work visa

    Syntricate Technologies

    New York, NY
    5 days ago
  • $127k - $249k

     ...The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    New York, NY
    1 day ago
  • $175k - $230k

     ...This role is critical to ensure Sage can live up to its mission to be a 24x7, highly available platform for elder care. As a Site Reliability Engineer, you'll partner with engineering teams across the organization to achieve four 9s of uptime for our platform.... 
    Apprenticeship
    Work at office
    Local area
    Remote work
    2 days per week

    Sage Group plc

    New York, NY
    4 days ago
  • $51.9 per hour

     ...Overview This job is responsible for the reliability, availability, and performance of...  ...operational efficiency. This role blends software engineering, clinical engineering, and security...  ...operations. Works cross-functionally with AHN site leaders and teams to navigate and to... 
    For contractors
    Local area

    Highmark Health

    New York, NY
    2 days ago
  •  ...Participate in an oncall rotation. Work with teams across the company to ensure we achieve the right balance of developer velocity, reliability and performance, and cost efficiency. What You’ll Bring 5+ years of experience Experience with containerization and orchestration... 

    SwiftCruit

    New York, NY
    1 day ago
  • $111k - $160k

     ...Join Mizuho as a Site Reliability Engineer! In this role you will play a crucial role in maintaining the reliability, scalability, and overall performance of our production systems. This position collaborates closely with development, operations, and product teams to automate... 
    Work at office
    Local area
    Remote work

    Mizuho Financial Group Inc

    New York, NY
    2 days ago
  • $131k - $164k

     ...Staff Site Reliability Engineer New York, New York, United States Position Overview We are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and automation frameworks, to join our global Infrastructure... 
    Work at office
    Local area
    Flexible hours

    Diligent

    New York, NY
    2 days ago
  •  ...Site Reliability Engineering As a Site Reliability Engineering at JPMorgan Chase within the Enterprise technology, liquidity risk team, you are the non-functional requirement owner and champion for the applications in your remit. You are a key influencer in your team... 

    Chase

    Jersey City, NJ
    1 day ago
  • $150k - $170k

     ...Senior Site Reliability Engineer – Zip Co Join to apply for the Senior Site Reliability Engineer role at Zip Co At Zip, we build cloud‑native software applications that serve millions of customers and process billions of dollars in payments. We’re looking for a seasoned... 
    Casual work
    Work at office
    Remote work
    Flexible hours

    ZIP

    New York, NY
    3 days ago
  •  ...Curated careers, resources, tips and trends from the DevOps World. The Site Reliability Engineer position at Remotive revolves around ensuring the reliability, availability, and performance of services. This role requires a combination of software engineering and system... 
    Remote work

    DevOpsChat

    New York, NY
    1 day ago
  • $185k - $227k

     ...united by this common purpose and we are hiring the world’s best engineers, scientists, designers, product managers, operations experts...  ...on for more details. ROLE AND RESPONSIBILITIES A Senior Site Reliability Engineer (SRE) is expected to own the operational stability... 
    Remote work

    JUUL Labs

    New York, NY
    1 day ago
  • $200k - $240k

     ...expertise across machine learning, UI/UX, large language models, and medicine. Job Description We’re hiring an experienced Site Reliability Engineer for our Boston or NYC office! You can expect to: Design, build, and maintain resilient, scalable, and secure... 
    Work at office

    Verana Health

    New York, NY
    3 days ago
  •  ...risk—the leading cause of cybersecurity breaches—and build safer, more resilient organizations. The Role: As a Senior Site Reliability Engineer (SRE) at Dune Security, you will play a critical role in ensuring our platform's stability, scalability, and security. You... 
    Full time
    Work at office

    Dune Security

    New York, NY
    4 days ago
  •  ...they are shifting towards Linux – (70% Windows, 30% Linux) Remote access technology protocols are a plus Job Description: Site Reliability Engineer Periodic updates and maintenance of Windows-based golden image for ESX & AWS. Patching of software, systems, appliances etc... 
    Remote work
    Shift work

    TechDigital Group

    New York, NY
    2 days ago
  • $7.5k

     ...and benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more. As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and monitor... 
    Work at office
    Local area

    The Voleon Group

    New York, NY
    1 day ago
  • $157.5k - $254.35k

     ...signature and contract lifecycle management (CLM). What you’ll do We are looking for a self‑motivated, driven and creative Senior Site Reliability Engineer to join the Site Reliability team. Metrics and analytics drive engineering at DocuSign and ensure that we are dedicating... 
    Contract work
    Work at office
    Local area
    Remote work

    DocuSign

    New York, NY
    2 days ago
  • $150k - $200k

     ...Join to apply for the Senior Site Reliability Engineer role at Gradle Inc. Develocity is a first‑of‑its‑kind toolchain observability and acceleration platform that helps software teams adopt and improve DORA capabilities (including continuous delivery) in order to achieve... 
    Full time
    Local area
    Remote work
    Work from home

    Gradle Inc.

    New York, NY
    1 day ago
  •  ...subscriptions at scale, combining the agility of a high-growth business with the backing of a global organization. As the Site Reliability Engineer, you will help ensure the reliability, scalability, and observability of CloudBlue’s multi-tenant SaaS platforms used by service... 
    Remote work
    Worldwide
    Flexible hours

    HostPapa

    New York, NY
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Director, Site Reliability Engineering. Be the first to apply!