Director, Site Reliability Engineering
$205k - $305kStellar
Interested in working on cutting-edge blockchain technology and creating equitable access to the global financial system? Since 2014, the mission-driven team at the Stellar Development Foundation (SDF) has helped fuel the tremendous growth of the Stellar blockchain network, an open-source platform that operates at high-scale today. Developers and companies around the world build on it, and the SDF team is expanding to support the rapidly growing and changing Stellar ecosystem.
SDF is looking for a Director of Site Reliability Engineering to lead a small, high-leverage SRE team and help shape how engineering teams own, operate, and improve production services.
This is a senior engineering leadership role reporting to the CTO. You will set the vision, operating model, and culture for SRE while owning the core infrastructure services that help SDF engineering teams build, deploy, observe, and operate software with confidence.
Engineering teams at SDF own the services they build. SRE provides the frameworks, standards, shared infrastructure, tooling, observability practices, and enablement model that make strong service ownership possible across engineering.
You will be successful here if you bring strong technical judgment, pragmatic leadership, and the ability to influence through trust, clarity, and execution. SDF is a small, mission-driven foundation with a broad technical surface area, so this role requires leverage, ownership, and a bias toward solving the right problems over creating processes for its own sake.
In this role, you will:
-
Lead, coach, and develop a distributed SRE team, setting a clear vision, charter, operating model, priorities, and success measures.
-
Define and roll out a Service Ownership & Maturity Framework across engineering, with expectations that vary appropriately by service criticality.
-
Own and improve core engineering infrastructure services, including cloud foundations, Kubernetes and compute patterns, CI/CD, observability, secrets management, GitHub workflows, and infrastructure automation.
-
Help engineering teams become stronger owners and operators of their services through better standards, dashboards, runbooks, alerting, escalation paths, operational readiness, and deployment practices.
-
Make reliability, operational maturity, infrastructure health, and developer productivity more measurable through trusted metrics and practical operational intelligence.
-
Improve deployment automation, resilience, self-healing patterns, disaster recovery readiness, and service reliability based on actual impact and risk.
-
Mature incident response, escalation, postmortems, and on-call health across a geographically distributed team.
-
Build paved paths and self-service infrastructure that reduce toil, lower cognitive load, and help engineering teams move faster while strengthening ownership and reliability.
-
Partner closely with Security, Compliance, Legal, Finance, Procurement, and Corporate IT where infrastructure, access management, cloud operations, vendor review, or controls intersect with engineering.
-
Pragmatically evaluate AI-assisted and agentic workflows where they can improve infrastructure operations, service ownership, developer workflows, or toil reduction.
You have:
-
10+ years of experience in SRE, infrastructure engineering, platform engineering, cloud infrastructure, production operations, or closely related engineering roles.
-
5+ years of experience leading, managing, or formally developing infrastructure, SRE, platform, or reliability engineers.
-
Strong experience defining team charters, operating models, roadmaps, success measures, and engineering practices for infrastructure or reliability teams.
-
Deep technical judgment across cloud infrastructure, production operations, distributed systems, reliability tradeoffs, automation, and operational risk.
-
3+ years of experience with modern cloud infrastructure in AWS, GCP, or similar environments.
-
3+ years of experience with Kubernetes, container orchestration, infrastructure-as-code, declarative systems, CI/CD, and deployment safety.
-
Strong experience with observability, monitoring, alerting, logging, dashboards, SLOs/SLIs, incident response, postmortems, and on-call practices.
-
Experience helping product or application engineering teams improve service ownership, operational readiness, and production accountability.
-
A pragmatic approach to tooling: you understand when to build, buy, adapt, simplify, or retire systems based on the actual engineering problem.
-
The ability to operate effectively in a small or mid-sized engineering organization where influence comes from credibility, judgment, and outcomes rather than bureaucracy.
-
Clear executive communication skills and the ability to partner directly with a CTO and senior engineering leaders.
Bonus Points if:
-
Experience leading SRE, infrastructure, or platform work in a lean, high-agency organization.
-
Experience supporting globally distributed teams or 24/7 operational coverage.
-
Experience improving developer productivity through paved paths, self-service infrastructure, automation, and reduced toil.
-
Experience with infrastructure security fundamentals, secrets management, access controls, cloud security practices, or compliance-related infrastructure controls.
-
Experience in financial services, regulated environments, blockchain, crypto, Web3, or other high-reliability technical ecosystems.
-
Experience evaluating vendors and infrastructure platforms with skepticism, technical rigor, and cost discipline.
-
Practical experience applying AI-assisted or agentic workflows to infrastructure, reliability, operations, observability, or developer productivity.
We offer competitive pay with a base salary range for this position of $205,000 - $305,000 depending on job-related knowledge, skills, experience, and location. In addition, we offer lumen-denominated grants along with the following perks and benefits:
USA Benefits/Perks:
-
Competitive health, dental & vision coverage with most plans covered at 100% for the employee + any dependents
-
Flexible time off + 15 company holidays including a company-wide holiday break
-
Generous paid parental leave for all parents, plus paid pregnancy disability leave for birthing parents
-
Gym reimbursement ($80 per month)
-
Life & ADD (up to $50K)
-
Short & Long term disability
-
401K with 4% match
-
Health & Dependent Care FSA Accounts
-
Commuter benefits with $250/month employer contribution
-
Health Savings Account (HSA) with monthly employer contribution
-
Family building benefits through Kindbody
-
Wellbeing benefits (One Medical, Rightway, Headspace)
-
L&D budget of $1,500/year
-
Daily lunch and snacks in office
-
Company retreats
#LI-Hybrid
About Stellar
Stellar is more than a blockchain. Powered by a decentralized, fast, scalable, and uniquely sustainable network made for financial products and services and a thriving and passionate ecosystem that includes a non-profit organization driven by a mission, Stellar is paving the path to unlock the world’s economic potential through blockchain technology. Built with speed and low costs in mind, the Stellar network provides builders and financial institutions worldwide a platform to issue assets, and to send and convert currencies in real time creating real world utility. Founded in 2014, the Stellar Development Foundation (SDF) supports the continued development and growth of the Stellar network and also serves the ecosystem of NGOs, corporations, universities, small businesses, governments, and solo entrepreneurs building on the Stellar network through tooling, funding and strategic collaborations. Together, Stellar is where blockchain meets the real world.
About the Stellar Development Foundation
The Stellar Development Foundation (SDF) is a non-profit organization focused on working with and supporting change-makers to create equitable access to the global financial system through blockchain technology. SDF provides grants, investments, funding, and other awards to builders and organizations. SDF also develops resources and tooling on the Stellar network to help unlock real world utility. As a nonprofit foundation, SDF puts the health of the Stellar network and the Stellar ecosystem and its mission above all else.
We look forward to hearing from you!
Privacy Policy
By submitting your application, you are agreeing to our use and processing of your data in accordance with our Privacy Policy.
SDF is committed to diversity in its workforce and is proud to be an equal opportunity employer. SDF does not make hiring or employment decisions on the basis of race, color, religion, creed, gender, national origin, age, disability, veteran status, marital status, pregnancy, sex, gender expression or identity, sexual orientation, citizenship, or any other basis protected by applicable local, state or federal law.
#J-18808-Ljbffr- ...The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also...SuggestedFull timeWork experience placementRemote work
$207k - $300k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Master's degree in Computer Science or Engineering.Experience mentoring engineers and cultivating... ...consensus across cross-functional teams.Site Reliability Engineering (SRE) combines...Suggested$171.6k - $223k
...development and learning. It allows us to scale easily, enabling our engineers to enhance attention on new features and capabilities. A key... ...of members all over the world. Peloton is looking for a Site Reliability Engineer to create the tooling and services which simplify...SuggestedTemporary workLocal area$100k - $250k
...systems) can be yours. What you’ll do Improve observability, reliability and availability by defining and measuring key metrics.... ...function improvements. Educate, mentor and hold accountable the engineering team to improve the reliability of our systems and make...SuggestedLocal area$100k - $135k
...Model-Based Manufacturing, where context-aware production planning informs design in real-time—eliminating the disconnect between engineering and manufacturing. Our platform automates what can be automated and captures tribal knowledge where automation falls short,...SuggestedWork at officeFlexible hours- ...Improve the reliability of mission critical solutions, applications, and platforms Software development for enterprises Continuous... ...Shell Languages Powershell and Bash Windows and Linux Years of Experience: 5 Years of Software Engineering #J-18808-LjbffrWork experience placement
$104k - $178k
## Sr. Site Reliability Engineer IApply: Hybrid: NYC Global HQ: Full time: Posted 12 Days Ago: JR00000779# ****Who We Are****DV is the leader in digital performance solutions, helping our advertiser and agency partners Verify the quality of their digital campaigns, Optimise...Full time$208.5k - $216.5k
...pipelines - including its signal quality and its cost.Drive reliability improvements using SLOs and telemetry data, closing observability... ...review. Typically 8+ years in SRE, DevOps or infrastructure engineering, though scope and impact weigh more than tenure.Technical...Full timeTemporary workFlexible hours- ...firm that has been operating at scale since 2014, with millions of global users and a reputation for rigorous engineering. The Role As Senior Site Reliability Engineer , you will own the infrastructure foundation that the entire engineering organization depends on....
$120k - $150k
...allows each person to achieve personal success and add value to our teams and communities. We are currently looking for a Site Reliability Engineer to join our Platform Engineering team in New York, NY. About The Role Join our Platform Engineering team as a Site...- ...DEPARTMENT: Product Engineering / Operational Readiness REPORTING TO: Senior Manager, System... ...Engineering team is responsible for the reliability, monitoring, automation, and operational... ...services. Role Overview: The Site Reliability Engineer II (SRE) is responsible...Full timeTemporary workWork at officeRemote workFlexible hours
- ...Ireland Full time Are you passionate about building reliable, scalable systems that power critical business solutions?... ...can learn more about LexisNexis Risk at the Role As a Site Reliability Engineer (SRE), you will bridge software development and IT operations...Full timeWork from home
$150k - $200k
...Defence and Government. They’re continuing their expansion of their New York (and Washington DC) engineering teams and looking for an Infrastructure Engineer / Site Reliability Engineer / Forward Deployed Infrastructure Engineer with a strong software engineering...- ...Job Description Chariot’s engineering hire will be responsible for taking the Chariot platform to the next level. You will lead the evolution of our banking product, DAF product and technology strategy, along with building a growing team of software engineers. You will...Work experience placementWork at officeWork from homeMonday to FridayMonday to Thursday
$200k - $240k
...problems and help health systems deliver better care, we'd love to meet you! About the role We're looking for a Senior Site Reliability Engineer to join our Infrastructure Engineering team and get their hands directly into the systems that keep our healthcare...Work at office3 days per week- ...Responsibilities Support and enhance the reliability, availability, and performance of... ...service improvements. Work closely with engineering teams to implement, test, and deploy... ...production environments, platform engineering, site reliability, DevOps, or infrastructure...
$150k - $170k
...Senior Site Reliability Engineer – Zip CoJoin to apply for the Senior Site Reliability Engineer role at Zip CoAt Zip, we build cloud-native software applications that serve millions of customers and process billions of dollars in payments. We're looking for a seasoned...Casual workWork at officeRemote work- ...our clients to succeed in an evolving digital landscape. Role Overview We are seeking an experienced Observability / Site Reliability Engineer (SRE) to design, scale, and maintain our enterprise monitoring and alerting ecosystems. In this role, you will bridge the...
$120k - $160k
...and benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more. As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and...Work at officeLocal area- ...Traders, Tower Research, PDT Partners, SIG, and more. We're looking for a midlevel or senior IC to join our Backend Engineering team as a Site Reliability Engineer. You'll own the uptime, performance, and observability of our platform, and help set the standard for how...Full timeLocal areaRemote work
- ...their own infrastructure, behind their own controls, with the reliability and operational clarity they would expect from any critical system... ..., Support, and TAMs to trust. Partner with product engineers on infrastructure requirements for new Retool products, especially...
- ...you’ll be building the future of financial infrastructure. As part of our global expansion, we’re looking for a hands‑on Site Reliability Engineer (SRE) to design, scale, and safeguard the reliability of our next‑generation financial platforms. This is a high‑impact...Remote work
- ...Plaid Inc is looking for a Staff Site Reliability Engineer to lead the reliability practices across product engineering. You will architect SLO and error-budget programs, ensuring new products are production-ready while promoting safety gates for high velocity. The...
$191k - $226k
...— and using AI to scale that impact further and faster than anyone else can. About the role: We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of the cloud infrastructure powering Garner's products and AI/ML workloads...Remote workWork visaFlexible hours- ...Job Description Job Description Location: New York, NY, USA Exp: 8-12 Years Client: Amex Job Description: SRE Engineer (This is not a Devops role, strictly need an SRE Engineer, who has great analytical skills and is a good incident manager as well)...
- ...Site Reliability Engineer Our Client, a multinational telecommunications technology company is seeking a Site Reliability Engineer (SRE I) to join our Video Platform Engineering Team. As a Level 1 SRE, you will work closely with senior engineers to respond to incidents...Temporary work
$120k - $180k
...people, and works with high-profile manufacturers including leaders in space and defense. You will be the first dedicated Site Reliability Engineer and own critical infrastructure end to end. This is a greenfield opportunity to architect the path from AWS to on-premises...Permanent employmentFull timeRelocation package- ...right care, at the right time, in the right setting. Role Description You'll join Tennr's Infrastructure team as a Site Reliability Engineer, focused on the systems that keep us reliable, observable, and secure. This is a hands-on role with real room to grow: you...Work at office
$182.8k - $247.3k
...to develop education for our half a billion (and growing!) learners around the world. About the role... As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure the company’s sophisticated distributed...Work experience placement- ...visible impact and put something genuinely rare on your CV, keep reading . About The Role We're looking for a Senior Site Reliability Engineer who's passionate about building reliable, scalable infrastructure that helps developers ship better software faster. You'...Remote workWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Director, Site Reliability Engineering. Be the first to apply!
- principal developer New York, NY
- principal security engineer New York, NY
- senior director engineering New York, NY
- director of product engineering New York, NY
- director sales engineering New York, NY
- data center chief engineer New York, NY
- engineering director New York, NY
- principal network engineer New York, NY
- hotel chief engineer New York, NY
- principal cloud engineer New York, NY



