Senior Manager, Site Reliability Engineering (SRE)
Nium
Nium provides global infrastructure for real-time cross-border payments. We were founded on the mission to deliver the global payments infrastructure of tomorrow, today. Our platform enables banks, fintechs, and global businesses to move money instantly, everywhere.
Co-headquartered in San Francisco and Singapore with offices in 14 markets worldwide, we are entering one of the most exciting chapters in our journey. In March 2026, we delivered the largest month in our 11-year history with record revenue, record volumes, and EBITDA profitability. Today, Nium moves nearly $60B in payments annually, almost entirely for enterprises, while continuing to strengthen an already healthy balance sheet.
It is an incredible time to join us, and we are only just getting started.
Our payout network spans 190+ countries and 100 currencies, with 100 + corridors in real time. We power seamless transfers to accounts, wallets, and cards, support local collections in 35 markets, and as a principal card issuer on Visa, Mastercard, Discover, and UATP, Nium issues over 50 million card tokens every year. Backed by regulatory licenses in 40+ markets, we make it simple for our partners to onboard, integrate, and scale globally. This scale and innovation have earned us recognition as one of CNBC’s World’s Top Fintech Companies 2025, winner of Best Cross-Border Payments Solution at the PayTech Awards, and inclusion in FXC Intelligence’s Top 100 Cross-Border Payments Companies list.
In 2024, we raised US$50 million in Series E funding at a US$1.4 billion valuation to accelerate network expansion, product innovation, and talent growth. With the B2B payments market projected to hit US$175 trillion by 2030, Nium offers ambitious builders the chance to shape the future of global money movement with the scale of a leader and the energy of a high-growth company.
The RoleNium is looking for a Senior Manager, Site Reliability Engineering to lead the teams responsible for the availability, performance, scalability, and operational excellence of our global payments platform. This is a hands-on leadership role: you will build and grow a team of SREs, define the reliability roadmap, and partner closely with product engineering, security, and infrastructure teams to ensure Nium's systems meet the always-on expectations of a regulated financial platform operating across 100+ markets.
You will own incident management, observability, capacity planning, and production readiness practices, while championing a culture of blameless postmortems, proactive risk reduction, and engineering-driven automation. This role sits at the intersection of engineering leadership and operational rigor, and reports into Nium's engineering leadership.
Responsibilities- Lead, mentor, and grow a team of SREs and reliability engineers across multiple time zones, setting clear goals, career paths, and performance expectations.
- Own Nium's reliability strategy — defining and driving SLIs/SLOs/error budgets across critical payment, card issuance, and compliance services.
- Drive incident management end-to-end: on-call structure, escalation paths, major incident response, and blameless postmortems that produce durable fixes, not just tickets.
- Partner with product engineering leaders to embed reliability, scalability, and operational readiness into the software development lifecycle from design through launch.
- Build and scale observability (metrics, logging, tracing, alerting) so that issues are detected and diagnosed before they impact customers or partner banks.
- Lead capacity planning and performance engineering for systems processing high-volume, real-time financial transactions across a global, multi-region infrastructure.
- Champion automation and self-healing systems to reduce toil, eliminate manual runbooks, and improve mean-time-to-detect and mean-time-to-resolve.
- Own disaster recovery, business continuity, and chaos engineering practices, running regular game days to validate resilience assumptions.
- Collaborate with Security and Compliance teams to ensure infrastructure practices meet regulatory and audit requirements (PCI-DSS, SOC 2, ISO 27001, and regional financial regulations).
- Manage the reliability budget: tooling investments, cloud cost/performance trade-offs, and staffing plans, in partnership with finance and engineering leadership.
- Represent SRE in executive reviews, translating technical risk and system health into business-relevant reporting for leadership and the board.
- Establish and continuously refine production readiness reviews, runbooks, and operational standards across all engineering teams.
- 10+ years of experience in software engineering, infrastructure, or site reliability engineering, with 4+ years in a people-management or technical leadership role leading SRE/DevOps/Infrastructure teams.
- Proven track record operating and scaling production systems for a high-availability, transaction-heavy platform — fintech, payments, banking, or e-commerce experience strongly preferred.
- Deep hands-on expertise with cloud infrastructure (AWS) Kubernetes, container orchestration, and infrastructure-as-code (Terraform, CloudFormation, or similar).
- Strong background in observability stacks (Prometheus, Grafana, Datadog, ELK/OpenSearch, or equivalent) and building alerting that reduces noise while catching real issues.
- Demonstrated experience defining and operationalizing SLOs/error budgets, and using them to drive engineering prioritization.
- Solid understanding of distributed systems, databases, caching, messaging queues, and API-driven microservice architectures at scale.
- Experience leading major incident response for critical, customer-facing systems, including postmortem processes that drive real change.
- Familiarity with security and compliance frameworks relevant to financial services (PCI-DSS, SOC 2, ISO 27001) and how they shape infrastructure and access practices.
- Excellent communication skills — able to translate technical reliability concepts for engineering peers, product leaders, and executives alike.
- A pragmatic, metrics-driven approach to engineering decisions: you measure before you optimize, and you test assumptions in production rather than in theory.
- Experience with scripting/automation (Python, Go, or Bash) and CI/CD pipelines (Jenkins, ArgoCD, GitLab CI, or similar).
- Experience operating systems subject to real-time payment rails, card networks, or banking core integrations.
- Prior experience running SRE or infrastructure functions through hypergrowth or rapid international expansion.
- Exposure to FinOps practices and cloud cost optimization at scale.
- Experience building or scaling an SRE function from the ground up within an existing engineering organization.
$185k - $200k
Back to All JobsSenior Site Reliability Engineer (SRE) Dayton, OH (Remote) full time Top Secret (TS)... ...Position Overview Metronome is seeking a Senior Site Reliability Engineer (SRE) to... ...-trust principles. Deploy and manage cloud infrastructure using Terraform...SeniorFull timeRemote work- ...About The Role: We're looking for a Senior Site Reliability Engineer to help us mature and scale the... ...infrastructure fully under Terraform-managed IaC, replacing manual and ad-hoc provisioning... ...You Bring: ~6+ years in SRE, DevOps, or infrastructure...SeniorRemote workFlexible hours
$149.4k - $202k
...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic... ...of cloud native systems. Our SREs don’t just manage infrastructure; they build it using Infrastructure...SeniorRemote work$140k - $180k
...more at Opportunity We’re looking for a Senior Site Reliability Engineer to help build and scale a high-impact SRE function. You’ll be a technical leader on a... ...Deep understanding of observability, incident management, and system performance ~ Proficiency in at...SeniorWork experience placementLocal areaRemote workVisa sponsorshipWork visa- ...Senior Site Reliability Engineer At Swile, we believe that good products can help reduce friction in daily professional life and boost employee satisfaction... .... Your role as a Senior Site Reliability Engineer (SRE) centers around creatively solving problems, ensuring a...SeniorRemote work
$139.7k - $232.9k
...Wilmington Center Wilmington, DE, location with the flexibility to work from home one day per week Overview The Site Reliability Engineering (SRE) Manager leads teams responsible for the reliability, availability, performance, and operational excellence of critical...Full timeWork from home1 day per week- ...Remote - United States Job Category Information Technology, Platform Engineering, Site Reliability Engineering Industry Computer Software, SaaS, National Security Employee Type FT Exempt Manage Others No Minimum Experience 5 Years Required Security...SeniorRemote work
$165k - $225k
...Sr. Site Reliability Engineer (SRE) Chicago, IL or Remote Moonlite delivers high-performance AI infrastructure for organizations running intensive... ...grade clusters from the ground up (not just deploying in managed environments). You'll ensure enterprise-grade reliability...SeniorRemote workFlexible hours$120k - $175k
...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's...SeniorFull timeRemote workWork visaFlexible hours$175k - $215k
...Sr. Manager, Site Reliability Engineer At Disney, we're storytellers. We make the impossible possible. The... ...Provide strategic leadership for multiple SRE teams, fostering a culture of... ...skills, with experience influencing senior stakeholders and driving cross-functional...SeniorLocal area- ...Sr Application Performance and Observability Engineer At Sequoia Connect, we are a Talent-First Technology Ecosystem that redefines... ...JDBC connection pools, application thread pools, and session management. Experience tracing transactions across legacy or partially...SeniorRemote workWorldwide
$500 per month
...researchers, designers, and product managers to quickly recruit... ...Spain or Portugal. Key engineering and product teams are based... ...Role We’re hiring a Senior Site Reliability Engineer to join our Platform... ...make an impact. As an SRE at Maze, you will:...SeniorRemote jobFull timeFlexible hours- ...Job Title: Site Reliability Engineer (Azure Government & Infrastructure) Pay Type : SALARIED EXEMPT... ...The Site Reliability Engineer (SRE) for Azure Government & Infrastructure... ...Government networking architecture, including management of NSGs, Azure Firewall, VNETs, and...Full timeRemote workMonday to Friday
$100k - $180k
...Site Reliability Engineer (SRE) - Remote Bright Vision Technologies is a technology consulting and software development company delivering... ...posture in collaboration with security teams, including patch management, vulnerability remediation, and secure-by-default...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...Site Reliability Engineer (SRE) Immediate need for a talented Site Reliability Engineer (SRE). This is a 12+ months contract opportunity with... ...resolution status (written and verbal) to project team and management ~ Provide reactive, break-fix support Our client...Contract workLocal areaImmediate start
- ...nervous system for a borderless global economy. As a Site Reliability Engineer (SRE) at Unlimit, you will help ensure the reliability, scalability... ...and supporting services. Own and improve incident management, including troubleshooting, escalation handling, and...Full timeLocal area
$80k - $95k
...applications, and web-based product offerings. In this role, the Site Reliability Engineer (SRE) will play a key role in maintaining resources at peak... ...** * Monitor systems for uptime, create and manage alerting * Triaging issues through log analysis and basic...Remote workVisa sponsorshipWork visa- ...We're seeking an SRE to ensure the reliability and performance of our clients' critical systems. You'... ...monitoring and observability solutions Manage incident response and post-mortems... ...to have Experience with chaos engineering Knowledge of distributed systems...Remote workFlexible hours
- ...Science, Information Technology, Engineering, or equivalent field ~3-5 years of experience in Site Reliability Engineering, Production... ...Demonstrated experience with incident management, production support, root... ...~ Understanding of SRE principles, including observability...Remote work
- ...cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-... ...workflows for accuracy and reliability. Work with AWS, Azure,... ...Cloud Infrastructure Site Reliability Engineering (SRE... ...developer workflow and project management platforms such as GitHub,...For contractorsRemote work
- ...Site Reliability Engineer (SRE) Location: Remote (Secaucus, NJ) Duration: Contract Experience: 7+ Years Job Description 4+ years of experience... ...Strong communication skills (written/verbal) Time management Analytic problem solver Self-starter Result...Contract workWork experience placementRemote work
- ...Site Reliability Engineer (SRE) Location: Remote Shift Timings: 5:30 PM to 3:00 AM IST to ensure support for global operations. Job Description... ...or something similar. Experience with configuration management tools like Ansible, Puppet, or Chef. Familiarity with...Remote workShift work
- ...Description At Rocket.net, reliability, performance, and... ...We are looking for a Site Reliability Engineer to help maintain the... ...issues. Act as a senior escalation point for... ...of experience in SRE, DevOps, Platform Engineering... ...supporting managed WordPress hosting platforms...Remote workFlexible hoursShift work
- ...Site Reliability Engineer (SRE) We are seeking an experienced Site Reliability Engineer (SRE) with strong expertise in Dynatrace to join our... ...with deep experience in monitoring, automation, incident management, and cloud-native technologies. Key Responsibilities...Remote work
$104k - $166k
Responsibilities Peraton is looking for a Senior Cloud Site Reliability Engineer (SRE) who will be responsible for designing and developing advanced... ...Infrastructure as Code (IaC) using Terraform to manage AWS resources related to SRE solutions, incorporating...SeniorContract workWork experience placementRemote workShift work$106.7k - $177.9k
...organization and help drive the reliability, resiliency, and performance of... ...depend on every day. As a Senior Software Engineer, Site Reliability Engineering (SRE), you will play a key role in supporting... ...efforts.• Utilize source code management tools to manage and deploy...SeniorPermanent employmentFull timeWork experience placement- ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes... ..., Datadog, OpenSearch, PagerDuty Requirements: Senior SRE with 5+ years AWS infra, 3+ years in senior/lead roles;...SeniorFull timeRemote work
- ...The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also...SeniorFull timeWork experience placementRemote work
- As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing AI... ..., Platform Engineering, Site Reliability Engineering (SRE), Systems Administration, or related technical roles.Experience...SeniorWork at officeImmediate startWorldwideShift work
- Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and... ...channels. This role applies Google-inspired SRE principles to balance feature velocity... ...reliability initiatives, and manage operational risk.Design and maintain observability...SeniorLocal areaRemote workFlexible hoursShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Manager, Site Reliability Engineering (SRE). Be the first to apply!
- site reliability engineer Remote
- site reliability engineer remote Remote
- site reliability engineer sre Remote
- senior living director Remote
- senior manager customer operations Remote
- senior support engineer Remote
- senior product manager mobile Remote
- senior c++ developer Remote
- senior java developer Remote
- senior development engineer Remote



