Manager, Site Reliability Engineering
$150k - $170kLinuxconfig
DriveWealth is on a mission to make investing easier. We believe that everyone should have the ability to control their financial future, and that access to financial markets should not be limited by geography, wealth, or legacy systems. We are a global B2B financial technology organization dedicated to democratizing access to financial independence around the world. Our mission is realized through an API‑based platform, empowering our partners to offer seamless investing and trading experiences to clients worldwide, all from their mobile devices. Our technology provides partners with a modern, extensible toolkit, enabling traditional investment workflows and innovative techniques like fractional share ownership. DriveWealth has evolved into a global platform offering trading of US equities, mutual funds, ETFs, fixed income, and options.
There’s never been a better time to build a category‑defining business and there has rarely been a team better positioned for this opportunity. Our culture blends the pace and agility of a fintech start‑up with the impact, stability, and discipline of Wall Street. We encourage creativity and experimentation while ensuring institutional‑grade execution and regulatory compliance in everything we do. Join us and help build the future of global investing!
About The Role
As the Manager of Site Reliability Engineering, you'll lead a team of SRE Automation Engineers while remaining a hands‑on technical authority for our Brokerage‑as‑a‑Service platform. This isn't a purely people‑management seat, you're expected to bring the same principal‑level SRE depth to automation design and engineering as an individual contributor, while also building the team, setting technical direction, and developing your engineers' careers.
This role is centered on reducing manual toil through engineering, applying Google's SRE principles: SLOs, error budgets, blameless postmortems, and systematic toil reduction, adapted to a regulated brokerage environment. You'll carry two responsibilities at once: driving the automation agenda, building and orchestrating workflows in Rundeck and Airflow to eliminate repetitive work, and growing your team of SRE Automation Engineers into a high‑functioning automation practice. You'll guide the design of internal SRE platforms, automate complex workflows, and ensure our Kubernetes‑based and colo ecosystems can handle the demands of global financial markets, while owning the people side of the team: mentorship, performance, and growth, and the day‑to‑day management of the team's Jira board. While this role includes participation in on‑call rotations supporting our 24/7 global operations, your primary mission is to build systems that make manual intervention obsolete, and a team capable of sustaining that mission.
What You'll Do
- Team Leadership & Development : Manage, mentor, and grow a team of SRE Automation Engineers—setting technical direction, running 1:1s, owning performance management and career development, and managing the team's Jira board to prioritize and track sprint work.
- Engineering & Automation : Lead the design and development of internal tooling and automation— including Rundeck and Airflow-based orchestration—to eliminate repetitive manual toil and improve developer velocity, staying hands‑on with the most complex, highest‑leverage automation work yourself.
- SRE Practice & Governance : Adapt Google's SRE principles to our environment—defining SLIs, SLOs, and error budgets, and using them to guide engineering and operational priorities.
- Infrastructure as Code : Set architectural standards for modular, reusable IaC using Terraform and oversee GitOps workflows via ArgoCD.
- Platform Governance : Review software architecture and Kubernetes metrics to ensure high availability, capacity planning, and cost‑optimization across AWS regions, and hold the team accountable to those standards.
- Incident Engineering : Lead incident response for critical events, drive complex root‑cause analysis (RCA), and champion a blameless post‑mortem culture across the organization.
- Collaboration & Stakeholder Management : Partner with engineering leadership to align SRE priorities with business goals, and foster adoption of new tools, security standards, and reliability best practices across teams.
You Bring
- People Leadership : Prior experience managing or leading SRE/DevOps engineers, ideally in a fintech or highly regulated environment. Able to flex between hands‑on principal‑level engineering and coaching and developing a team.
- Google SRE Fundamentals : Working knowledge of Google's SRE practices—SLIs/SLOs, error budgets, toil reduction, and blameless postmortems—and experience adapting them to a regulated environment.
- Linux & Networking Mastery : Proficient in Linux administration with a deep understanding of the TCP/IP stack, OSI model, DNS, and network troubleshooting.
- FinTech Background : Experience working in highly regulated financial environments or with FIX/API connectivity.
- Production Kubernetes : Hands‑on experience managing production‑grade clusters, including RBAC, autoscaling, Helm, and multi‑cluster patterns.
- Cloud Native Expertise (AWS) : Strong grasp of AWS core services, security, and high‑availability patterns. Proficiency with boto3 and AWS CLI for automation.
- Modern CI/CD & GitOps : Experience building secure, automated delivery pipelines and operating GitOps workflows (ArgoCD).
- Code Proficiency : Strong scripting and development skills in Python or Golang, along with Bash and Ansible.
- Observability : Experience with Grafana/Similar tools, Prometheus, Understanding of logs shipping, management and metric first alerting.
- Security Mindset : Experience with secrets management, vulnerability scanning, and securing the software supply chain.
- AI & Prompt Engineering : Familiarity with using LLMs, Public MCPs, or Bedrock Agent Core to enhance SRE workflows.
- Data & Middleware & Orchestration : Hands‑on experience with Rundeck and Airflow for job orchestration and automation, plus experience managing Kafka, MQ, or SQS.
Location
This role is open to candidates in the following locations: Chicago, IL - Hybrid
- This role is expected to come into the office on a cadence set by the Hiring Manager/Team.
- If you're not based in the location listed above, this role is not a fit, and we cannot accommodate remote work outside these locations.
- Applicants must be authorized to work for any employer in the U.S. DriveWealth does not sponsor or take over sponsorship of an employment visa at this time.
Pay Range : $150,000 – $170,000 USD
Working at DriveWealth
We do our best work when we're in the same room. To maintain the speed our partners expect, our New York, Chicago, and Lithuania teams work in office on a hybrid schedule. We've found that being physically side‑by‑side is the only way to solve complex problems in real‑time and stay truly accountable to the products we ship. When you're here, you're working directly with the people making the decisions.
To support that work, we provide competitive compensation, equity, and a 401(k) match. We also offer Medical insurance, Dental insurance, Vision insurance, Disability insurance, and Paid Parental Leave, along with a wellness reimbursement, a company‑provided phone, and a personal development allowance. Finally, we value the time you spend away from the office with generous Paid Time Off (PTO) and observed holidays.
Work Authorization
Applicants must possess the legal right to work in the country where the position is located at the time of application. DriveWealth requires all employees to provide original documentation verifying their work authorization on or before their first day of employment.
For U.S‑based roles: Applicants must be currently authorized to work in the United States on a full‑time basis without the need for current or future visa sponsorship. DriveWealth does not provide visa sponsorship or support for employment authorization, including transfers, at this time. Offers of employment are strictly contingent upon an individual’s ability to secure and maintain the legal right to work at the Company.
How We Think About AI
We leverage AI to work smarter and move faster. We seek AI‑curious talent who are proactive about using emerging tools to increase signal quality, reduce friction, and improve outcomes to deliver products faster, provide better service to our partners, and to streamline processes. Your ability to leverage our internal tools and technology to drive results is as important to us as your core domain expertise.
Compensation
Pay is generally based on the level, complexity, responsibility, location, and job duties/requirements of the specific position. We then source candidates with the requisite skills, expertise, training, and experience. If you are selected for an interview, please feel welcome to speak to a recruiter about our compensation philosophy and other available benefits. This role is eligible for base, bonus, equity, 401(k) match, and heavily subsidized benefits and perks.
To build technology and products that are used and loved by people and solve real‑world problems, we need to build a team with many different perspectives and experiences. We are an equal opportunity employer. We do not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. We encourage candidates from all backgrounds to apply.
#J-18808-Ljbffr$180k - $230k
...re looking for a Senior SRE to own the reliability, scalability, and observability of our... ...ll work closely with platform and data engineering to keep high-throughput, data-intensive... ...toil — deployment pipelines, capacity management, self-healing systems Partner with engineering...SuggestedWork at officeLocal areaImmediate startRemote work3 days per week- ...Engineering, Product, Design, and Marketing Engineering Compensation ~ Zone 1 Base... ...collaborative workspaces, Mail’s inbox management, and Go, the proactive AI assistant that... ...for building software to ensure the reliability of our back-end systems, working with engineers...SuggestedWorldwideHome officeFlexible hours
$150.4k - $277.6k
...Platforms SRE team under the Apple Service Engineering division is one of the most exciting... ...of services that ingest, transform and manage all media content across products like... ...years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure...SuggestedRelocationDay shift$182.8k - $247.3k
...to develop education for our half a billion (and growing!) learners around the world. About the role... As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed systems...SuggestedWork experience placement- ...A senior Site Reliability Engineer will join an established infrastructure function responsible for highly available, security-conscious cloud... ...core component of infrastructure provisioning and lifecycle management. The position blends hands-on production engineering with...SuggestedFull timeRemote work
$114k - $148k
...using objective, job-related criteria. Summary As a Site Reliability Engineer, you will focus on ensuring the platform and services customers... ...members as needed. You will interact with internal staff, managers, and customers to implement and maintain operations. A...Work experience placement- ...Quarterhill is seeking a Senior Site Reliability Engineer (SRE) to join our growing team. This role is an exciting opportunity to contribute... ...reliability metrics for revenue-critical services. Incident Management: Lead incident response — perform root cause analysis,...Local area
$107.9k - $195.05k
...The Digital Sector at Leidos currently has an opening for a Site Reliability Engineer (SRE) / Senior Cloud Engineer to work in our Baltimore,... ...environments. Partner with the DevOps Lead/Configuration Manager to build and maintain CI/CD pipelines, ensuring automated testing...Contract workWork at office- ...SLOs, incident response, and production reliability for a system that processes millions of... ...structured logging Implement chaos engineering practices to proactively identify failure... ...team members including your potential manager and cross-functional partners. We...Remote work
$148.5k - $223.9k
...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the... ...toil and improve operational efficiency. Incident Management : Lead the coordinated response to incidents as an Incident...WorldwideWeekend work- ...As a Senior Site Reliability Engineer on our cloud engineering team, you'll keep our production environment healthy, secure, and running smoothly. This is an operations-focused role: you'll own the day-to-day administration of our AWS accounts and databases, backup posture...Work experience placement
$110k - $145k
...will liaise with product and engineering teams to ensure applications... ...for platform and product reliability. The ideal candidate is a... ...Collaboration and Stakeholder Management: Collaborate closely with cross... ...as a platform engineer, site reliability engineer, systems...Work experience placement- ...and connect with 28,396 DevOps professionals. The Senior Site Reliability Engineer role at Jobicy is designed for experienced professionals... ...development teams to improve software development practices, and managing incident response to maintain system uptime. Candidates...Remote workFlexible hours
- ...NestJS, Datadog, TypeScript, React, SQL Position: Senior Site Reliability Engineer Engagement period: Ongoing Interview timeline: ASAP... ...limiting CI/CD pipelines, GitHub Actions Secrets management, least-privilege IAM, patching discipline Data & Runtime...Contract workImmediate start
$115.5k - $164.8k
...where you matter. Your Impact As an engineer on the APX SRE CloudOps team, you will... ...previously required human intervention with reliable, tested automation. You will also... ...of applicable experience. ~ Experience managing cloud platforms such as Azure, AWS, or similar...Work experience placementWork at officeRemote work$150k - $220k
...Senior Site Reliability EngineerJob detailsDepartment / EngineeringRemoteFull-time$150,000 USD... ...About The RoleAs a Senior Site Reliability Engineer, you own significant pieces of our... ...speed of getting changes to production.* Manage upgrades across infrastructure and the...Full timeRemote work$140k - $195k
...Improve reliability, observability, service health, incident response, and operational... ...and operational clarity matter. The Site Reliability Engineer role helps turn business needs into... ...Kubernetes, Observability, Incident Management, AWS. ~ Ability to work remotely...Remote work$110k - $145k
...across the U.S., Canada, and India. We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and... ...scalability, performance, and operational health. Define and manage SLIs and SLOs, using error budgets to guidedelivery...Flexible hours- ...arc of the patient journey. The Opportunity: Machine Learning Engineer Patients count on our platform 24/7. You'll build and... ...of microservices, databases and real-time model endpoints. Manage backups, disaster-recovery plans, secrets and certificate lifecycles...
- ...tooling company whose product is used by engineering teams at thousands of software... ...for application monitoring and incident management. Their infrastructure runs across three... ...of nine. As Senior SRE you will lead reliability initiatives across the platform — from...
- ...Zof AI is seeking a Site Reliability Engineer to run the infrastructure that lets fleets of sandboxed agents execute customer code safely and... ...Automate provisioning, deployment, rollback, and environment management. Partner with engineers to make infrastructure fast and...Full time
- ...in revenue and is the leading specialized storage cloud – managing over three billion gigabytes of data storage for 500K+ customers... ..., and individuals. About the Role We are seeking a Site Reliability Engineer II (SRE II) to help ensure the stability, scalability, and...
$185 per hour
...available, secure, and easy to operate. We are hiring a Senior Site Reliability Engineer to join our Security and Site Reliability team. You will... ...standards, and drift control. ~ Improve environment management, cluster practices, and deployment reliability. ~ Make...Work at office$180k - $200k
...Come join tastytrade, part of IG Group, as we build the reliability practice behind the brokerage platform that active options... ...equities traders rely on every market day. As our first Senior Site Reliability Engineer, you'll define what reliability means at tastytrade, from...Work at office3 days per week$104k - $178k
## Sr. Site Reliability Engineer IApply: Hybrid: NYC Global HQ: Full time: Posted 12 Days Ago: JR00000779# ****Who We Are****DV is the leader... ...environments.* Respond to incidents and drive them to resolution, managing Sev1/Sev2 situations.* Reduce MTTR (mean time to...Full time$135.2k - $181.2k
...mechanical, and sensor-based systems to ensure reliability and performance. Configure,... ...to achieve optimal system operation. Manage and administer Linux and Windows environments... ...including an interest in emerging data engineering tools and methodologies. Preferred...Worldwide- ...match. The role We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi-... ...using AI-assisted development workflows Partner closely with engineering on reliability reviews and architecture decisions ~5-8...
- ...join us on our mission of providing humankind access to the galaxy beyond our planet. About the Role We are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site...Full timeWork at office
- ...services for the legal industry — making eDiscovery, case management, and litigation prep simple, fluid, and affordable for law... ...tolerance for firms in active litigation. We're looking for a Site Reliability Engineer to help maintain the reliability, scalability, and...
$95k - $171k
...infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's... ...inference platform. As an Site Reliability Engineer II, you will be responsible for:... ...integrating into Akamai's existing incident management processes Contributing to SLO tracking...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
- site reliability engineer Eastern, KY
- site safety Eastern, KY
- on-site clinical research associate (traveling/remote) Eastern, KY
- construction site safety Eastern, KY
- junior website developer Eastern, KY
- historic site Eastern, KY
- IT site lead Eastern, KY
- site leader Eastern, KY
- official site Eastern, KY
- junior site reliability engineer

