Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Manager, Site Reliability Engineering

$150k - $170k

DriveWealth

Manager, Site Reliability Engineering

Office - Chicago

DriveWealth is on a mission to make investing easier. We believe that everyone should have the ability to control their financial future, and that access to financial markets should not be limited by geography, wealth, or legacy systems. We are a global B2B financial technology organization dedicated to democratizing access to financial independence around the world. Our mission is realized through an API-based platform, empowering our partners to offer seamless investing and trading experiences to clients worldwide, all from their mobile devices. Our technology provides partners with a modern, extensible toolkit, enabling traditional investment workflows and innovative techniques like fractional share ownership. DriveWealth has evolved into a global platform offering trading of US equities, mutual funds, ETFs, fixed income, and options.

There's never been a better time to build a category-defining business and there has rarely been a team better positioned for this opportunity. Our culture blends the pace and agility of a fintech start-up with the impact, stability, and discipline of Wall Street. We encourage creativity and experimentation while ensuring institutional-grade execution and regulatory compliance in everything we do. Join us and help build the future of global investing!

About The Role

As the Manager of Site Reliability Engineering, you'll lead a team of SRE Automation Engineers while remaining a hands-on technical authority for our Brokerage-as-a-Service platform. This isn't a purely people-management seat, you're expected to bring the same principal-level SRE depth to automation design and engineering as an individual contributor, while also building the team, setting technical direction, and developing your engineers' careers.

This role is centered on reducing manual toil through engineering, applying Google's SRE principles: SLOs, error budgets, blameless postmortems, and systematic toil reduction, adapted to a regulated brokerage environment. You'll carry two responsibilities at once: driving the automation agenda, building and orchestrating workflows in Rundeck and Airflow to eliminate repetitive work, and growing your team of SRE Automation Engineers into a high-functioning automation practice. You'll guide the design of internal SRE platforms, automate complex workflows, and ensure our Kubernetes-based and colo ecosystems can handle the demands of global financial markets, while owning the people side of the team: mentorship, performance, and growth, and the day-to-day management of the team's Jira board. While this role includes participation in on-call rotations supporting our 24/7 global operations, your primary mission is to build systems that make manual intervention obsolete, and a team capable of sustaining that mission.

What You'll Do
  • Team Leadership & Development: Manage, mentor, and grow a team of SRE Automation Engineers—setting technical direction, running 1:1s, owning performance management and career development, and managing the team's Jira board to prioritize and track sprint work.
  • Engineering & Automation: Lead the design and development of internal tooling and automation—including Rundeck and Airflow-based orchestration—to eliminate repetitive manual toil and improve developer velocity, staying hands-on with the most complex, highest-leverage automation work yourself.
  • SRE Practice & Governance: Adapt Google's SRE principles to our environment—defining SLIs, SLOs, and error budgets, and using them to guide engineering and operational priorities.
  • Infrastructure as Code: Set architectural standards for modular, reusable IaC using Terraform and oversee GitOps workflows via ArgoCD.
  • Platform Governance: Review software architecture and Kubernetes metrics to ensure high availability, capacity planning, and cost-optimization across AWS regions, and hold the team accountable to those standards.
  • Incident Engineering: Lead incident response for critical events, drive complex root-cause analysis (RCA), and champion a blameless post-mortem culture across the organization.
  • Collaboration & Stakeholder Management: Partner with engineering leadership to align SRE priorities with business goals, and foster adoption of new tools, security standards, and reliability best practices across teams.
You Bring
  • People Leadership: Prior experience managing or leading SRE/DevOps engineers, ideally in a fintech or highly regulated environment. Able to flex between hands-on principal-level engineering and coaching and developing a team.
  • Google SRE Fundamentals: Working knowledge of Google's SRE practices—SLIs/SLOs, error budgets, toil reduction, and blameless postmortems—and experience adapting them to a regulated environment.
  • Linux & Networking Mastery: Proficient in Linux administration with a deep understanding of the TCP/IP stack, OSI model, DNS, and network troubleshooting.
  • FinTech Background: Experience working in highly regulated financial environments or with FIX/API connectivity.
  • Production Kubernetes: Hands-on experience managing production-grade clusters, including RBAC, autoscaling, Helm, and multi-cluster patterns.
  • Cloud Native Expertise (AWS): Strong grasp of AWS core services, security, and high-availability patterns. Proficiency with boto3 and AWS CLI for automation.
  • Modern CI/CD & GitOps: Experience building secure, automated delivery pipelines and operating GitOps workflows (ArgoCD).
  • Code Proficiency: Strong scripting and development skills in Python or Golang, along with Bash and Ansible.
  • Observability: Experience with Grafana/Similar tools, Prometheus, Understanding of logs shipping, management and metric first alerting.
  • Security Mindset: Experience with secrets management, vulnerability scanning, and securing the software supply chain.
  • AI & Prompt Engineering: Familiarity with using LLMs, Public MCPs, or Bedrock Agent Core to enhance SRE workflows.
  • Data & Middleware & Orchestration: Hands-on experience with Rundeck and Airflow for job orchestration and automation, plus experience managing Kafka, MQ, or SQS.
Location

This role is open to candidates in the following locations: Chicago, IL - Hybrid

  • This role is expected to come into the office on a cadence set by the Hiring Manager/Team.
  • If you're not based in the location listed above, this role is not a fit, and we cannot accommodate remote work outside these locations.
  • Applicants must be authorized to work for any employer in the U.S. DriveWealth does not sponsor or take over sponsorship of an employment visa at this time.

Pay Range: $150,000 – $170,000 USD

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Manager, Site Reliability Engineering in Chicago, IL vacancy
  •  ...Site Reliability Engineering Manager MIDWEST IL - CHICAGOThe Performance Engineering practice within Technology is focused on optimizing the performance and scalability of enterprise applications through the combination of testing, diagnostics & monitoring, performance... 
    Suggested
    Work experience placement

    ClifyX

    Chicago, IL
    5 days ago
  • $130k - $225k

     ...expectations, integrity, innovation and a willingness to challenge consensus.The Algorithmic Trading Team is looking for a Site Reliability Engineer for our Chicago office. The SRE team is critical to the success of our trading - ensuring that our production trading... 
    Suggested
    Temporary work
    Work at office
    Flexible hours

    DRW

    Chicago, IL
    4 days ago
  • Qualifications: 8+ years of Software Engineering experience, or equivalent...  ...and maintain scalable and reliable infrastructure on Google...  ...effectively with the client, IT management and staff, and other groups in...  ...resources Willingness to work on-site at stated location in the job... 
    Suggested
    Contract work
    For contractors
    Work experience placement

    Cedent Consulting

    Chicago, IL
    4 days ago
  •  ...HudsonEmail: ****@*****.***: (***) ***-****Job Title: Site Reliability Engineer (Infrastructure & Systems)Location: Chicago, IL (Greater...  ...(AWS or Google Cloud Platform / GCP).Familiarity with managed container orchestrators such as Amazon EKS or Google GKE.Exposure... 
    Suggested
    Local area

    Objective Paradigm

    Chicago, IL
    5 days ago
  • $62 - $80 per hour

    Chicago, IllinoisRemote LocalContract$62/hr - $80/hrA senior Site Reliability Engineer will join an established infrastructure function...  ...core component of infrastructure provisioning and lifecycle management. The position blends hands-on production engineering with... 
    Suggested
    Full time
    Temporary work
    Remote work
    Flexible hours

    Motion Recruitment

    Chicago, IL
    1 day ago
  • $130k - $150k

     ...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are...  ...including Windows Server, VMware vSphere, VMware Site Recovery Manager (SRM), SAN technologies, and the Rubrik ecosystem, with the... 
    Work at office
    Work from home
    3 days per week

    CRA International

    Chicago, IL
    3 days ago
  • $91.2k - $136.8k

    Reliability Engineer - IE08GEWe’re determined to make a difference and are proud to be an insurance company that goes well beyond coverages and...  ...field.3+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), or DevOps.Hands-on experience... 
    Full time
    Temporary work
    Work at office
    3 days per week

    The Hartford Financial Services Group

    Chicago, IL
    4 days ago
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and...  ...role.Strong knowledge of SRE best practices and incident management protocolsDeep experience using and/or configuring New Relic... 
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Chicago, IL
    4 days ago
  • $108.08k - $172.5k

    Work with development and platform engineering teams to migrate and maintain applications in Google Cloud. Apply Observability concepts...  ...rotation support for production systems, facilitate incident management and conduct post-incident reviews. Drive, contribute and... 
    Full time
    Remote work
    Worldwide

    CME- Group

    Chicago, IL
    4 days ago
  •  ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment...  ...banking, financial transaction processing and asset management. We offer a competitive total rewards package including base... 

    JP Morgan Chase

    Chicago, IL
    5 days ago
  • Play a key role in ensuring system reliability at one of the world’s most iconic and...  ...largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial...  ...transaction processing and asset management. We offer a competitive total... 

    JP Morgan Chase

    Chicago, IL
    5 days ago
  • $130k - $180k

     ...belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a...  ...rotations. You’ll be a key voice in observability, change management, and service scalability, providing guidance during complex... 
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to Friday
    Flexible hours

    Imanage

    Chicago, IL
    1 day ago
  • $106k - $130k

     ...ineligible for employment Visa sponsorship.Role Summary The Senior Site Reliability Engineer applies software engineering and systems engineering...  ...as Code, automation, testing, incident response, capacity management, resilience, and operational readiness. Identify recurring... 
    Hourly pay
    Full time
    Immediate start
    Visa sponsorship
    Work visa
    Flexible hours

    Early Warning

    Chicago, IL
    5 days ago
  • $158.5k - $172k

     ...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will...  ...ecosystems. Our team is responsible for managing our centralized Enterprise Logging...  ...high-impact position driving continuous reliability, deep system optimization, and automation... 
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    Chicago, IL
    4 days ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical...  ...and alerting systems.The Fleet Management team provides the core runtime environment...  ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Chicago, IL
    1 day ago
  • $150k - $200k

     ...healthcare organization, creating unique engineering challenges around scale, reliability, security, real-time communication...  ...NOCD is looking for a Senior Site Reliability Engineer (SRE) to help...  ...with DevSecOps, IAM, secrets management, and encryption . Experience building... 
    Full time
    Work at office

    NOCD

    Chicago, IL
    3 days ago
  •  ...We are seeking a Staff Site Reliability Engineer to serve as the foundational Technical Lead for our Platform Engineering SRE organization. In this role, you will be the primary architect and visionary for the core technology foundations. As the technical lead for all... 
    Full time

    Informatic Technologies, Inc.

    Chicago, IL
    36 minutes ago
  • $132.1k - $220.1k

    Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week on-site...  ...GitOps: Mastery of Terraform module design and ArgoCD for managing immutable infrastructure at an enterprise scale.Distributed... 
    Full time
    Work at office
    Local area
    Worldwide
    2 days per week

    CME- Group

    Chicago, IL
    1 day ago
  • $127k - $249k

     ...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas...  ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Chicago, IL
    2 days ago
  • $160k - $210k

     ...you'll do:Join our Platform Engineering team, where you'll ensure the...  ...mentoring engineers across reliability initiativesAnalyze, troubleshoot...  ...provisioning, scaling, and management across all...  ...years of experience in DevOps, Site Reliability Engineering, or... 
    Work at office
    Worldwide
    Monday to Friday
    Flexible hours

    NinjaTrader Group

    Chicago, IL
    4 days ago
  • $194k - $267k

     ...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk...  ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Chicago, IL
    4 days ago
  • $55k - $151.47k

     ...LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in...  ...data integrity and accessibility- Leading incident management and resolution efforts to maintain operational continuityWhat... 
    Full time
    H1b

    PwC

    Chicago, IL
    3 days ago
  •  ...and companies, alikeKlover’s engineering team powers one of the fastest...  ...systems that prioritize reliability, security, and performance, and...  ...candidateAbout the RoleAs a Senior/Staff Site Reliability Engineer, you...  ...metrics to our Google-managed Prometheus instance and build... 
    Work at office
    Immediate start
    Remote work

    Attain Data

    Chicago, IL
    4 days ago
  • $250k - $350k

     ...where quantitative researchers, engineers, traders, and operational...  ...boost stability, throughput, and reliability Qualifications Minimum of 3...  ...in production support, site reliability, or infrastructure...  ...and Bash Hands-on experience managing Kubernetes in a production setting... 
    Full time

    Engtal Inc

    Chicago, IL
    2 days ago
  • $194k - $267k

     ..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Chicago, IL
    5 days ago
  • $204k - $306k

     ...We're all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity,...  ...week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    Okta

    Chicago, IL
    4 days ago
  •  ...Senior Site Reliability Engineer About The Position We are looking for a Senior Reliability Engineer to join our Platform team. In this...  ...engineering organization to build automated processes and tools for managing application and service deployments Own and support... 
    Temporary work
    Flexible hours

    Talentify.io

    Chicago, IL
    4 days ago
  •  ...services for the legal industry — making eDiscovery, case management, and litigation prep simple, fluid, and affordable for law...  ...tolerance for firms in active litigation. We're looking for a Site Reliability Engineer to help maintain the reliability, scalability, and... 

    Nextpoint

    Chicago, IL
    2 days ago
  • $152k - $205k

     ...Are you a systems-minded engineer who is happiest when production...  ...designed for? Do you want to own reliability for a platform that answers...  ...We’re looking for a Senior Site Reliability Engineer to join...  ...that make shipping boring. Manage infrastructure through code and... 
    Local area
    Remote work
    Work from home
    Visa sponsorship

    Fingerprint

    Chicago, IL
    3 days ago
  • $180k - $200k

     ...Come join tastytrade, part of IG Group, as we build the reliability practice behind the brokerage platform that active options...  ...equities traders rely on every market day. As our first Senior Site Reliability Engineer, you'll define what reliability means at tastytrade, from... 
    Work at office
    3 days per week

    tastyworks

    Chicago, IL
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!