Manager, Site Reliability Engineering
$150k - $170kDriveWealth
Manager, Site Reliability Engineering
Office - Chicago
DriveWealth is on a mission to make investing easier. We believe that everyone should have the ability to control their financial future, and that access to financial markets should not be limited by geography, wealth, or legacy systems. We are a global B2B financial technology organization dedicated to democratizing access to financial independence around the world. Our mission is realized through an API-based platform, empowering our partners to offer seamless investing and trading experiences to clients worldwide, all from their mobile devices. Our technology provides partners with a modern, extensible toolkit, enabling traditional investment workflows and innovative techniques like fractional share ownership. DriveWealth has evolved into a global platform offering trading of US equities, mutual funds, ETFs, fixed income, and options.
There's never been a better time to build a category-defining business and there has rarely been a team better positioned for this opportunity. Our culture blends the pace and agility of a fintech start-up with the impact, stability, and discipline of Wall Street. We encourage creativity and experimentation while ensuring institutional-grade execution and regulatory compliance in everything we do. Join us and help build the future of global investing!
About The Role
As the Manager of Site Reliability Engineering, you'll lead a team of SRE Automation Engineers while remaining a hands-on technical authority for our Brokerage-as-a-Service platform. This isn't a purely people-management seat, you're expected to bring the same principal-level SRE depth to automation design and engineering as an individual contributor, while also building the team, setting technical direction, and developing your engineers' careers.
This role is centered on reducing manual toil through engineering, applying Google's SRE principles: SLOs, error budgets, blameless postmortems, and systematic toil reduction, adapted to a regulated brokerage environment. You'll carry two responsibilities at once: driving the automation agenda, building and orchestrating workflows in Rundeck and Airflow to eliminate repetitive work, and growing your team of SRE Automation Engineers into a high-functioning automation practice. You'll guide the design of internal SRE platforms, automate complex workflows, and ensure our Kubernetes-based and colo ecosystems can handle the demands of global financial markets, while owning the people side of the team: mentorship, performance, and growth, and the day-to-day management of the team's Jira board. While this role includes participation in on-call rotations supporting our 24/7 global operations, your primary mission is to build systems that make manual intervention obsolete, and a team capable of sustaining that mission.
What You'll Do
- Team Leadership & Development: Manage, mentor, and grow a team of SRE Automation Engineers—setting technical direction, running 1:1s, owning performance management and career development, and managing the team's Jira board to prioritize and track sprint work.
- Engineering & Automation: Lead the design and development of internal tooling and automation—including Rundeck and Airflow-based orchestration—to eliminate repetitive manual toil and improve developer velocity, staying hands-on with the most complex, highest-leverage automation work yourself.
- SRE Practice & Governance: Adapt Google's SRE principles to our environment—defining SLIs, SLOs, and error budgets, and using them to guide engineering and operational priorities.
- Infrastructure as Code: Set architectural standards for modular, reusable IaC using Terraform and oversee GitOps workflows via ArgoCD.
- Platform Governance: Review software architecture and Kubernetes metrics to ensure high availability, capacity planning, and cost-optimization across AWS regions, and hold the team accountable to those standards.
- Incident Engineering: Lead incident response for critical events, drive complex root-cause analysis (RCA), and champion a blameless post-mortem culture across the organization.
- Collaboration & Stakeholder Management: Partner with engineering leadership to align SRE priorities with business goals, and foster adoption of new tools, security standards, and reliability best practices across teams.
You Bring
- People Leadership: Prior experience managing or leading SRE/DevOps engineers, ideally in a fintech or highly regulated environment. Able to flex between hands-on principal-level engineering and coaching and developing a team.
- Google SRE Fundamentals: Working knowledge of Google's SRE practices—SLIs/SLOs, error budgets, toil reduction, and blameless postmortems—and experience adapting them to a regulated environment.
- Linux & Networking Mastery: Proficient in Linux administration with a deep understanding of the TCP/IP stack, OSI model, DNS, and network troubleshooting.
- FinTech Background: Experience working in highly regulated financial environments or with FIX/API connectivity.
- Production Kubernetes: Hands-on experience managing production-grade clusters, including RBAC, autoscaling, Helm, and multi-cluster patterns.
- Cloud Native Expertise (AWS): Strong grasp of AWS core services, security, and high-availability patterns. Proficiency with boto3 and AWS CLI for automation.
- Modern CI/CD & GitOps: Experience building secure, automated delivery pipelines and operating GitOps workflows (ArgoCD).
- Code Proficiency: Strong scripting and development skills in Python or Golang, along with Bash and Ansible.
- Observability: Experience with Grafana/Similar tools, Prometheus, Understanding of logs shipping, management and metric first alerting.
- Security Mindset: Experience with secrets management, vulnerability scanning, and securing the software supply chain.
- AI & Prompt Engineering: Familiarity with using LLMs, Public MCPs, or Bedrock Agent Core to enhance SRE workflows.
- Data & Middleware & Orchestration: Hands-on experience with Rundeck and Airflow for job orchestration and automation, plus experience managing Kafka, MQ, or SQS.
Location
This role is open to candidates in the following locations: Chicago, IL - Hybrid
- This role is expected to come into the office on a cadence set by the Hiring Manager/Team.
- If you're not based in the location listed above, this role is not a fit, and we cannot accommodate remote work outside these locations.
- Applicants must be authorized to work for any employer in the U.S. DriveWealth does not sponsor or take over sponsorship of an employment visa at this time.
Pay Range: $150,000 – $170,000 USD
- ...Site Reliability Engineering Manager MIDWEST IL - CHICAGOThe Performance Engineering practice within Technology is focused on optimizing the performance and scalability of enterprise applications through the combination of testing, diagnostics & monitoring, performance...SuggestedWork experience placement
$130k - $225k
...expectations, integrity, innovation and a willingness to challenge consensus.The Algorithmic Trading Team is looking for a Site Reliability Engineer for our Chicago office. The SRE team is critical to the success of our trading - ensuring that our production trading...SuggestedTemporary workWork at officeFlexible hours- Qualifications: 8+ years of Software Engineering experience, or equivalent... ...and maintain scalable and reliable infrastructure on Google... ...effectively with the client, IT management and staff, and other groups in... ...resources Willingness to work on-site at stated location in the job...SuggestedContract workFor contractorsWork experience placement
- ...HudsonEmail: ****@*****.***: (***) ***-****Job Title: Site Reliability Engineer (Infrastructure & Systems)Location: Chicago, IL (Greater... ...(AWS or Google Cloud Platform / GCP).Familiarity with managed container orchestrators such as Amazon EKS or Google GKE.Exposure...SuggestedLocal area
$62 - $80 per hour
Chicago, IllinoisRemote LocalContract$62/hr - $80/hrA senior Site Reliability Engineer will join an established infrastructure function... ...core component of infrastructure provisioning and lifecycle management. The position blends hands-on production engineering with...SuggestedFull timeTemporary workRemote workFlexible hours$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...including Windows Server, VMware vSphere, VMware Site Recovery Manager (SRM), SAN technologies, and the Rubrik ecosystem, with the...Work at officeWork from home3 days per week$91.2k - $136.8k
Reliability Engineer - IE08GEWe’re determined to make a difference and are proud to be an insurance company that goes well beyond coverages and... ...field.3+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), or DevOps.Hands-on experience...Full timeTemporary workWork at office3 days per week$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and... ...role.Strong knowledge of SRE best practices and incident management protocolsDeep experience using and/or configuring New Relic...Full timeTemporary workWork experience placementFlexible hours$108.08k - $172.5k
Work with development and platform engineering teams to migrate and maintain applications in Google Cloud. Apply Observability concepts... ...rotation support for production systems, facilitate incident management and conduct post-incident reviews. Drive, contribute and...Full timeRemote workWorldwide- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment... ...banking, financial transaction processing and asset management. We offer a competitive total rewards package including base...
- Play a key role in ensuring system reliability at one of the world’s most iconic and... ...largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial... ...transaction processing and asset management. We offer a competitive total...
$130k - $180k
...belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a... ...rotations. You’ll be a key voice in observability, change management, and service scalability, providing guidance during complex...Work at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$106k - $130k
...ineligible for employment Visa sponsorship.Role Summary The Senior Site Reliability Engineer applies software engineering and systems engineering... ...as Code, automation, testing, incident response, capacity management, resilience, and operational readiness. Identify recurring...Hourly payFull timeImmediate startVisa sponsorshipWork visaFlexible hours$158.5k - $172k
...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will... ...ecosystems. Our team is responsible for managing our centralized Enterprise Logging... ...high-impact position driving continuous reliability, deep system optimization, and automation...Full timeTemporary workWork at officeFlexible hours3 days per week$127k - $249k
Platform Engineering is the department within SRE that is responsible for a range of critical... ...and alerting systems.The Fleet Management team provides the core runtime environment... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...Work at officeLocal areaRemote workWorldwideFlexible hours$150k - $200k
...healthcare organization, creating unique engineering challenges around scale, reliability, security, real-time communication... ...NOCD is looking for a Senior Site Reliability Engineer (SRE) to help... ...with DevSecOps, IAM, secrets management, and encryption . Experience building...Full timeWork at office- ...We are seeking a Staff Site Reliability Engineer to serve as the foundational Technical Lead for our Platform Engineering SRE organization. In this role, you will be the primary architect and visionary for the core technology foundations. As the technical lead for all...Full time
$132.1k - $220.1k
Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week on-site... ...GitOps: Mastery of Terraform module design and ArgoCD for managing immutable infrastructure at an enterprise scale.Distributed...Full timeWork at officeLocal areaWorldwide2 days per week$127k - $249k
...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Local areaRemote workWorldwideFlexible hours$160k - $210k
...you'll do:Join our Platform Engineering team, where you'll ensure the... ...mentoring engineers across reliability initiativesAnalyze, troubleshoot... ...provisioning, scaling, and management across all... ...years of experience in DevOps, Site Reliability Engineering, or...Work at officeWorldwideMonday to FridayFlexible hours$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$55k - $151.47k
...LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in... ...data integrity and accessibility- Leading incident management and resolution efforts to maintain operational continuityWhat...Full timeH1b- ...and companies, alikeKlover’s engineering team powers one of the fastest... ...systems that prioritize reliability, security, and performance, and... ...candidateAbout the RoleAs a Senior/Staff Site Reliability Engineer, you... ...metrics to our Google-managed Prometheus instance and build...Work at officeImmediate startRemote work
$250k - $350k
...where quantitative researchers, engineers, traders, and operational... ...boost stability, throughput, and reliability Qualifications Minimum of 3... ...in production support, site reliability, or infrastructure... ...and Bash Hands-on experience managing Kubernetes in a production setting...Full time$194k - $267k
..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$204k - $306k
...We're all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity,... ...week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions...Permanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week- ...Senior Site Reliability Engineer About The Position We are looking for a Senior Reliability Engineer to join our Platform team. In this... ...engineering organization to build automated processes and tools for managing application and service deployments Own and support...Temporary workFlexible hours
- ...services for the legal industry — making eDiscovery, case management, and litigation prep simple, fluid, and affordable for law... ...tolerance for firms in active litigation. We're looking for a Site Reliability Engineer to help maintain the reliability, scalability, and...
$152k - $205k
...Are you a systems-minded engineer who is happiest when production... ...designed for? Do you want to own reliability for a platform that answers... ...We’re looking for a Senior Site Reliability Engineer to join... ...that make shipping boring. Manage infrastructure through code and...Local areaRemote workWork from homeVisa sponsorship$180k - $200k
...Come join tastytrade, part of IG Group, as we build the reliability practice behind the brokerage platform that active options... ...equities traders rely on every market day. As our first Senior Site Reliability Engineer, you'll define what reliability means at tastytrade, from...Work at office3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
- site reliability engineer sre Chicago, IL
- site reliability engineer Chicago, IL
- site reliability engineer remote Chicago, IL
- official site Chicago, IL
- site services specialist Chicago, IL
- construction site safety Chicago, IL
- IT site lead Chicago, IL
- site recruiter Chicago, IL
- site leader Chicago, IL
- site safety Chicago, IL



