Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Manager, Site Reliability Engineering

$204k - $306k

Okta, Inc.

Secure Every Identity, from AI to Human

Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Manager, Site Reliability Engineering

San Francisco, California

Secure Every Identity, from AI to Human

Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.

**This position requires 2 days a week in our San Francisco Office.

The IDaaS Site Reliability Engineering Group

Okta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput and 99.999 availability. We're looking for a technical leader to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and tooling.

As the Manager of Infrastructure Platform and Shared Services, you will oversee multiple teams focused on Edge networking, K8s platform, CI/CD, Observability, automation platform & tooling.

What you'll be doing

  • Managing a team of SRE's supporting various workloads and teams that support our IDaaS platform.
  • Drive the microservice journey, DevOps maturity, and workload reliability in tandem with architects and teams across the organization.
  • Accelerate the velocity of SRE and product engineering by developing powerful tooling, intuitive self-service capabilities, and robust self-healing patterns.
  • Lead, mentor, and grow a high-performing team of engineers and managers across platform, infrastructure, and shared services domains.
  • Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints.
  • Improve SDLC processes for Cloud infrastructure as a code, including the maturity of CI/CD pipelines, change and release management
  • Manage service and business expectations and prioritize resource allocation
  • Maintain a deep knowledge of industry best practices, evolving trends, and technologies


What you'll bring to the role

  • 3+ years of experience in technical leadership & people management
  • Extensive experience using Agile and DevOps methodologies to build product infrastructure and shared service at scale
  • Experience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS. Experience supporting a multi-Cloud environment will be a plus.
  • Strong expertise in cloud-native architectures, containerization (Kubernetes), IaC (Terraform), and CI/CD pipelines
  • Strong background and hands-on experience in SW development, PaaS and automation
  • Deep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment.
  • Effective verbal, written communication and interpersonal skills
  • Computer Science Degree or related degree or equivalent experience


Additional requirements:

  • This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.


#LI-Hybrid

P24518_3462184

Below is the annual base salary range for candidates located in San Francisco Bay Area. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit:

The annual base salary range for this position for candidates located in the San Francisco Bay area is between: $204,000—$306,000 USD

The Okta Experience

  • Supporting Your Well-Being
  • Driving Social Impact
  • Developing Talent and Fostering Connection + Community


We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.

Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.

If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding pleaseuse this Form to request an accommodation.

Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, pleaseclick here to view our full NYC AEDT Notice.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Manager, Site Reliability Engineering in Chicago, IL vacancy
  •  ...Site Reliability Engineering Manager MIDWEST IL - CHICAGOThe Performance Engineering practice within Technology is focused on optimizing the performance and scalability of enterprise applications through the combination of testing, diagnostics & monitoring, performance... 
    Suggested
    Work experience placement

    ClifyX

    Chicago, IL
    2 days ago
  • $130k - $150k

     ...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are...  ...including Windows Server, VMware vSphere, VMware Site Recovery Manager (SRM), SAN technologies, and the Rubrik ecosystem, with the... 
    Suggested
    Work at office
    Work from home
    3 days per week

    CRA International

    Chicago, IL
    1 day ago
  •  ...treatments for the right patients, at the right time.The Site Reliability Engineering team works with all departments and business units to provide...  ...skills in all of these areas.You have experience managing cloud infrastructure in AWS, GCP, or AzureYou have built and... 
    Suggested
    Full time

    Tempus

    Chicago, IL
    4 days ago
  • $158.5k - $172k

     ...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will...  ...ecosystems. Our team is responsible for managing our centralized Enterprise Logging...  ...high-impact position driving continuous reliability, deep system optimization, and automation... 
    Suggested
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    Chicago, IL
    2 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range...  ...observability and alerting systems.The Fleet Management team provides the core runtime...  ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager... 
    Suggested
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Chicago, IL
    4 days ago
  • Qualifications: 8+ years of Software Engineering experience, or equivalent...  ...and maintain scalable and reliable infrastructure on Google...  ...effectively with the client, IT management and staff, and other groups in...  ...resources Willingness to work on-site at stated location in the job... 
    Contract work
    For contractors
    Work experience placement

    Cedent Consulting

    Chicago, IL
    2 days ago
  • $130k - $225k

     ...expectations, integrity, innovation and a willingness to challenge consensus.The Algorithmic Trading Team is looking for a Site Reliability Engineer for our Chicago office. The SRE team is critical to the success of our trading - ensuring that our production trading... 
    Temporary work
    Work at office
    Flexible hours

    DRW

    Chicago, IL
    2 days ago
  •  ...Contact: Mike LaTulipEmail: ****@*****.*** Title: Site Reliability Engineer (Infrastructure & Systems)Location: Chicago, IL (Greater...  ...(AWS or Google Cloud Platform / GCP).Familiarity with managed container orchestrators such as Amazon EKS or Google GKE.Exposure... 
    Local area

    Objective Paradigm

    Chicago, IL
    2 days ago
  •  ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment...  ...banking, financial transaction processing and asset management. We offer a competitive total rewards package including base... 

    JP Morgan Chase

    Chicago, IL
    3 days ago
  • Play a key role in ensuring system reliability at one of the world’s most iconic and...  ...largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial...  ...transaction processing and asset management. We offer a competitive total... 

    JP Morgan Chase

    Chicago, IL
    3 days ago
  • $130k - $180k

     ...belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a...  ...rotations. You’ll be a key voice in observability, change management, and service scalability, providing guidance during complex... 
    Work at office
    Local area
    Remote work
    Worldwide
    Monday to Friday
    Flexible hours

    Imanage

    Chicago, IL
    4 days ago
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and...  ...role.Strong knowledge of SRE best practices and incident management protocolsDeep experience using and/or configuring New Relic... 
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Chicago, IL
    1 day ago
  • $108.08k - $172.5k

    Work with development and platform engineering teams to migrate and maintain applications in Google Cloud. Apply Observability concepts...  ...rotation support for production systems, facilitate incident management and conduct post-incident reviews. Drive, contribute and... 
    Full time
    Remote work
    Worldwide

    CME- Group

    Chicago, IL
    1 day ago
  • $100.7k - $167.8k

    Job SummaryThe Site Reliability Engineer III is a pivotal architect of stability for CME Clearing & Risk. You will engineer secure, scalable,...  ...gap between development and operations, you ensure our risk management services remain resilient and high-performing for... 
    Full time
    Worldwide

    CME- Group

    Chicago, IL
    2 days ago
  • $160k - $210k

     ...you'll do:Join our Platform Engineering team, where you'll ensure the...  ...mentoring engineers across reliability initiativesAnalyze, troubleshoot...  ...provisioning, scaling, and management across all...  ...years of experience in DevOps, Site Reliability Engineering, or... 
    Work at office
    Worldwide
    Monday to Friday
    Flexible hours

    NinjaTrader Group

    Chicago, IL
    1 day ago
  • $127k - $249k

     ...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas...  ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Chicago, IL
    7 hours ago
  • $194k - $267k

     ..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Chicago, IL
    4 days ago
  • $112.5k - $187.5k

     ...NoticePersonal Information We CollectYour Privacy ChoicesTeam OverviewAt TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Flexible hours
    2 days per week

    TransUnion

    Chicago, IL
    7 hours ago
  • $132.1k - $220.1k

    Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week on-site...  ...GitOps: Mastery of Terraform module design and ArgoCD for managing immutable infrastructure at an enterprise scale.Distributed... 
    Full time
    Work at office
    Local area
    Worldwide
    2 days per week

    CME- Group

    Chicago, IL
    4 days ago
  •  ...the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology team...  ...hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-... 

    JP Morgan Chase

    Chicago, IL
    4 days ago
  •  ...and companies, alikeKlover’s engineering team powers one of the fastest...  ...systems that prioritize reliability, security, and performance, and...  ...candidateAbout the RoleAs a Senior/Staff Site Reliability Engineer, you...  ...metrics to our Google-managed Prometheus instance and build... 
    Work at office
    Immediate start
    Remote work

    Attain Data

    Chicago, IL
    2 days ago
  • $194k - $267k

     ...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk...  ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    Chicago, IL
    2 days ago
  • $118.3k - $219.8k

    Are you excited to lead Site Reliability Engineering teams that keep mission-critical, 24/7 services running reliably and securely?Do you enjoy...  ...teams to deliver proactive monitoring, effective incident management, and continuous optimisation. Through strong alignment... 
    Full time
    Local area

    RELX Group

    Chicago, IL
    4 days ago
  •  ...A leading quantitative trading firm is seeking a Head of Site Reliability Engineering to help scale one of its most critical infrastructure organisations...  ...leadership roles within the organisation. What You'll Do Manage a team of Site Reliability Engineers. Drive reliability,... 
    Immediate start

    Acquire Me

    Chicago, IL
    2 days ago
  •  ...Edward Jones Site Reliability Engineer 100% remote Initial contract is 6 months, but will be a multi year engagement. Position Overview...  ...of our systems. You will be responsible for incident management, root cause analysis, and implementing postmortem processes... 
    Contract work
    Remote work

    HCL Global Systems

    Chicago, IL
    4 days ago
  •  ...building and running systems that must perform reliably under real-time market conditions. The culture is highly collaborative, engineering-driven, and focused on continuous...  ...a related field 3+ years of experience in site reliability, systems engineering, or technical... 

    Fintal Partners

    Chicago, IL
    1 day ago
  •  ...Site Reliability Engineer II Play a key role in ensuring system reliability at one of the world's most iconic and largest financial institutions. As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology... 
    Local area

    hackajob

    Chicago, IL
    3 days ago
  • $190.8k - $267.1k

     ...influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your...  ...more. In this role, you will also take ownership of risk management, ensuring the reliability and performance of our systems.... 
    Work experience placement
    Home office
    Flexible hours

    Alien Blue

    Chicago, IL
    5 days ago
  • $125.04k - $187.56k

     ...Digital and E-commerce, Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible for ensuring the scalability, reliability...  ...(SLOs) and service level indicators (SLIs). Build and manage microservices-based platforms leveraging Spring Boot, Java,... 
    Full time
    Work at office
    Remote work
    Flexible hours

    ViziRecruiter

    Chicago, IL
    3 days ago
  • $150k - $155k

     ...Site Reliability Engineer Hybrid (3 days onsite, 2 days remote) full‑time. No visa sponsorship. Base pay: $150,000 – $155,000 per year, subject...  ...large‑scale distributed systems Experience managing infrastructure in public cloud environments like AWS (preferred... 
    Full time
    Work experience placement
    Remote work
    Visa sponsorship

    Request Technology

    Chicago, IL
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!