Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Manager, Site Reliability Engineering - Infrastructure Platform

$232k - $319k

Okta

Secure Every Identity, from AI to Human

Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.

The Infrastructure Platform and Shared Services Team

Okta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput and 99.999 availability. We're looking for a technical leader to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and tooling. 

As the Sr. Manager of Infrastructure Platform and Shared Services, you will oversee multiple teams focused on Edge networking, K8s platform, Observability, automation platform & tooling. 

What you’ll be doing 

  • Lead the Infra platform and shared services org and various initiatives across SRE & Infrastructure organization.
  • Build a world-class observability platform and monitoring capabilities enabled with self-service
  • Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and intuitive self-service capabilities.
  • Own the design and operation of scalable, self-service Cloud infrastructure platforms (e.g. Observability Platform, SRE Productivity, deployments, and Edge Infrastructure)
  • Lead, mentor, and grow a high-performing team of engineers and managers across SRE and infrastructure shared services domains.
  • Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints.
  • Improve SDLC processes for Cloud infrastructure as a code, including the maturity of product deployements, change and release management 
  • Manage service and business expectations and prioritize resource allocation
  • Maintain a deep knowledge of industry best practices, evolving trends, and technologies

What you’ll bring to the role

  • 6+ years of experience in technical leadership & people management 
  • 3+ years of experience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS. Experience supporting a multi-Cloud environment will be a plus.
  • Strong expertise in cloud-native architectures, Edge infrastructure (WAF, ALB, NLB, Apache, Nginx), IaC (Terraform), Splunk, Grafana and CI/CD pipelines.
  • Strong background and hands-on experience in SRE automation& tooling 
  • Deep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment.
  • Demonstrated ability to lead cross-functional teams and manage large-scale programs
  • Effective verbal, written communication and interpersonal skills
  • Computer Science Degree or related degree or equivalent experience 

Additional requirements:

  • This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing the U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.

#LI-Hybrid

P8841_3271246

Below is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: .

The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between: $232,000—$319,000 USD

The Okta Experience

We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.

Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.

If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.

Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.
Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Senior Manager, Site Reliability Engineering - Infrastructure Platform in San Francisco, CA vacancy
  • $232k - $319k

     ...the trusted, neutral infrastructure that enables organizations...  ....The Infrastructure Platform and Shared Services...  ...great people and reliable, cost-effective, and...  ...tooling. As the Sr. Manager of Infrastructure Platform...  ...of SRE and product engineering by developing robust... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    3 days ago
  • $127k - $249k

     ...looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team...  ...of our cloud-based infrastructure. As a Staff SRE, you will...  ...controls that reinforce the platform’s security posture.This is...  ...compute security, identity management, and cloud security posture... 
    Senior
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    5 days ago
  • $196k - $220.5k

     ...everyone does on our platform: play video games....  ...games. Our Platform Infrastructure teams are responsible...  ...ensuring Discord remains reliable, efficient, and scalable. As a Senior Software Engineer on these teams, you...  .... Write code and manage infrastructure to... 
    Senior
    Full time
    Relocation
    Relocation package

    Discord

    San Francisco, CA
    1 day ago
  • $245k - $295k

     ...the only vertically integrated AI infrastructure company built from the ground up...  ....About the RoleWe are seeking a Senior Manager, Infrastructure Platform Engineering to lead a team building core...  ...scale compute infrastructure into reliable, secure, and efficiently allocatable... 
    Senior
    Temporary work
    Immediate start

    Crusoe

    San Francisco, CA
    3 days ago
  •  ...re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability...  ..., automate, and maintain the infrastructure that powers our core platform—...  ...HaveExperience with cloud and managed services (e.g. AWS)Experience... 
    Senior

    Alembic

    San Francisco, CA
    5 days ago
  • We are looking for a Senior or Staff level Site Reliability Engineer to strengthen the reliability...  ...maturity of our platform in San Francisco, California...  ..., and support incident management, debugging, and...  ...planning.• Contribute to infrastructure and delivery workflows... 
    Senior

    Robert Half

    San Francisco, CA
    5 days ago
  • $117k - $209.33k

     ...help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build...  ..., security, compliance, platform, and infrastructure teams to ensure services are reliable...  ...production readiness, incident management, observability, resilience testing... 
    Senior
    Full time
    For contractors

    Autodesk

    San Francisco, CA
    6 days ago
  • $152.5k - $205k

     ...leading internet financial platform companies, building the...  ...programmable blockchain infrastructure. Circle’s platform includes...  ...be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team...  ...authoring reusable modules, managing state and environments,... 
    Senior
    Flexible hours

    Circle

    San Francisco, CA
    4 days ago
  • $152.5k - $205k

     ...leading internet financial platform companies, building the...  ...and programmable blockchain infrastructure. Circle’s platform includes...  ...’ll be responsible forThe Site Reliability Engineer builds and maintains...  ...controls, auditability, and cost management. You will troubleshoot... 
    Senior
    Flexible hours

    Circle

    San Francisco, CA
    5 days ago
  • $190.8k - $267.1k

     ...Reddit's advertising platform, enabling...  ...its business. The reliability of our Ads systems...  ...closely with Ads Engineering teams to improve...  ...looking for a Staff Site Reliability Engineer...  ...future of Ads infrastructure at Reddit.What you...  ..., incident management, and performance... 
    Senior
    For contractors
    Work experience placement
    Remote work
    Flexible hours

    Reddit

    San Francisco, CA
    4 days ago
  • $165k - $225.6k

     ...the trusted, neutral infrastructure that enables...  ...infrastructure to enterprise platforms, we partner across...  ...functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    5 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within...  ...a range of critical infrastructure and operational functions...  ...systems.The Fleet Management team provides the core...  ...that ensure cluster reliability and security (e.g., CoreDNS...  ...cloud infrastructure platforms, including AWS, GCP,... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    7 days ago
  • $165k - $227k

     ...the trusted, neutral infrastructure that enables...  ...too, let's talk.The Engineering OpportunityWe are looking...  ...for an experienced Senior Site Reliability Engineer to join...  ...continuously invest in platform engineering,...  ...networking, and traffic management.Experience with... 
    Senior
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    5 days ago
  •  ...payments and financial platform for global...  ...combination of proprietary infrastructure and software, we...  ...solutions to manage everything from...  ...the teamThe Engineering team at Airwallex...  ...build scalable, reliable, and secure products...  ...What you’ll doAs a Senior Site Reliability... 
    Senior
    Temporary work
    Local area
    Worldwide

    Airwallex

    San Francisco, CA
    5 days ago
  • $148.5k - $223.9k

     ...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San...  ...with counterparts in the Infrastructure and R&D organizations,...  ..., and AI-powered platforms. You will not only respond...  ...efficiency.Incident Management: Lead the coordinated... 
    Senior
    Full time
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    4 days ago
  • $167.7k - $245.2k

     ...our US GovCloud platform. This team is responsible...  ...Federal region’s infrastructure and operations,...  ..., change management, monitoring, emergency...  ...for talented engineers with a software or...  ...teams to ensure the reliability, performance and...  ...Cisco careers site to discover more... 
    Senior
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    San Francisco, CA
    4 days ago
  • $150k - $220k

     ...in-class technology infrastructure to power a global private...  ...Role: As an engineering organization, we...  ...activity. Engineering managers enable engineers to...  ...purpose. The Manager, Site Reliability Engineering will...  ...partnering closely with Platform, Engineering,... 
    Local area

    Forge Global

    San Francisco, CA
    4 days ago
  • $180.5k - $236.91k

     ...Oscar. We're hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering...  ...full stack technology platform and a relentless focus...  ...such as DevOps, site reliability, and cloud best practices...  ...with partners, product managers, and designers to solve... 
    Senior
    Remote job
    Full time
    Work at office

    Oscar Health

    San Francisco, CA
    1 day ago
  •  ...remote containers ) that we manage. We’re revenue‑generating and...  ...and growing. You’ll own reliability, performance, and security for...  ...secure, multi‑tenant container infrastructure with fast startup and smart...  .... Why Julius Small, senior team; massive impact surface... 
    Senior
    Full time
    Remote work

    Julius Ai

    San Francisco, CA
    1 day ago
  •  ...pioneering Causal AI platform. We help the world's...  ...built on Grace Blackwell infrastructure — one of the fastest...  ...real-world scale, reliability, and security demands...  ...we're looking for an engineer who wants to own the...  ...device configuration management end to end, ensuring... 
    Senior

    Alembic

    San Francisco, CA
    7 days ago
  • $100k - $300k

     ...Abnormal AI, Zscaler Preeminent research labs like Deepmind and SAIL About the Role We're hiring a Senior Storage Infrastructure Engineer on our Core Platform team to own how we store, protect, and operate data at scale. You'll build the backup, observability, and... 
    Senior
    Full time

    Cogent Security

    San Francisco, CA
    1 day ago
  • $167.7k - $245.2k

     ...Digital Experience Assurance platform that empowers...  ...ImpactWe are seeking a skilled Senior Site Reliability Engineer (SRE) in Production Engineering...  .... You will design and manage large-scale, highly...  ...Manage a rapidly growing infrastructure capable of handling substantial... 
    Senior
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    San Francisco, CA
    7 days ago
  • $190k - $280k

     ...informative deployment pipeline. As an engineer on the Infrastructure Engineering team, you’ll help...  ...and maintain internal software and platform capabilities that reduce the cognitive...  ...and developer tooling. You’ll create reliable, repeatable abstractions that help... 
    Senior
    Hourly pay
    Full time
    Work at office

    Sentry

    San Francisco, CA
    1 day ago
  • $160k - $220k

     ...employs a team of 450 engineers and entrepreneurs....  ...California, USA. Senior Software Engineer - Infrastructure As a Senior...  ...satellites.  Speed and reliability aren’t luxuries...  ...environments Flaky test management and infrastructure...  ...PTO, and free on-site catered meals.... 
    Senior
    Permanent employment
    Full time
    Remote work
    Flexible hours

    Astranis

    San Francisco, CA
    1 day ago
  • $130k - $175k

     ...groundbreaking educational platform that promotes student...  ...curriculum management functionality, Kiddom...  ...and empower Kiddom’s engineering by building a scalable...  ...services. Practicing Infrastructure as Code (IaC) wherever...  ..., past experience, seniority, and demonstrated role... 
    Senior
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work

    Kiddom

    San Francisco, CA
    1 day ago
  • $180k - $240k

     ...FranciscoInfrastructure - Cloud Infrastructure /Full time /HybridWanna join...  ...access to space by building reliable, shareable satellites that...  ...SeniorCloud Infrastructure Engineer on our Cloud Infrastructure...  ...benefits with financial supportOff-sites and many social events and... 
    Senior
    Full time
    Temporary work

    Loft Orbital

    San Francisco, CA
    6 days ago
  •  ...real-world challenges. The Infrastructure Engineering team is crucial to the...  ...-Region Cloud: Design and manage globally distributed, multi...  ...peak performance, maximum reliability, and cost-efficiency across...  ...modeling best practices in site reliability, proactive system... 
    Senior
    Full time
    Shift work

    Hayden Ai

    San Francisco, CA
    1 day ago
  • $160k - $195k

     ...vertically integrated AI infrastructure company built from...  ...’re seeking a Senior Cloud Infrastructure Engineer to own the design,...  ...implementation, and lifecycle management of Crusoe’s...  ...IT, Security, and site-specific operations to keep systems reliable, secure, and ready... 
    Senior
    Temporary work
    Work at office

    Crusoe

    San Francisco, CA
    7 days ago
  • $174k - $252k

     ...years of experience developing large-scale infrastructure, distributed systems or networks, or...  ...technologies. Google's software engineers develop the next-generation technologies...  ...With your technical expertise you will manage project priorities, deadlines, and deliverables... 
    Senior

    Google

    San Francisco, CA
    5 days ago
  • $137.1k - $201.6k

     ...DoorDash’s GenAI Platform team sits within Machine...  ...and builds the shared infrastructure that helps DoorDash,...  ...serving and inference engines, fine-tuning and training...  ...role is ideal for a senior engineer who enjoys...  ...-source models with reliability, fallback, observability... 
    Senior
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Manager, Site Reliability Engineering - Infrastructure Platform. Be the first to apply!