Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Manager, Site Reliability Engineering - Infrastructure Platform

$232k - $319k

Okta

Secure Every Identity, from AI to Human

Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.

The Infrastructure Platform and Shared Services Team

Okta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput and 99.999 availability. We're looking for a technical leader to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and tooling. 

As the Sr. Manager of Infrastructure Platform and Shared Services, you will oversee multiple teams focused on Edge networking, K8s platform, Observability, automation platform & tooling. 

What you’ll be doing 

  • Lead the Infra platform and shared services org and various initiatives across SRE & Infrastructure organization.
  • Build a world-class observability platform and monitoring capabilities enabled with self-service
  • Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and intuitive self-service capabilities.
  • Own the design and operation of scalable, self-service Cloud infrastructure platforms (e.g. Observability Platform, SRE Productivity, deployments, and Edge Infrastructure)
  • Lead, mentor, and grow a high-performing team of engineers and managers across SRE and infrastructure shared services domains.
  • Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints.
  • Improve SDLC processes for Cloud infrastructure as a code, including the maturity of product deployements, change and release management 
  • Manage service and business expectations and prioritize resource allocation
  • Maintain a deep knowledge of industry best practices, evolving trends, and technologies

What you’ll bring to the role

  • 6+ years of experience in technical leadership & people management 
  • 3+ years of experience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS. Experience supporting a multi-Cloud environment will be a plus.
  • Strong expertise in cloud-native architectures, Edge infrastructure (WAF, ALB, NLB, Apache, Nginx), IaC (Terraform), Splunk, Grafana and CI/CD pipelines.
  • Strong background and hands-on experience in SRE automation& tooling 
  • Deep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment.
  • Demonstrated ability to lead cross-functional teams and manage large-scale programs
  • Effective verbal, written communication and interpersonal skills
  • Computer Science Degree or related degree or equivalent experience 

Additional requirements:

  • This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing the U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.

#LI-Hybrid

P8841_3271246

Below is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: .   

The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between: $232,000—$319,000 USD

The Okta Experience

We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.

Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.

If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please  use this Form to request an accommodation.

Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please  click here to view our full NYC AEDT Notice.
Vacancy posted 4 hours ago
Similar jobs that could be interesting for youBased on the Senior Manager, Site Reliability Engineering - Infrastructure Platform in Washington DC vacancy
  •  ...Director to support heavy civil highway and infrastructure projects. The ideal candidate must have a Bachelor's Degree in Civil Engineering and over 15 years of experience in heavy...  ...construction. Responsibilities include managing budgeting, overseeing project claims, and... 
    Senior

    The Lane Construction Corporation

    Bethesda, MD
    2 days ago
  • Scientific Research Corporation is seeking an experienced Program Manager in Alexandria, Virginia, to oversee Test and Evaluation Infrastructure Projects. The ideal candidate has over 15 years of project management experience, with at least 3 in a lead role, and proficiency... 
    Senior
    Contract work

    Scientific Research Corporation

    Alexandria, VA
    5 days ago
  •  ...Senior Infrastructure Project Manager We are seeking a Senior Infrastructure Project Manager to join the Infrastructure Platform services department-specific PMO of our prestigious Randstad Digital...  ...and maintain SharePoint project sites within the EPM system. Ensure... 
    Senior
    Long term contract
    For contractors
    Work experience placement
    For subcontractor

    Samprasoft

    Washington DC
    2 days ago
  •  ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability...  ...in building a sustainable platform by developing systems for...  ...teams to ensure seamless infrastructure and application integration...  ...Docker orchestration and management. ~ Experience with Kubernetes... 
    Senior
    Work experience placement

    Samprasoft

    Washington DC
    2 days ago
  • $175k - $250k

     ...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable...  .... Modality: On‑Site only. Must live...  ...flexibility. Their platform allows users to connect...  ...scalability, performance, and reliability across environments....  ...workloads at scale Manage and automate GPU... 
    Senior
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    Washington DC
    1 day ago
  • $168k - $200k

     ...is the data collaboration platform trusted for healthcare....  ...We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform...  ...with deep cloud infrastructure and data platform experience...  ...workspace setup, cluster/job management, and integration with CI/... 
    Senior
    Remote work

    Datavant

    Washington DC
    1 day ago
  • $166k - $220k

     ...expectations. Our systems integration engineers internalize the nuances of...  ...We are looking for a Site Reliability Engineer (SRE) to join AGD,...  ...of Kubernetes cloud infrastructure, DevOps, CI/CD and improving...  ...developer experience. You will be managing cloud deployments in AWS,... 
    Senior
    Full time
    Work experience placement

    Mosaic

    Washington DC
    6 hours ago
  • $210k - $230k

     ...is currently hiring for a Senior Site Reliability Engineer (SRE) to design,...  ...scalable, and resilient infrastructure systems. The ideal candidate...  ...Design, deploy, and manage cloud infrastructure using...  ...Build self-service tools and platforms to enable development teams... 
    Senior
    Currently hiring
    Remote work

    GovCIO

    Arlington, VA
    1 day ago
  •  ...and intellectually curious Senior Site Reliability Engineer to join our team working...  ...of critical application infrastructure for a federal financial agency...  ...and hands-on experience managing infrastructure with tools...  ...continual improvement of platform operational processes,... 
    Senior
    Remote work

    Elevate Government Solutions

    Washington DC
    3 days ago
  • $149.4k - $202k

     ...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering...  ...systems. Our SREs don’t just manage infrastructure; they build it using...  ...across multiple services or entire platforms, ensuring alignment with business... 
    Senior
    Remote work

    Noctua Technology

    Washington DC
    3 days ago
  •  ...payments and financial platform for global...  ...combination of proprietary infrastructure and software, we...  ...solutions to manage everything from...  ...software engineers into superheroes....  ...You’ll Do As a Senior Software Engineer...  ...autonomous systems reliable in production infrastructure... 
    Senior
    Full time
    Worldwide

    Airwallex

    Washington DC
    more than 2 months ago
  • $136.2k - $214.01k

     ...execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a...  ...team player who cares about the infrastructure, remains calm in crisis, collaborates...  ..., etc. • Experience automating management and operational tasks using Node.Js... 
    Senior
    Full time
    Flexible hours

    Proofpoint

    Laurel, MD
    2 days ago
  • $121.4k - $218.6k

     ...Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud...  ...optimization. As a Senior Site Reliability Engineer...  ...Investigating and analyzing platform components to identify... 
    Senior
    Work experience placement
    Work at office

    Akamai

    Washington DC
    5 days ago
  • $155.4k - $210.2k

     ...AWS has provided a highly reliable, scalable, low-cost infrastructure platform that powers hundreds of...  ...collaborate with and manage multi-disciplinary internal...  ...closing of real estate sites; this includes build to...  ...teams, including Design Engineering, Legal, Economic Development... 
    Senior
    Local area
    Flexible hours

    Amazon

    Arlington, VA
    3 days ago
  • $150k - $180k

     ...and Mission Solutions (the platforms).Together, our teams...  ...seeking an experienced Senior Site Reliability Engineer to help design, build, operate...  ...- and business-critical infrastructure that powers Umbra's...  ...functional teams, product managers, and stakeholders to align... 
    Senior
    Permanent employment
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    Umbra

    Arlington, VA
    2 days ago
  • $81.1k - $187k

     ...Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud...  ...will work closely with development, infrastructure, security, and operations teams to monitor...  ...Capacity Ingestion and Management: -Takes proactive steps to design... 
    Senior
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Washington DC
    1 day ago
  • $126k - $189k

     ...offices in Seattle. On-site requirements vary...  ...expert systems engineer who occupies the...  ...that world-class infrastructure should be a...  ...the rigor of a  Senior Software Engineer...  ...automate cluster health management. Performance...  ...or "Site Reliability Engineering" (SRE... 
    Senior
    Full time
    Contract work
    Work experience placement
    Work at office
    Flexible hours
    Weekend work

    The Allen Institute For Artificial Intelligence

    Washington DC
    more than 2 months ago
  •  ...Federal Credit Union in Suitland, MD, seeks an experienced IT Infrastructure Engineer to configure, secure, and optimize our Microsoft 365 and...  ...baselines. Lead AD architecture and GPOs, migrate tools to Azure, manage Exchange Online and Teams governance, and partner with... 
    Senior

    Andrews-Federal-Credit-Union

    Suitland, MD
    4 days ago
  • A leading cybersecurity engineering company is seeking a Project Leader 2 for Critical Infrastructure Solutions in McLean, Virginia. This executive leadership role demands expertise in managing multi-million dollar electrical projects, offering leadership to project teams... 
    Senior

    M.C. Dean, Inc.

    Mc Lean, VA
    1 day ago
  • $165k - $170k

     ...Senior AWS Cloud Infrastructure Engineer You’re Opportunity: At PLACE, we’re building...  ...you’ll partner with our Platform & Infrastructure...  ...systems secure, scalable, and reliable across every engineering...  ...policies, AWS KMS, Secrets Manager, and least-privilege access... 
    Senior
    Full time
    Work at office
    Work from home

    PLACE Corporate Careers

    Bethesda, MD
    2 days ago
  • An established industry player is seeking a Structural Engineer 3 to join their dynamic team. This role involves designing and evaluating...  ...for engineering and a desire to make a significant impact on infrastructure, this opportunity is perfect for you. Join a team committed... 
    Senior
    For contractors

    M.C. Dean, Inc.

    Mc Lean, VA
    2 days ago
  • $165k - $265k

     ...human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD) At...  ...'s software and GPU infrastructure, you will design, operate...  ...Operations, and GPU platforms. You will develop...  ...automation to deploy and manage on-premise compute...  ...engineersAs a senior engineer you must lead... 
    Senior
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    11 hours ago
  • $229.9k - $262.4k

     ...Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good...  ...technical program managers, and product managers...  ...available through this site. Capital One... 
    Senior
    Full time
    Part time
    Local area

    Capital One

    McLean, VA
    3 days ago
  •  ...We build and operate the platform that our engineers and our customers rely on...  ...environments. Develop infrastructure as code using Ansible, Terraform...  ...Improve observability, reliability, security, and...  ...Containers PKI and certificate management Secrets management... 
    Senior
    Full time
    Temporary work
    Part time
    Worldwide

    Shield AI

    Washington DC
    a month ago
  • Senior Lead Infrastructure Engineer We're looking for a talented, senior engineering...  ...Infrastructure Platforms team, you will...  ...controls, and reliability enabling product teams...  ...planning, performance management, placement...  ...care coverage, on-site health and wellness... 
    Senior

    Chase

    Mc Lean, VA
    5 days ago
  •  ...Combinator’s #1 GovTech startup. About the Role We want a Platform/Infrastructure engineer to help shape how Promise delivers and runs software, and...  ...Platform experience (our primary cloud) Experience managing ArgoCD Typescript, Golang Fintech, payments, or... 
    Permanent employment
    Full time
    Local area
    Flexible hours

    Promise

    Washington DC
    more than 2 months ago
  • $167.85k - $209.75k

     ...Senior Software Engineer, AI/ML Platforms and Infrastructure – Brain Health Accelerator The Allen Institut...  ...are secure, reliable, compliant, and cost-effective...  ...tuning and cost management Provide AI workflow...  ...Occasional travel to partner sites required Occasional... 
    Senior
    Work at office
    Local area
    Remote work
    Visa sponsorship
    Work visa
    Relocation package

    Allen Institute

    Washington DC
    more than 2 months ago
  • $100.1k - $150.4k

     ...business and financial challenges. Builds and manages long term customer relationships/...  ....Maximizes assigned Project Development Engineering resources effectively and efficiently. Ensures...  ...Us tab on the Johnson Controls Careers site at Controls International plc. is an... 
    Senior
    Full time
    Work at office

    Johnson Controls

    Washington DC
    2 days ago
  •  ...Summary The Cloud Platform Foundations team builds...  ...operates the core infrastructure behind Temporal Cloud, with a focus on reliability, scalability, and automation...  ...and namespace management , and the customer-facing...  ...ll work closely with engineers and product managers... 
    Senior
    Full time

    Temporal Technologies

    Washington DC
    19 days ago
  •  ...hosted and operated across managed GPU clouds and...  ...our agent-sandboxing platform: hardware-isolated microVMs...  ...— building the fork engine, the guest agent, and...  ...open-weight LLM serving infrastructure (vLLM/SGLang or similar...  ...packaging models into reliable, metered production... 
    Senior
    Full time
    Remote work
    Work visa
    Flexible hours
    Day shift

    Azx Inc

    Washington DC
    28 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Manager, Site Reliability Engineering - Infrastructure Platform. Be the first to apply!