Manager, Site Reliability Engineering
$204k - $306kOkta, Inc.
Secure Every Identity, from AI to Human
Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Manager, Site Reliability EngineeringSan Francisco, California
Secure Every Identity, from AI to Human
Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.**This position requires 2 days a week in our San Francisco Office.
The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput and 99.999 availability. We're looking for a technical leader to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and tooling.
As the Manager of Infrastructure Platform and Shared Services, you will oversee multiple teams focused on Edge networking, K8s platform, CI/CD, Observability, automation platform & tooling.
What you'll be doing
- Managing a team of SRE's supporting various workloads and teams that support our IDaaS platform.
- Drive the microservice journey, DevOps maturity, and workload reliability in tandem with architects and teams across the organization.
- Accelerate the velocity of SRE and product engineering by developing powerful tooling, intuitive self-service capabilities, and robust self-healing patterns.
- Lead, mentor, and grow a high-performing team of engineers and managers across platform, infrastructure, and shared services domains.
- Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints.
- Improve SDLC processes for Cloud infrastructure as a code, including the maturity of CI/CD pipelines, change and release management
- Manage service and business expectations and prioritize resource allocation
- Maintain a deep knowledge of industry best practices, evolving trends, and technologies
What you'll bring to the role
- 3+ years of experience in technical leadership & people management
- Extensive experience using Agile and DevOps methodologies to build product infrastructure and shared service at scale
- Experience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS. Experience supporting a multi-Cloud environment will be a plus.
- Strong expertise in cloud-native architectures, containerization (Kubernetes), IaC (Terraform), and CI/CD pipelines
- Strong background and hands-on experience in SW development, PaaS and automation
- Deep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment.
- Effective verbal, written communication and interpersonal skills
- Computer Science Degree or related degree or equivalent experience
Additional requirements:
- This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.
#LI-Hybrid
P24518_3462184
Below is the annual base salary range for candidates located in San Francisco Bay Area. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit:
The annual base salary range for this position for candidates located in the San Francisco Bay area is between: $204,000—$306,000 USDThe Okta Experience
- Supporting Your Well-Being
- Driving Social Impact
- Developing Talent and Fostering Connection + Community
We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.
Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws. If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding pleaseuse this Form to request an accommodation. Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, pleaseclick here to view our full NYC AEDT Notice.- ...Site Reliability Engineering Manager MIDWEST IL - CHICAGOThe Performance Engineering practice within Technology is focused on optimizing the performance and scalability of enterprise applications through the combination of testing, diagnostics & monitoring, performance...SuggestedWork experience placement
$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...including Windows Server, VMware vSphere, VMware Site Recovery Manager (SRM), SAN technologies, and the Rubrik ecosystem, with the...SuggestedWork at officeWork from home3 days per week- ...treatments for the right patients, at the right time.The Site Reliability Engineering team works with all departments and business units to provide... ...skills in all of these areas.You have experience managing cloud infrastructure in AWS, GCP, or AzureYou have built and...SuggestedFull time
$158.5k - $172k
...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will... ...ecosystems. Our team is responsible for managing our centralized Enterprise Logging... ...high-impact position driving continuous reliability, deep system optimization, and automation...SuggestedFull timeTemporary workWork at officeFlexible hours3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range... ...observability and alerting systems.The Fleet Management team provides the core runtime... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...SuggestedWork at officeLocal areaRemote workWorldwideFlexible hours- Qualifications: 8+ years of Software Engineering experience, or equivalent... ...and maintain scalable and reliable infrastructure on Google... ...effectively with the client, IT management and staff, and other groups in... ...resources Willingness to work on-site at stated location in the job...Contract workFor contractorsWork experience placement
$130k - $225k
...expectations, integrity, innovation and a willingness to challenge consensus.The Algorithmic Trading Team is looking for a Site Reliability Engineer for our Chicago office. The SRE team is critical to the success of our trading - ensuring that our production trading...Temporary workWork at officeFlexible hours- ...Contact: Mike LaTulipEmail: ****@*****.*** Title: Site Reliability Engineer (Infrastructure & Systems)Location: Chicago, IL (Greater... ...(AWS or Google Cloud Platform / GCP).Familiarity with managed container orchestrators such as Amazon EKS or Google GKE.Exposure...Local area
- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment... ...banking, financial transaction processing and asset management. We offer a competitive total rewards package including base...
- Play a key role in ensuring system reliability at one of the world’s most iconic and... ...largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial... ...transaction processing and asset management. We offer a competitive total...
$130k - $180k
...belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a... ...rotations. You’ll be a key voice in observability, change management, and service scalability, providing guidance during complex...Work at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and... ...role.Strong knowledge of SRE best practices and incident management protocolsDeep experience using and/or configuring New Relic...Full timeTemporary workWork experience placementFlexible hours$108.08k - $172.5k
Work with development and platform engineering teams to migrate and maintain applications in Google Cloud. Apply Observability concepts... ...rotation support for production systems, facilitate incident management and conduct post-incident reviews. Drive, contribute and...Full timeRemote workWorldwide$100.7k - $167.8k
Job SummaryThe Site Reliability Engineer III is a pivotal architect of stability for CME Clearing & Risk. You will engineer secure, scalable,... ...gap between development and operations, you ensure our risk management services remain resilient and high-performing for...Full timeWorldwide$160k - $210k
...you'll do:Join our Platform Engineering team, where you'll ensure the... ...mentoring engineers across reliability initiativesAnalyze, troubleshoot... ...provisioning, scaling, and management across all... ...years of experience in DevOps, Site Reliability Engineering, or...Work at officeWorldwideMonday to FridayFlexible hours$127k - $249k
...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Local areaRemote workWorldwideFlexible hours$194k - $267k
..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$112.5k - $187.5k
...NoticePersonal Information We CollectYour Privacy ChoicesTeam OverviewAt TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$132.1k - $220.1k
Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week on-site... ...GitOps: Mastery of Terraform module design and ArgoCD for managing immutable infrastructure at an enterprise scale.Distributed...Full timeWork at officeLocal areaWorldwide2 days per week- ...the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology team... ...hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-...
- ...and companies, alikeKlover’s engineering team powers one of the fastest... ...systems that prioritize reliability, security, and performance, and... ...candidateAbout the RoleAs a Senior/Staff Site Reliability Engineer, you... ...metrics to our Google-managed Prometheus instance and build...Work at officeImmediate startRemote work
$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$118.3k - $219.8k
Are you excited to lead Site Reliability Engineering teams that keep mission-critical, 24/7 services running reliably and securely?Do you enjoy... ...teams to deliver proactive monitoring, effective incident management, and continuous optimisation. Through strong alignment...Full timeLocal area- ...A leading quantitative trading firm is seeking a Head of Site Reliability Engineering to help scale one of its most critical infrastructure organisations... ...leadership roles within the organisation. What You'll Do Manage a team of Site Reliability Engineers. Drive reliability,...Immediate start
- ...Edward Jones Site Reliability Engineer 100% remote Initial contract is 6 months, but will be a multi year engagement. Position Overview... ...of our systems. You will be responsible for incident management, root cause analysis, and implementing postmortem processes...Contract workRemote work
- ...building and running systems that must perform reliably under real-time market conditions. The culture is highly collaborative, engineering-driven, and focused on continuous... ...a related field 3+ years of experience in site reliability, systems engineering, or technical...
- ...Site Reliability Engineer II Play a key role in ensuring system reliability at one of the world's most iconic and largest financial institutions. As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology...Local area
$190.8k - $267.1k
...influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your... ...more. In this role, you will also take ownership of risk management, ensuring the reliability and performance of our systems....Work experience placementHome officeFlexible hours$125.04k - $187.56k
...Digital and E-commerce, Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible for ensuring the scalability, reliability... ...(SLOs) and service level indicators (SLIs). Build and manage microservices-based platforms leveraging Spring Boot, Java,...Full timeWork at officeRemote workFlexible hours$150k - $155k
...Site Reliability Engineer Hybrid (3 days onsite, 2 days remote) full‑time. No visa sponsorship. Base pay: $150,000 – $155,000 per year, subject... ...large‑scale distributed systems Experience managing infrastructure in public cloud environments like AWS (preferred...Full timeWork experience placementRemote workVisa sponsorship
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
- site reliability engineer remote Chicago, IL
- site reliability engineer Chicago, IL
- site reliability engineer sre Chicago, IL
- junior website developer Chicago, IL
- website content developer Chicago, IL
- on site coordinator Chicago, IL
- website coordinator Chicago, IL
- site leader Chicago, IL
- site recruiter Chicago, IL
- historic site Chicago, IL

