Manager, Site Reliability Engineering
$204k - $306kOkta
Secure Every Identity, from AI to HumanIdentity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity, from AI to HumanIdentity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.**This position requires 2 days a week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput and 99.999 availability. We're looking for a technical leader to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and tooling. As the Manager of Infrastructure Platform and Shared Services, you will oversee multiple teams focused on Edge networking, K8s platform, CI/CD, Observability, automation platform & tooling. What you’ll be doing Managing a team of SRE’s supporting various workloads and teams that support our IDaaS platform.Drive the microservice journey, DevOps maturity, and workload reliability in tandem with architects and teams across the organization.Accelerate the velocity of SRE and product engineering by developing powerful tooling, intuitive self-service capabilities, and robust self-healing patterns.Lead, mentor, and grow a high-performing team of engineers and managers across platform, infrastructure, and shared services domains.Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints.Improve SDLC processes for Cloud infrastructure as a code, including the maturity of CI/CD pipelines, change and release management Manage service and business expectations and prioritize resource allocationMaintain a deep knowledge of industry best practices, evolving trends, and technologiesWhat you’ll bring to the role3+ years of experience in technical leadership & people management Extensive experience using Agile and DevOps methodologies to build product infrastructure and shared service at scaleExperience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS. Experience supporting a multi-Cloud environment will be a plus.Strong expertise in cloud-native architectures, containerization (Kubernetes), IaC (Terraform), and CI/CD pipelinesStrong background and hands-on experience in SW development, PaaS and automationDeep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment.Effective verbal, written communication and interpersonal skillsComputer Science Degree or related degree or equivalent experience Additional requirements:This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.#LI-HybridP24518_3462184Below is the annual base salary range for candidates located in San Francisco Bay Area. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: The annual base salary range for this position for candidates located in the San Francisco Bay area is between: $204,000—$306,000 USDThe Okta ExperienceSupporting Your Well-BeingDriving Social Impact Developing Talent and Fostering Connection + CommunityWe are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.
$150k - $220k
...innovators in this way. The Role: As an engineering organization, we pride ourselves on... ...engineering as a creative activity. Engineering managers enable engineers to do their best work... ..., mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team...SuggestedLocal area$210.38k - $243.21k
Manager, Site Reliability Engineer (Hybrid in South San Francisco)About the RoleWe are seeking an experienced and hands-on Site Reliability Engineering (SRE) Manager to lead our Site Operations and infrastructure initiatives. This role is responsible for ensuring the reliability...Suggested- ...builds the platforms and tooling that help engineering teams develop, deploy, and operate... ...default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll... ...habits and tooling.Architect and manage the SLO and error-budget framework, empowering...SuggestedPermanent employmentWork experience placementWork at officeLocal area
$113.4k - $162k
...conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd,... ...GitHub, Terraform, Ansible, or similar tools to build and manage cloud infrastructure efficiently.Incident Management Expert...SuggestedTemporary work$117k - $209.33k
...Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable,... ...such as SLOs/SLIs, production readiness, incident management, observability, resilience testing, and toil reduction. Success...SuggestedFull timeFor contractors$114.3k - $235.32k
...advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-... ...and HelmSupporting infrastructure provisioning and change management through Terraform/TerragruntBuilding and supporting CI/CD...Work at officeLocal areaRelocationRelocation package$127k - $249k
The TeamPlatform Engineering sits within SRE and builds the core infrastructure... ...Engineering, the Fabric team manages the global network substrate... ...role in engineering the reliable, globally connected, multi-... ...seeking a talented Senior Site Reliability Engineer (SRE) with...Local areaRemote workWorldwideFlexible hours$152.5k - $205k
...everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common... ..., observability, access controls, auditability, and cost management. You will troubleshoot production issues, document operational...Flexible hours$190.8k - $267.1k
...while helping Reddit grow its business. The reliability of our Ads systems directly impacts... ...Reliability team partners closely with Ads Engineering to improve reliability, scalability,... ...advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate,...For contractorsWork experience placement$106k - $130k
...ineligible for employment Visa sponsorship.Role Summary The Senior Site Reliability Engineer applies software engineering and systems engineering... ...as Code, automation, testing, incident response, capacity management, resilience, and operational readiness. Identify recurring...Hourly payFull timeImmediate startVisa sponsorshipWork visaFlexible hours$127k - $249k
Platform Engineering is the department within SRE that is responsible for a range of critical... ...and alerting systems.The Fleet Management team provides the core runtime environment... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...Work at officeLocal areaRemote workWorldwideFlexible hours- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and... ...cross-functionallyNice-to-HaveExperience with cloud and managed services (e.g. AWS)Experience supporting data-intensive platforms...
$148.5k - $223.9k
...future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with... ...eliminate toil and improve operational efficiency.Incident Management: Lead the coordinated response to incidents as an...Full timeWorldwideWeekend work$152.5k - $205k
...a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate... ...experience, including authoring reusable modules, managing state and environments, and delivering infrastructure changes...Flexible hours- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology,... ...exposure (training can be provided)Cloud/SaaS experienceMemory management and dump analysis (Java heap dump analysis preferred)ITSM/...
- ...principles to see it in full.About the teamThe Engineering team at Airwallex is a diverse group of... ..., working together to build scalable, reliable, and secure products that empower... ...Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work closely...Temporary workLocal area
$55k - $151.47k
...LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in... ...data integrity and accessibility- Leading incident management and resolution efforts to maintain operational continuityWhat...Full timeH1b$194k - $267k
..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco, CA.... ...Remote unavailable. Modality: On-Site only. Must live within... ...scalability, performance, and reliability across environments. What You... ...powers AI workloads at scale Manage and automate GPU compute clusters...Full timeRemote workRelocationRelocation package$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$195k - $257.5k
...is a stakeholder.What you’ll be responsible for:As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and... ...on:Operate and scale production blockchain infrastructure, managing full nodes across networks such as Arc, Ethereum, Solana,...Flexible hours- ...culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of... ...~6+ years of experience in Site Reliability Engineering, managing infrastructure and services at scale ~ History of end-to...Immediate startRemote workWorldwide
$210k - $240k
...Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual... ..., deployment automation, rollback mechanisms, and config management Implement and maintain monitoring, alerting, and incident...Full time- ...daily users while enabling our engineering teams to ship fast. You'll... ...automation and tooling that improves reliability and partnering with... ...reliability best practices Manage and optimize our infrastructure... ...you'll bring ~5+ years in Site Reliability Engineering, DevOps...Work at officeWork from home
$153k - $191.3k
...manufacturing, data processing, and software engineering, our office is a truly inspiring mix of... ...environments, to guarantee the reliability, scalability, and availability of our services... ..., particularly resource optimization, management, and cluster tuning in a constrained...Full timeTemporary workFor contractorsWork at officeLocal areaRemote workHome office3 days per week- ...and the U.S. Special Forces. The Role We're hiring a Site Reliability Engineer to own the operational health of our connected sensor... ...Systems Builder — Close the Loop Build and maintain fleet management systems: OTA update pipelines, device health tracking, remote...Remote work
$189k - $283.6k
...Afterpay is transforming the way customers manage their spending over time. TIDAL is a... ...proactively and reactively improve the reliability of Block's platform and critical infrastructure... ...strong desire to perform and grow as an engineer ~5+ years of software development...Full timeRelocation packageFlexible hoursShift work- ...for As an SRE at Wordbricks, you will keep our systems fast, reliable, and boring. You'll own the infrastructure and operations... ...systems Build and maintain CI/CD, observability, and alerting Manage cloud infrastructure across Cloudflare, AWS, and Vercel Lead...Remote workFlexible hours
- ...infrastructure, behind their own controls, with the reliability and operational clarity they would... ...that make this possible: Retool Cloud, managed single tenant environments, BYOC (bring-... ...TAMs to trust. Partner with product engineers on infrastructure requirements for new...
$98.58k - $138.02k
...office locations: Austin, TX; Irvine, CA; or Akron, OH. Site Reliability Engineer II will be responsible for supporting, enhancing, and... ...Terraform, Ansible, or CloudFormation. Work within change management protocols to provide maximum uptime for production systems...Work at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
- site reliability engineer sre San Francisco, CA
- site reliability engineer San Francisco, CA
- site reliability engineer remote San Francisco, CA
- official site San Francisco, CA
- site services specialist San Francisco, CA
- construction site safety San Francisco, CA
- IT site lead San Francisco, CA
- site recruiter San Francisco, CA
- site leader San Francisco, CA
- site safety San Francisco, CA


