Manager, Site Reliability Engineering
$204k - $306kOkta
Secure Every Identity, from AI to HumanIdentity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity, from AI to HumanIdentity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.**This position requires 2 days a week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput and 99.999 availability. We're looking for a technical leader to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and tooling. As the Manager of Infrastructure Platform and Shared Services, you will oversee multiple teams focused on Edge networking, K8s platform, CI/CD, Observability, automation platform & tooling. What you’ll be doing Managing a team of SRE’s supporting various workloads and teams that support our IDaaS platform.Drive the microservice journey, DevOps maturity, and workload reliability in tandem with architects and teams across the organization.Accelerate the velocity of SRE and product engineering by developing powerful tooling, intuitive self-service capabilities, and robust self-healing patterns.Lead, mentor, and grow a high-performing team of engineers and managers across platform, infrastructure, and shared services domains.Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints.Improve SDLC processes for Cloud infrastructure as a code, including the maturity of CI/CD pipelines, change and release management Manage service and business expectations and prioritize resource allocationMaintain a deep knowledge of industry best practices, evolving trends, and technologiesWhat you’ll bring to the role3+ years of experience in technical leadership & people management Extensive experience using Agile and DevOps methodologies to build product infrastructure and shared service at scaleExperience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS. Experience supporting a multi-Cloud environment will be a plus.Strong expertise in cloud-native architectures, containerization (Kubernetes), IaC (Terraform), and CI/CD pipelinesStrong background and hands-on experience in SW development, PaaS and automationDeep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment.Effective verbal, written communication and interpersonal skillsComputer Science Degree or related degree or equivalent experience Additional requirements:This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.#LI-HybridP24518_3462184Below is the annual base salary range for candidates located in San Francisco Bay Area. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: The annual base salary range for this position for candidates located in the San Francisco Bay area is between: $204,000—$306,000 USDThe Okta ExperienceSupporting Your Well-BeingDriving Social Impact Developing Talent and Fostering Connection + CommunityWe are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.
$147k - $202.4k
...let's talk.Our company is seeking a highly skilled Senior Site Reliability Engineer to join our team. We are a SaaS company specializing in securing... ...IaC): Deep experience with Terraform for provisioning and managing cloud infrastructure and services.Continuous Delivery:...SuggestedWork at officeLocal areaWorldwideFlexible hoursShift work$165k - $225.6k
...infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build, improve, and...SuggestedPermanent employmentLocal areaWorldwideFlexible hours$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at...SuggestedPermanent employmentWork at officeLocal areaWorldwideFlexible hours$194k - $267k
...let's talk.The TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission... ...balancing, ingress, TLS, service networking, and traffic management.Strategic experience designing comprehensive observability...SuggestedLocal areaWorldwideFlexible hours$194k - $267k
..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses...SuggestedPermanent employmentWork at officeLocal areaWorldwideFlexible hours- ...Robinhood is seeking a Staff Software Engineer for its Storage Platform in Bellevue, WA. You will design and evolve core storage infrastructure... ...databases, key‑value systems, and caching—while driving reliability, performance, and cost efficiency across multi‑region...
- ...System Reliability Engineer (SRE)At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we... ...and systems behind all of T-Mobile's IT services, including management of scalability, availability, latency, performance, security...Full timeTemporary workPart timeWork experience placementFlexible hours
- Technical/Functional Skills Windows Servers, Digital: Microsoft Azure Windows Powershell, Digital: DevOps Roles & Responsibilities Windows Server 2012 -2019 Administration Microsoft Azure Azure AAD DFSR, DHCP DNS, KMS, WSUS TCP/IP Hyper-V High Availability Clusters ...
$232k - $319k
...scale the service with great people and reliable, cost-effective, and efficient infrastructure... ..., processes, and tooling. As the Sr. Manager of Infrastructure Platform and Shared... ...serviceAccelerate the velocity of SRE and product engineering by developing robust platforms, powerful...Permanent employmentLocal areaWorldwideFlexible hours$160k - $210k
...Dynamic Deals through their preferred DSP, leveraging our managed service DSP, or utilizing our industry-first ContextGPT product... .... Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management...Work at officeImmediate startRemote workWork from home- ...UNSTOPPABLE for our employees!Job Overview:T-Mobile's Forward Deployment Engineering practice embeds engineers directly inside our business units to ship production agentic AI solutions. The Sr Manager, Forward Deployment Engineering leads the management layer of that...Full timeTemporary workPart timeWork experience placementLocal areaFlexible hours
$170k - $220k
Who We're Looking ForWe’re looking for a hands-on, high-agency Site Reliability Engineer to help shape and scale the reliability layer of our stack. You'll own the release pipeline end-to-end — managing daily releases, weekly deploys, and hotfixes — while also automating...- ...and private sectors, and the Department of Defense. Our services span all aspects of business, providing a holistic approach for managing an organization.Job DescriptionResponsible for the monitoring, provisioning, resiliency, and customer interactions. Working extensively...
$119.8k - $234.7k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole... ...EngineeringDiscipline: Site Reliability EngineeringCompany:... ...As a Senior Site Reliability Engineer, you will lead reliability improvements... ...eligibility requirements.For manager-level roles, a Tier 5 (T5)...Ongoing contractLocal area3 days per week$134.25k - $214.8k
...matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on... ...cloud-native distributed systems.Employ technical project management skills, with the ability to properly scope, plan, and define...Work experience placementWork at officeRemote workFlexible hours$143k - $191k
...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental... ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and...Full timeTemporary workWork experience placementImmediate start$102.1k - $202.2k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewAre... ...no further than the Microsoft Defender engineering team. We are looking for a Site...Ongoing contractLocal area3 days per week$165k - $230k
...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology... .... RESPONSIBILITIES: Develop automation to deploy and manage compute resources both on-premises and in the cloudDeploy and...Permanent employmentTemporary workWork at officeImmediate startMonday to FridayWeekend work$167.7k - $245.2k
...maintaining our FedRAMP offering. Your ImpactAs a FedRAMP Site Reliability Engineer(SRE), you will lead the operations and architecture of our... ...will implement an "everything-as-code" strategy to design and manage resilient AWS cloud-native services while maintaining...Full timeTemporary workWork at officeLocal areaFlexible hours2 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range... ...observability and alerting systems.The Fleet Management team provides the core runtime... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...Work at officeLocal areaRemote workWorldwideFlexible hours$102.1k - $202.2k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole... ...EngineeringDiscipline: Site Reliability EngineeringCompany:... ...secure, resilient, and easier to manage for customers operating in highly... ..., you will collaborate with engineers across disciplines to deliver...Ongoing contractWork experience placementLocal areaRemote work3 days per week$165k - $270k
...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building... ...infrastructure to support a multi-region environment Manage petabyte scale bare metal compute clusters Closely collaborate...Permanent employmentTemporary workWorldwideWeekend work$165k - $230k
...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging... ...alerting infrastructure to support a multi-region environment Manage petabyte scale bare metal compute clusters Closely...Permanent employmentTemporary workWorldwideWeekend work$134.25k - $214.8k
...where you matter.Your ImpactAre you an engineer who gets excited about the challenge... ...the Observability team within Axon's Site Reliability organization — a focused team responsible... ...ArgoCD, and Helm — including capacity management, cybersecurity requirements and...Work experience placementWork at officeRemote work- ...limits of what's possible.As a Lead Software Engineer at JPMorganChase within the Enterprise... ...transaction processing and asset management. We offer a competitive total rewards package... ...comprehensive health care coverage, on-site health and wellness centers, a...
- .... We are responsible for the reliability of all the company's major data... ..., services, and query engines. We serve business needs across... ...utilized effectively.- Incident Management: Lead efforts to troubleshoot... ...emerging technologies related to site reliability and...
$127k - $249k
We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure... ...Azure, GCP), including network and compute security, identity management, and cloud security posture management (CSPM)Automation and...Local areaRemote workWorldwideFlexible hours- ...Site Reliability Engineer Comtech is a woman-owned small business founded in 1998 and headquartered in Reston, VA. We offer IT solutions across the disciplines of program/project management, applications development, infrastructure, Cyber security, and enterprise content...
$94k - $142.3k
...customer and partner enablement, applications engineering, infrastructure, collaboration,... ...delivered globally, at scale, sustainably.As a Site Reliability Operations Engineer you'll be part of... ...’ll Actually Be Doing...Respond to and manage major incidents affecting internal...Full timeShift work- ...JPMorganChase in Seattle seeks a Lead Software Engineer to join the Enterprise Technology, Infrastructure Platforms team. You will act as a core technical contributor, delivering trusted, scalable technology across multiple domains, while guiding AI-assisted engineering...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
- site reliability engineer sre Bellevue, WA
- site reliability engineer Bellevue, WA
- site services specialist Bellevue, WA
- junior website developer Bellevue, WA
- official site Bellevue, WA
- on site coordinator Bellevue, WA
- site leader Bellevue, WA
- historic site Bellevue, WA
- construction site safety Bellevue, WA
- website coordinator Bellevue, WA


