Staff Site Reliability Engineer - Kubernetes
$194k - $267kOkta for Developers
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. We are looking for builders and owners who operate with speed, urgency, and execution excellence. This is an opportunity to do career‑defining work. We're all in on this mission. If you are too, let’s talk. Okta Workforce Identity Cloud (WIC) provides easy, secure access for your workforce so you can focus on other strategic priorities—like reducing costs and doing more for your customers. If you like to be challenged and have a passion for solving large‑scale automation, testing, and tuning problems, we would love to hear from you. The ideal candidate exemplifies the ethic of “If you have to do something more than once, automate it” and can rapidly self‑educate on new concepts and tools. Position Overview The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud‑native applications and services. This position focuses on architecting and managing reliable, scalable, and secure Kubernetes‑based platforms on AWS, ensuring high availability and performance while optimizing costs and automation. The ideal candidate will have hands‑on experience with AWS infrastructure, Kubernetes platform creation, Helm charts, Karpenter scaling, and Istio service mesh. Key Responsibilities Kubernetes Platform Creation: Design, implement, and maintain highly available, scalable, and fault‑tolerant Kubernetes platforms. AWS Infrastructure Management: Build, manage, and optimize AWS cloud infrastructure, including EKS, ECS, S3, VPCs, RDS, IAM, and more. Helm Management: Utilize Helm to automate and streamline the deployment of applications and services to Kubernetes clusters. Karpenter Implementation: Implement and manage Karpenter to dynamically scale Kubernetes clusters in response to workload demands. Istio Service Mesh Management: Configure and manage Istio to provide service‑to‑service communication, security, and observability within Kubernetes clusters. Platform Automation & Scaling: Automate the deployment, scaling, and management of infrastructure and applications with CI/CD pipelines. Incident Management & Troubleshooting: Respond to incidents, troubleshoot, and resolve system issues related to performance, availability, and security. Security & Compliance: Design and implement secure cloud infrastructure with appropriate access controls and compliance frameworks. Documentation & Knowledge Sharing: Create and maintain detailed documentation for Kubernetes platform setup, operational procedures, and best practices. Required Qualifications 4+ years of experience with Kubernetes/Helm. 4+ years of experience with Terraform. 5+ years of experience with AWS. Experience with multi‑region cloud environments. Proven experience with AWS services (EC2, RDS, S3, CloudFormation, IAM, etc.) and solid understanding of cloud-native architectures. Strong expertise in Kubernetes platform creation, management, and optimisation. Hands‑on experience with Helm for Kubernetes application deployment and management. Practical experience with Karpenter for dynamic scaling of Kubernetes clusters and optimising resource usage. Expertise in managing and securing Istio for service mesh, including traffic management, security, and observability. Proficiency in CI/CD pipelines and automation tools (Jenkins, GitLab, CircleCI, Terraform, Ansible, Spinnaker). Strong scripting skills in Python, Bash, or Go. Experience with monitoring, logging, and alerting tools such as Prometheus, Grafana, CloudWatch, and ELK Stack. Preferred Qualifications Understanding of security best practices for cloud platforms and Kubernetes (RBAC, encryption, compliance frameworks). Familiarity with Docker and containerisation principles. Bachelor’s degree in Computer Science, Engineering, or related field (or equivalent professional experience). Certifications (preferred): CKA, CKAD, or AWS Certified DevOps Engineer. Additional Requirements Access to federal environments and/or protected federal data, with required U.S. Person status documentation. Requires in‑person onboarding and travel to the San Francisco, CA HQ office or Chicago office during the first week of employment. Salary and Benefits Annual base salary range for this position in the San Francisco Bay area: $194,000—$267,000 USD. For other locations (California excluding Bay Area, Colorado, Illinois, New York, and Washington): $174,000—$214,000 USD. Salary is based on skills, qualifications, experience, and work location. Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (PTO and parental leave) in accordance with applicable plans and policies. The Okta Experience Supporting Your Well‑Being Driving Social Impact Developing Talent and Fostering Connection & Community Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and conviction records, consistent with applicable laws. If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding, please use this Form to request an accommodation. Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT) as defined in New York City Local Law 144. For more information, see our NYC AEDT Notice. Okta is committed to complying with applicable data privacy and security laws and regulations. For more information, see our Personnel and Job Candidate Privacy Notice at #J-18808-Ljbffr Okta for Developers
$40 per hour
A technology solutions provider is seeking a remote Junior SRE/DevOps Engineer. The ideal candidate should have foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes. Responsibilities include gaining experience in a DevOps-driven environment. Applicants...SuggestedRemote jobLong term contractInternship- ...world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics... ...modern technologies such as Databricks, Snowflake, AWS, and Kubernetes Collaborates with other software engineers and teams to...SuggestedWork at office
- ...JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining... ...position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability... ...of containerization (Docker, Kubernetes) and orchestration frameworks Experience...SuggestedWork at office
- ...Site Reliability Engineer (SRE) The successful applicant may be performing work in FedRAMP High or IL-5 environments, and therefore, must... ...experience operating production services using Docker and Kubernetes in cloud or hybrid environments. Proficiency in one or...SuggestedPermanent employmentWorldwideShift work
- ...Mandatory Skills: AWS/Azure/GCP (GCP is not used very much ). Kubernetes /Helm,Docker,Gitlab,Grafana,Cyberark/Hashicorp Vault, Terraform etc. Experience utilizing Java, Perl, Python, Go and scripting experience in Shell and Perl to automate reports and monitor...Suggested
$100k - $115k
...Internal Developer Platform Engineer Analytic Partners is a global... ...customers and optimizing for reliability, usability, and delivery... ...experience in Platform Engineering, Site Reliability Engineering,... ...platforms such as Docker and Kubernetes. ~ In-depth experience with...Temporary work- ...Site Reliability Engineer Location- Wilmington De, Washington DC, Dallas, TX (Onsite Position) Full time position Minimum Qualifications... ..., Terraform, CloudFormation, Ansible, Docker, Packer, and Kubernetes. Strong understanding of cloud platforms (AWS preferred...Full time
- ...Site Reliability Engineer We are looking for a Site Reliability Engineer for our client location in Dallas TX with the following skills: Java Spring boot, Kubernetes, and eCommerce experience required. Key responsibilities include working with the applications, engineering...Work at office
- ...Job Position:- Site Reliability Engineer Duration:- Long Term Client:- UPS This is a Hybrid Work Model (3x a week... ...Proficiency in Google Cloud services (Compute Engine, Kubernetes Engine, Cloud Storage, BigQuery, Pub/Sub, etc.). Familiarity...
- ...Senior Site Reliability Engineer Come join a growing bank at the heart of the innovation, technology... ...both technical and non-technical staff Familiar with system hardening and... ...high-growth environment ~ Hands on Kubernetes skills is nice to have ~ Cybersecurity...
- ...interview process. Lantern is seeking an experienced Senior Site Reliability Engineer to champion the reliability, availability, and performance... ...and reliability testing Experience with Azure Kubernetes Service and containerized workloads Relevant certifications...Temporary workFlexible hours
- ...applications are highly available, reliable, and performant at a global... ...Computer Science or related Engineering field required. Master's... ...year of experience in Mesos, Kubernetes, OpenShift and/or Deis or... ...1 year of lead experience of site reliability engineering team...Contract workWork at office
$72.1k - $158.62k
...DevOps Engineer We're building a world of health around every individual — shaping... ...to improve deployment velocity, system reliability, and operational efficiency through DevOps... ...(Docker) and orchestration (Kubernetes/AKS) ~ Strong understanding of version...Hourly payFull timeTemporary work- ...infrastructure and applications with high reliability, resiliency, performance & quality, and... .../playbooks; and, Using Chaos Engineering to test the robustness of the systems and... ...technologies and orchestration (Docker, Kubernetes-AKS, EKS, GKE) ~3+ year implementing...
- ...DevOps Engineer With Kubernetes Experience is Must Sonsoft, Inc. is a USA based corporation duly organized under the laws of the Commonwealth of Georgia. Sonsoft Inc. is growing at a steady pace specializing in the fields of Software Development, Software Consultancy...Contract workH1b
- ...financial services organization, is seeking an experienced Senior Site Reliability Engineers (SRE) to join a newly formed Application Reliability team.... ...runbooks Deep experience with container orchestration (Kubernetes/EKS highly preferred) and cloud‑native ecosystems Track...Contract work
- Senior Site Reliability Engineer Lantern is seeking an experienced Senior Site Reliability Engineer to champion the reliability, availability... ...engineering and reliability testing. Experience with Azure Kubernetes Service and containerized workloads. Relevant...Temporary workFlexible hours3 days per week
$152k - $195k
...Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes‑based infrastructure and CI/CD systems. You will also own the infrastructure...$130k
...Distributed Systems Software Engineer, Python / Go Join to apply for the Distributed Systems... ...based on Juju, Terraform, OpenStack, Kubernetes when deployed under highly diverse... ...approaches and infrastructure for validating reliability, performance, and resilience of cloud...Full timeLocal areaRemote workWorldwide$85 - $90 per hour
...Role: Senior SRE Engineer Location: Dallas / Fort Worth, Texas Rate: up to $85-$9... ...Structure: 8 Month contract *** 4 days on-site *** -- We have a great new... ...container orchestration platforms such as Kubernetes. ~ Experience using IAC tools such as...Hourly payContract workWork experience placement$116.36k - $155.15k
...and experience in system architecture and engineering disciplines. Specific technical... ...Cloud Platform. Support and troubleshoot Kubernetes clusters and containers. Support, troubleshoot... ...due diligence activities including site surveys, design, design review, bill of...Full timeTemporary workRemote work1 day per week$140k - $200k
...– Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and... ...: Proficiency in deploying high availability applications on Kubernetes What We Offer A dynamic environment where your contributions...Work at office- ...assessment capabilities in the most demanding environments. Join a global team of 35 000 engineers, software developers, and cyber experts who turn complex challenges into reliable, next generation systems that keep warfighters ahead of emerging threats. Your talent will...WorldwideFlexible hours
- ...generative AI and cloud-native platforms to advanced release engineering practices, our teams are redefining how financial technology... ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages...Full timeH1bWork at officeRemote workVisa sponsorshipFlexible hours2 days per week3 days per week
- .... Collaborate with cross-functional teams to identify reliability risks and improve system architecture. Develop and enhance... ...environments. Experience in customer-facing roles. Certifications in Site Reliability Engineering, DevOps, or Performance Engineering....
- ...Role: Site Reliability Engineer 6+ months Contract role Remote About the Role We are looking for a dynamic and accomplished Site Reliability Engineer (SRE) who excels at solving complex reliability challenges and thrives in high-impact environments....Contract workRemote work
$57k - $113k
...Site Reliability Engineer Huntington will not sponsor applicants for this position for immigration benefits, including but not limited to assisting with obtaining work permission for F-1 students, H-1B professionals, O-1 workers, TN workers, E-3 workers, among other...Full timeH1bWork at officeRemote workWork from homeFlexible hours$79.31k - $158.62k
...Software Development Engineer, Site Reliability Engineering (SRE) We're building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you'll be surrounded by passionate colleagues who...Hourly payFull timeTemporary work- ...Job Title: Site Reliability Engineer Location: Dallas TX (HYBRID) Duration :Full Time Job Description: Skill: Site Reliability... ...-wide applications. • ssists in the development staff in understanding the software products and any enhancements...Full timeWork at office
$50 - $53 per hour
...area onsite at the project, significantly reducing and/or eliminating the demands to travel. Key Responsibilities: Site Reliability Engineers are expected to be able to drive technology triage efforts to completion by assisting with restoral steps, identifying root...Hourly payLive inWork at officeLocal areaFlexible hours3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer - Kubernetes. Be the first to apply!
- IT site lead Syracuse, NY
- junior website developer Syracuse, NY
- site safety Syracuse, NY
- site services specialist Syracuse, NY
- site leader Syracuse, NY
- on-site clinical research associate (traveling/remote) Syracuse, NY
- assistant chief engineer
- engineering administrative assistant
- staff chemical engineer
- staff process engineer



