Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Site Reliability Engineer - Kubernetes

$194k - $267k

Okta for Developers

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. We are looking for builders and owners who operate with speed, urgency, and execution excellence. This is an opportunity to do career‑defining work. We're all in on this mission. If you are too, let’s talk. Okta Workforce Identity Cloud (WIC) provides easy, secure access for your workforce so you can focus on other strategic priorities—like reducing costs and doing more for your customers. If you like to be challenged and have a passion for solving large‑scale automation, testing, and tuning problems, we would love to hear from you. The ideal candidate exemplifies the ethic of “If you have to do something more than once, automate it” and can rapidly self‑educate on new concepts and tools. Position Overview The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud‑native applications and services. This position focuses on architecting and managing reliable, scalable, and secure Kubernetes‑based platforms on AWS, ensuring high availability and performance while optimizing costs and automation. The ideal candidate will have hands‑on experience with AWS infrastructure, Kubernetes platform creation, Helm charts, Karpenter scaling, and Istio service mesh. Key Responsibilities Kubernetes Platform Creation: Design, implement, and maintain highly available, scalable, and fault‑tolerant Kubernetes platforms. AWS Infrastructure Management: Build, manage, and optimize AWS cloud infrastructure, including EKS, ECS, S3, VPCs, RDS, IAM, and more. Helm Management: Utilize Helm to automate and streamline the deployment of applications and services to Kubernetes clusters. Karpenter Implementation: Implement and manage Karpenter to dynamically scale Kubernetes clusters in response to workload demands. Istio Service Mesh Management: Configure and manage Istio to provide service‑to‑service communication, security, and observability within Kubernetes clusters. Platform Automation & Scaling: Automate the deployment, scaling, and management of infrastructure and applications with CI/CD pipelines. Incident Management & Troubleshooting: Respond to incidents, troubleshoot, and resolve system issues related to performance, availability, and security. Security & Compliance: Design and implement secure cloud infrastructure with appropriate access controls and compliance frameworks. Documentation & Knowledge Sharing: Create and maintain detailed documentation for Kubernetes platform setup, operational procedures, and best practices. Required Qualifications 4+ years of experience with Kubernetes/Helm. 4+ years of experience with Terraform. 5+ years of experience with AWS. Experience with multi‑region cloud environments. Proven experience with AWS services (EC2, RDS, S3, CloudFormation, IAM, etc.) and solid understanding of cloud-native architectures. Strong expertise in Kubernetes platform creation, management, and optimisation. Hands‑on experience with Helm for Kubernetes application deployment and management. Practical experience with Karpenter for dynamic scaling of Kubernetes clusters and optimising resource usage. Expertise in managing and securing Istio for service mesh, including traffic management, security, and observability. Proficiency in CI/CD pipelines and automation tools (Jenkins, GitLab, CircleCI, Terraform, Ansible, Spinnaker). Strong scripting skills in Python, Bash, or Go. Experience with monitoring, logging, and alerting tools such as Prometheus, Grafana, CloudWatch, and ELK Stack. Preferred Qualifications Understanding of security best practices for cloud platforms and Kubernetes (RBAC, encryption, compliance frameworks). Familiarity with Docker and containerisation principles. Bachelor’s degree in Computer Science, Engineering, or related field (or equivalent professional experience). Certifications (preferred): CKA, CKAD, or AWS Certified DevOps Engineer. Additional Requirements Access to federal environments and/or protected federal data, with required U.S. Person status documentation. Requires in‑person onboarding and travel to the San Francisco, CA HQ office or Chicago office during the first week of employment. Salary and Benefits Annual base salary range for this position in the San Francisco Bay area: $194,000—$267,000 USD. For other locations (California excluding Bay Area, Colorado, Illinois, New York, and Washington): $174,000—$214,000 USD. Salary is based on skills, qualifications, experience, and work location. Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (PTO and parental leave) in accordance with applicable plans and policies. The Okta Experience Supporting Your Well‑Being Driving Social Impact Developing Talent and Fostering Connection & Community Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and conviction records, consistent with applicable laws. If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding, please use this Form to request an accommodation. Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT) as defined in New York City Local Law 144. For more information, see our NYC AEDT Notice. Okta is committed to complying with applicable data privacy and security laws and regulations. For more information, see our Personnel and Job Candidate Privacy Notice at #J-18808-Ljbffr Okta for Developers

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Site Reliability Engineer - Kubernetes in Syracuse, NY vacancy
  • $40 per hour

    A technology solutions provider is seeking a remote Junior SRE/DevOps Engineer. The ideal candidate should have foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes. Responsibilities include gaining experience in a DevOps-driven environment. Applicants... 
    Suggested
    Remote job
    Long term contract
    Internship

    BayOne Solutions

    Richardson, TX
    2 days ago
  •  ...world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics...  ...modern technologies such as Databricks, Snowflake, AWS, and Kubernetes Collaborates with other software engineers and teams to... 
    Suggested
    Work at office

    J.P. Morgan

    Dallas, TX
    22 days ago
  •  ...JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining...  ...position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability...  ...of containerization (Docker, Kubernetes) and orchestration frameworks Experience... 
    Suggested
    Work at office

    J.P. Morgan

    Dallas, TX
    22 days ago
  •  ...Site Reliability Engineer (SRE) The successful applicant may be performing work in FedRAMP High or IL-5 environments, and therefore, must...  ...experience operating production services using Docker and Kubernetes in cloud or hybrid environments. Proficiency in one or... 
    Suggested
    Permanent employment
    Worldwide
    Shift work

    Webex Events (formerly Socio)

    Richardson, TX
    4 days ago
  •  ...Mandatory Skills: AWS/Azure/GCP (GCP is not used very much ). Kubernetes /Helm,Docker,Gitlab,Grafana,Cyberark/Hashicorp Vault, Terraform etc. Experience utilizing Java, Perl, Python, Go and scripting experience in Shell and Perl to automate reports and monitor... 
    Suggested

    Omni Inclusive

    Dallas, TX
    4 days ago
  • $100k - $115k

     ...Internal Developer Platform Engineer Analytic Partners is a global...  ...customers and optimizing for reliability, usability, and delivery...  ...experience in Platform Engineering, Site Reliability Engineering,...  ...platforms such as Docker and Kubernetes. ~ In-depth experience with... 
    Temporary work

    Analytic Partners

    Dallas, TX
    1 day ago
  •  ...Site Reliability Engineer Location- Wilmington De, Washington DC, Dallas, TX (Onsite Position) Full time position Minimum Qualifications...  ..., Terraform, CloudFormation, Ansible, Docker, Packer, and Kubernetes. Strong understanding of cloud platforms (AWS preferred... 
    Full time

    Yochana

    Dallas, TX
    3 days ago
  •  ...Site Reliability Engineer We are looking for a Site Reliability Engineer for our client location in Dallas TX with the following skills: Java Spring boot, Kubernetes, and eCommerce experience required. Key responsibilities include working with the applications, engineering... 
    Work at office

    STIAOS Technologies

    Dallas, TX
    2 days ago
  •  ...Job Position:- Site Reliability Engineer Duration:- Long Term Client:- UPS This is a Hybrid Work Model (3x a week...  ...Proficiency in Google Cloud services (Compute Engine, Kubernetes Engine, Cloud Storage, BigQuery, Pub/Sub, etc.). Familiarity... 

    Sparktek

    Farmers Branch, TX
    1 day ago
  •  ...Senior Site Reliability Engineer Come join a growing bank at the heart of the innovation, technology...  ...both technical and non-technical staff Familiar with system hardening and...  ...high-growth environment ~ Hands on Kubernetes skills is nice to have ~ Cybersecurity... 

    Professional Recruiters

    Dallas, TX
    1 day ago
  •  ...interview process. Lantern is seeking an experienced Senior Site Reliability Engineer to champion the reliability, availability, and performance...  ...and reliability testing Experience with Azure Kubernetes Service and containerized workloads Relevant certifications... 
    Temporary work
    Flexible hours

    Lantern

    Dallas, TX
    4 days ago
  •  ...applications are highly available, reliable, and performant at a global...  ...Computer Science or related Engineering field required. Master's...  ...year of experience in Mesos, Kubernetes, OpenShift and/or Deis or...  ...1 year of lead experience of site reliability engineering team... 
    Contract work
    Work at office

    3B Staffing LLC

    Irving, TX
    5 days ago
  • $72.1k - $158.62k

     ...DevOps Engineer We're building a world of health around every individual — shaping...  ...to improve deployment velocity, system reliability, and operational efficiency through DevOps...  ...(Docker) and orchestration (Kubernetes/AKS) ~ Strong understanding of version... 
    Hourly pay
    Full time
    Temporary work

    Oak St. Health

    Garland, TX
    5 days ago
  •  ...infrastructure and applications with high reliability, resiliency, performance & quality, and...  .../playbooks; and, Using Chaos Engineering to test the robustness of the systems and...  ...technologies and orchestration (Docker, Kubernetes-AKS, EKS, GKE) ~3+ year implementing... 

    Software Technology Inc

    Dallas, TX
    5 hours ago
  •  ...DevOps Engineer With Kubernetes Experience is Must Sonsoft, Inc. is a USA based corporation duly organized under the laws of the Commonwealth of Georgia. Sonsoft Inc. is growing at a steady pace specializing in the fields of Software Development, Software Consultancy... 
    Contract work
    H1b

    SonSoft

    Dallas, TX
    23 hours ago
  •  ...financial services organization, is seeking an experienced Senior Site Reliability Engineers (SRE) to join a newly formed Application Reliability team....  ...runbooks Deep experience with container orchestration (Kubernetes/EKS highly preferred) and cloud‑native ecosystems Track... 
    Contract work

    Optomi

    Dallas, TX
    2 days ago
  • Senior Site Reliability Engineer Lantern is seeking an experienced Senior Site Reliability Engineer to champion the reliability, availability...  ...engineering and reliability testing. Experience with Azure Kubernetes Service and containerized workloads. Relevant... 
    Temporary work
    Flexible hours
    3 days per week

    Lantern

    Dallas, TX
    2 days ago
  • $152k - $195k

     ...Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes‑based infrastructure and CI/CD systems. You will also own the infrastructure... 

    SecurityScorecard

    Dallas, TX
    5 days ago
  • $130k

     ...Distributed Systems Software Engineer, Python / Go Join to apply for the Distributed Systems...  ...based on Juju, Terraform, OpenStack, Kubernetes when deployed under highly diverse...  ...approaches and infrastructure for validating reliability, performance, and resilience of cloud... 
    Full time
    Local area
    Remote work
    Worldwide

    Canonical

    Syracuse, NY
    4 days ago
  • $85 - $90 per hour

     ...Role:  Senior SRE Engineer  Location: Dallas / Fort Worth, Texas Rate: up to $85-$9...  ...Structure: 8 Month contract *** 4 days on-site *** -- We have a great new...  ...container orchestration platforms such as Kubernetes. ~ Experience using IAC tools such as... 
    Hourly pay
    Contract work
    Work experience placement

    CorGTA

    Dallas, TX
    12 days ago
  • $116.36k - $155.15k

     ...and experience in system architecture and engineering disciplines. Specific technical...  ...Cloud Platform. Support and troubleshoot Kubernetes clusters and containers. Support, troubleshoot...  ...due diligence activities including site surveys, design, design review, bill of... 
    Full time
    Temporary work
    Remote work
    1 day per week

    Lumen

    Syracuse, NY
    4 days ago
  • $140k - $200k

     ...– Speechify has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and...  ...: Proficiency in deploying high availability applications on Kubernetes What We Offer A dynamic environment where your contributions... 
    Work at office

    Speechify

    Syracuse, NY
    4 days ago
  •  ...assessment capabilities in the most demanding environments. Join a global team of 35 000 engineers, software developers, and cyber experts who turn complex challenges into reliable, next generation systems that keep warfighters ahead of emerging threats. Your talent will... 
    Worldwide
    Flexible hours

    Lockheed Martin

    Syracuse, NY
    1 day ago
  •  ...generative AI and cloud-native platforms to advanced release engineering practices, our teams are redefining how financial technology...  ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages... 
    Full time
    H1b
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours
    2 days per week
    3 days per week

    GMAC Financial Services

    Irving, TX
    2 days ago
  •  .... Collaborate with cross-functional teams to identify reliability risks and improve system architecture. Develop and enhance...  ...environments. Experience in customer-facing roles. Certifications in Site Reliability Engineering, DevOps, or Performance Engineering.... 

    Cynet Systems

    Dallas, TX
    1 day ago
  •  ...Role: Site Reliability Engineer 6+ months Contract role Remote About the Role We are looking for a dynamic and accomplished Site Reliability Engineer (SRE) who excels at solving complex reliability challenges and thrives in high-impact environments.... 
    Contract work
    Remote work

    ECHO IT SOLUTIONS INC .

    Farmers Branch, TX
    1 day ago
  • $57k - $113k

     ...Site Reliability Engineer Huntington will not sponsor applicants for this position for immigration benefits, including but not limited to assisting with obtaining work permission for F-1 students, H-1B professionals, O-1 workers, TN workers, E-3 workers, among other... 
    Full time
    H1b
    Work at office
    Remote work
    Work from home
    Flexible hours

    Huntington

    Dallas, TX
    2 days ago
  • $79.31k - $158.62k

     ...Software Development Engineer, Site Reliability Engineering (SRE) We're building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you'll be surrounded by passionate colleagues who... 
    Hourly pay
    Full time
    Temporary work

    Oak St. Health

    Richardson, TX
    4 days ago
  •  ...Job Title: Site Reliability Engineer Location: Dallas TX (HYBRID) Duration :Full Time Job Description: Skill: Site Reliability...  ...-wide applications. • ssists in the development staff in understanding the software products and any enhancements... 
    Full time
    Work at office

    Syntricate Technologies

    Dallas, TX
    4 days ago
  • $50 - $53 per hour

     ...area onsite at the project, significantly reducing and/or eliminating the demands to travel. Key Responsibilities: Site Reliability Engineers are expected to be able to drive technology triage efforts to completion by assisting with restoral steps, identifying root... 
    Hourly pay
    Live in
    Work at office
    Local area
    Flexible hours
    3 days per week

    Accenture

    Irving, TX
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Site Reliability Engineer - Kubernetes. Be the first to apply!