Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior DevOps Engineer, Infrastructure & Reliability

Worth Ai

Worth AI, a leader in the computer software industry, is looking for a Senior DevOps Engineer to join our Infrastructure team with a singular mission: to make our systems faster, more reliable, and more resilient while making life dramatically easier for engineers shipping software. 

This is a hands-on build role. You will spend most of your time writing Terraform, tuning Kubernetes workloads, automating things that are currently manual, and shipping infrastructure changes to production. You'll join a small platform team with an established roadmap and existing patterns, and a strong voice in how the work gets built.

  • Implement scalable Infrastructure-as-Code patterns using tools like Terraform to standardize cloud provisioning and reduce configuration drift.
  • Own and evolve our Kubernetes platform (EKS or self-managed), ensuring workloads are secure, scalable, and resilient by default.
  • Optimize CI/CD pipelines to improve deployment frequency, reduce lead time, and increase confidence in releases.
  • Design and enforce secure networking, IAM, and secrets management strategies across environments.
  • Improve observability by refining metrics, logs, and tracing using tools like DataDog, ensuring actionable insight into system health.
  • Optimize cloud cost efficiency through rightsizing, autoscaling strategies, and architectural improvements.
  • Implement disaster recovery planning, backup strategies, and multi-region resilience initiatives.
  • Refactor brittle or manually managed infrastructure into automated, testable, and reproducible systems.
  • Introduce new infrastructure tooling or architectural shifts and drive adoption through documentation, workshops, and hands-on support.
  • Partner with engineering teams to eliminate friction in CI/CD, deployments, and cloud environments.
  • Communicate technical trade-offs clearly across engineering and product stakeholders, balancing speed with safety.
Technology Stack
  • Cloud & Infrastructure: AWS (EKS, RDS, MSK, S3, Lambda, IAM, VPC)
    Containerization & Orchestration: Kubernetes, ArgoCD
    Infrastructure-as-Code: Terraform
    CI/CD: GitHub Actions
    Monitoring & Observability: DataDog
    Data & Messaging: PostgreSQL, Kafka, Redis
    Languages (as needed): Bash, Python, TypeScript, JavaScript

Requirements

  • 8+ years in DevOps, SRE, or infrastructure engineering.
  • Proven experience designing and operating production Kubernetes environments at scale.
  • Deep hands-on expertise with AWS infrastructure and cloud networking.
  • Strong experience building and maintaining Terraform modules across large cloud environments.
  • Demonstrated ownership of CI/CD systems and measurable improvement of DORA metrics.
  • Experience leading incident response processes and driving meaningful postmortem outcomes.
  • Strong understanding of distributed systems, event-driven architectures (Kafka), and database performance (PostgreSQL).
  • Proven ability to modernize legacy infrastructure and eliminate manual operational toil.
  • Track record of taking a scoped infrastructure project from an ambiguous starting point to production without needing daily direction.
  • Demonstrated ability to build trust across teams while raising the reliability bar.
Success Metrics
  • System Reliability: Maintain or exceed defined SLO/SLA targets with reduced incident frequency and duration.
  • Infrastructure Stability: Reduce production incidents caused by misconfiguration, manual processes, or infrastructure drift.
  • Operational Efficiency: Increase the percentage of infrastructure managed through code and automation.
  • Cost Optimization: Improve cloud cost efficiency without sacrificing reliability or performance.
Bonus Points (Nice to Have)
  • Experience coding applications
  • Experience operating high-throughput Kafka clusters (MSK or self-managed).
  • Strong background in database performance tuning (PostgreSQL, Redis).
  • Experience implementing autoscaling strategies for high-traffic systems.
  • Familiarity with service mesh technologies.
  • Experience building internal developer platforms (IDP).
  • Background in security best practices (zero-trust networking, policy-as-code).
  • Experience with multi-region or globally distributed systems.
  • Experience introducing platform-wide reliability frameworks (SLOs, error budgets, chaos testing).

All Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando.

Benefits

  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k)
  • Life Insurance
  • Flexible Paid Time Off
  • 9 paid Holidays
  • Family Leave
  • Remote
  • Hybrid work (for Orlando Associates)
  • Free Food & Snacks (Orlando)
  • Wellness Resources
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior DevOps Engineer, Infrastructure & Reliability in United States vacancy
  •  ...Senior DevOps EngineerWorth AI, a leader in the computer software industry, is looking for a Senior DevOps Engineer to join our Infrastructure team with a singular mission: to make our systems faster, more reliable, and more resilient while making life dramatically easier... 
    Senior
    Shift work

    Worth AI

    Orlando, FL
    4 days ago
  • United States Digital Space LLC is seeking a Senior Software Engineer, Infrastructure (Release Engineering) based in San Francisco. This role involves...  ...processes. With a strong emphasis on performance and reliability, successful candidates will have at least 7 years of experience... 
    Senior
    Work at office
    Remote work

    United States Digital Space LLC

    San Francisco, CA
    2 days ago
  •  ...innovative AI solutions company is seeking a Senior DevOps Engineer to architect and maintain the core infrastructure supporting cutting-edge AI applications. The role...  ..., and championing best practices in system reliability. Ideal candidates should have over 7 years of... 
    Senior
    Remote job
    Full time
    Flexible hours

    New Code Inc

    Palo Alto, CA
    2 days ago
  •  ...will work closely with the other software engineering teams with a focus on software development and infrastructure design providing the expertise in performance...  ....We are looking for an experienced Senior Site Reliability Engineer that brings a broad set of technical... 
    Senior
    Work experience placement

    Kaizen Gaming (Stoiximan/Betano)

    Athens, TX
    4 days ago
  • Pivotal Health in San Francisco is seeking a Senior Platform Engineer to design, scale, and harden the foundation powering our platform. You...  ...with engineering teams to evolve cloud architecture, improve reliability and security, and ensure scalable, event-driven systems.... 
    Senior

    Pivotal Health

    San Francisco, CA
    2 days ago
  • Omnissa, a leading AI-driven digital work platform, seeks a DevOps Lead for Cloud Automation & Reliability in a hybrid setup. You will guide a senior SRE team, drive automation, and partner with engineering to scale cloud-native services across multiple regions. Ideal... 
    Senior

    Omnissa, LLC in

    Atlanta, GA
    1 day ago
  • A technology services company is seeking a Senior Site Reliability Engineer / DevOps Engineer in Sunnyvale, CA. The ideal candidate will have over 8 years of experience in DevOps, expertise in Docker and Kubernetes, and proficiency with Terraform or Ansible. Responsibilities... 
    Senior

    Donato Technologies Inc

    Sunnyvale, CA
    5 days ago
  •  ...SRI Tech SolutionsJob Title: Senior DevOps Cloud EngineerLocation: San...  ...Senior DevOps Cloud Engineer to design, implement, and maintain...  ...scalable, and secure cloud infrastructure. This role focuses on...  ...ensuring high performance and reliability.Manage and optimize Linux-based... 
    Senior

    SRI Tech

    San Diego, CA
    4 days ago
  •  ...transforming financial services. As a Cloud Infrastructure Engineer, you will design and manage scalable,...  ...expert who thrives on building reliable, high-performance cloud environments and...  ...to cutting-edge AI workloads, modern DevOps practices, and real-world impact on enterprise... 
    Senior
    Full time
    Remote work

    Motion Recruitment

    New York, NY
    2 days ago
  • $232k - $319k

     ...secures AI by building the trusted, neutral infrastructure that enables organizations to safely...  ...the service with great people and reliable, cost-effective, and efficient infrastructure...  ...the velocity of SRE and product engineering by developing robust platforms, powerful... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    2 days ago
  • $85.8k - $180.2k

    Job Title: Senior Cloud DevOps Engineer - Identity ManagementJob Category: Information TechnologyTime...  ...hybrid cloud and on-premises infrastructures.• Drive the modernization of authentication...  ..., and optimize performance and reliability of identity systems and associated... 
    Senior
    Contract work
    Work experience placement
    Flexible hours

    CACI International

    San Antonio, TX
    1 day ago
  •  ...Position Overview We are seeking an experienced Senior DevOps / Site Reliability Engineer (SRE) to support a mission-critical national security...  ...AWS cloud environments, Linux systems administration, infrastructure automation, and software development. The successful... 
    Senior
    Full time
    Visa sponsorship

    Omniscius Consulting

    Washington DC
    18 days ago
  • $176k - $276k

     ...!NVIDIA invites applications for a Senior DevOps Platform Engineer skilled in Platform and Release Engineering...  ..., and maintaining foundational infrastructure and CI/CD systems that run AI/...  ...foster engineering rigor by setting up reliable release workflows, automation... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...9031472Reference26-02146Job Title: Senior Azure DevOps Engineer Location: Phoenix AZ - Onsite Experience...  ..., Kubernetes, CI/CD automation, Infrastructure as Code (IaC), and cloud-native...  ...improve deployment efficiency and reliability.Troubleshoot production issues and... 
    Senior

    AgreeYa Solutions

    Phoenix, AZ
    1 day ago
  • Everpure in Santa Clara is seeking a Senior Software Engineer in Production Engineering to own the CI and test-orchestration platform that enables...  ...systems behind CI, automate quality gates, and improve reliability with observability and data-driven insights. #J-18808-... 
    Senior

    Everpure

    Santa Clara, CA
    5 days ago
  • $136.6k - $184.8k

    AWS Infrastructure Services owns the design, planning, delivery, and operation...  ..., hardware, and network engineers, supply chain specialists,...  ...from our vendors. As a Senior Supplier Quality Engineer you...  ...cross-functional teams such as reliability and design engineering teams... 
    Senior
    Flexible hours

    Amazon

    Herndon, VA
    2 days ago
  • $110.3k - $165.5k

    SUMMARYWe are seeking an experienced Senior DevOps Engineer that will be responsible for the governance...  ...the Oracle Cloud Platforms are reliable and deployments are repeatable, auditable...  ....Working knowledge of Oracle Cloud Infrastructure and Oracle Fusion Cloud Applications.... 
    Senior
    H1b
    Work at office

    Mortenson

    Minneapolis, MN
    1 day ago
  •  ...Tech SolutionsJob Title: Senior DevSecOps Cloud...  ...Senior DevSecOps Cloud Engineer to design, implement,...  ...scalable, and secure cloud infrastructure. This role emphasizes...  ...IaC) tools, and modern DevOps practices with a...  ...high performance and reliability • Manage and optimize... 
    Senior

    SRI Tech

    San Diego, CA
    4 days ago
  •  ...scalable application, and high-performing engineering team is infrastructure built to perform under pressure.As our Senior DevOps Developer, you'll play a critical role in...  ...closely with engineering teams to improve reliability, accelerate deployments, strengthen security... 
    Senior
    Full time

    Floor & Decor

    Atlanta, GA
    2 days ago
  • $113k - $188k

     ...TrustWe are seeking a highly skilled Senior DevOps / Cloud Engineer to support and enhance our existing...  ...deep hands-on expertise in cloud infrastructure, CI/CD, automation, application deployment...  ...end, and driving implementation of reliable, secure, scalable, and compliant... 
    Senior
    Full time
    Flexible hours

    Guidehouse

    Bethesda, MD
    4 days ago
  • $131k - $271.6k

     ...national security and critical infrastructure customers.Must be a U.S....  ...Cloud Development Operations Engineer who possesses a strong background...  ..., ensuring scalability, reliability, and security of cloud-based...  ...such as AWS Certified DevOps Engineer or AWS Certified Solutions... 
    Senior
    Permanent employment
    Full time
    Work experience placement
    Worldwide
    Flexible hours

    SAP

    Newtown Square, PA
    8 hours ago
  •  ...modernization effort across its infrastructure, cloud platforms, and...  ...technologies, automation, and modern DevOps practices, they are looking for an experienced DevOps Engineer who enjoys building scalable...  ...Engineering, or Site Reliability Engineering Deep Linux administration... 
    Senior
    Full time

    Motion Recruitment

    Massachusetts
    1 day ago
  •  ...SaaS product used globally. Own infrastructure strategy, reliability, and automation while influencing engineering at scale, reporting to the Cloud DevOps Team Leader. About Innocraft & Matomo...  .... We’re looking for a Senior Infrastructure DevOps Engineer to take... 
    Senior
    Full time
    Work at office
    Remote work

    Matomo

    Remote
    18 days ago
  • $120k - $160k

     ...pride ourselves on offering a reliable lifeline when needed....  ...solutions in a collaborative engineering environment, then you will be excited about the Senior Cloud DevSecOps Engineer opportunity...  ...applications and deployment infrastructure supporting mission-critical systems... 
    Senior
    Full time
    Work at office
    Remote work
    3 days per week

    Iridium

    Chandler, AZ
    2 days ago
  • $126k - $204.5k

    Role Description We're looking for an Infrastructure Engineer to build developer tooling that enables developer velocity and reliability for the entire engineering organization. Our team owns the full end-to-end development lifecycle, from local development tools, CI/CD... 
    Senior
    Full time
    Local area

    Palo Alto Networks

    Remote
    2 days ago
  •  ...We are seeking a highly skilled Senior DevOps / Cloud Engineer to support and enhance our existing AWS...  ...deep hands-on expertise in cloud infrastructure, CI/CD, automation, application deployment...  ...end, and driving implementation of reliable, secure, scalable, and compliant... 
    Senior

    TalTeam

    Bethesda, MD
    3 days ago
  • Dexian has been engaged to find a resourceful Senior Infrastructure and Azure DevOps Engineer who can demonstrate his/her understanding of the workings of...  ...CI/CD pipeline.Assist with the design and building of reliable, fault tolerant private cloud infrastructure... 
    Senior
    Work experience placement
    Relocation
    2 days per week

    Cedent Consulting

    Pleasanton, CA
    4 days ago
  • $87.95k - $162.88k

     ...apply now.We are currently seeking a Senior AI DevOps Engineer (AI Ops / Platform Engineering) (...  ...accelerate software delivery while improving reliability, observability, and operational...  ...LLMs with engineering tools, cloud infrastructure, and operational platforms. Build... 
    Senior
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours
    3 days per week

    NTT DATA

    Atlanta, GA
    2 days ago
  • $166k - $220k

     ...such, it is critical that Anduril services are reliable and maintainable. This means that all services & infrastructure across the fleet need to be able to be observed...  ...infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build &... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    1 day ago
  • The Data Infrastructure SRE team is responsible for the reliability, scalability, and efficiency of the core data services...  ...about building features, but about engineering the resilience and performance...  ...practices while working alongside senior engineers to solve challenging... 
    Senior

    TikTok

    Seattle, WA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior DevOps Engineer, Infrastructure & Reliability. Be the first to apply!