Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer 2

Kong

Are you ready to unlock intelligence?If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.Are you ready to unlock intelligence?If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS platform that powers the world’s API connectivity.You’ll design, automate, and run production systems serving thousands of customers across AWS, GCP, and Azure. You’ll work on everything from multi-region Kubernetes clusters to service mesh and gateway architectures, ensuring the reliability, scalability, and security of Kong’s SaaS offerings.This is a hands-on role ideal for engineers who thrive on running production SaaS systems at scale, automating operations, and continuously improving performance, resilience, and deployment pipelines.What You’ll Do:Operate and scale Kong’s global SaaS platform (Konnect), ensuring reliability, availability, and performance across regions and clouds.Build, automate, and maintain Kubernetes-based infrastructure and deployment workflows using Terraform/Terragrunt, Helm, and ArgoCD.Design, maintain, and optimize multi-region data and caching layers — including PostgreSQL, Redis, ClickHouse, and Druid — for high availability and low latency.Operate and improve Kong Gateway and Kong Mesh environments supporting hybrid and distributed architectures.Develop and maintain CI/CD pipelines and GitOps workflows to automate service delivery and ensure consistent infrastructure changes.Enhance observability and incident response readiness through systems like Datadog, Prometheus, Grafana, and Thanos, defining and tracking SLOs.Collaborate closely with development and security teams to ensure smooth operation of SaaS services in compliance with reliability, security, and regulatory standards.Participate in a global 24/7 on-call rotation and drive continuous improvement of operational playbooks and postmortem practices.Lead and contribute to scaling initiatives that improve elasticity, reliability, and cost-efficiency across the SaaS platform.What You’ll Bring:BS in Computer Science or equivalent practical experience.Proven experience managing SaaS or PaaS systems at enterprise scale (multi-region, multi-tenant, secure environments).Deep expertise in Kubernetes, including debugging cluster/networking issues and designing for fault tolerance and scalability.Strong proficiency with Infrastructure as Code tools like Terraform or Terragrunt.Experience with CI/CD pipelines and GitOps workflows (ArgoCD, Atlantis, Helm).Proficiency in one or more programming languages (Go, Python, Bash) for automation and tooling.Solid understanding of Linux/Unix systems, networking (DNS, TLS/SSL, load balancers and distributed systems.Experiencing working with API gateway and service mesh technologiesFamiliarity with streaming systems like Kafka and observability platforms (Datadog, Prometheus, Grafana).Experience working in a 24/7/365 production support environment.Bonus Points:Hands-on experience with Kong Gateway, Kong Mesh, or similar service connectivity technologies.Experience operating ClickHouse, Druid, or other time-series and analytics databases.Experience managing PostgreSQL and Redis in multi-region configurations.Working knowledge of AWS networking (PrivateLink, Transit Gateway, VPC Peering, Firewalls), Azure VNet, or GCP NCC.Strong understanding of disaster recovery, resiliency testing, and compliance-driven reliability practices.#LI-KC1About Kong:Kong Inc., the AI Connectivity Company, is building the connectivity layer of AI. Trusted by the Fortune 500 and AI-native startups alike, Kong’s unified API and AI platform enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI traffic — on any model, any cloud. For more information, visit .Compensation Range: $123K - $150KLocationWashington, United StatesEmployment TypeFull timeLocation TypeRemoteDepartmentAll Cost CenterR&DENGCompensation$123K – $150KCompensation decisions are based on various factors, including, but not limited to, experience, skills, certifications, location, and business needs. Employees may be eligible for additional benefits, including health, dental, and vision insurance, paid holidays and PTO; and professional development opportunities and tools.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer 2 in Washington DC vacancy
  •  ...new job is posted. Sign in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259,000.00 1 day ago...  ...Engineer - New Shepard Seattle, WA $157,053.75-$239,127.00 2 weeks ago Seattle, WA $151,300.00-$261,500.00 2 weeks ago... 
    Suggested
    Contract work
    Remote work

    Signature IT World Inc

    Washington DC
    23 hours ago
  • $210k - $230k

     ...GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient...  ...for GitOps) Familiarity with compliance frameworks (SOC 2, HIPAA, FedRAMP) Previous experience in a DevOps or... 
    Suggested
    Currently hiring
    Remote work

    GovCIO

    Arlington, VA
    3 days ago
  •  ...Description Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the...  ...workloads during model training and high-volume inference. 2. MLOps & AI Infrastructure Model Serving Reliability: Ensure... 
    Suggested
    Local area

    Tiger Analytics Inc.

    Washington DC
    a month ago
  • $20k

     ...possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (STARSHIELD) At SpaceX we’re leveraging our experience in...  ...BENEFITS:Pay Range:Level 1: $125,000.00 - $160,000.00Level 2: $145,000.00 - $195,000.00Your actual level and base salary... 
    Suggested
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    4 days ago
  • $140k - $210k

     ...Mission As the world’s number 1 job site*, our mission is to help people get...  ...Visits, March 2026) Day to Day As an Engineering Manager in Site Reliability Engineering at Indeed, you will...  ...0,000 - 210,000 USD per year Tier 2 - United States of America 155,000 -... 
    Suggested
    Work experience placement
    Local area

    Indeed

    Washington DC
    3 days ago
  •  ...position. Position Summary: ISI is looking for a Project Engineer Level 2 to provide Owner's Representative construction management...  ...environments. Responsibilities: · Assist the Government in site evaluations, field surveys, and site visits to assess... 
    Permanent employment
    For contractors
    Work experience placement
    Work at office
    Monday to Friday

    ISI Professional Services

    Arlington, VA
    27 days ago
  • $174k - $238k

     ...Federal SRE Team We are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging Products Group...  ...States, as defined in Federal Acquisition Regulation (FAR) 2.101 U.S. Security Clearance status - the employee must be... 
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    4 days ago
  • [Position Overview] ~ Job Title: Software Test Engineer [Chinese Bilingual] ~ Education: Bachelor's Degree in Electrical Engineering, Computer Science, or related field ~2+ years of hands-on experience in software testing for mobile communication terminal products... 
    Hourly pay
    Full time
    H1b
    Local area
    Immediate start
    Flexible hours
    Night shift

    JND

    Washington DC
    more than 2 months ago
  • $115.5k - $164.8k

     ...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant...  ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations... 
    Work experience placement
    Work at office
    Remote work

    Axon

    Washington DC
    3 days ago
  • $75.06 - $107.61 per hour

     ...experience that is rare within our industry. The Cloud Software Engineer Level 2 Engineer provides cloud software research, development, and...  .... Build and maintain automated CI/CD pipelines for rapid, reliable software delivery. Implement comprehensive testing... 
    Hourly pay
    Contract work
    Work experience placement

    Wyetech

    Laurel, MD
    25 days ago
  • $87.1k - $157.45k

     ...USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep...  ...help us build systems that stay reliable when things get complicated. We need a Site Reliability Engineer who has experience...  .../programming experience of 2+ years. • Ability to debug complex... 
    Local area
    Immediate start
    Work from home
    Flexible hours

    Leidos

    Springfield, VA
    2 days ago
  •  ...Washington D.C., District of Columbia, United States About the job Sr. Site Reliability Engineer Our Client is currently hiring a full-time Sr. Site Reliability Engineer (SRE), who will play a vital role in continuously driving improvements in observability, performance... 
    Full time
    Currently hiring
    3 days per week

    CruitZi

    Washington DC
    1 day ago
  • $230k - $250k

     ...GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is... 
    Remote work

    GovCIO

    Arlington, VA
    3 days ago
  • $141.8k - $195k

     ...best work, grow fast, and bring their full selves to the herd. Why You'll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl... 
    Temporary work
    Remote work

    Cribl

    Washington DC
    2 days ago
  • $175k - $250k

     ...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or...  ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design, build... 
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    Washington DC
    23 hours ago
  • $120k - $155k

     ...200B+ market, we’re excited to talk to you! What will the Site Reliability Engineer do? We're looking for a Site Reliability Engineer who's passionate...  ...us to share more about Coterie and the position. Phase 2 Selected candidates will be invited to meet with our Hiring... 
    Full time
    Immediate start
    Remote work
    Visa sponsorship
    Flexible hours

    Cintrifuse

    Springfield, VA
    4 days ago
  • $166k - $220k

     ...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the...  ...solutions we ship. ABOUT THE JOB We are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine... 
    Full time
    Work experience placement

    Mosaic

    Washington DC
    4 days ago
  • $103.5k - $150k

     ...experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure...  ...days per week onsite. Qualifications Minimum Qualifications 2+ years of experience in Site Reliability Engineering,... 
    Temporary work
    Work experience placement
    Local area
    3 days per week

    Medallia

    McLean, VA
    4 days ago
  •  ...Site Reliability Engineer III (AI Platform) Location: Mount Laurel, NJ (Onsite) Duration: Contract Experience: 4+ years About the Role We are seeking a Site Reliability Engineer (SRE) III to support a cutting-edge AI Platform Engineering team responsible... 
    Contract work

    GCS Recruitment

    Laurel, MD
    4 days ago
  •  ...ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running...  ...of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident... 
    Full time
    Contract work
    Remote work
    Flexible hours

    Twenty Inc.

    Arlington, VA
    3 days ago
  •  ...Job Description Job Description Description: Onsite in Washington, DC   our client seeks a Sr. Site Reliability Engineer III to design, automate, and operate mission-critical systems for federal environments. The role focuses on Kubernetes or VMWare platforms,... 
    Hourly pay
    Permanent employment
    Full time
    Local area
    Immediate start

    Eliassen Group

    Washington DC
    a month ago
  • Role Summary The Senior Site Reliability Engineer (SRE) is a hands-on role responsible for the availability, performance, and end-to-end observability of QSR digital platforms across Mobile (iOS/Android), Web, and POS systems. This role is part of the Observability... 
    Flexible hours

    Donato Technologies, Inc

    Washington DC
    2 days ago
  •  ...Join the Site Reliability Engineering (SRE) team to support high-engagement multimodal applications. You will be responsible for managing infrastructure, observability solutions, and platform automation while ensuring the reliability and security of hosted applications... 

    NextGen | GTA: A Kelly Telecom Company

    Laurel, MD
    1 day ago
  •  ...technology solutions using a tailored Agile methodology. We are seeking a highly motivated and intellectually curious Senior Site Reliability Engineer to join our team working with a Federal client. The position will be a remote role open to US citizens residing in the... 
    Remote work

    VALID8 Financial

    Washington DC
    4 days ago
  • $130k - $160k

     ...customers. Our customers and partners trust us to deliver reliable, first-to-market solutions and safeguard the data we receive...  ...pioneers of market-changing solutions. We are seeking a Site Reliability Engineer to design, build, and maintain highly available systems and... 
    Local area

    LE038 Second Sight Solutions, LLC

    Washington DC
    6 hours ago
  • $160k - $180k

     ...Site Reliability Engineer Location: Hybrid – Washington DC/Virginia/Maryland metro with the ability to travel to Patuxent River, MD, as needed (up to 20% of the time). Compensation: $160,000 - 180,000 per year, depending on experience and qualifications. Employment... 
    Full time
    Temporary work
    Local area
    Remote work
    Flexible hours

    RiseMe

    Washington DC
    4 days ago
  •  ...Washington, District of Columbia, United States Contractor | On-site Job Description We are seeking an experienced Site Reliability Engineer (SRE) to help build and maintain highly reliable, scalable, and secure technology platforms. The SRE will combine software... 
    For contractors

    Mybridge

    Washington DC
    1 day ago
  • $100k - $160k

     ...innovative and trusted results. We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team! The Site...  ...in Computer Science or related field A minimum of two (2) years of experience working with on-premise and off-... 
    Temporary work

    Cathexis

    McLean, VA
    4 days ago
  • $107k - $220k

     ...The Site Reliability Engineer (SRE) will ensure the reliability, performance, and scalability of the WDP System. This person will define and track Key Performance Indicators (KPIs) and Service Level Objectives (SLOs), identify and resolve performance bottlenecks, and perform... 
    Full time
    Contract work
    Temporary work
    Work at office
    Visa sponsorship
    Work visa

    Avalore, LLC

    Arlington, VA
    3 days ago
  •  ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable platform by developing systems for analyzing environments, predicting, and resolving issues, and supporting the production environment... 
    Work experience placement

    Samprasoft

    Washington DC
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer 2. Be the first to apply!