Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer 2

Kong

Are you ready to unlock intelligence?If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.Are you ready to unlock intelligence?If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS platform that powers the world’s API connectivity.You’ll design, automate, and run production systems serving thousands of customers across AWS, GCP, and Azure. You’ll work on everything from multi-region Kubernetes clusters to service mesh and gateway architectures, ensuring the reliability, scalability, and security of Kong’s SaaS offerings.This is a hands-on role ideal for engineers who thrive on running production SaaS systems at scale, automating operations, and continuously improving performance, resilience, and deployment pipelines.What You’ll Do:Operate and scale Kong’s global SaaS platform (Konnect), ensuring reliability, availability, and performance across regions and clouds.Build, automate, and maintain Kubernetes-based infrastructure and deployment workflows using Terraform/Terragrunt, Helm, and ArgoCD.Design, maintain, and optimize multi-region data and caching layers — including PostgreSQL, Redis, ClickHouse, and Druid — for high availability and low latency.Operate and improve Kong Gateway and Kong Mesh environments supporting hybrid and distributed architectures.Develop and maintain CI/CD pipelines and GitOps workflows to automate service delivery and ensure consistent infrastructure changes.Enhance observability and incident response readiness through systems like Datadog, Prometheus, Grafana, and Thanos, defining and tracking SLOs.Collaborate closely with development and security teams to ensure smooth operation of SaaS services in compliance with reliability, security, and regulatory standards.Participate in a global 24/7 on-call rotation and drive continuous improvement of operational playbooks and postmortem practices.Lead and contribute to scaling initiatives that improve elasticity, reliability, and cost-efficiency across the SaaS platform.What You’ll Bring:BS in Computer Science or equivalent practical experience.Proven experience managing SaaS or PaaS systems at enterprise scale (multi-region, multi-tenant, secure environments).Deep expertise in Kubernetes, including debugging cluster/networking issues and designing for fault tolerance and scalability.Strong proficiency with Infrastructure as Code tools like Terraform or Terragrunt.Experience with CI/CD pipelines and GitOps workflows (ArgoCD, Atlantis, Helm).Proficiency in one or more programming languages (Go, Python, Bash) for automation and tooling.Solid understanding of Linux/Unix systems, networking (DNS, TLS/SSL, load balancers and distributed systems.Experiencing working with API gateway and service mesh technologiesFamiliarity with streaming systems like Kafka and observability platforms (Datadog, Prometheus, Grafana).Experience working in a 24/7/365 production support environment.Bonus Points:Hands-on experience with Kong Gateway, Kong Mesh, or similar service connectivity technologies.Experience operating ClickHouse, Druid, or other time-series and analytics databases.Experience managing PostgreSQL and Redis in multi-region configurations.Working knowledge of AWS networking (PrivateLink, Transit Gateway, VPC Peering, Firewalls), Azure VNet, or GCP NCC.Strong understanding of disaster recovery, resiliency testing, and compliance-driven reliability practices.#LI-KC1About Kong:Kong Inc., the AI Connectivity Company, is building the connectivity layer of AI. Trusted by the Fortune 500 and AI-native startups alike, Kong’s unified API and AI platform enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI traffic — on any model, any cloud. For more information, visit .Compensation Range: $123K - $150KLocationWashington, United StatesEmployment TypeFull timeLocation TypeRemoteDepartmentAll Cost CenterR&DENGCompensation$123K – $150KKong has different base pay ranges for different work locations within the United States and Canada, which allows us to pay employees competitively and consistently in different geographic markets. Compensation varies depending on a wide array of factors, including but not limited to specific candidate location, role, skill set and level of experience. Certain roles are eligible for additional rewards including sales incentives depending on the terms of the applicable plan and role. Benefits may vary depending on location. US based employees are typically offered access to healthcare benefits, a 401(k) plan, short and long term disability benefits, basic life and AD&D insurance, among others.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer 2 in Washington DC vacancy
  • $210k - $230k

    GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient...  ...for GitOps)• Familiarity with compliance frameworks (SOC 2, HIPAA, FedRAMP)• Previous experience in a DevOps or Platform... 
    Suggested
    Currently hiring
    Remote work

    Govcio

    Arlington, VA
    3 days ago
  •  ...Site Reliability Engineer (SRE) Randstad is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our client in the Washington...  ...Required Experience & Technical Skills ~2+ years of hands-on experience in a Site Reliability Engineering... 
    Suggested

    Software Technology Inc

    Washington DC
    4 days ago
  •  ...Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian...  ...workloads during model training and high-volume inference. 2. MLOps & AI Infrastructure Model Serving... 
    Suggested
    Local area

    Tiger Analytics

    Washington DC
    3 days ago
  • $125k - $185k

     ...children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-...  ...assistance• Take what you need paid time off, not accrual based• 2 weeks paid time off built into the end of each year (subject... 
    Suggested
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    1 day ago
  • $125k - $185k

    Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world...  ..., and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate...  ...need paid time off, not accrual based• 2 weeks paid time off built into the end... 
    Suggested
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    3 days ago
  • $174k - $238k

     ...The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging Products Group...  ...States, as defined in Federal Acquisition Regulation (FAR) 2.101U.S. Security Clearance status - the employee must be... 
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    4 days ago
  • Job TitleSoftware Engineer Level 2LocationWashington, DC 20032 US (Primary)CategoryResearch, Development, and EngineeringJob TypeFull-TimeCareer...  ...DescriptionPrescient Edge is seeking a Software Engineer Level 2 to support a federal government client.Please note that the... 
    Contract work

    Prescient Edge

    Washington DC
    1 day ago
  • $112k - $218.4k

     ...: Our team is looking for a Senior Active Directory Site Reliability Engineer. Our mission is to improve the availability, latency, performance and...  ...Science, Information Technology, or related field AND 2+ years technical experience in software engineering, network... 
    Full time
    Local area

    Microsoft

    Washington DC
    7 days ago
  •  ...Black Cape  Title: Spectacular Full Stack Software Engineer (2+ years - Senior Level) Locations: Arlington, VA, Reston, VA and areas throughout the DMV area Onsite Expectations: MUST be willing to go onsite 2 - 3 days per week; could be up to 5 days per week... 
    Full time
    For subcontractor
    2 days per week
    3 days per week

    Black Cape

    Arlington, VA
    8 hours ago
  •  ...Black Cape Title: FrontEnd Software Engineer (2+ years to Senior Levels)  Locations: Arlington, VA, Reston, VA and areas throughout...  ...environments or willingness to work at classified government sites as needed Experience developing on MacOSX, Windows, or Linux... 
    Full time
    For subcontractor

    Black Cape

    Arlington, VA
    8 hours ago
  • $130k - $138k

     ...Software Engineer Level 2 Position: Software Engineer Level 2 (Visualization Developer) Location: Laurel, MD (On-site) Category: Software Engineering Schedule: Standard Day Shift, Monday–Friday Clearance Requirement: Active TS/SCI with Polygraph Experience Requirement... 
    Temporary work
    Monday to Friday
    Flexible hours
    Day shift

    ClearanceJobs

    Laurel, MD
    2 days ago
  • $160k - $210k

     ...industry. Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service...  ...hybrid work schedule of 3 days in office (Mon/Tue/Wed) and 2 days remote (Thursday/Friday). Responsibilities Design... 
    Work at office
    Immediate start
    Remote work
    Work from home

    Cognitiv

    Washington DC
    11 days ago
  • $98.8k - $164.6k

     ...Communications, Customer Care, Engineering & Product, Finance, Human...  ...engineering decisions, and help build reliable, maintainable systems that...  ...equivalent practical experience.2+ years of professional...  ...and business travel, we work on-site five days a week.Compensation... 
    Full time
    Part time

    Washington Post

    Washington DC
    8 hours ago
  • $115.5k - $164.8k

     ...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant...  ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations... 
    Work experience placement
    Work at office
    Remote work

    Axon

    Washington DC
    3 days ago
  • $185k - $230k

    As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services... 
    Full time
    Local area
    Immediate start

    MetroStar Systems

    Washington DC
    1 day ago
  • $166k - $220k

     ...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the...  ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine... 
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    8 hours ago
  • $165k - $230k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts.... 
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    3 days ago
  • $230k - $250k

    GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is... 
    Remote work

    Govcio

    Arlington, VA
    1 day ago
  • $86.8k - $198k

    Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether...  ...environments, including in Amazon Web Services (AWS)2+ years of experience with monitoring and observability... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    McLean, VA
    1 day ago
  • $112k - $179k

     ...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems... 
    Contract work
    Worldwide
    Shift work

    Peraton Corporation

    Washington DC
    1 day ago
  • $150k - $180k

     ...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business...  ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or... 
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work
    Worldwide

    Umbra

    Arlington, VA
    1 day ago
  •  ...position. Position Summary:  ISI is looking for a Project Engineer Level 2 to provide Owner's Representative construction management...  ...environments. Responsibilities: Assist the Government in site evaluations, field surveys, and site visits to assess... 
    Permanent employment
    For contractors
    Work experience placement
    Work at office
    Monday to Friday

    ISI Professional Services

    Arlington, VA
    14 days ago
  •  ...Associate or Mid-Level Flight Deck Design Engineer to join a team of highly skilled...  ...objectives and requirements and impacts on data reliability and integrity, fault trees analysis (FTAs...  ...Experience: \tBachelor's degree and typically 2 or more years' experience in an... 
    Work experience placement

    Systemart LLC

    Washington DC
    22 days ago
  •  ...solutions using a tailored Agile methodology. We are seeking a highly motivated and intellectually curious Senior Site Reliability Engineer to join our team working with a Federal client. The position will be a remote role open to US citizens residing in the... 
    Remote work

    Elevate Government Solutions

    Washington DC
    2 days ago
  •  ...Site Reliability Engineer ValidaTek is building teams of Site Reliability Engineers (SRE's) to support internal and external engineering and operations of a large scale and world-wide Enterprise IT environment that covers application hosting and support, enterprise... 

    ClearanceJobs

    Washington DC
    4 days ago
  •  ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable platform by developing systems for analyzing environments, predicting, and resolving issues, and supporting the production environment... 
    Work experience placement

    Samprasoft

    Washington DC
    1 day ago
  • $75.7k - $136.3k

     ...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and... 
    Work experience placement
    Work at office

    Akamai

    Washington DC
    4 days ago
  • $107k - $220k

     ...The Site Reliability Engineer (SRE) will ensure the reliability, performance, and scalability of the WDP System. This person will define and track Key Performance Indicators (KPIs) and Service Level Objectives (SLOs), identify and resolve performance bottlenecks, and perform... 
    Full time
    Contract work
    Temporary work
    Work at office
    Visa sponsorship
    Work visa

    Avalore, LLC

    Arlington, VA
    3 days ago
  •  ...Catalyst, and In-Q-Tel. Mission | On Site | Full Time | Active TS/SCI with Full...  ...government customer site, ensuring the reliability and performance of Twenty's mission-critical...  ...technical ownership and customer-facing engineering: you'll define how we measure... 
    Full time
    Contract work
    Remote work
    Flexible hours

    Twenty Technologies

    Arlington, VA
    3 days ago
  • OB SUMMARYThe Systems Engineer - Site Reliability Engineering (SRE) is responsible for the reliability, scalability, and performance of mission-critical cloud and on-prem services that support millions of Marriot customers globally. This role involves overseeing incident... 
    Full time
    For contractors
    Work at office
    Remote work
    Flexible hours
    Shift work

    Marriott International

    Bethesda, MD
    8 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer 2. Be the first to apply!