Site Reliability Engineer 2
Kong
Are you ready to unlock intelligence?If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.Are you ready to unlock intelligence?If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS platform that powers the world’s API connectivity.You’ll design, automate, and run production systems serving thousands of customers across AWS, GCP, and Azure. You’ll work on everything from multi-region Kubernetes clusters to service mesh and gateway architectures, ensuring the reliability, scalability, and security of Kong’s SaaS offerings.This is a hands-on role ideal for engineers who thrive on running production SaaS systems at scale, automating operations, and continuously improving performance, resilience, and deployment pipelines.What You’ll Do:Operate and scale Kong’s global SaaS platform (Konnect), ensuring reliability, availability, and performance across regions and clouds.Build, automate, and maintain Kubernetes-based infrastructure and deployment workflows using Terraform/Terragrunt, Helm, and ArgoCD.Design, maintain, and optimize multi-region data and caching layers — including PostgreSQL, Redis, ClickHouse, and Druid — for high availability and low latency.Operate and improve Kong Gateway and Kong Mesh environments supporting hybrid and distributed architectures.Develop and maintain CI/CD pipelines and GitOps workflows to automate service delivery and ensure consistent infrastructure changes.Enhance observability and incident response readiness through systems like Datadog, Prometheus, Grafana, and Thanos, defining and tracking SLOs.Collaborate closely with development and security teams to ensure smooth operation of SaaS services in compliance with reliability, security, and regulatory standards.Participate in a global 24/7 on-call rotation and drive continuous improvement of operational playbooks and postmortem practices.Lead and contribute to scaling initiatives that improve elasticity, reliability, and cost-efficiency across the SaaS platform.What You’ll Bring:BS in Computer Science or equivalent practical experience.Proven experience managing SaaS or PaaS systems at enterprise scale (multi-region, multi-tenant, secure environments).Deep expertise in Kubernetes, including debugging cluster/networking issues and designing for fault tolerance and scalability.Strong proficiency with Infrastructure as Code tools like Terraform or Terragrunt.Experience with CI/CD pipelines and GitOps workflows (ArgoCD, Atlantis, Helm).Proficiency in one or more programming languages (Go, Python, Bash) for automation and tooling.Solid understanding of Linux/Unix systems, networking (DNS, TLS/SSL, load balancers and distributed systems.Experiencing working with API gateway and service mesh technologiesFamiliarity with streaming systems like Kafka and observability platforms (Datadog, Prometheus, Grafana).Experience working in a 24/7/365 production support environment.Bonus Points:Hands-on experience with Kong Gateway, Kong Mesh, or similar service connectivity technologies.Experience operating ClickHouse, Druid, or other time-series and analytics databases.Experience managing PostgreSQL and Redis in multi-region configurations.Working knowledge of AWS networking (PrivateLink, Transit Gateway, VPC Peering, Firewalls), Azure VNet, or GCP NCC.Strong understanding of disaster recovery, resiliency testing, and compliance-driven reliability practices.#LI-KC1About Kong:Kong Inc., the AI Connectivity Company, is building the connectivity layer of AI. Trusted by the Fortune 500 and AI-native startups alike, Kong’s unified API and AI platform enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI traffic — on any model, any cloud. For more information, visit .Compensation Range: $123K - $150KLocationWashington, United StatesEmployment TypeFull timeLocation TypeRemoteDepartmentAll Cost CenterR&DENGCompensation$123K – $150KKong has different base pay ranges for different work locations within the United States and Canada, which allows us to pay employees competitively and consistently in different geographic markets. Compensation varies depending on a wide array of factors, including but not limited to specific candidate location, role, skill set and level of experience. Certain roles are eligible for additional rewards including sales incentives depending on the terms of the applicable plan and role. Benefits may vary depending on location. US based employees are typically offered access to healthcare benefits, a 401(k) plan, short and long term disability benefits, basic life and AD&D insurance, among others.
$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient... ...for GitOps)• Familiarity with compliance frameworks (SOC 2, HIPAA, FedRAMP)• Previous experience in a DevOps or Platform...SuggestedCurrently hiringRemote work- ...Site Reliability Engineer (SRE) Randstad is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our client in the Washington... ...Required Experience & Technical Skills ~2+ years of hands-on experience in a Site Reliability Engineering...Suggested
- ...Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian... ...workloads during model training and high-volume inference. 2. MLOps & AI Infrastructure Model Serving...SuggestedLocal area
$125k - $185k
...children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-... ...assistance• Take what you need paid time off, not accrual based• 2 weeks paid time off built into the end of each year (subject...SuggestedFull timeWork experience placementWork at officeRemote workWork from homeRelocation package$125k - $185k
Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world... ..., and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate... ...need paid time off, not accrual based• 2 weeks paid time off built into the end...SuggestedFull timeWork experience placementWork at officeRemote workWork from homeRelocation package$174k - $238k
...The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging Products Group... ...States, as defined in Federal Acquisition Regulation (FAR) 2.101U.S. Security Clearance status - the employee must be...Local areaWorldwideFlexible hours- Job TitleSoftware Engineer Level 2LocationWashington, DC 20032 US (Primary)CategoryResearch, Development, and EngineeringJob TypeFull-TimeCareer... ...DescriptionPrescient Edge is seeking a Software Engineer Level 2 to support a federal government client.Please note that the...Contract work
$112k - $218.4k
...: Our team is looking for a Senior Active Directory Site Reliability Engineer. Our mission is to improve the availability, latency, performance and... ...Science, Information Technology, or related field AND 2+ years technical experience in software engineering, network...Full timeLocal area- ...Black Cape Title: Spectacular Full Stack Software Engineer (2+ years - Senior Level) Locations: Arlington, VA, Reston, VA and areas throughout the DMV area Onsite Expectations: MUST be willing to go onsite 2 - 3 days per week; could be up to 5 days per week...Full timeFor subcontractor2 days per week3 days per week
- ...Black Cape Title: FrontEnd Software Engineer (2+ years to Senior Levels) Locations: Arlington, VA, Reston, VA and areas throughout... ...environments or willingness to work at classified government sites as needed Experience developing on MacOSX, Windows, or Linux...Full timeFor subcontractor
$130k - $138k
...Software Engineer Level 2 Position: Software Engineer Level 2 (Visualization Developer) Location: Laurel, MD (On-site) Category: Software Engineering Schedule: Standard Day Shift, Monday–Friday Clearance Requirement: Active TS/SCI with Polygraph Experience Requirement...Temporary workMonday to FridayFlexible hoursDay shift$160k - $210k
...industry. Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service... ...hybrid work schedule of 3 days in office (Mon/Tue/Wed) and 2 days remote (Thursday/Friday). Responsibilities Design...Work at officeImmediate startRemote workWork from home$98.8k - $164.6k
...Communications, Customer Care, Engineering & Product, Finance, Human... ...engineering decisions, and help build reliable, maintainable systems that... ...equivalent practical experience.2+ years of professional... ...and business travel, we work on-site five days a week.Compensation...Full timePart time$115.5k - $164.8k
...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant... ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations...Work experience placementWork at officeRemote work$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...Full timeLocal areaImmediate start$166k - $220k
...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the... ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine...Full timeWork experience placementImmediate start$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....Permanent employmentTemporary workImmediate startWeekend work$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...Remote work$86.8k - $198k
Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether... ...environments, including in Amazon Web Services (AWS)2+ years of experience with monitoring and observability...Full timeContract workPart timeWork at officeLocal areaRemote work$112k - $179k
...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems...Contract workWorldwideShift work$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...Permanent employmentFull timeWork at officeLocal areaRemote workWorldwide- ...position. Position Summary: ISI is looking for a Project Engineer Level 2 to provide Owner's Representative construction management... ...environments. Responsibilities: Assist the Government in site evaluations, field surveys, and site visits to assess...Permanent employmentFor contractorsWork experience placementWork at officeMonday to Friday
- ...Associate or Mid-Level Flight Deck Design Engineer to join a team of highly skilled... ...objectives and requirements and impacts on data reliability and integrity, fault trees analysis (FTAs... ...Experience: \tBachelor's degree and typically 2 or more years' experience in an...Work experience placement
- ...solutions using a tailored Agile methodology. We are seeking a highly motivated and intellectually curious Senior Site Reliability Engineer to join our team working with a Federal client. The position will be a remote role open to US citizens residing in the...Remote work
- ...Site Reliability Engineer ValidaTek is building teams of Site Reliability Engineers (SRE's) to support internal and external engineering and operations of a large scale and world-wide Enterprise IT environment that covers application hosting and support, enterprise...
- ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable platform by developing systems for analyzing environments, predicting, and resolving issues, and supporting the production environment...Work experience placement
$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office$107k - $220k
...The Site Reliability Engineer (SRE) will ensure the reliability, performance, and scalability of the WDP System. This person will define and track Key Performance Indicators (KPIs) and Service Level Objectives (SLOs), identify and resolve performance bottlenecks, and perform...Full timeContract workTemporary workWork at officeVisa sponsorshipWork visa- ...Catalyst, and In-Q-Tel. Mission | On Site | Full Time | Active TS/SCI with Full... ...government customer site, ensuring the reliability and performance of Twenty's mission-critical... ...technical ownership and customer-facing engineering: you'll define how we measure...Full timeContract workRemote workFlexible hours
- OB SUMMARYThe Systems Engineer - Site Reliability Engineering (SRE) is responsible for the reliability, scalability, and performance of mission-critical cloud and on-prem services that support millions of Marriot customers globally. This role involves overseeing incident...Full timeFor contractorsWork at officeRemote workFlexible hoursShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer 2. Be the first to apply!
- site reliability engineer Washington DC
- site reliability engineer remote Washington DC
- site reliability engineer sre Washington DC
- site recruiter Washington DC
- junior website developer Washington DC
- on site coordinator Washington DC
- construction site safety Washington DC
- site services specialist Washington DC
- website content developer Washington DC
- website coordinator Washington DC




