Site Reliability Engineer
VDart Inc
Job Title: Site Reliability Engineer
Location: Bellevue, WA & Frisco, TX
Duration: / Term: Contract
Experience Desired: 7+ Years
Job Description:
This role serves as a vendor-provided Site Reliability Engineer responsible for improving and protecting the reliability, scalability, and performance of a platform currently hosted on Microsoft Azure that will be migrated to Tencent Kubernetes Engine (TKE) in the future. It manages availability, latency, performance, security, and capacity while enabling efficient, automated software delivery across both the current Azure environment and the upcoming TKE platform. The role differentiates by combining deep Azure cloud infrastructure expertise with Kubernetes-based container orchestration, positioning the team for a smooth cloud-to-TKE migration. Success is measured by improved system uptime, faster incident resolution, and a reliable, well-supported migration path. The work directly impacts IT service quality, operational resilience, and customer experience through the transition and beyond.
Main Responsibilities
- Monitor, troubleshoot, and resolve incidents affecting availability, latency, and performance of current Azure-hosted workloads
- Provision, configure, and manage Azure infrastructure (VMs, networking, storage, IAM) to support production and non-production environments
- Support planning and execution of the platform's migration from Azure to Tencent Kubernetes Engine (TKE), including workload containerization and cutover activities
- Design, build, and maintain CI/CD pipelines that support automated deployment and testing today on Azure and going forward on TKE
- Build and maintain observability tooling - dashboards, alerts, logging, and health checks - to proactively identify and address system risks across both environments
- Drive automation and infrastructure-as-code practices to reduce manual toil and improve deployment consistency
- Collaborate with internal engineering teams and stakeholders to support incident response, capacity planning, and migration readiness
Required Qualifications
Education: Bachelor's degree in Computer Science, Engineering, or related field, or equivalent practical experience.
Experience: 4+ years in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles, including hands-on production experience with Microsoft Azure (compute, networking, storage, IAM). Experience with Tencent Kubernetes Engine (TKE) or comparable Kubernetes platforms strongly preferred, as the platform will migrate to TKE.
Technical Skills: Proficiency with CI/CD tooling (e.g., Azure DevOps, Jenkins, GitLab CI), containerization and orchestration (Kubernetes, Docker, TKE), infrastructure-as-code (Terraform, ARM/Bicep), scripting/automation (Python, Bash, PowerShell), and monitoring/observability platforms (e.g., Grafana, Prometheus, Azure Monitor).
Other: Strong troubleshooting and incident-response skills; experience supporting cloud platform migrations a plus; ability to work effectively as an embedded vendor resource within a client engineering team; on-call availability as required.
Preferred Qualifications
Azure certifications (e.g., AZ-104, AZ-400, AZ-500) and/or Kubernetes certifications (CKA/CKAD). Prior experience migrating workloads from a cloud VM-based platform to Kubernetes/TKE, including containerization of legacy services.
Key Skills:
SRE, Azure, Kubernetes, Terraform, observability
$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...SuggestedPermanent employmentFull timeWork at officeLocal areaRemote workWorldwide$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...SuggestedRemote work$112k - $179k
...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems...SuggestedContract workWorldwideShift work$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...SuggestedFull timeLocal areaImmediate start$166k - $220k
...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the... ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine...SuggestedFull timeWork experience placementImmediate start- ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS...Temporary work
$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...Currently hiringRemote work$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....Permanent employmentTemporary workImmediate startWeekend work$115.5k - $164.8k
...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant... ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations...Work experience placementWork at officeRemote work$125k - $174.33k
...Learn more on Life at Coupa blog and hear from our employees about their experiences working at Coupa. The Impact of a Lead Site Reliability Engineer at Coupa: Joining the team as a Lead Site Reliability Engineer - Active Directory on the Coupa Cloud Operations team will...- ...Site Reliability Engineer Qualifications: ~10+ years of overall experience in IT including, with hands-on Development and Systems engineering background ~3-5 years of experience in a Site Reliability Engineering role ~ Experience with Enterprise Cloud transformation...Temporary workImmediate start
$106.3k - $221.1k
...more. Join us to drive positive, lasting change that moves missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability of the Client System. The engineer will define and track Key...Live inWork at officeLocal area- ...ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running... ...of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident...Full timeContract workRemote workFlexible hours
$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...Work experience placementWork at office$190k - $225k
...Job TitleSRE + Release Pipeline EngineerJob DescriptionSRE + Release Pipeline Engineer Clearance: Top Secret clearance Location: Remote Role Framing Owns the path from local development to deployed AWS cluster across DCSA's GovCloud (IL2/IL5) and classified (IL6/Secret...Full timePart timeWork experience placementLocal areaRemote work$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$125k - $185k
...lifesaving drugs, forecast supply chain disruptions, locate missing children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-performance, scalable, and reliable services for our production...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package- OB SUMMARYThe Systems Engineer - Site Reliability Engineering (SRE) is responsible for the reliability, scalability, and performance of mission-critical cloud and on-prem services that support millions of Marriot customers globally. This role involves overseeing incident...Full timeFor contractorsWork at officeRemote workFlexible hoursShift work
$125k - $185k
Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions... ...locate missing children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$182k - $250.8k
...Team at Okta is the backbone of our platform's reliability and operational excellence. We are a forward-thinking group of engineers and leaders who believe that great... ...for millions of users worldwide. As a Manager, Site Reliability Engineer, you'll lead this team with...Permanent employmentLocal areaRemote workWorldwideFlexible hoursWeekend workWeekday work$174k - $238k
...work. We're all in on this mission. If you are too, let's talk.The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging Products Group (EPG). Our mission is to build highly reliable, scalable,...Local areaWorldwideFlexible hours$207k - $284.9k
...on this mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from AI to HumanIdentity is... ...mission. If you are too, let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team supports government customers...Permanent employmentLocal areaWorldwideFlexible hoursDay shift$174k - $239k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc. is looking for an experienced Staff TDI Site Reliability...Work experience placementLocal areaWorldwideFlexible hours$175k - $250k
Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or... ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design,...Full timeRemote workRelocationRelocation package- ...Job Description Job Description Description: Onsite in Washington, DC our client seeks a Sr. Site Reliability Engineer III to design, automate, and operate mission-critical systems for federal environments. The role focuses on Kubernetes or VMWare platforms,...Hourly payPermanent employmentFull timeLocal areaImmediate start
$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering , this role will...Permanent employmentLocal areaWorldwideFlexible hours$147k - $202.4k
...-defining work. We're all in on this mission. If you are too, let's talk. Our company is seeking a highly skilled Senior Site Reliability Engineer to join our team. We are a SaaS company specializing in securing large-scale systems. This role is a blend of software engineering...Work at officeLocal areaWorldwideFlexible hoursShift work$135k - $155k
...Responsibilities Own assigned infrastructure and software reliability problem statements through completion with guidance from senior engineers. Write, improve, and document efficient code and existing systems. Develop, configure, and maintain AWS accounts,...Full time$112k - $218.4k
...: Our team is looking for a Senior Active Directory Site Reliability Engineer. Our mission is to improve the availability, latency, performance and security of the Identity systems behind Microsoft's cloud. Like traditional operations, we keep important revenue-critical...Full timeLocal area- GovCIO LLC seeks a Senior Site Reliability Engineer to ensure reliability, scalability, and performance of mission-critical federal IT systems. You will design and automate cloud infrastructure, build robust CI/CD pipelines, and implement monitoring, alerting, and incident...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer Washington DC
- site reliability engineer remote Washington DC
- site reliability engineer sre Washington DC
- site recruiter Washington DC
- junior website developer Washington DC
- on site coordinator Washington DC
- construction site safety Washington DC
- site services specialist Washington DC
- website content developer Washington DC
- website coordinator Washington DC


