Sr. Site Reliability Engineer
Apptad Inc
SRE Frisco, TX & Bellevue, WA
Key Responsibilities
- Monitor, troubleshoot, and resolve production incidents impacting availability, latency, and performance .
- Provision and manage Azure infrastructure, including VMs, networking, storage, and IAM .
- Support the Azure-to-TKE migration , including containerization, Kubernetes deployments, testing, and cutover activities.
- Design and maintain CI/CD pipelines for automated deployment and testing.
- Build and maintain monitoring, dashboards, alerting, logging, and health-check solutions.
- Implement Infrastructure as Code (IaC) and automation to reduce manual operational effort.
- Support capacity planning, performance optimization, reliability, and security initiatives.
- Collaborate with engineering and infrastructure teams on incident response and migration readiness .
- Participate in on-call support as required.
Required Qualifications
- Bachelor's degree in Computer Science, Engineering, IT, or related field, or equivalent practical experience.
- 4+ years of experience in SRE, DevOps, Cloud Infrastructure, or related roles.
- Strong production experience with Microsoft Azure .
- Hands-on experience with Kubernetes and Docker .
- Strong troubleshooting and incident-response capabilities.
- Experience supporting cloud infrastructure and production workloads.
Required Technical Skills
Cloud:
- Microsoft Azure
- Azure VMs
- Azure Networking
- Azure Storage
- Azure IAM
- Azure Monitor
Containers & Kubernetes:
- Kubernetes
- Docker
- Tencent Kubernetes Engine (TKE) preferred
- Containerization and Kubernetes deployments
CI/CD:
- Azure DevOps
- Jenkins
- GitLab CI/CD
Infrastructure as Code:
- Terraform
- ARM Templates
- Bicep
Scripting & Automation:
- Python
- Bash
- PowerShell
Monitoring & Observability:
- Prometheus
- Grafana
- Azure Monitor
- Logging, alerting, dashboards, and health checks
Preferred Qualifications
- Experience with Tencent Kubernetes Engine (TKE) or other managed Kubernetes platforms.
- Experience migrating workloads from cloud VMs to Kubernetes .
- Experience containerizing legacy applications.
- Experience with Azure-to-Kubernetes cloud migration projects.
- Azure certifications such as AZ-104, AZ-400, or AZ-500 .
- Kubernetes certifications such as CKA or CKAD .
- Experience with capacity planning, high availability, disaster recovery, and performance optimization.
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Sr. Site Reliability Engineer in Washington DC vacancy
$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...SeniorFull timeLocal areaImmediate start$165k - $230k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts...SeniorPermanent employmentTemporary workImmediate startWeekend work$132.23k - $176.31k
...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...SeniorTemporary workRemote work- ...Job Description Job Description Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian of our production ecosystems, ensuring that our complex, data-driven AI platforms remain resilient...SeniorLocal area
- ...Job Description Job Description Description: Onsite in Washington, DC our client seeks a Sr. Site Reliability Engineer III to design, automate, and operate mission-critical systems for federal environments. The role focuses on Kubernetes or VMWare platforms,...SeniorHourly payPermanent employmentFull timeLocal areaImmediate start
$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...SeniorCurrently hiringRemote work$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...SeniorPermanent employmentFull timeWork at officeLocal areaRemote workWorldwide$166k - $220k
...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the... ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine...SeniorFull timeWork experience placementImmediate start$207k - $284.9k
...on this mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from AI to HumanIdentity is... ...mission. If you are too, let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team supports government customers...SeniorPermanent employmentLocal areaWorldwideFlexible hoursDay shift- ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable platform by developing systems for analyzing environments, predicting, and resolving issues, and supporting the production environment...SeniorWork experience placement
$106.3k - $221.1k
...more. Join us to drive positive, lasting change that moves missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability of the Client System. The engineer will define and track Key...SeniorLive inWork at officeLocal area$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...SeniorWork experience placementWork at office- ...Excellent analytical and problem-solving skills with a proactive approach. AI/ML experience or a strong interest in applying AI/ML to reliability, security, or operational efficiency is a plus. Benefits Health, dental, and vision insurance; 401(k); flexible spending...SeniorFull timeWork at officeFlexible hours
- ...Position: Senior Site Reliability Engineer (SRE) Location: Redmond WA (Onsite) Duration: Fulltime Job Description We are seeking a highly skilled and hands-on Senior Site Reliability Engineer (SRE) to design, build, automate, and operate mission...SeniorFull time
$147k - $202.4k
...-defining work. We're all in on this mission. If you are too, let's talk. Our company is seeking a highly skilled Senior Site Reliability Engineer to join our team. We are a SaaS company specializing in securing large-scale systems. This role is a blend of software engineering...SeniorWork at officeLocal areaWorldwideFlexible hoursShift work$112k - $218.4k
...: Our team is looking for a Senior Active Directory Site Reliability Engineer. Our mission is to improve the availability, latency, performance and security of the Identity systems behind Microsoft's cloud. Like traditional operations, we keep important revenue-critical...SeniorFull timeLocal area$175k - $250k
Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or... ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design,...SeniorFull timeRemote workRelocationRelocation package$160k - $210k
...change and achieving remarkable growth in a rapidly evolving industry. Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management across Cognitiv. Our immediate challenge is to scale...SeniorWork at officeImmediate startRemote workWork from home$125k - $185k
Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions... ...locate missing children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package- OB SUMMARYThe Systems Engineer - Site Reliability Engineering (SRE) is responsible for the reliability, scalability, and performance of mission-critical cloud and on-prem services that support millions of Marriot customers globally. This role involves overseeing incident...Full timeFor contractorsWork at officeRemote workFlexible hoursShift work
$112k - $179k
...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems...Contract workWorldwideShift work$125k - $185k
...lifesaving drugs, forecast supply chain disruptions, locate missing children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-performance, scalable, and reliable services for our production...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...Remote work$115.5k - $164.8k
...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant... ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations...Work experience placementWork at officeRemote work- ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS...Temporary work
$182k - $250.8k
...Team at Okta is the backbone of our platform's reliability and operational excellence. We are a forward-thinking group of engineers and leaders who believe that great... ...for millions of users worldwide. As a Manager, Site Reliability Engineer, you'll lead this team with...Permanent employmentLocal areaRemote workWorldwideFlexible hoursWeekend workWeekday work$174k - $239k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc. is looking for an experienced Staff TDI Site Reliability...Local areaWorldwideFlexible hours$174k - $238k
...work. We're all in on this mission. If you are too, let's talk.The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging Products Group (EPG). Our mission is to build highly reliable, scalable,...Local areaWorldwideFlexible hours- ...ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running... ...of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident...Full timeContract workRemote workFlexible hours
- ...Site Reliability Engineer Qualifications: ~10+ years of overall experience in IT including, with hands-on Development and Systems engineering background ~3-5 years of experience in a Site Reliability Engineering role ~ Experience with Enterprise Cloud transformation...Temporary workImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr. Site Reliability Engineer. Be the first to apply!
Related searches
- site reliability engineer Washington DC
- site reliability engineer remote Washington DC
- site reliability engineer sre Washington DC
- senior technical analyst Washington DC
- senior associate attorney Washington DC
- senior developer Washington DC
- sr industrial security specialist Washington DC
- senior aws cloud engineer Washington DC
- senior manager business development Washington DC
- remote senior salesforce administrator Washington DC


