Senior Site Reliability Engineer
$210k - $230kGovcio
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation, reliability, and performance optimization across multi-cloud environments. This position is located in Arlington, VA and is a hybrid remote/onsite position.ResponsibilitiesKey Responsibilities:Infrastructure & Automation• Design, deploy, and manage cloud infrastructure using Infrastructure as Code (IaC) principles• Develop and maintain Terraform modules for AWS and Azure environments• Create and manage Ansible playbooks for configuration management and application deployment• Implement CI/CD pipelines using GitHub Actions to automate build, test, and deployment processes• Implement GitOps workflows for declarative infrastructure and application delivery• Build self-service tools and platforms to enable development teamsReliability & Performance• Establish and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs)• Implement comprehensive monitoring, logging, and alerting solutions• Conduct capacity planning and performance tuning• Perform root cause analysis and implement preventive measures• Design and execute chaos engineering experiments to validate system resilienceDisaster Recovery & Business Continuity• Design and implement disaster recovery strategies across multi-cloud environments• Develop and maintain backup and restore procedures• Create and test business continuity plans• Implement automated failover mechanisms• Document recovery time objectives (RTO) and recovery point objectives (RPO)Cloud Operations• Manage and optimize AWS services (EC2, S3, RDS, Lambda, ECS, EKS, CloudWatch, etc.)• Manage and optimize Azure services (VMs, Storage, SQL Database, AKS, Monitor, etc.)• Implement cost optimization strategies and resource tagging• Ensure security best practices and compliance requirements• Manage identity and access management (IAM) policiesCollaboration & Leadership• Participate in on-call rotation and incident response• Collaborate with development teams on architecture and design decisions• Mentor team members on SRE practices and tools• Document systems, processes, and runbooks• Drive continuous improvement initiativesQualificationsRequired Education and Experience• Bachelor’s Degree with 12+ yrs experience• Clearance Level: Active Secret with the ability to obtain and hold DEA suitabilityTechnical Skills• Cloud Platforms: 3+ years of hands-on experience with AWS and Azure• Infrastructure as Code: Expert-level proficiency with Terraform• Configuration Management: Strong experience with Ansible• Scripting: Proficiency in Python, Bash, or PowerShell• Containerization: Experience with Docker and Kubernetes• Version Control: Strong Git and GitHub workflow knowledge• GitOps: Experience implementing GitOps practices and workflows• Monitoring Tools: Experience with Prometheus, Grafana, ELK Stack, or similar• CI/CD: Hands-on experience with GitHub Actions, Jenkins, GitLab CI, or Azure DevOpsCore Competencies• Deep understanding of Microsoft/Linux systems administration• Strong networking knowledge (TCP/IP, DNS, load balancing, VPN)• Experience with database administration (PostgreSQL, MySQL, SQL Server)• Knowledge of security best practices and compliance frameworks• Understanding of microservices architecture and distributed systems• Experience with disaster recovery planning and executionSoft Skills• Excellent problem-solving and analytical abilities• Strong communication skills, both written and verbal• Ability to work independently and in team environments• Customer-focused mindset with emphasis on reliability• Adaptability to rapidly changing technologies and requirementsPreferred Qualifications• AWS Certified Solutions Architect or SysOps Administrator• Azure Administrator or Solutions Architect certification• Certified Kubernetes Administrator (CKA)• HashiCorp Certified: Terraform Associate• GitHub Certified or demonstrated expertise with GitHub Enterprise• Experience with service mesh technologies (Istio, Linkerd)• Knowledge of observability platforms (Datadog, New Relic, Dynatrace)• Experience with GitOps tools and practices (ArgoCD, Flux, GitHub Actions for GitOps)• Familiarity with compliance frameworks (SOC 2, HIPAA, FedRAMP)• Previous experience in a DevOps or Platform Engineering rolePosted Salary RangeUSD $210,000.00 - USD $230,000.00 /Yr.
- ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS...SeniorTemporary work
$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...SeniorPermanent employmentFull timeWork at officeLocal areaRemote workWorldwide$166k - $220k
...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the... ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Costa...SeniorFull timeWork experience placementImmediate start$207k - $284.9k
...We're all in on this mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from AI to... ...mission. If you are too, let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team supports government...SeniorPermanent employmentLocal areaWorldwideFlexible hoursDay shift$166k - $220k
...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &... ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril...SeniorFull timeWork experience placementImmediate start$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to... ...ensuring scalability, performance, and reliability across environments. What You’ll Do Design...SeniorFull timeRemote workRelocationRelocation package$168k - $200k
...that is passionate about creating transformative change in healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the forefront of building and operating a resilient, observable, and...Senior- ...Azure, Oracle, Cassandra, SQL Server, My SQL and Mongo DB Seniority level Seniority level Mid-Senior level Employment type Employment... ...new job is posted. Sign in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259,000.00 1 day ago...SeniorContract workRemote work
$149.4k - $202k
...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing...SeniorRemote work$106.3k - $221.1k
...more. Join us to drive positive, lasting change that moves missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability of the Client System. The engineer will define and track Key...SeniorLive inWork at officeLocal area$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SeniorWork experience placementWork at office$91.4k - $187k
...best practices, and ability to develop tools that automate incident management. Description We are looking for a Senior Site Reliability Engineer to join our OCI team. This role is part of a globally distributed team responsible for detecting, triaging, and...SeniorTemporary workWork experience placementFlexible hours$153k - $185k
...Senior Site Reliability Engineer El Segundo, California, United States About Varda Low Earth orbit is open for business. Varda is accelerating the development of commercial space infrastructure, from in-orbit pharmaceutical processing to reliable and economical...SeniorPermanent employmentFull timeImmediate startRelocation packageFlexible hoursWeekend work$149.4k - $202k
Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on...SeniorRemote work$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering , this role will...SeniorPermanent employmentLocal areaWorldwideFlexible hours- Senior Site Reliability Engineer - Network Operations (Remote) Fastly 15 August 2025 SRE DevOps Automation Networking BGP Fastly is seeking a Senior Site Reliability Engineer (Networking) to join our Technical Operations (TechOps) team. You will be responsible for building...SeniorRemote jobLocal areaFlexible hoursNight shift
$147k - $202.4k
...excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Senior Site Reliability Engineer (SRE) - Security and Data Systems Our company is seeking a highly skilled Senior Site Reliability Engineer to join...SeniorPermanent employmentWork at officeLocal areaWorldwideFlexible hoursShift work$232k - $319k
...to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and... ...with self-service Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and...SeniorPermanent employmentLocal areaWorldwideFlexible hours$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...SeniorFull timeLocal areaImmediate start$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....SeniorPermanent employmentTemporary workImmediate startWeekend work$81.1k - $187k
...service according to terms for reliability and functionality.- Assists... ...deployments.- Gains basic knowledge of site reliability trends and shares... ...and escalate issues to senior team members. Collects and... ...a skilled Site Reliability Engineer to design, build, operate, and...SeniorTemporary workImmediate startFlexible hoursShift work$121.5k - $264.1k
...guidance on practices and terms for reliability and functionality.-... ...and Resolution:- Serves as a senior management escalation point for... ...and maintaining knowledge of site reliability trends and sharing... ...years of experience in software engineering, infrastructure management,...SeniorTemporary workImmediate startFlexible hours$118k - $177k
...Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation Continuous Diagnostics and Mitigation (CDM) Cyber data solution...SeniorRemote work$81.1k - $187k
.... You'll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... .... Responsibilities Escalation points for junior Site Reliability Engineers during complex or high-impact incidents. Manage...SeniorTemporary workWork experience placementMonday to FridayFlexible hoursShift workNight shift$81.1k - $187k
...Infrastructure Engineer Takes proactive steps to design and architect infrastructure and service to ensure reliability and functionality. Forecasts demands and responds to capacity... ...potential impact and develops knowledge of site reliability trends. Key...SeniorTemporary workFlexible hours- ...Job Description Job Description Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian of our production ecosystems, ensuring that our complex, data-driven AI platforms remain resilient...SeniorLocal area
- ...Job Description Job Description Description: Onsite in Washington, DC our client seeks a Sr. Site Reliability Engineer III to design, automate, and operate mission-critical systems for federal environments. The role focuses on Kubernetes or VMWare platforms,...SeniorHourly payPermanent employmentFull timeLocal areaImmediate start
- ...Manager is backfilling a principal engineer and long-tenured VMware/Kubernetes expert... ...Job Description- Position Title: Senior SRE/DevOps Engineer Location Information... ...Terraform and Ansible, ensuring rapid, reliable, and auditable platform provisioning and...SeniorContract workRemote work
$133k - $190k
Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. Direct Hire Role **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is...Full time$125k - $185k
Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions... ...locate missing children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- srs distribution Arlington, VA
- senior associate architect Arlington, VA
- senior dynamics crm developer Arlington, VA
- senior application security Arlington, VA
- senior account director Arlington, VA
- senior supervisor Arlington, VA
- senior plumbing designer Arlington, VA
- senior advisor Arlington, VA
- senior cloud data engineer Arlington, VA
- senior technical analyst Arlington, VA


