Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$210k - $230k

Govcio LLC

Overview:

GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation, reliability, and performance optimization across multi-cloud environments. This position is located in Arlington, VA and is a hybrid remote/onsite position.

Responsibilities:

Key Responsibilities:


Infrastructure & Automation
• Design, deploy, and manage cloud infrastructure using Infrastructure as Code (IaC) principles
• Develop and maintain Terraform modules for AWS and Azure environments
• Create and manage Ansible playbooks for configuration management and application deployment
• Implement CI/CD pipelines using GitHub Actions to automate build, test, and deployment processes
• Implement GitOps workflows for declarative infrastructure and application delivery
• Build self-service tools and platforms to enable development teams


Reliability & Performance
• Establish and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs)
• Implement comprehensive monitoring, logging, and alerting solutions
• Conduct capacity planning and performance tuning
• Perform root cause analysis and implement preventive measures
• Design and execute chaos engineering experiments to validate system resilience


Disaster Recovery & Business Continuity
• Design and implement disaster recovery strategies across multi-cloud environments
• Develop and maintain backup and restore procedures
• Create and test business continuity plans
• Implement automated failover mechanisms
• Document recovery time objectives (RTO) and recovery point objectives (RPO)


Cloud Operations
• Manage and optimize AWS services (EC2, S3, RDS, Lambda, ECS, EKS, CloudWatch, etc.)
• Manage and optimize Azure services (VMs, Storage, SQL Database, AKS, Monitor, etc.)
• Implement cost optimization strategies and resource tagging
• Ensure security best practices and compliance requirements
• Manage identity and access management (IAM) policies


Collaboration & Leadership
• Participate in on-call rotation and incident response
• Collaborate with development teams on architecture and design decisions
• Mentor team members on SRE practices and tools
• Document systems, processes, and runbooks
• Drive continuous improvement initiatives

Qualifications:

Required Education and Experience
• Bachelor’s Degree with 12+ yrs experience
• Clearance Level: Active Secret with the ability to obtain and hold DEA suitability

Technical Skills
• Cloud Platforms: 3+ years of hands-on experience with AWS and Azure
• Infrastructure as Code: Expert-level proficiency with Terraform
• Configuration Management: Strong experience with Ansible
• Scripting: Proficiency in Python, Bash, or PowerShell
• Containerization: Experience with Docker and Kubernetes
• Version Control: Strong Git and GitHub workflow knowledge
• GitOps: Experience implementing GitOps practices and workflows
• Monitoring Tools: Experience with Prometheus, Grafana, ELK Stack, or similar
• CI/CD: Hands-on experience with GitHub Actions, Jenkins, GitLab CI, or Azure DevOps


Core Competencies
• Deep understanding of Microsoft/Linux systems administration
• Strong networking knowledge (TCP/IP, DNS, load balancing, VPN)
• Experience with database administration (PostgreSQL, MySQL, SQL Server)
• Knowledge of security best practices and compliance frameworks
• Understanding of microservices architecture and distributed systems
• Experience with disaster recovery planning and execution


Soft Skills
• Excellent problem-solving and analytical abilities
• Strong communication skills, both written and verbal
• Ability to work independently and in team environments
• Customer-focused mindset with emphasis on reliability
• Adaptability to rapidly changing technologies and requirements

Preferred Qualifications
• AWS Certified Solutions Architect or SysOps Administrator
• Azure Administrator or Solutions Architect certification
• Certified Kubernetes Administrator (CKA)
• HashiCorp Certified: Terraform Associate
• GitHub Certified or demonstrated expertise with GitHub Enterprise
• Experience with service mesh technologies (Istio, Linkerd)
• Knowledge of observability platforms (Datadog, New Relic, Dynatrace)
• Experience with GitOps tools and practices (ArgoCD, Flux, GitHub Actions for GitOps)
• Familiarity with compliance frameworks (SOC 2, HIPAA, FedRAMP)
• Previous experience in a DevOps or Platform Engineering role

Posted Salary Range: USD $210,000.00 - USD $230,000.00 /Yr.
Vacancy posted 6 hours ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Arlington, VA vacancy
  •  ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS... 
    Senior
    Temporary work

    Kong

    Washington DC
    3 days ago
  • $150k - $180k

     ...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business...  ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or... 
    Senior
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work
    Worldwide

    Umbra

    Arlington, VA
    3 days ago
  • $166k - $220k

     ...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the...  ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    3 days ago
  • $207k - $284.9k

     ...We're all in on this mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from AI to...  ...mission. If you are too, let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team supports government... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours
    Day shift

    Okta

    Washington DC
    1 day ago
  • $166k - $220k

     ...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &...  ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    1 day ago
  • $149.4k - $202k

     ...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing... 
    Senior
    Remote work

    Noctua Technology

    Washington DC
    13 hours ago
  •  ...foundation of success and bringing it to the digital space - ready to join us? What’s the position? We are looking for a Senior Site Reliability Engineer who combines deep infrastructure expertise with a forward-thinking approach to AI-driven operations. In this role you... 
    Senior
    Remote work
    Flexible hours
    Night shift

    GrabJobs

    Washington DC
    1 day ago
  • $82.3k - $228.8k

     ...an inclusive environment, empowering our employees to be their authentic selves. We are seeking a highly experienced Senior Site Reliability Engineer – Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud... 
    Senior
    Temporary work
    Work at office
    Remote work
    Worldwide
    3 days per week

    GrabJobs

    Washington DC
    4 days ago
  • $175k - $250k

    Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to...  ...ensuring scalability, performance, and reliability across environments. What You’ll Do... 
    Senior
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    Washington DC
    2 days ago
  •  ...Azure, Oracle, Cassandra, SQL Server, My SQL and Mongo DB Seniority level Seniority level Mid-Senior level Employment type Employment...  ...new job is posted. Sign in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259,000.00 1 day ago... 
    Senior
    Contract work
    Remote work

    Signature IT World Inc

    Washington DC
    2 days ago
  •  ...Candidates eligible for polygraph upgrade will be considered. Position Overview We are seeking an experienced Senior DevOps / Site Reliability Engineer (SRE) to support a mission-critical national security program. This position is ideal for a hands-on engineer with... 
    Senior
    Full time
    Visa sponsorship

    Omniscius Consulting

    Washington DC
    8 days ago
  • $185k - $230k

    As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services... 
    Senior
    Full time
    Local area
    Immediate start

    MetroStar Systems

    Washington DC
    4 days ago
  • $119.8k - $234.7k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole...  ...EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...at a rapid pace. We empower engineers to deliver creative solutions...  ...industry. We’re looking for a Senior Site Reliability Engineer and... 
    Senior
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Reston, VA
    4 days ago
  • $165k - $230k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts.... 
    Senior
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    1 day ago
  • $81.1k - $187k

     ...service according to terms for reliability and functionality.- Assists...  ...deployments.- Gains basic knowledge of site reliability trends and shares...  ...and escalate issues to senior team members. Collects and...  ...a skilled Site Reliability Engineer to design, build, operate, and... 
    Senior
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle Corporation

    Reston, VA
    2 days ago
  • $121.5k - $264.1k

     ...guidance on practices and terms for reliability and functionality.-...  ...and Resolution:- Serves as a senior management escalation point for...  ...and maintaining knowledge of site reliability trends and sharing...  ...years of experience in software engineering, infrastructure management,... 
    Senior
    Temporary work
    Immediate start
    Flexible hours

    Oracle Corporation

    Reston, VA
    2 days ago
  •  ...Job Description Job Description Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian of our production ecosystems, ensuring that our complex, data-driven AI platforms remain resilient... 
    Senior
    Local area

    Tiger Analytics Inc.

    Washington DC
    more than 2 months ago
  •  ...Job Description Job Description Description: Onsite in Washington, DC   our client seeks a Sr. Site Reliability Engineer III to design, automate, and operate mission-critical systems for federal environments. The role focuses on Kubernetes or VMWare platforms,... 
    Senior
    Hourly pay
    Permanent employment
    Full time
    Local area
    Immediate start

    Eliassen Group

    Washington DC
    23 days ago
  • $165k - $225.6k

     ...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering , this role will... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    9 days ago
  • $133k - $190k

    Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. Direct Hire Role **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is... 
    Full time

    SOC Support Services

    McLean, VA
    1 day ago
  • $166k - $220k

     ...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the...  ...end solutions we ship. ABOUT THE JOB We are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    2 days ago
  • $230k - $250k

    GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is... 
    Remote work

    Govcio

    Arlington, VA
    4 days ago
  • $125k - $185k

     ...lifesaving drugs, forecast supply chain disruptions, locate missing children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-performance, scalable, and reliable services for our production... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    4 days ago
  • $80k - $133k

     ...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and... 
    Permanent employment
    Full time
    Contract work
    Remote work
    Flexible hours

    Guidehouse

    McLean, VA
    1 day ago
  • $115.5k - $164.8k

     ...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant...  ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations... 
    Work experience placement
    Work at office
    Remote work

    Axon

    Washington DC
    1 day ago
  • $112k - $179k

     ...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems... 
    Contract work
    Worldwide
    Shift work

    Peraton Corporation

    Washington DC
    4 days ago
  • $125k - $185k

    Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions...  ...locate missing children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    1 day ago
  • $174k - $239k

     ...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc. is looking for an experienced Staff TDI Site Reliability... 
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    3 days ago
  • $174k - $238k

     ...work. We're all in on this mission. If you are too, let's talk.The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging Products Group (EPG). Our mission is to build highly reliable, scalable,... 
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    1 day ago
  •  ...SRE Support Engineer - Observability While this position is not currently open, we are interviewing strong candidates for upcoming opportunities...  ...support across Slack and tickets, improving monitoring reliability, and reducing incident impact through better triage,... 
    Remote work

    GrabJobs

    Washington DC
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!