Site Reliability Engineer
Software Technology Inc
Site Reliability Engineer (SRE)
Randstad is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our client in the Washington D.C. area, focusing on optimizing the availability, performance, and scalability of critical production services. The ideal candidate will bridge the gap between development and operations by applying software engineering principles to infrastructure and operational problems. This role requires a strong background in CI/CD pipeline development, infrastructure automation using Infrastructure-as-Code (IaC), incident response, and deep experience with cloud platforms, preferably AWS. The SRE will collaborate across engineering teams to drive automation, enhance observability, and ensure the continuous, secure delivery of high-quality software.
Responsibilities
- Deployment & Automation: Design, build, and maintain robust Continuous Integration/Continuous Delivery (CI/CD) pipelines utilizing tools suchs as GitHub Actions, Jenkins, or AWS CodePipeline.
- Infrastructure-as-Code (IaC): Automate the provisioning and management of cloud infrastructure using IaC tools like Terraform, CloudFormation, or AWS CDK.
- Monitoring & Observability: Develop comprehensive monitoring dashboards, alerting rules, and logging configurations using platforms such as AppDynamics, CloudWatch, or Dynatrace to proactively ensure systems meet defined Service Level Objectives (SLOs).
- Incident Response & Remediation: Participate in a rotating on-call schedule, triage and resolve high-priority incidents, and conduct blameless postmortem reviews to identify and implement root cause remediations.
- Security & Compliance: Contribute to a DevSecOps culture by assisting with secrets management and integrating security scanning tools (e.g., AWS ECR, Checkmarx, Synk) directly into CI/CD pipelines.
- Documentation & Knowledge Sharing: Create and maintain high-quality technical documentation, runbooks, and escalation procedures to ensure system readiness and operational efficiency.
- Cross-Functional Collaboration: Partner with application developers, infrastructure engineers, and security teams to successfully deploy and sustain production-grade services.
- Database Management: Apply knowledge of relational (MySQL, PostgreSQL) and NoSQL (MongoDB) databases to optimize database structures and contribute to data modeling efforts.
Qualifications
Required Experience & Technical Skills
- 2+ years of hands-on experience in a Site Reliability Engineering, DevOps, or Infrastructure support role.
- Proficiency with at least one major cloud platform (AWS experience is strongly preferred).
- Experience with building and managing CI/CD pipelines (e.g., Jenkins, GitHub Actions, AWS CodePipeline).
- Proficiency in automating infrastructure with an IaC tool (e.g., Terraform, CloudFormation).
- Strong working knowledge of Linux-based systems and shell scripting.
- Familiarity with version control systems, particularly Git.
- Understanding of core monitoring and alerting principles and experience with common observability tools.
- Basic understanding of core cloud services (e.g., AWS S3, EFS, Kinesis) and basic troubleshooting techniques.
- Willingness to participate in on-call rotations and take ownership of service reliability.
Education & Soft Skills
- Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related technical field—or equivalent hands-on professional experience.
- Strong desire to learn and continually grow expertise in automation, observability, and SRE best practices.
- Excellent problem-solving, analytical, and communication skills to work effectively with diverse teams.
$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...SuggestedRemote work$112k - $179k
...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems...SuggestedContract workWorldwideShift work- ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS...SuggestedTemporary work
$166k - $220k
...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the... ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine...SuggestedFull timeWork experience placementImmediate start$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...SuggestedFull timeLocal areaImmediate start$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....Permanent employmentTemporary workImmediate startWeekend work$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...Currently hiringRemote work$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...Permanent employmentFull timeWork at officeLocal areaRemote workWorldwide- ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable platform by developing systems for analyzing environments, predicting, and resolving issues, and supporting the production environment...Work experience placement
$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office- ...Site Reliability Engineer Location- Wilmington De, Washington DC, Dallas, TX (Onsite Position) Full time position Minimum Qualifications Bachelor’s degree in computer science, Engineering, or a related technical field. Minimum of 5 years of experience...Full time
$125k - $185k
...lifesaving drugs, forecast supply chain disruptions, locate missing children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-performance, scalable, and reliable services for our production...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$106.3k - $221.1k
...Senior Site Reliability Engineer At Accenture Federal Services, nothing matters more than helping the US federal government make the nation stronger and safer and life better for people. Our 13,000+ people are united in a shared purpose to pursue the limitless potential...Work at officeLocal area$125k - $185k
Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions... ...locate missing children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package- ...Site Reliability Engineer ValidaTek is building teams of Site Reliability Engineers (SRE's) to support internal and external engineering and operations of a large scale and world-wide Enterprise IT environment that covers application hosting and support, enterprise...
- ...ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running... ...of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident...Full timeContract workRemote workFlexible hours
- ...Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian of our production ecosystems, ensuring that our complex, data-driven AI platforms remain resilient, scalable, and highly performant...Local area
$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...Work experience placementWork at office$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$153k - $185k
...Senior Site Reliability Engineer El Segundo, California, United States About Varda Low Earth orbit is open for business. Varda is accelerating the development of commercial space infrastructure, from in-orbit pharmaceutical processing to reliable and economical...Permanent employmentFull timeImmediate startRelocation packageFlexible hoursWeekend work- ...Site Reliability Engineer Qualifications: ~10+ years of overall experience in IT including, with hands-on Development and Systems engineering background ~3-5 years of experience in a Site Reliability Engineering role ~ Experience with Enterprise Cloud transformation...Temporary workImmediate start
$174k - $239k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc. is looking for an experienced Staff TDI Site Reliability...Local areaWorldwideFlexible hours$207k - $284.9k
...is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Senior Manager, Site Reliability Engineering Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI...Permanent employmentFull timeLocal areaWorldwideFlexible hoursDay shift$174k - $238k
...work. We're all in on this mission. If you are too, let's talk.The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging Products Group (EPG). Our mission is to build highly reliable, scalable,...Local areaWorldwideFlexible hours$174k - $267k
...defining work. We're all in on this mission. If you are too, let's talk. The Team We are looking for an experienced Staff Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable, scalable, and secure cloud services...Full timeLocal areaWorldwideFlexible hours$166k - $220k
...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &... ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril...Full timeWork experience placementImmediate start$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering , this role will...Permanent employmentLocal areaWorldwideFlexible hours- ...bottlenecks, and improve system health—utilization, performance, and reliability—across our infrastructure Understand how systems fail and... ..., documentation, and code review—that automates reliability engineering work: Deployment tooling Fault-injection/chaos...Full timeCasual workLocal areaWorldwideFlexible hoursShift work
$175k - $250k
Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or... ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design,...Full timeRemote workRelocationRelocation package- ...certificates. Qualifications & Requirements ~ Bachelor’s degree in Computer Science, Engineering, or a related technical field. ~3+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, or infrastructure-focused roles. ~ Hands-on...Temporary work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer remote Washington DC
- site reliability engineer Washington DC
- site reliability engineer sre Washington DC
- junior website developer Washington DC
- website content developer Washington DC
- on site coordinator Washington DC
- website coordinator Washington DC
- site leader Washington DC
- site recruiter Washington DC
- historic site Washington DC


