Software Engineer- Site Reliability Engineering (SRE)
$106.5k - $177.5kNoctua Technology
Software Engineer- Site Reliability Engineering (SRE)
DC, MD, VA, CA
The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on the seamless integration, scalability, and long-term reliability of cloud native systems. Our SREs don’t just manage infrastructure; they build it using Infrastructure as Code (IaC), monitor it through advanced observability stacks, and protect it by engineering for failure. We work closely with clients to bridge the gap between development and operations.
We are seeking a motivated Site Reliability Engineer (SRE) to join our dynamic team. As a key contributor, you will apply software engineering principles to operations, focusing on the reliability, scalability, and performance of production systems. You will play a crucial role in reducing toil through automation, defining and monitoring Service Level Objectives (SLOs), and implementing best practices for system stability and incident response. This role requires working with modern cloud technologies to ensure the high availability and efficiency of applications and infrastructure.
- Location : Primarily Remote. Candidates must be based in CA or DC Metro Area for proximity to project and client teams.
- Security Clearance Requirement : Applicants must be US citizens and eligible to obtain and maintain an active Secret security clearance or above.
Key Responsibilities
Site Reliability Engineering
- Define, measure, and report on Service Level Indicators (SLIs) and Service Level Objectives (SLOs) to ensure system reliability and uptime.
- Develop and deploy Infrastructure as Code (IaC) using Terraform, CloudFormation, or similar tools, with an emphasis on repeatability and change management.
- Implement and manage containerized and serverless architectures using Docker, Kubernetes, and cloud-native services, focusing on performance and error budgets.
- Build and maintain reliable and self-healing CI/CD pipelines to automate deployments and improve development workflows.
Toil Reduction and Incident Management
- Implement and refine comprehensive monitoring, alerting, and logging to detect and address performance and availability issues proactively.
- Eliminate toil by extensively automating operational tasks, including provisioning, patching, and deployments, using scripting and configuration management tools such as Python, Bash, or Ansible.
- Conduct post-incident reviews (blameless postmortems) to drive continuous improvement in system reliability and operational processes.
Testing and Service Resiliency
- Implement cloud security best practices, including identity and access management (IAM), encryption, and compliance controls.
- Proactively identify and address system weaknesses and ensure performance under stress.
- Support disaster recovery and high availability strategies through backup and failover planning.
Collaboration and Knowledge Sharing
- Collaborate with development teams to improve the operability and production readiness of applications from design through deployment.
- Create and maintain documentation for cloud architectures, deployment processes, and best practices.
- Contribute to internal knowledge-sharing initiatives, ensuring continuous learning within the team.
Stakeholder Communication
- Provide technical guidance and support to clients and internal teams on cloud infrastructure and reliability best practices, with a focus on defining Service Level Agreements (SLAs).
- Act on client feedback to refine and enhance cloud solutions.
- Conduct training and knowledge-sharing sessions to help clients manage their cloud environments effectively.
Continuous Learning and Innovation
- Stay updated on the latest developments in cloud infrastructure and technology trends.
- Drive innovation by proposing and implementing new techniques and technologies.
Qualifications
- 1-5 years of experience in site reliability engineering, cloud engineering, or related fields.
- Strong software engineering skills with an emphasis on writing clean, modular, and maintainable code, specifically for automation and system management.
- Proficiency in Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
- Experience with containerization and orchestration tools like Docker and Kubernetes.
- Knowledge of networking concepts, cloud security best practices, and identity management.
- Experience with programming or scripting languages such as Python, Bash, or Go.
- Familiarity with CI/CD pipelines and DevOps methodologies.
- Strong problem-solving skills and the ability to troubleshoot complex cloud environments.
- Effective communication skills and a willingness to learn and collaborate.
Preferred Qualifications
- Bachelor's or advanced degree in Computer Science or a related field.
- Any of the below cloud certifications:
- Google Cloud Professional Cloud Architect
- Google Cloud Professional Cloud DevOps Engineer
- AWS Certified Solutions Architect
- AWS Certified Developer
- AWS Certified SysOps Administrator
- Azure Solutions Architect Expert
- CompTIA Security+ certification or an equivalent DoD 8140/8570 IAT Level II baseline certification.
Salary Range: $106,500 - $177,500
#J-18808-Ljbffr- ...ITIL-based processes. Define and monitor SRE metrics including SLIs, SLOs, and error... ...~ Bachelor’s degree in Computer Science, Engineering, or a related technical field. ~3+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, or...SuggestedTemporary work
- ...Overview We are seeking an experienced Senior DevOps / Site Reliability Engineer (SRE) to support a mission-critical national security program.... ...systems administration, infrastructure automation, and software development. The successful candidate will help design, build...SuggestedFull timeVisa sponsorship
- ...Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public Trust clearance. What You Will Do We are seeking a Site Reliability Engineer (SRE) to support the SBA Disaster Lending Platform modernization effort in a remote...SuggestedFull timeLocal areaRemote work
- ...SRE Engineer Location: Washington, DC (Onsite) Duration: 08-17-2026 - 07-30-2027 Key Responsibilities Observability & Monitoring... ...(RCA), and author comprehensive knowledge base articles. Reliability Engineering: Champion SRE metrics including Service Level...Suggested
$90k - $150k
...Workplaces honoree, is seeking a SRE Engineer to support our growing team... ...for improving the reliability, availability, performance,... ...customers. Work Environment: On-site Key Responsibilities:... ...related field. ~5 years of software engineering, 3 years site reliability...SuggestedPermanent employmentFull timeContract work$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...Currently hiringRemote work$166k - $220k
...unique combinations of hardware and software tailored to different mission requirements... .... Our systems integration engineers internalize the nuances of each deployment... ....ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing...Full timeWork experience placementImmediate start- ...particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS platform that powers the...Temporary work
$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver... ...and configuration management workflows to support reliable software delivery and operational observability across development...Full timeLocal areaImmediate start$115.5k - $164.8k
...with our ecosystem of devices and cloud software. Like our products, we work better... ...company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a... ...previously required human intervention with reliable, tested automation. You will also...Work experience placementWork at officeRemote work$112k - $179k
...integration for development of hardware and software solutions, and task support for the... ...seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC... ...applications and infrastructure. The SRE will drive automation initiatives,...Contract workWorldwideShift work$174k - $238k
...mission. If you are too, let's talk.The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the... ...EPG SRE organization, partnering closely with software engineers, architects, and product teams to design...Local areaWorldwideFlexible hours$166k - $220k
...that Anduril services are reliable and maintainable. This means... ...infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team,... ...as well as our software delivery plane for edge systems... ...as interfacing with other SRE and software teams across the...Full timeWork experience placementImmediate start$207k - $284.9k
...you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from... ...too, let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team supports... ...leader who understands both the SRE discipline and the unique demands of federal...Permanent employmentLocal areaWorldwideFlexible hoursDay shift$174k - $239k
...across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc.... ...service and advocate for SRE and DevOps practices across teams... ..., while leveraging secure software development practices....Local areaWorldwideFlexible hours$145k - $175k
...better care. Requirements As a Senior Site Reliability Engineer at Commence, you will own the reliability... ...they become habits. Collaborate with software engineers to establish reliability-... ...Qualifications 7+ years of experience in SRE, platform engineering, or DevOps roles...Full timeRemote work- ...Description Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the... ...scalable, and highly performant. This role is a hybrid of software engineering and systems architecture, with a specialized...Local area
$165k - $225.6k
...across functions to drive scale, reliability, and innovation through technology. The Senior Site Reliability Engineer Opportunity Reporting to the... ...Collaboration & Advocacy: Partner with software engineering teams to champion DevOps and SRE best practices, deliver...Permanent employmentLocal areaWorldwideFlexible hours$82.3k - $228.8k
...leading provider of cloud contact center software, bringing the power of cloud... ...are seeking a highly experienced Senior Site Reliability Engineer – Compute Platforms to design, implement... ...will also collaborate with platform and SRE teams to maintain secure, performant,...Temporary workWork at officeRemote workWorldwide3 days per week- ...A leading security infrastructure firm in Washington, D.C. is seeking a hands-on Site Reliability Engineer (SRE) with expertise in Kubernetes and cloud infrastructure. The role emphasizes total ownership of security infrastructure while defending against advanced threats...
$100k - $110k
...and company events. We are seeking an operational-focused Site Reliability Engineer (SRE) to maximize the availability, performance, and... ...preferred. Education & Experience: Minimum of 4 years of software/systems experience, with at least 1-2 years focused on live...Permanent employmentRemote workFlexible hours- ...Site Reliability Engineer ValidaTek is building teams of Site Reliability Engineers (SRE's) to support internal and external engineering and operations of a large scale and... ...and module writing (Chef, Ansible, Salt) Software development, preferred in Python, Ruby or...
- ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable platform by developing systems for analyzing environments, predicting, and resolving issues, and supporting the production environment...Work experience placement
$163.62k - $212.71k
...are looking for an experienced SRE leader with the skills and passion... ...and processes that improve our engineering teams' productivity and streamline the software development lifecycle. Your... ...seasoned and strategic Lead/Principal Site Reliability Engineer to drive the...Full timePart timeWork experience placementWork at officeLocal areaImmediate startRemote workWork from homeFlexible hoursShift work3 days per week1 day per week$194k - $267k
...more than once, automate it” and who can rapidly self-educate on new concepts and tools. Position Overview: The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$194k - $267k
...We are seeking a highly technical Staff Observability Site Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem... ..., scalable Observability Platform that enables our SRE teams and business partners. You will treat infrastructure...Permanent employmentWork at officeLocal areaWorldwideFlexible hours- ...Inc Production support expertise with SRE Observability experience : Proactive issue... ...Sign in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259,... ...WA/ Santa Clara/San Diego, CA) Senior Software Engineer - Optical Network Agents &...Contract workRemote work
- Senior Site Reliability Engineer - Network Operations (Remote) Fastly 15 August 2025 SRE DevOps Automation Networking BGP Fastly is seeking a Senior Site Reliability Engineer... ...engineering teams to shape roadmaps and software solutions. * Mentor team members on global routing...Remote jobLocal areaFlexible hoursNight shift
$165k - $230k
...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology... ...national security space and commercial opportunities. Software engineering and innovation is at the core of these programs...Permanent employmentTemporary workImmediate startWeekend work$80k - $133k
Job Family:Software Development & SupportTravel Required:Up to 10%Clearance Required:Ability... ...partners to establish and maintain SRE practice in an Agile Scrum framework.Participate... ...in IT administration, software engineering, or platform engineering, with a focus on...Permanent employmentFull timeContract workRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Software Engineer- Site Reliability Engineering (SRE). Be the first to apply!
- agile software developer Washington DC
- software developer internship no experience Washington DC
- intermediate software engineer Washington DC
- software engineer staff Washington DC
- experienced software developer Washington DC
- work from home software developer Washington DC
- software developer fintech Washington DC
- software data engineer Washington DC
- financial software developer Washington DC
- software developer internship Washington DC


