Senior Software Engineer- Site Reliability Engineering (SRE)
$149.4k - $202kNoctua Technology
Senior Software Engineer- Site Reliability Engineering (SRE)
DC, MD, VA, CA
The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on the seamless integration, scalability, and long-term reliability of cloud native systems. Our SREs don’t just manage infrastructure; they build it using Infrastructure as Code (IaC), monitor it through advanced observability stacks, and protect it by engineering for failure. We work closely with clients to bridge the gap between development and operations.
We are seeking a highly experienced and autonomous Senior Site Reliability Engineer (SRE) to join our dynamic team. As a technical leader, you will define the strategy and apply advanced software engineering principles to operations, focusing on the architecture, reliability, and long-term performance of large-scale production systems. You will play a crucial role in reducing toil through automation, defining and monitoring Service Level Objectives (SLOs), and implementing best practices for system stability and incident response. This role requires working with modern cloud technologies to ensure the high availability and efficiency of applications and infrastructure.
- Location : Primarily Remote. Candidates must be based in CA or DC Metro Area for proximity to project and client teams.
- Security Clearance Requirement : Applicants must be US citizens and eligible to obtain and maintain an active Secret security clearance or above.
Key Responsibilities
Site Reliability Engineering
- Drive the definition and adoption of SLIs and SLOs across multiple services or entire platforms, ensuring alignment with business goals.
- Design and architect Infrastructure as Code (IaC) solutions for large-scale, complex environments, establishing standards and best practices.
- Implement and manage containerized and serverless architectures using Docker, Kubernetes, and cloud-native services, focusing on performance and error budgets.
- Build and maintain reliable and self-healing CI/CD pipelines to automate deployments and improve development workflows.
Toil Reduction and Incident Management
- Implement and refine comprehensive monitoring, alerting, and logging to detect and address performance and availability issues proactively.
- Lead the strategic effort to eliminate toil, identifying and championing major automation projects that deliver significant organizational efficiency.
- Lead high-severity incident response and coordinate blameless postmortems for major outages, driving the resulting remediation and systemic improvements.
Testing and Service Resiliency
- Implement cloud security best practices, including identity and access management (IAM), encryption, and compliance controls.
- Proactively identify and address system weaknesses and ensure performance under stress.
- Support disaster recovery and high availability strategies through backup and failover planning.
Collaboration and Knowledge Sharing
- Serve as a primary SRE liaison for development teams, influencing application architecture and design to meet reliability and scalability targets from inception.
- Create and maintain documentation for cloud architectures, deployment processes, and best practices.
- Contribute to internal knowledge-sharing initiatives, ensuring continuous learning within the team.
Stakeholder Communication
- Act as a subject matter expert and trusted advisor to clients and internal leadership on cloud infrastructure, reliability strategy, and Service Level Agreement (SLA) negotiations.
- Act on client feedback to refine and enhance cloud solutions.
- Conduct training and knowledge-sharing sessions to help clients manage their cloud environments effectively.
Continuous Learning and Innovation
- Stay updated on the latest developments in cloud infrastructure and technology trends.
- Drive innovation by proposing and implementing new techniques and technologies.
Qualifications
- 5+ years of experience in site reliability engineering, cloud engineering, or related fields.
- Strong software engineering skills with an emphasis on writing clean, modular, and maintainable code, specifically for automation and system management.
- Deep experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
- Deep experience with containerization and orchestration tools like Docker and Kubernetes.
- Deep knowledge of networking concepts, cloud security best practices, and identity management.
- Experience with programming or scripting languages such as Python, Bash, or Go.
- Experience with CI/CD pipelines and DevOps methodologies.
- Strong problem-solving skills and the ability to troubleshoot complex cloud environments.
- Demonstrated ability to influence technical decision-making across organizational boundaries.
Preferred qualifications
- Bachelor's or advanced degree in Computer Science or a related field.
- Any of the below cloud certifications:
- Google Cloud Professional Cloud Architect
- Google Cloud Professional Cloud DevOps Engineer
- AWS Certified Solutions Architect
- AWS Certified Developer
- AWS Certified SysOps Administrator
- CompTIA Security+ certification or an equivalent DoD 8140/8570 IAT Level II baseline certification.
Salary Range : $149,400 - $202,000
#J-18808-Ljbffr$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...SeniorCurrently hiringRemote work$166k - $220k
...unique combinations of hardware and software tailored to different mission requirements... .... Our systems integration engineers internalize the nuances of each deployment... ....ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing...SeniorFull timeWork experience placementImmediate start$166k - $220k
...unique combinations of hardware and software tailored to different mission requirements... .... Our systems integration engineers internalize the nuances of each deployment... ....ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing...SeniorFull timeWork experience placementImmediate start- ...Production support expertise with SRE Observability experience :... ..., My SQL and Mongo DB Seniority level ~ Seniority... ...set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $20... .../San Diego, CA) Senior Software Engineer - Optical Network...SeniorContract workRemote work
$207k - $284.9k
...mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from... ..., let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team... ...leader who understands both the SRE discipline and the unique demands of...SeniorPermanent employmentLocal areaWorldwideFlexible hoursDay shift$130k - $180k
Job DescriptionEverforth ECS is seeking a Cloud Site Reliability Engineer (SRE)to work in our Arlington, VA office/remotely. Our Philosophy We believe... ...to engineer the cloud to run itself. That means writing software and automation that lets systems detect and recover from...For contractorsWork at officeRemote workShift work$160k - $210k
...'re growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure... ...closely with our datacenter-focused SRE, so while deep, hands-on AWS expertise... ...+ years of experience in operations, software engineering, or as an SRE. ~ Working...SeniorWork at officeImmediate startRemote workWork from home$150k - $180k
...are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale... ...organization.This position is based on-site in either our Arlington, VA office,... ...methodologies.Expertise in infrastructure and software architecture, capable of designing and...SeniorPermanent employmentFull timeWork at officeLocal areaRemote workWorldwide$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to... ...ensuring scalability, performance, and reliability across environments. What You’ll Do Design...SeniorFull timeRemote workRelocationRelocation package- ...SRE Engineer Location: Washington, DC (Onsite) Duration: 08-17-2026 - 07-30-2027 Key Responsibilities Observability & Monitoring... ...(RCA), and author comprehensive knowledge base articles. Reliability Engineering: Champion SRE metrics including Service Level...
$106.3k - $221.1k
...missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability... ...Science, Information Systems, Information Technology, or Software Engineering. # Equivalent Training: Completion of one...SeniorLive inWork at officeLocal area- ...Description Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the... ...scalable, and highly performant. This role is a hybrid of software engineering and systems architecture, with a specialized...SeniorLocal area
- ...Excellent analytical and problem-solving skills with a proactive approach. AI/ML experience or a strong interest in applying AI/ML to reliability, security, or operational efficiency is a plus. Benefits Health, dental, and vision insurance; 401(k); flexible spending...SeniorFull timeWork at officeFlexible hours
$55k - $187k
...Not Applicable Specialism IFS - Internal Firm Services - Other Management Level Senior Associate Job Description & Summary The Opportunity As a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability,...SeniorFull timeH1b$114.6k - $252.1k
Job Title: SRE Platform EngineerJob Category: Information TechnologyTime Type: Full timeMinimum Clearance Required to... ...* * * The Opportunity: CACI is seeking a seasoned Site Reliability (SRE) Platform Engineer to support the Department of Homeland Security (DHS) Office...Full timeContract workWork experience placementWork at officeFlexible hours$232k - $319k
...scale the service with great people and reliable, cost-effective, and efficient infrastructure... ...org and various initiatives across SRE & Infrastructure organization. Build... ...Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful...SeniorPermanent employmentLocal areaWorldwideFlexible hours$115.5k - $164.8k
...with our ecosystem of devices and cloud software. Like our products, we work better... ...company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a... ...previously required human intervention with reliable, tested automation. You will also...Work experience placementWork at officeRemote work$112k - $179k
...integration for development of hardware and software solutions, and task support for the... ...seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC... ...applications and infrastructure. The SRE will drive automation initiatives,...Contract workWorldwideShift work$165k - $270k
...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology... ...national security space and commercial opportunities. Software engineering and innovation is at the core of these programs...SeniorPermanent employmentTemporary workImmediate startWeekend work$182k - $250.8k
...too, let's talk.The SRE Leadership TeamThe... ...of our platform's reliability and operational... ...thinking group of engineers and leaders who believe... .... As a Manager, Site Reliability... ..., resilience, and software engineering rigor... ...reliability as a senior technical leader in...Permanent employmentLocal areaRemote workWorldwideFlexible hoursWeekend workWeekday work$174k - $239k
...across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc.... ...service and advocate for SRE and DevOps practices across teams... ..., while leveraging secure software development practices....Work experience placementLocal areaWorldwideFlexible hours$174k - $238k
...mission. If you are too, let's talk.The Federal SRE TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the... ...EPG SRE organization, partnering closely with software engineers, architects, and product teams to design...Local areaWorldwideFlexible hours- ...Site Reliability Engineer Qualifications: ~10+ years of overall experience in IT including, with... ...efforts ~ Experience with SRE principles and transformation ~3+ years... .../CD, etc.) ~ Solid understanding of Software coding techniques and experience with...Temporary workImmediate start
$160k - $180k
...Site Reliability Engineer Location: Hybrid – Washington DC/Virginia/Maryland metro with the ability to travel to Patuxent River, MD, as needed... ...Minimum Qualifications: ~4-5 years hands-on experience in a SRE/DevOps role supporting production systems ~ Production...Full timeTemporary workLocal areaRemote workFlexible hours$95k - $171k
...infrastructure? Do you want to build your SRE career on one of the most exciting... ..., Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$105.79k - $141.05k
...demand networking at scale. As Lead SRE, you'll own the reliability of that platform — partnering with operations... ...'ll coordinate across architecture, engineering, and systems development... ...networking fundamentals, cloud platforms, software development and troubleshooting...Temporary workRemote work- ...motivated candidate to join our talented Team. Job Title: SRE / DevOps Engineer Job Location: Mclean, VA Duration: 3-month... ...possibility of extension Job Description: We are seeking a Site Reliability Engineer (SRE) with strong expertise in the client...
- ...SRE/DevOps Engineer Location: McLean, VA (5 Days mandatory) - Only locals/nearby F2F interview mandatory Developing appropriate DevOps... ...Establishing a continuous build environment to accelerate software deployment and development processes. Providing a DevOps...Local area
- ...we’re ready to bring on an engineer to help make it happen. The... .... We’re looking for a Senior Software Engineer to own that outcome... ...and media plane itself, an SRE who owns the platform underneath... ..., and the PM who owns reliability. Your job is the outcome across...SeniorFull timeDay shift
$180k - $200k
...Your trust is important to us. Stay safe and informed. Senior Software Engineer Nexxen SSP | Hybrid — 3 days/week in office Join... ...and in-memory stores (Redis, Kafka, Aerospike) • Apply an SRE mindset: observability, incident response, capacity planning...SeniorFull timeWork at officeFlexible hours3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Software Engineer- Site Reliability Engineering (SRE). Be the first to apply!
- software developer fintech Washington DC
- startup software engineer Washington DC
- financial software developer Washington DC
- junior software developer remote Washington DC
- software engineer Washington DC
- software data engineer Washington DC
- software developer internship no experience Washington DC
- part time software developer Washington DC
- software engineer entry level Washington DC
- software engineer mainframe Washington DC



