Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Software Engineer- Site Reliability Engineering (SRE)

$149.4k - $202k

Noctua Technology

Senior Software Engineer- Site Reliability Engineering (SRE)

DC, MD, VA, CA

The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on the seamless integration, scalability, and long-term reliability of cloud native systems. Our SREs don’t just manage infrastructure; they build it using Infrastructure as Code (IaC), monitor it through advanced observability stacks, and protect it by engineering for failure. We work closely with clients to bridge the gap between development and operations.

We are seeking a highly experienced and autonomous Senior Site Reliability Engineer (SRE) to join our dynamic team. As a technical leader, you will define the strategy and apply advanced software engineering principles to operations, focusing on the architecture, reliability, and long-term performance of large-scale production systems. You will play a crucial role in reducing toil through automation, defining and monitoring Service Level Objectives (SLOs), and implementing best practices for system stability and incident response. This role requires working with modern cloud technologies to ensure the high availability and efficiency of applications and infrastructure.

  • Location : Primarily Remote. Candidates must be based in CA or DC Metro Area for proximity to project and client teams.
  • Security Clearance Requirement : Applicants must be US citizens and eligible to obtain and maintain an active Secret security clearance or above.
Key Responsibilities
Site Reliability Engineering
  • Drive the definition and adoption of SLIs and SLOs across multiple services or entire platforms, ensuring alignment with business goals.
  • Design and architect Infrastructure as Code (IaC) solutions for large-scale, complex environments, establishing standards and best practices.
  • Implement and manage containerized and serverless architectures using Docker, Kubernetes, and cloud-native services, focusing on performance and error budgets.
  • Build and maintain reliable and self-healing CI/CD pipelines to automate deployments and improve development workflows.
Toil Reduction and Incident Management
  • Implement and refine comprehensive monitoring, alerting, and logging to detect and address performance and availability issues proactively.
  • Lead the strategic effort to eliminate toil, identifying and championing major automation projects that deliver significant organizational efficiency.
  • Lead high-severity incident response and coordinate blameless postmortems for major outages, driving the resulting remediation and systemic improvements.
Testing and Service Resiliency
  • Implement cloud security best practices, including identity and access management (IAM), encryption, and compliance controls.
  • Proactively identify and address system weaknesses and ensure performance under stress.
  • Support disaster recovery and high availability strategies through backup and failover planning.
Collaboration and Knowledge Sharing
  • Serve as a primary SRE liaison for development teams, influencing application architecture and design to meet reliability and scalability targets from inception.
  • Create and maintain documentation for cloud architectures, deployment processes, and best practices.
  • Contribute to internal knowledge-sharing initiatives, ensuring continuous learning within the team.
Stakeholder Communication
  • Act as a subject matter expert and trusted advisor to clients and internal leadership on cloud infrastructure, reliability strategy, and Service Level Agreement (SLA) negotiations.
  • Act on client feedback to refine and enhance cloud solutions.
  • Conduct training and knowledge-sharing sessions to help clients manage their cloud environments effectively.
Continuous Learning and Innovation
  • Stay updated on the latest developments in cloud infrastructure and technology trends.
  • Drive innovation by proposing and implementing new techniques and technologies.
Qualifications
  • 5+ years of experience in site reliability engineering, cloud engineering, or related fields.
  • Strong software engineering skills with an emphasis on writing clean, modular, and maintainable code, specifically for automation and system management.
  • Deep experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
  • Deep experience with containerization and orchestration tools like Docker and Kubernetes.
  • Deep knowledge of networking concepts, cloud security best practices, and identity management.
  • Experience with programming or scripting languages such as Python, Bash, or Go.
  • Experience with CI/CD pipelines and DevOps methodologies.
  • Strong problem-solving skills and the ability to troubleshoot complex cloud environments.
  • Demonstrated ability to influence technical decision-making across organizational boundaries.
Preferred qualifications
  • Bachelor's or advanced degree in Computer Science or a related field.
  • Any of the below cloud certifications:
    • Google Cloud Professional Cloud Architect
    • Google Cloud Professional Cloud DevOps Engineer
    • AWS Certified Solutions Architect
    • AWS Certified Developer
    • AWS Certified SysOps Administrator
  • CompTIA Security+ certification or an equivalent DoD 8140/8570 IAT Level II baseline certification.

Salary Range : $149,400 - $202,000

#J-18808-Ljbffr
Vacancy posted 9 hours ago
Similar jobs that could be interesting for youBased on the Senior Software Engineer- Site Reliability Engineering (SRE) in Washington DC vacancy
  • $149.4k - $202k

    Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on the... 
    Senior
    Remote work

    Noctua Technology

    Washington DC
    5 days ago
  •  ...Manager is backfilling a principal engineer and long-tenured VMware/Kubernetes expert...  ...Job Description- Position Title: Senior SRE/DevOps Engineer Location Information...  ...Terraform and Ansible, ensuring rapid, reliable, and auditable platform provisioning and... 
    Senior
    Contract work
    Remote work

    Stellent IT LLC

    Washington DC
    2 days ago
  • $210k - $230k

    GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation... 
    Senior
    Currently hiring
    Remote work

    Govcio

    Arlington, VA
    5 days ago
  •  ...particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS platform that powers the... 
    Senior
    Temporary work

    Kong

    Washington DC
    4 days ago
  • $166k - $220k

     ...unique combinations of hardware and software tailored to different mission requirements...  .... Our systems integration engineers internalize the nuances of each deployment...  ....ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    1 day ago
  • $207k - $284.9k

     ...mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from...  ..., let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team...  ...leader who understands both the SRE discipline and the unique demands of... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours
    Day shift

    Okta

    Washington DC
    1 day ago
  • $166k - $220k

     ...that Anduril services are reliable and maintainable. This means...  ...infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team,...  ...as well as our software delivery plane for edge systems...  ...as interfacing with other SRE and software teams across the... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Washington DC
    1 day ago
  •  ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable platform by developing systems for analyzing environments, predicting, and resolving issues, and supporting the production environment... 
    Senior
    Work experience placement

    Samprasoft

    Washington DC
    4 days ago
  • $121.4k - $218.6k

     ...our critical AI Hardware SRE Team! The AI Hardware SRE...  ...best-in-class uptime and reliability of our AI hardware infrastructure...  ...high-density hardware and software infrastructure spanning...  ...are breached. As a Senior Site Reliability Engineer, you will be responsible for... 
    Senior
    Work experience placement
    Work at office

    Akamai

    Washington DC
    2 days ago
  • $153k - $185k

     ...Senior Site Reliability Engineer El Segundo, California, United States About Varda Low Earth orbit is open for...  ...who applies first-principles thinking to both software delivery (DevOps) and production reliability (SRE), and thrives in complex, mission-critical environments... 
    Senior
    Permanent employment
    Full time
    Immediate start
    Relocation package
    Flexible hours
    Weekend work

    Varda Space Industries

    Washington DC
    5 days ago
  •  ...Production support expertise with SRE Observability experience :...  ..., My SQL and Mongo DB Seniority level Seniority level Mid-...  ...set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,0...  ...Clara/San Diego, CA) Senior Software Engineer - Optical Network... 
    Senior
    Contract work
    Remote work

    Signature IT World Inc

    Washington DC
    2 days ago
  •  ...ITIL-based processes. Define and monitor SRE metrics including SLIs, SLOs, and error...  ...~ Bachelor’s degree in Computer Science, Engineering, or a related technical field. ~3+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, or... 
    Temporary work

    2T Consulting

    Washington DC
    16 days ago
  • $165k - $225.6k

     ...across functions to drive scale, reliability, and innovation through technology. The Senior Site Reliability Engineer Opportunity Reporting to the...  ...Collaboration & Advocacy: Partner with software engineering teams to champion DevOps and SRE best practices, deliver... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    25 days ago
  • Senior Site Reliability Engineer - Network Operations (Remote) Fastly 15 August 2025 SRE DevOps Automation Networking BGP Fastly is seeking a Senior Site Reliability Engineer...  ...with engineering teams to shape roadmaps and software solutions. * Mentor team members on global... 
    Senior
    Remote job
    Local area
    Flexible hours
    Night shift

    Fastly

    Bethesda, MD
    5 days ago
  • $149.4k - $202k

    Noctua Technology is seeking a Senior Software Engineer specializing in Site Reliability Engineering to join their team. This role focuses on the reliability and performance of cloud-native applications, emphasizing Infrastructure as Code and automation. The ideal candidate... 
    Senior
    Remote job

    Noctua Technology

    Washington DC
    5 days ago
  • $147k - $202.4k

     ...re all in on this mission. If you are too, let's talk. Senior Site Reliability Engineer (SRE) - Security and Data Systems Our company is seeking...  ...securing large-scale systems. This role is a blend of software engineering and systems administration, where you'll be... 
    Senior
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours
    Shift work

    Okta

    Washington DC
    4 days ago
  • $232k - $319k

     ...scale the service with great people and reliable, cost-effective, and efficient infrastructure...  ...org and various initiatives across SRE & Infrastructure organization. Build...  ...Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    15 days ago
  • $185k - $230k

    As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver...  ...and configuration management workflows to support reliable software delivery and operational observability across development... 
    Senior
    Full time
    Local area
    Immediate start

    MetroStar Systems

    Washington DC
    3 days ago
  • $150k - $180k

     ...are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale...  ...organization.This position is based on-site in either our Arlington, VA office,...  ...methodologies.Expertise in infrastructure and software architecture, capable of designing and... 
    Senior
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work
    Worldwide

    Umbra

    Arlington, VA
    4 days ago
  • $130.5k - $171k

     ...At Spire, the Space Reliability Engineering team's mission is to...  ...ground stations, and software through vigilant monitoring...  .... We embrace the Site Reliability Engineering...  ..., devops, or SRE experience Demonstrated...  ...tool). Work details Seniority level: Mid-Senior level... 
    Senior
    Full time
    Work at office
    Flexible hours
    3 days per week

    Spire

    Washington DC
    5 days ago
  • $106.3k - $221.1k

     ...Senior Site Reliability Engineer At Accenture Federal Services, nothing matters more than helping the US federal government make the nation stronger...  ...Science, Information Systems, Information Technology, or Software Engineering. Equivalent Training: Completion of one of... 
    Senior
    Work at office
    Local area

    Accenture Federal Services

    Arlington, VA
    3 days ago
  •  ...Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian...  ...scalable, and highly performant. This role is a hybrid of software engineering and systems architecture, with a specialized... 
    Senior
    Local area

    Tiger Analytics

    Washington DC
    1 day ago
  •  ...SRE Engineer Location: Washington, DC (Onsite) Duration: 08-17-2026 - 07-30-2027 Key Responsibilities Observability & Monitoring...  ...(RCA), and author comprehensive knowledge base articles. Reliability Engineering: Champion SRE metrics including Service Level Indicators... 

    Georgia IT Inc

    Washington DC
    3 days ago
  •  ...Job Description Strong experience in SRE / Release Engineering / Platform Engineering roles. Proven experience managing and operating workloads deployed on Kubernetes, on-premise (bare metal / VMs), and PCF, with good-to-have exposure to public cloud platforms (AWS... 

    Pacer Group

    Washington DC
    1 day ago
  • $90k - $150k

     ...Workplaces honoree, is seeking a SRE Engineer to support our growing team...  ...for improving the reliability, availability, performance,...  ...customers. Work Environment: On-site Key Responsibilities:...  ...related field. ~5 years of software engineering, 3 years site reliability... 
    Permanent employment
    Full time
    Contract work

    Spatial Front

    Arlington, VA
    4 days ago
  • $120k - $150k

     ...marketers and event professionals and offers software solutions to hotels, special event...  ..., we'd love to meet you. As a Senior Site Reliability Engineer, you will use your advanced...  ...smarter, self-healing systems. As a Cvent SRE you will be a force for positive change... 
    Senior
    Work experience placement
    Work at office
    Worldwide

    Cvent

    McLean, VA
    1 day ago
  • $128.5k - $190k

     ...The Role and Team The Site Reliability Engineering organization at Medallia brings...  ...SaaS platform. As a Senior Site Reliability Engineer,...  ...platforms. Partner with software engineering teams to improve...  ...practices. Drive adoption of SRE principles, reliability... 
    Senior
    Temporary work
    Work experience placement
    Local area

    Medallia

    McLean, VA
    5 days ago
  • $175k - $250k

    Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to...  ...ensuring scalability, performance, and reliability across environments. What You’ll Do... 
    Senior
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    Washington DC
    2 days ago
  • $112k - $179k

     ...integration for development of hardware and software solutions, and task support for the...  ...seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC...  ...applications and infrastructure. The SRE will drive automation initiatives,... 
    Contract work
    Worldwide
    Shift work

    Peraton Corporation

    Washington DC
    4 days ago
  • $165k - $230k

     ...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology...  ...national security space and commercial opportunities. Software engineering and innovation is at the core of these programs... 
    Senior
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Software Engineer- Site Reliability Engineering (SRE). Be the first to apply!