Senior Software Engineer- Site Reliability Engineering (SRE)
$149.4k - $202kNoctua Technology
Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on the seamless integration, scalability, and long-term reliability of cloud native systems. Our SREs don’t just manage infrastructure; they build it using Infrastructure as Code (IaC), monitor it through advanced observability stacks, and protect it by engineering for failure. We work closely with clients to bridge the gap between development and operations. We are seeking a highly experienced and autonomous Senior Site Reliability Engineer (SRE) to join our dynamic team. As a technical leader, you will define the strategy and apply advanced software engineering principles to operations, focusing on the architecture, reliability, and long-term performance of large-scale production systems. You will play a crucial role in reducing toil through automation, defining and monitoring Service Level Objectives (SLOs), and implementing best practices for system stability and incident response. This role requires working with modern cloud technologies to ensure the high availability and efficiency of applications and infrastructure. Location : Primarily Remote. Candidates must be based in CA or DC Metro Area for proximity to project and client teams. Security Clearance Requirement : Applicants must be US citizens and eligible to obtain and maintain an active Secret security clearance or above. Key Responsibilities Site Reliability Engineering Drive the definition and adoption of SLIs and SLOs across multiple services or entire platforms, ensuring alignment with business goals. Design and architect Infrastructure as Code (IaC) solutions for large-scale, complex environments, establishing standards and best practices. Implement and manage containerized and serverless architectures using Docker, Kubernetes, and cloud-native services, focusing on performance and error budgets. Build and maintain reliable and self-healing CI/CD pipelines to automate deployments and improve development workflows. Toil Reduction and Incident Management Implement and refine comprehensive monitoring, alerting, and logging to detect and address performance and availability issues proactively. Lead the strategic effort to eliminate toil, identifying and championing major automation projects that deliver significant organizational efficiency. Lead high-severity incident response and coordinate blameless postmortems for major outages, driving the resulting remediation and systemic improvements. Testing and Service Resiliency Implement cloud security best practices, including identity and access management (IAM), encryption, and compliance controls. Proactively identify and address system weaknesses and ensure performance under stress. Support disaster recovery and high availability strategies through backup and failover planning. Collaboration and Knowledge Sharing Serve as a primary SRE liaison for development teams, influencing application architecture and design to meet reliability and scalability targets from inception. Create and maintain documentation for cloud architectures, deployment processes, and best practices. Contribute to internal knowledge-sharing initiatives, ensuring continuous learning within the team. Stakeholder Communication Act as a subject matter expert and trusted advisor to clients and internal leadership on cloud infrastructure, reliability strategy, and Service Level Agreement (SLA) negotiations. Act on client feedback to refine and enhance cloud solutions. Conduct training and knowledge-sharing sessions to help clients manage their cloud environments effectively. Continuous Learning and Innovation Stay updated on the latest developments in cloud infrastructure and technology trends. Drive innovation by proposing and implementing new techniques and technologies. Qualifications 5+ years of experience in site reliability engineering, cloud engineering, or related fields. Strong software engineering skills with an emphasis on writing clean, modular, and maintainable code, specifically for automation and system management. Deep experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation. Deep experience with containerization and orchestration tools like Docker and Kubernetes. Deep knowledge of networking concepts, cloud security best practices, and identity management. Experience with programming or scripting languages such as Python, Bash, or Go. Experience with CI/CD pipelines and DevOps methodologies. Strong problem-solving skills and the ability to troubleshoot complex cloud environments. Demonstrated ability to influence technical decision-making across organizational boundaries. Preferred qualifications Bachelor's or advanced degree in Computer Science or a related field. Any of the below cloud certifications: Google Cloud Professional Cloud Architect Google Cloud Professional Cloud DevOps Engineer AWS Certified Solutions Architect AWS Certified Developer AWS Certified SysOps Administrator CompTIA Security+ certification or an equivalent DoD 8140/8570 IAT Level II baseline certification. Salary Range : $149,400 - $202,000 #J-18808-Ljbffr
- ...Production support expertise with SRE Observability experience :... ..., My SQL and Mongo DB Seniority level Seniority level Mid-... ...set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,0... ...Clara/San Diego, CA) Senior Software Engineer - Optical Network...SeniorContract workRemote work
$131k - $227.13k
...Description: The 1LMX MES COE is seeking an engineer who will own infrastructure‑as‑code, cloud platform, and reliability for the Apriso environment on AWS. This... ...full‑stack development, DevOps, and Site Reliability Engineering (SRE) practices to deliver a production‑...SuggestedFull timeTemporary workWork experience placementWork at officeRemote workRelocationFlexible hoursShift work3 days per week$121.4k - $218.6k
...Join our highly skilled Site Reliability Engineering team! Our team designs, develops... ...products and services. Our SRE teams solve reliability,... ...the Akamai Cloud. As a Senior Site Reliability Engineer,... ...practices Collaborating with software engineering, infrastructure...SeniorWork experience placementWork at office$166k - $220k
...Senior Site Reliability Engineer Anduril Industries is a defense technology company with a mission to transform... ...unique combinations of hardware and software tailored to different mission... ...looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing...SeniorFull timeWork experience placementImmediate start$153k - $185k
...Senior Site Reliability Engineer El Segundo, California, United States About Varda Low Earth orbit is open for... ...who applies first-principles thinking to both software delivery (DevOps) and production reliability (SRE), and thrives in complex, mission-critical environments...SeniorPermanent employmentFull timeImmediate startRelocation packageFlexible hoursWeekend work$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to... ...ensuring scalability, performance, and reliability across environments. What You’ll Do Design...SeniorFull timeRemote workRelocationRelocation package- ...SRE Engineer Location: Washington, DC (Onsite) Duration: 08-17-2026 - 07-30-2027 Key Responsibilities Observability & Monitoring... ...(RCA), and author comprehensive knowledge base articles. Reliability Engineering: Champion SRE metrics including Service Level...
$135k - $150k
Senior Site Reliability Engineer Job number: 884 This is a remote position. Ad Hoc is a technology company that empowers organizations... ...with our partners to solve the right problems and deliver software that works. The Veterans Affairs business unit helps...SeniorRemote workFlexible hours- ...Title: Sr. IT Application Solutions Architect /SRE Engineer Important Note : We have shifted to adopting SAFe and 1. Encourage... ...have their own equipment. Access to a virtual desktop set up (software) will be provided by Lumen's client, allowing the user access...For contractorsRemote workShift work
$106.3k - $221.1k
...missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability... ...Science, Information Systems, Information Technology, or Software Engineering. # Equivalent Training: Completion of one...SeniorLive inWork at officeLocal area$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The... ...demands and responds to capacity needs. Collaborates with software development teams to develop reliable and scalable infrastructures...SeniorTemporary workImmediate startFlexible hoursShift work$147k - $202k
...TechOps) team, we live this mission by building the most reliable and performant systems on the planet. We empower... ...need. The Role We are looking for an experienced Senior Site Reliability Engineer (SRE) who thrives on the challenge of managing large-scale cloud...SeniorPermanent employmentLocal areaWorldwideFlexible hours$207k - $284.9k
...re all in on this mission. If you are too, let's talk. Senior Manager, Site Reliability Engineering Secure Every Identity, from AI to Human Identity is... ...for a technical leader who understands both the SRE discipline and the unique demands of federal customer relationships...SeniorPermanent employmentLocal areaWorldwideFlexible hoursDay shift$84.9k - $209.5k
...unencumbered and will need your contribution to make it a special engineering center with the focus on excellence. Health Data... ...What You'll Do Service Ownership –You will be part of the SRE team, whose mission is the shared full stack ownership of a collection...Temporary workImmediate startFlexible hours$95k - $171k
...infrastructure? Do you want to build your SRE career on one of the most exciting... ..., Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...hands-on Development and Systems engineering background ~3-5 years of experience in a Site Reliability Engineering role ~... ...transformation efforts ~ Experience with SRE principles and transformation... ....) ~ Solid understanding of Software coding techniques and...Temporary workImmediate start
$112k - $179k
...integration for development of hardware and software solutions, and task support for the... ...seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC... ...applications and infrastructure. The SRE will drive automation initiatives,...Contract workWorldwideShift work$165k - $230k
...the ultimate goal of enabling human life on Mars. SR. SITE RELIABILITY ENGINEER (STARSHIELD) Starshield leverages SpaceX’s Starlink technology... ...national security space and commercial opportunities. Software engineering and innovation is at the core of these...SeniorPermanent employmentTemporary workImmediate startWeekend work$51.9 per hour
...job is responsible for the reliability, availability, and performance... .... This role blends software engineering, clinical engineering, and security... ...cross-functionally with AHN site leaders and teams to navigate... ...maintaining system health for SRE practitioners (e.g., latency...For contractorsLocal area- ...motivated candidate to join our talented Team. Job Title: SRE / DevOps Engineer Job Location: Mclean, VA Duration: 3-month... ...of extension Job Description: We are seeking a Site Reliability Engineer (SRE) with strong expertise in the client ecosystem...
$131k - $164k
...Staff Site Reliability Engineer New York, New York, United States Position... ...team. This role is a hands-on senior engineering position responsible... ...teams (Network, Security, SRE, and DevOps) to ensure the... ...only want to help build the software company of the future, but who...Work at officeLocal areaFlexible hours- ...SRE/DevOps Engineer Location: McLean, VA (5 Days mandatory) - Only locals/nearby F2F interview mandatory Developing appropriate DevOps... ...Establishing a continuous build environment to accelerate software deployment and development processes. Providing a DevOps...Local area
$220k - $250k
...Staff Site Reliability Engineer Yugabyte is the company behind YugabyteDB, the AI-ready, multi-modal... ...We are looking for a strong Staff SRE who exemplifies collaboration, teamwork... ...strategies Requirements ~ Strong software design and implementation skills in...H1bLocal areaWorldwideVisa sponsorship- ...Description Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the... ...scalable, and highly performant. This role is a hybrid of software engineering and systems architecture, with a specialized...SeniorLocal area
- ...Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public Trust clearance. What You Will Do We are seeking a Site Reliability Engineer (SRE) to support the SBA Disaster Lending Platform modernization effort in a remote...Contract workLocal areaRemote work
$75 per hour
job summary: Randstad is partnering with a premier client in the Washington, D.C. area to find a talented Site Reliability Engineer (SRE) to champion system availability, performance, and automation across their enterprise cloud infrastructure. In this role, you will...Hourly payContract workTemporary workWork experience placement$160k - $185k
...motivated Sr. Infrastructure Engineer to join our Hardware... ...of highly performant and reliable infrastructure solutions.... ...under the guidance of more senior engineers. Help document... ...experience in cloud operations, site reliability engineering (SRE), or related technical...SeniorFull timeTemporary workCasual workWork at officeRemote workFlexible hours$126k - $248k
...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB... ...execution, strengthen production reliability practices, and coordinate cross-... ...technical role partnering with software engineering teams ~ Proven track...Local areaRemote workWorldwideFlexible hours$147k - $202k
...looking for an Observability Engineer to help ensure that our... ...continuing to rapidly ship software that our customers love.... ...have experience within the Site Reliability Engineering (SRE) field or as a Development... ...development in these areas. As a Senior Engineer on this team, you...SeniorLocal areaFlexible hours$160k - $210k
...) is searching for a Senior DevSecOps Engineer. How you will contribute... ...with a need to be on-site and remote, depending... ...for technology, software, and DevSecOps.... ...and manage the site reliability of the systems Help... ..., DevOps, DevSecOps, SRE, or Infrastructure Engineer...SeniorContract workRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Software Engineer- Site Reliability Engineering (SRE). Be the first to apply!
- software engineer full time Washington DC
- software system engineer Washington DC
- consulting software engineer Washington DC
- software engineer travel Washington DC
- software engineer mainframe Washington DC
- real time software engineer Washington DC
- network software engineer Washington DC
- senior software engineer remote Washington DC
- entry level software engineer remote Washington DC
- software engineer intern Washington DC


