Site Reliability Engineer - CTJ - POLY
$119.8k - $234.7kMicrosoft Corporation
Job ID: 200028305Posted: 2026-02-20Location: United States, Virginia, RestonSalary: USD $119,800 - $234,700 per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual ContributorTravel: Less than 25%Profession: Software EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewMicrosoft has an exciting opportunity for a Senior Site Reliability Engineer (SRE) to join the Azure Silver and Sovereign Team as part of the Azure Data Transfer (ADT) team. Azure Data Transfer enables secure access and data transfer between enclaves and supports multiple transfer and access patterns for highly regulated industries. In this role, you will apply SRE principles—availability, latency, performance, efficiency, change management, and incident response—to help ensure ADT is dependable at scale.We are looking for engineers to join a fast-paced team and solve complex reliability challenges in mission-critical distributed systems spanning data transmission across clouds. Our team works across all facets of isolated system engineering and is deeply involved in defining and improving service health through SLIs/SLOs and error budgets, building automation to reduce toil, strengthening observability (logs, metrics, traces), reducing systemic latency, validating and transforming data, and optimizing throughput and capacity. You will build, deploy, and operate systems that enable a broad set of Azure services to be consumed by customers in highly secured and regulated environments, meeting strict security policy and assurance requirements for public and private sector customers.Microsoft’s mission is to empower every person and every organization on the planet to achieve more. As employees we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.ResponsibilitiesOwns reliability architecture and end-to-end service understanding (dependencies, failure modes, and customer journeys) for distributed systems at scale. Defines and improves service health via SLIs/SLOs, error budgets, and well-defined operational readiness criteria. Drives cross-team reliability reviews and recommends design changes, runbooks, and safe rollout/rollback strategies that improve availability, latency, performance, and efficiency while managing cost.Maintains deep, current expertise in cloud reliability practices and the evolving technology landscape. Drives adoption of new platform capabilities and operational patterns (e.g., progressive delivery, resilience testing, chaos engineering where appropriate). Mentors engineers through design reviews, incident walkthroughs, and knowledge sharing to raise the reliability bar across related services.Implements reliable, scalable, and high-performance changes using SRE practices (progressive delivery, feature flags where applicable, safe rollouts/rollbacks). Owns implementation and rollback plans, validates operational readiness, and reduces toil through automation, self-healing, and standardized playbooks.Leverages telemetry and production signals to identify reliability risks and recurring failure patterns, then ships configuration changes, code fixes, or automation to address root causes. Expands infrastructure-as-code and operational tooling so teams can manage platforms and services safely and repeatably through code and policy.Builds and improves observability (metrics, logs, traces, dashboards, alerts) and uses it to detect, diagnose, and prevent incidents. Defines actionable alerting, reduces noise, and ensures instrumentation supports SLO reporting and rapid troubleshooting. Develops automation to validate telemetry pipelines and to enable automated mitigation and safer incident response.Participates in on-call rotations and leads response for complex, high-impact incidents by establishing incident command, assessing impact, coordinating responders, and driving mitigations to restore service within SLOs. Produces and contributes to blameless postmortems with corrective and preventative actions (CPAs), tracks them to completion, and implements automation and guardrails to prevent recurrence.Applies secure-by-design and compliance requirements to operations, monitoring, and automation (least privilege, auditability, change control, and data handling). Partners with security, privacy, and compliance teams to identify gaps, prioritize fixes, and implement automated controls and detection to prevent repeated violationsEmbody our culture and valuesQualificationsRequired / Minimum Qualifications:Master's Degree in Computer Science, Information Technology, or related field AND 2+ years technical experience in software engineering, network engineering, or systems administration OR Bachelor's Degree in Computer Science, Information Technology, or related field AND 4+ years technical experience in software engineering, network engineering, or systems administration OR equivalent experience.Other Requirements:Security Clearance Requirements: Candidates must be able to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include, but are not limited to the following specialized security screenings: The successful candidate must have an active U.S. Government Top Secret Clearance with access to Sensitive Compartmented Information (SCI) based on a Single Scope Background Investigation (SSBI) with Polygraph. Ability to meet Microsoft, customer and/or government security screening requirements are required pre-offer and post-hire for this role. Failure to maintain or obtain the appropriate U.S. Government clearance and/or customer screening requirements may result in employment action up to and including termination.Clearance Verification: This position requires successful verification of the stated security clearance to meet federal government customer requirements. You will be asked to provide clearance verification information prior to an offer of employment.Microsoft Cloud Background Check: This position will be required to pass the Microsoft Cloud background check upon hire/transfer and every two years thereafter. Citizenship & Citizenship Verification: This position requires verification of U.S. citizenship due to citizenship-based legal restrictions. Specifically, this position supports United States federal, state, and/or local United States government agency customer and is subject to certain citizenship-based restrictions where required or permitted by applicable law. To meet this legal requirement, citizenship will be verified via a valid passport, or other approved documents, or verified US government ClearancePreferred Qualifications:Bachelor's Degree in Computer Science, Information Technology, or related field AND 8+ years technical experience in software engineering, network engineering, service engineering, or systems engineeringOR equivalent experience.3+ years technical experience working with large-scale cloud or distributed systemsExperience building automation with Ansible and developing/operating CI/CD pipelines (e.g., Azure DevOps, GitHub Actions) to deliver reliable, repeatable deployments.Expertise in problem solving and analyzing distributed systems and critical production service environmentsExpertise in Linux, specifically Rocky 9, Redhat, Mariner or similar in throughput management, troubleshooting and security hardeningSite Reliability Engineering IC4 - The typical base pay range for this role across the U.S. is USD $119,800 - $234,700 per year. There is a different range applicable to specific work locations, within the San Francisco Bay area and New York City metropolitan area, and the base pay range for this role in those locations is USD $160,200 - $261,000 per year. Certain roles may be eligible for benefits and other compensation. Find additional benefits and pay information here:This position will be open for a minimum of 5 days, with applications accepted on an ongoing basis until the position is filled.Microsoft is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with religious accommodations and/or a reasonable accommodation due to a disability during the application process, read more about requesting accommodations.
- ...most critical customers? We’re looking for a Software Engineer with the right mix of software development, on-line... ...highest expectations for feature quality, security, reliability, availability, and performance.The Site Reliability Engineering (SRE) team provides...SuggestedInternshipWork at office
$119.8k - $234.7k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...opportunity for a Senior Software Engineer in the Cloud+Artificial Intelligence (AI... ...air-gapped environments, with a focus on reliability, security, and compliance.· Ability to build...SuggestedOngoing contractLocal area3 days per week$85.4k - $168.1k
...per yearEmployment type: Full-TimeWork site: Fully on-siteRole type: Individual ContributorTravel... ...looking for an early in career Software Engineer to join the Authentication team. This... ...that will improve the availability, reliability, efficiency, observability, and...SuggestedOngoing contractLocal area$119.8k - $234.7k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...no further than the Microsoft Defender engineering team. You will be building and... ...improve services to be scalable and highly reliable.Help deliver and improve engineering systems...SuggestedOngoing contractLocal area3 days per week$142.8k - $274.8k
...USD $142,800 - $274,800 per yearEmployment type: Full-TimeWork site: Fully on-siteRole type: People ManagerTravel: Less than 25%Profession... ...U.S. Government and sovereign cloud customers. As a Software Engineering Manager on the MSC team, you will partner with engineering...SuggestedOngoing contractLocal area$91.4k - $187k
Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand... ...- IC3Must Possess U.S. CITIZENSHIP and ACTIVE TS/SCI W/POLY SECURITY CLEARANCE To Be Eligible for Consideration. Technical...Temporary workFlexible hoursShift workWeekend work$142.8k - $274.8k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: People... ...exciting opportunity for a Principal Software Engineer Manager to lead engineering teams... ...design and deployment through live site reliability, operational excellence, and customer impact...Ongoing contractLocal area3 days per week$84.9k - $209.5k
...customer support, service owners, and engineering teams around the globe to ensure high-quality... ...- IC4Escalation points for junior site reliability engineers during complex or high-impact... ...Citizenship and possess and maintain TS/SCI w/Poly security clearance, reside in Austin,...Temporary workMonday to FridayFlexible hoursShift workNight shift- ...As a Software Development Engineer with a focus on Identity and Access Management (IDAM), you will leverage your expertise in both software... ...Active TS/SCI clearance and ability to obtain and maintain a CI poly. Must meet DoD 8570 IAT Level II requirements including one...Full timeTemporary workRelocation package
- ...technologies to ensure that scientists and engineers will be able to fully utilize modern HPC... ...deployment and coordination of initial site support for DC-MLS deployments which include... ...• Security+ Certification•TS/SCI with CI Poly security clearance required to...Full timeContract workWork at officeRemote workMonday to FridayFlexible hours
$164.38k - $274.52k
...Systems Engineer SME (TS/SCI with Poly Required) Job Category: IT Infrastructure and Operations Requisition... ...Type: Full-Time Work Location: On-site Salary Range: $164,382.40 USD to $274... ...building and maintaining trusted and reliable partnerships with our customers and industry...Full timeWork experience placementFlexible hours$81.1k - $187k
...infrastructure and/or service according to terms for reliability and functionality.- Assists team members... ...deployments.- Gains basic knowledge of site reliability trends and shares relevant... ...are seeking a skilled Site Reliability Engineer to design, build, operate, and automate...Temporary workImmediate startFlexible hoursShift work$135.8k - $183.8k
...dynamic and flexible work environment with competitive benefits and the ability to grow your career.We are looking for a Site Reliability Engineer to support our team responsible for building, managing, maintaining, deploying, and securing mission-critical services to...Work at officeFlexible hours$122k - $253k
...community, defense, civil, and commercial markets.Job Title: Site Reliability EngineerLocation: Sterling, VAClearance: TS/SCI Poly**This position is CONTINGENT upon contract award**The Site Reliability Engineer (SRE) collaboratively works closely with the contract...Full timeContract work- ...We are seeking one (1) Software Engineer to provide software development support for MOON 1998 . This position is in Chantilly .... ...using AWS Cloud Development Kits (CDK) Active TS/SCI w/ FS Poly Demonstrated experience with Systems Administration...Full time
$140k - $205k
Senior Technology Site Reliability EngineerCooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operations team.Position summary: The Senior Technology Site Reliability Engineer (“SRE”) is responsible for ensuring the reliability...Full timeTemporary workWork at officeFlexible hoursWeekend work$109.18k - $163.77k
...channels. As a global company, we have offices in nine countries and can insert advertisements around the world.Job SummaryThe Site Reliability Engineering team is responsible for managing the critical infrastructure that powers FreeWheel's Streaming Hub platform. Streaming...Full time$146k - $194k
...focused on positioning Anduril as a lead provider of specialized engineering and products for Intelligence Community (IC) customers. We... ...pressing national security requirements.ABOUT THE JOBAs a Site Reliability Engineer, your primary mission is to ensure the health,...Full timeWork experience placementImmediate startRemote work$138.1k - $198.2k
...more intuitive with technology that simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments... ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$121.5k - $264.1k
...sharing guidance on practices and terms for reliability and functionality.- Supervises team... ...developing and maintaining knowledge of site reliability trends and sharing valuable... ...Experience:9 years of experience in software engineering, infrastructure management, or related...Temporary workImmediate startFlexible hours$102k - $234.6k
...members in designing and architecting infrastructure and service for reliability and functionality. Provides day-to-day direction to help... ...to experiment with new technology, execute improvements, build site reliability knowledge, and provide clear data.Only Oracle brings...Temporary workImmediate startFlexible hours- ...timeDescriptionTake the next step in your professional career with Solerity. As a recognized leader in providing Information Technology, Engineering Services, Program Management, and Consulting Services to the U.S. Federal Government and Intelligence Community, we deliver...For contractorsRemote workFlexible hours
- ...Job Description: THE WORK This senior role fosters collaboration with other senior engineers for the development of advanced data analytics solutions and agile development projects in support of a high-visibility mission. This position involves providing technical...
$87.1k - $157.45k
...Qualifications by level -Mid-Level Software Development EngineerBachelor’s degree in Computer Science, Information Systems, Systems Engineering, or related field with 4+ years of experience (or Master’s with 2+ years)2+ years supporting Intelligence Community or DoD...Full timeFor contractors- ...global infrastructure and delivery of solutions that drive influence operations. The Sponsor requires specialized skills in cloud engineering and full stack development. WORK REQUIREMENTS: Full Stack Development Support – HRR: Yes • The Contractor shall participate in...Full timeFor contractorsWork at officeImmediate startRemote work
$130k - $200k
Summary Position Title: Site Reliability Engineer Position ID: TA247 Location(s): On-site; Aurora, CO; Herndon, VA Application Deadline:... ...Bitbucket and GitLab. ~ An active Top Secret/SCI clearance with Poly. ~ DOD 8570 or 8140 IAT Level II Certification, such as...Full timeTemporary workLocal area$128.83k - $193.25k
...you will be responsible for ensuring the reliability, scalability, and performance of our data systems. Working closely with data engineers and other operation sub-teams, you will manage... ...and benefits summary on our careers site for more details.EducationBachelor's DegreeWhile...Full time- ...EngineerDepartment: Govt Customer-ChantillyLocation: Chantilly, VATENICA is looking to hire a Cyber Splunk Systems Engineer. Must have active TS/SCI with CI poly. Position Description:The Cyber Systems Engineer Project Management Technical Support provides support to the...Contract workFor contractors
- ...The Contractor team shall provide systems engineering, custom application development, database administration and operations and maintenance... ...(cloud) components to maintain sustainability and reliability of the Sponsor’s system. Software engineering # Analyze complex...Full timeFor contractors
$62k - $141k
Site Reliability Engineer The Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you...Full timeContract workPart timeWork at officeLocal areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer - CTJ - POLY. Be the first to apply!
- site services specialist Reston, VA
- construction site safety Reston, VA
- site leader Reston, VA
- official site Reston, VA
- website content developer Reston, VA
- IT site lead Reston, VA
- site safety Reston, VA
- junior website developer Reston, VA
- on-site clinical research associate (traveling/remote) Reston, VA
- site reliability engineering manager


