Site Reliability Engineer
$110k - $125kASM Research, An Accenture Federal Services Company
The Site Reliability Engineer maintains and improves the availability, performance, resilience, and operational recoverability of enterprise identity, credential, and access-management services. This role supports cloud identity and directory platforms, including Microsoft Entra ID, Active Directory, and integrated authentication services, by applying site reliability engineering practices to measure service health, identify degradation, reduce operational toil, and strengthen reliability for mission-critical authentication and access services.
The engineer implements monitoring, alerting, dashboards, log-analysis capabilities, automation, and infrastructure-as-code solutions that support repeatable operations and recovery. Working with senior engineers, identity engineers, and operations teams, the Site Reliability Engineer supports incident response, root-cause analysis, preventive remediation, observability standards, and continuous service improvements across the enterprise identity environment.
Key Responsibilities
Monitor service-level indicators, service-level objectives, service-level agreements, availability, latency, authentication success rates, provisioning outcomes, directory synchronization status, endpoint availability, and service dependency health for enterprise identity services.
Implement and maintain observability solutions for Microsoft Entra ID, Active Directory, and integrated identity services using Azure Monitor, Log Analytics, Splunk, or comparable monitoring and security-event platforms.
Collect, query, correlate, and analyze identity audit, sign-in, provisioning, directory, and infrastructure logs to investigate service failures, identify abnormal trends, validate remediation, and support compliance-oriented operational reporting.
Support incident response for identity-service disruptions by assessing scope and business impact, executing documented runbooks, communicating status, engaging appropriate escalation teams, preserving diagnostic evidence, and validating service recovery.
Perform root-cause analysis for recurring authentication, authorization, provisioning, directory synchronization, monitoring, capacity, and configuration failures; document corrective and preventive actions that improve long-term service reliability.
Develop and support automated remediation workflows that detect known service conditions, validate safeguards, execute approved recovery actions, record results, and escalate when automation does not restore service health.
Use infrastructure-as-code tools, including Terraform, Ansible, ARM templates, or comparable technologies, to define, version, review, deploy, and maintain repeatable cloud and identity-supporting infrastructure configurations.
Apply Git-based version-control practices, including branching, pull requests, peer review, release tagging, change history, configuration rollback, and controlled promotion of infrastructure code across environments.
Support Microsoft Entra ID, Active Directory, or comparable enterprise identity platforms by troubleshooting authentication flows, conditional access effects, application integrations, directory objects, service principals, synchronization dependencies, and access-policy impacts.
Create and maintain operational runbooks, alert-response procedures, dashboards, recovery guides, post-incident records, and reliability-improvement backlogs that enable consistent support and continuous improvement.
Required Qualifications
High school diploma or equivalent and at least 3 years of experience in identity, credential, and access-management engineering, infrastructure operations, cloud operations, DevOps, site reliability engineering, or related enterprise IT roles.
At least 1 year of hands-on experience supporting Microsoft Entra ID, Active Directory, or another enterprise cloud identity or directory platform.
Working knowledge of site reliability engineering concepts, including service-level indicators, service-level objectives, service-level agreements, alert thresholds, availability monitoring, latency measurement, on-call response, reliability reporting, and operational toil reduction.
Experience implementing or supporting monitoring, alerting, log analysis, dashboards, or observability capabilities for cloud, directory, identity, or integrated authentication services.
Experience using infrastructure-as-code tools, such as Terraform, Ansible, ARM templates, or comparable technologies, to define, deploy, and maintain repeatable infrastructure configurations.
Experience using Git-based version control, including branching, pull requests, peer review, change history, tagging, rollback, and controlled promotion of code or configuration across environments.
Demonstrated ability to support incident response, technical triage, root-cause analysis, corrective actions, preventive measures, and recovery validation for service disruptions.
U.S. citizenship and ability to obtain and maintain a Public Trust background investigation.
Preferred Qualifications
Microsoft certification in identity and access administration, cloud administration, security operations, or Azure infrastructure.
HashiCorp Terraform Associate certification or demonstrated experience developing and operating production infrastructure-as-code pipelines.
Experience using Splunk, Azure Monitor, Log Analytics, Microsoft Sentinel, or comparable observability and security-event platforms.
Experience with automated orchestration or scripting using PowerShell, Python, Bash, or comparable technologies.
Compensation Ranges
Compensation ranges for ASM Research positions vary depending on multiple factors; including but not limited to, location, skill set, level of education, certifications, client requirements, contract-specific affordability, government clearance and investigation level, and years of experience. The compensation displayed for this role is a general guideline based on these factors and is unique to each role. Monetary compensation is one component of ASM's overall compensation and benefits package for employees.
EEO Requirements
It is the policy of ASM that an individual's race, color, religion, sex, disability, age, sexual orientation or national origin are not and will not be considered in any personnel or management decisions. We affirm our commitment to these fundamental policies.
All recruiting, hiring, training, and promoting for all job classifications is done without regard to race, color, religion, sex, disability, or age. All decisions on employment are made to abide by the principle of equal employment.
Physical Requirements
The physical requirements described in "Knowledge, Skills and Abilities" above are representative of those which must be met by an employee to successfully perform the primary functions of this job. (For example, "light office duties' or "lifting up to 50 pounds" or "some travel" required.) Reasonable accommodations may be made to enable individuals with qualifying disabilities, who are otherwise qualified, to perform the primary functions.
Disclaimer
The preceding job description has been designed to indicate the general nature and level of work performed by employees within this classification. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications required of employees assigned to this job.
$110,000-$125,000
EEO Requirements
It is the policy of ASM that an individual's race, color, religion, sex, disability, age, gender identity, veteran status, sexual orientation or national origin are not and will not be considered in any personnel or management decisions. We affirm our commitment to these fundamental policies.
All recruiting, hiring, training, and promoting for all job classifications is done without regard to race, color, religion, sex, veteran status, disability, gender identity, or age. All decisions on employment are made to abide by the principle of equal employment.
- ...The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper). As...SuggestedWork at officeLocal areaRemote workWorldwide
- ...and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing... ...problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes...SuggestedWork at officeLocal areaRemote workSleeping nights
$168k - $200k
...is passionate about creating transformative change in healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the forefront of building and operating a resilient, observable, and...Suggested$75.7k - $136.3k
...Join Our Compute Site Reliability Engineering Team Our team is responsible for improving the reliability, performance, and scalability of our Compute products and platforms. We solve complex problems, improve how our systems operate, and build automation that makes...SuggestedWork experience placementWork at office- Job Description Job Description Mandatory 1. Technical and Windows Application Architecture – performance evaluation and tuning of applications 2. Cloud knowledge incl. IaaS, PaaS etc. 3. Code as Service – experience working with containers like Docker...Suggested
$169.3k - $304.7k
...in building and maintaining fast, efficient, scalable, and reliable routing software and infrastructure that is responsible... ...growth and stability of our global platform. As a Principal Site Reliability Engineer - Network, you will be responsible for: Architecting,...Work experience placementWork at office$105.79k - $141.05k
...delivers on-demand networking at scale. As Lead SRE, you'll own the reliability of that platform — partnering with operations teams and... ..., and automation, and you'll coordinate across architecture, engineering, and systems development organizations to measurably improve...Temporary workRemote work- ...Director, Site Reliability Engineering (SRE) Are you passionate about leading global engineering teams that keep mission-critical enterprise products reliable, scalable, and resilient? Join Thomson Reuters as Director, Site Reliability Engineering (SRE) within Service...Work at officeLocal areaFlexible hours2 days per week3 days per week
$70.8k - $131.4k
Job DescriptionThomson Reuters is strengthening its Site Reliability Engineering capability to help engineering and operations teams build, operate, and improve reliable production services.The Site Reliability Engineer will support the tools, processes, and operational...Full timeWork at officeLocal areaFlexible hours$80k - $140k
Job DescriptionRBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and...Full timeFlexible hoursShift work$112.3k - $160.6k
...and organization with whom we partner and serve. Does this opportunity interest you? Western National is seeking a Site Reliability Engineer III to join our team! The individual in this role will have the opportunity to design, implement, manage, and monitor on...Full timeWork at officeLocal areaRemote workWork visaFlexible hours$158.9k - $295.1k
Are you passionate about leading global engineering teams that keep mission-critical enterprise products reliable, scalable, and resilient? Join Thomson Reuters as Director, Site Reliability Engineering (SRE) within Service Management in our CIO organization, supporting...Full timeWork at officeLocal areaFlexible hours$55k - $187k
...Internal Firm Services - Other Management Level Senior Associate Job Description & Summary The Opportunity As a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...Full timeH1b- ...associated with pace makers and leads, supportings manufacturing sites and working cross functionally with teams to get... ...years of experience? 1-5 years experienceJob Summary The Reliability Engineer II is responsible for ensuring the safety, reliability, and...
- ...Job Title: Reliability Engineer II Job ID: 27318 Location: 8200 Coral Sea Street NE Mounds View Minnesota 55112 Duration: 08 Months Pay Rate: $45.00-50.00/hr. on W2 Must-Have Skills ~ Bachelor's degree in Engineering (Mechanical, Electrical, Biomedical...
- ...more information: What you would do in this job The Reliability Engineer preforms specialized professional reliability engineering,... ...environments. Positions may require travel between the primary work site and other Council facilities or external sites. When visiting...For contractorsWork at office
$114.6k - $234.6k
...creates durable fixes and preventive controls. Designs performance, reliability, and fault-tolerance improvements for drivers, services, and... ...development lifecycle; provides guidance and coaching to engineers to drive improvements. Utilizes advanced knowledge to develop...Temporary workFlexible hoursShift work- ...We are seeking a highly skilled and motivated Lead Systems Engineer to join our team, focusing on the design, development, and optimization... ...Engineering (MBSE) methodologies to enhance the performance, reliability, and safety of advanced hypersonic test systems. The ideal...For subcontractorLocal area
$85k
...your skills and career. Here, you’ll be supported in progressing – whatever your ambitions. Position SummaryAs a forward deployed AI Engineer, you will serve as an embedded technical expert supporting the development, deployment, and sustainment of AI-enabled solutions...Hourly payContract workShift work$98.4k - $199k
...holding, within Old National Bank's Product Engineering organization, and sets the technical... ...AI and agentic delivery to ship secure, reliable software faster.This is an experienced individual... ...Type: Regular Full-TimeRequisition ID: 2026-19615Workplace Type: On Site$72k - $134k
*At Securian Financial, the internal job title for this role is Engineering Sr. Analyst.Position OverviewSecurian Financial is seeking a software developer to join the Bills, Beneficiary and Beyond (BBB) team within Employee Benefit Solutions Technology (EBST). BBB is responsible...Full timeWork at officeFlexible hours3 days per week$100.4k - $203k
...core values.ResponsibilitiesThe AWS Platform Engineering Manager is a hands-on technical leader responsible for the reliability, security, automation, and evolution of the organization... ...TechnologyPosition Type: Regular Full-TimeRequisition ID: 2026-20655Workplace Type: On SiteTemporary work$100k - $130k
...Job Description Job Description Software Engineer II – Embedded Systems Minneapolis, MN | Direct Hire | $100,000–$130,000 (... ...software implementations while improving system performance and reliability. Integrate software solutions with hardware systems and...Flexible hours$92.82k - $109.2k
...adhering to architectural best practices; considers scalability, reliability and performance of systems/contexts affected when defining... ...meet standardsConducts code reviews to provide guidance on engineering best practices and compliance with development proceduresAccountable...Full timeWork experience placementLocal area3 days per week- ...Senior Integration Software Engineer The Metropolitan Council, the regional government for the seven-county Twin Cities metropolitan area, is seeking a Senior Integration Software Engineer. The Council plans 30 years ahead for the future of the metropolitan area and...
$166.9k - $230.9k
...products and lending partners. The team focuses on improving reliability, observability, and scalability across origination workflows while... ...Upstart’s broader platform evolution. As a Senior Software Engineer (L5) at Upstart, you will design and build scalable backend...Summer workCurrently hiringLocal areaRemote workWork from home- ...Job Overview We are seeking an experienced Senior Software Engineer to join our team to architect, develop, and enhance our secure... ...deployment processes, handle complex troubleshooting, and ensure reliable system operations Drive continuous improvement within the team...Flexible hours
- ...Software Engineer Insight Global is hiring two mid-level Software Engineers to support a large healthcare client, Solventum (formerly... ...discrete microservices. Respond to QA/tester-identified issues and site-reported issues. Participate in an on-call support rotation (...Night shiftRotating shiftDay shiftAfternoon shift
$110k
...testing solutions for the aerospace industry. We deliver highly engineered facilities, electro-mechanical systems, and software... ...schedule available (every-other Friday off)! This position is on-site at our newly renovatedSt. Paul location. Every single thing...Temporary workLocal areaFlexible hours- ...ll Make in this RoleAs a Senior Software Engineer, you will have the opportunity to tap... ...Troubleshooting and optimizing software performance, reliability, and interoperability across hardware... ...Work location: This role follows an on-site working model, requiring the employee to...Full timeH1b
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- IT site lead Saint Paul, MN
- site safety Saint Paul, MN
- site leader Saint Paul, MN
- on-site clinical research associate (traveling/remote) Saint Paul, MN
- junior website developer Saint Paul, MN
- historic site Saint Paul, MN
- on site coordinator Saint Paul, MN
- construction site safety Saint Paul, MN
- official site Saint Paul, MN
- website coordinator Saint Paul, MN


