Site Reliability Engineer
$110k - $125kASM Research, An Accenture Federal Services Company
The Site Reliability Engineer maintains and improves the availability, performance, resilience, and operational recoverability of enterprise identity, credential, and access-management services. This role supports cloud identity and directory platforms, including Microsoft Entra ID, Active Directory, and integrated authentication services, by applying site reliability engineering practices to measure service health, identify degradation, reduce operational toil, and strengthen reliability for mission-critical authentication and access services.
The engineer implements monitoring, alerting, dashboards, log-analysis capabilities, automation, and infrastructure-as-code solutions that support repeatable operations and recovery. Working with senior engineers, identity engineers, and operations teams, the Site Reliability Engineer supports incident response, root-cause analysis, preventive remediation, observability standards, and continuous service improvements across the enterprise identity environment.
Key Responsibilities
- Monitor service-level indicators, service-level objectives, service-level agreements, availability, latency, authentication success rates, provisioning outcomes, directory synchronization status, endpoint availability, and service dependency health for enterprise identity services.
- Implement and maintain observability solutions for Microsoft Entra ID, Active Directory, and integrated identity services using Azure Monitor, Log Analytics, Splunk, or comparable monitoring and security-event platforms.
- Collect, query, correlate, and analyze identity audit, sign-in, provisioning, directory, and infrastructure logs to investigate service failures, identify abnormal trends, validate remediation, and support compliance-oriented operational reporting.
- Support incident response for identity-service disruptions by assessing scope and business impact, executing documented runbooks, communicating status, engaging appropriate escalation teams, preserving diagnostic evidence, and validating service recovery.
- Perform root-cause analysis for recurring authentication, authorization, provisioning, directory synchronization, monitoring, capacity, and configuration failures; document corrective and preventive actions that improve long-term service reliability.
- Develop and support automated remediation workflows that detect known service conditions, validate safeguards, execute approved recovery actions, record results, and escalates when automation does not restore service health.
- Use infrastructure-as-code tools, including Terraform, Ansible, ARM templates, or comparable technologies, to define, version, review, deploy, and maintain repeatable cloud and identity-supporting infrastructure configurations.
- Apply Git-based version-control practices, including branching, pull requests, peer review, release tagging, change history, configuration rollback, and controlled promotion of infrastructure code across environments.
- Support Microsoft Entra ID, Active Directory or comparable enterprise identity platforms by troubleshooting authentication flows, conditional access effects, application integrations, directory objects, service principals, synchronization dependencies, and access-policy impacts.
- Create and maintain operational runbooks, alert-response procedures, dashboards, recovery guides, post-incident records, and reliability-improvement backlogs that enable consistent support and continuous improvement.
Required Qualifications
- High school diploma or equivalent and at least 3 years of experience in identity, credential, and access-management engineering, infrastructure operations, cloud operations, DevOps, site reliability engineering, or related enterprise IT roles.
- At least 1 year of hands-on experience supporting Microsoft Entra ID, Active Directory, or another enterprise cloud identity or directory platform.
- Working knowledge of site reliability engineering concepts, including service-level indicators, service-level objectives, service-level agreements, alert thresholds, availability monitoring, latency measurement, on-call response, reliability reporting, and operational toil reduction.
- Experience implementing or supporting monitoring, alerting, log analysis, dashboards, or observability capabilities for cloud, directory, identity, or integrated authentication services.
- Experience using infrastructure-as-code tools, such as Terraform, Ansible, ARM templates, or comparable technologies, to define, deploy, and maintain repeatable infrastructure configurations.
- Experience using Git-based version control, including branching, pull requests, peer review, change history, tagging, rollback, and controlled promotion of code or configuration across environments.
- Demonstrated ability to support incident response, technical triage, root-cause analysis, corrective actions, preventive measures, and recovery validation for service disruptions.
- U.S. citizenship and ability to obtain and maintain a Public Trust background investigation.
Preferred Qualifications
- Microsoft certification in identity and access administration, cloud administration, security operations, or Azure infrastructure.
- HashiCorp Terraform Associate certification or demonstrated experience developing and operating production infrastructure-as-code pipelines.
- Experience using Splunk, Azure Monitor, Log Analytics, Microsoft Sentinel, or comparable observability and security-event platforms.
- Experience with automated orchestration or scripting using PowerShell, Python, Bash, or comparable technologies.
Compensation Ranges
Compensation ranges for ASM Research positions vary depending on multiple factors; including but not limited to, location, skill set, level of education, certifications, client requirements, contract-specific affordability, government clearance and investigation level, and years of experience. The compensation displayed for this role is a general guideline based on these factors and is unique to each role. Monetary compensation is one component of ASM’s overall compensation and benefits package for employees.
EEO Requirements
It is the policy of ASM that an individual’s race, color, religion, sex, disability, age, sexual orientation or national origin are not and will not be considered in any personnel or management decisions. We affirm our commitment to these fundamental policies.
All recruiting, hiring, training, and promoting for all job classifications is done without regard to race, color, religion, sex, disability, or age. All decisions on employment are made to abide by the principle of equal employment.
Physical Requirements
The physical requirements described in “Knowledge, Skills and Abilities” above are representative of those which must be met by an employee to successfully perform the primary functions of this job. (For example, “light office duties’ or “lifting up to 50 pounds” or “some travel” required.) Reasonable accommodations may be made to enable individuals with qualifying disabilities, who are otherwise qualified, to perform the primary functions.
Disclaimer
The preceding job description has been designed to indicate the general nature and level of work performed by employees within this classification. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications required of employees assigned to this job.
$110,000–$125,000
EEO Requirements
It is the policy of ASM that an individual’s race, color, religion, sex, disability, age, gender identity, veteran status, sexual orientation or national origin are not and will not be considered in any personnel or management decisions. We affirm our commitment to these fundamental policies.
All recruiting, hiring, training, and promoting for all job classifications is done without regard to race, color, religion, sex, veteran status, disability, gender identity, or age. All decisions on employment are made to abide by the principle of equal employment.
#J-18808-Ljbffr$141.8k - $195k
...best work, grow fast, and bring their full selves to the herd. Why You'll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl...SuggestedTemporary workRemote work- ...About the Opportunity Help ensure healthcare professionals can reliably access the applications they depend on to deliver patient care. Oracle Health is seeking a Principal Site Reliability Engineer to strengthen the reliability, performance, security, and...Suggested
- ...Job Description Job Description Job Title: Senior AWS Site Reliability Engineer (SRE) Location: Birmingham, Alabama Type: Contract To Hire Work Model: Onsite – onsite Hours: 40.0 Security Clearance: Overview Responsibilities Implement and improve...SuggestedContract workLocal area
- ..., process efficiency, and plant wide safety Provide technical assistance in the scope of engineering, preventative and predictive maintenance, and other facets of reliability engineering Update existing plant and maintenance equipment records and processes Monitor...SuggestedWork experience placementWork at officeLocal areaImmediate start
- ...Our client is seeking a Reliability Engineer to support equipment reliability, maintenance, and process improvement within a large-scale steel... ...Mode and Effects Analysis. The Environment This is an on-site individual contributor role based in Cayce, South Carolina....SuggestedWork experience placementWork at officeRelocation package
- ...Reliability Engineer The Reliability Engineer is responsible for overall evaluation and improvement of product reliability, through collecting and analyzing failure data, recommending design and or process improvement, support MTBF modeling activities, interfacing...
$293.9k - $406.8k
...comprehensive security outcomes, as a Distinguished Engineer. The team delivers secure, scalable... ...networking, with a strong emphasis on reliability, interoperability, and long-term... ...insurance. Please see the Cisco careers site to discover more benefits and perks. Employees...Full timeTemporary workLocal areaRemote workFlexible hours- ...Job Description Job Description #LI-DNI Senior Software Engineer - Modeling and Simulation Location: Onsite in Columbia,... ...create a safer world by translating scientific discoveries into reliable products that address urgent national security needs... at the...Work at officeLocal areaImmediate startRemote workWork from homeRelocationRelocation packageMonday to Friday3 days per week
$170k - $260k
...Are you a Senior Signals Software Engineer who is ready for a new challenge that will launch your career to the next level? Tired of... ...Cyber Security solution spaces. We excel at delivering stable and reliable software solutions using Agile Software Development principles....Full timeContract workRemote workWork from homeRelocation package$115k - $192.9k
...enterprise uses to run a highly Profitable, Critical and Complex Parts & Service business In this position.... As a software engineer on our global team, you'll work with Google Cloud Platform, modern IDEs, and GenAI tools to write code — using a tech stack that's...Immediate startRemote workFlexible hours$110k - $135k
...Software Engineer Company: Dedham Group Location: Remote, United States Date Posted: Sep 28, 2026 Employment Type: Full Time Job ID: R-2174 Description About The Dedham Group: The Dedham Group is proud to be a part of Norstella, an organization that...Full timeTemporary workLocal areaRemote workFlexible hoursShift work$235k - $300k
...The Vice President, Software Engineering is a pivotal role within Vontier's Convenience Retail business, spearheading the global engineering organization dedicated to retail solutions technologies. This encompasses a broad spectrum of products and systems, from cloud IoT...Local areaWorldwide$102.3k - $209.5k
...software, firmware, and hardware layers, working in DevOps and incident response paradigms to maintain reliability and availability. Work closely with platform engineering, operations, firmware development, silicon/board vendors, and data center teams to support and...Temporary workImmediate startFlexible hoursShift work$109k - $203k
Software Engineer - AI - CoCounsel Forward Deployed EngineeringAre you excited about building AI solutions that help legal professionals... ...will help translate legal workflows and business needs into reliable technical solutions that combine foundation models, proprietary...Full timeContract workWork at officeLocal areaFlexible hours- ...SC at the start of the contract. If you cant relocate and be on site 50% of the time do not submit your resume. Our direct... ...Grid, and others). • Standardizing and documenting design and engineering patterns, processes, and solutions. • Azure applications...Contract workRelocation
- ...Job Description Job Description AWS DevOps Engineer Job Category: Software Development / Engineering Job Type: Permanent... ..., and partnering with application teams to improve deployment reliability and efficiency. Required Qualifications ~5+ years of hands...Permanent employmentFull time
$83k - $166.1k
...maintaining backend systems, and ensuring the long-term scalability, reliability, and sustainability of the platform. The ideal candidate... ...Bachelor’s degree in Computer Science, Information Systems, Engineering, Healthcare Informatics, or a related field. Master’s degree...Temporary workWork experience placementImmediate startFlexible hours- Microsoft Power Platform Developer Michelin North America is hiring a Microsoft Power Platform Developer to support business teams by building and improving digital solutions. You will work onsite with internal users and technical teams to design tools that improve efficiency...Long term contract
- A leading software solutions provider is seeking a Software Engineer, Platform in Columbia, SC, to design and maintain scalable cloud-native applications. The ideal candidate will have at least 6 years of application development experience, strong skills in AWS or Azure...Full time
- ...Job Description: We are seeking a skilled and versatile Systems Engineer to support our G-TEAD (Global Tactical Edge Acquisition Directorate) effort located an Shaw AFB in Sumter County SC. The ideal candidate will provide combined systems engineering and program integration...Full timeTemporary workPart timeFor contractorsFor subcontractorSeasonal workWork at officeFlexible hours
- ...Enforcement Division (SLED) is seeking a highly skilled Senior Systems Engineer to provide advanced-level design, implementation, and support... ..., and backup platforms to ensure maximum performance and reliability. Serve as the technical lead for infrastructure...Local areaNight shift
- ...Description ThisWay Global is looking for a Distributed Systems Engineer in a remote role within the United States. ThisWay Global,... ..., distributed architecture, fault tolerance, and HPC-grade reliability. Location: Remote – United States Department:...Full timeRemote work
$95k - $115k
Senior Salesforce Developer (Remote) Ateko is seeking a Senior Salesforce Developer to work with our clients; understanding their business needs and translating them into technical solutions. In this role you will demonstrate in-depth knowledge of Salesforce development...Full timeWork at officeRemote workFlexible hours- ...Bureau of IT Infrastructure & Operations, Richland County. This is an in-office role and not a telecommute or remote position. Systems Engineer II We are looking for a Systems Engineer II who, under minimal supervision, is responsible for the design, implementation,...Work at officeWeekend work
- ...more at later. com . About this position: We are seeking a Staff Engineer with a strong background in modern web development... ...coding, vulnerability assessments, secrets management). - Ensure reliability and scalability through test automation, performance profiling...Full time
$100k
...created new and exciting opportunities in transportation that Hertz is uniquely positioned to capitalize on. We're looking for software engineers who will modernize Hertz's tech stack and, in the process, ship delightful products to meet the ever-increasing demands of our...Work experience placementWorldwide- Senior Software Engineer Opportunity Cognito is a fast-growing SaaS company empowering users to quickly build forms - and form-driven... ...delivering high-availability SaaS services with scalability, reliability, and maintainability in mind Experience with Azure or similar...Flexible hours
$125k - $150k
...across cloud and on-premises environments while working with technologies from NVIDIA, Microsoft, and Google. This role is ideal for engineers who love solving complex problems, building production software, and learning new technologies in a fast-moving AI landscape....Full timeRemote workWorldwideFlexible hours$79.2k - $209.5k
...feature and subsystem design, including tradeoffs around security, reliability, scalability, maintainability, performance, and change... ...standards, and reduce technical debt. Improve the team through engineering practices, operational practices, development process,...Temporary workFlexible hours$100k - $120k
...partnerships, while our focus on creativity and innovative solutions empowers our customer communities to thrive. The Senior Software Engineer on the New Ventures team designs, creates, maintains, audits, and improves software applications by performing coding, debugging,...Full timeTemporary workLocal areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site safety Columbia, SC
- website coordinator Columbia, SC
- on-site clinical research associate (traveling/remote) Columbia, SC
- site services specialist Columbia, SC
- on site coordinator Columbia, SC
- construction site safety Columbia, SC
- junior website developer Columbia, SC
- historic site Columbia, SC
- IT site lead Columbia, SC
- site leader Columbia, SC




