Site Reliability Engineer
$110k - $125kASM Research, An Accenture Federal Services Company
The Site Reliability Engineer maintains and improves the availability, performance, resilience, and operational recoverability of enterprise identity, credential, and access-management services. This role supports cloud identity and directory platforms, including Microsoft Entra ID, Active Directory, and integrated authentication services, by applying site reliability engineering practices to measure service health, identify degradation, reduce operational toil, and strengthen reliability for mission-critical authentication and access services.
The engineer implements monitoring, alerting, dashboards, log-analysis capabilities, automation, and infrastructure-as-code solutions that support repeatable operations and recovery. Working with senior engineers, identity engineers, and operations teams, the Site Reliability Engineer supports incident response, root-cause analysis, preventive remediation, observability standards, and continuous service improvements across the enterprise identity environment.
Key Responsibilities
Monitor service-level indicators, service-level objectives, service-level agreements, availability, latency, authentication success rates, provisioning outcomes, directory synchronization status, endpoint availability, and service dependency health for enterprise identity services.
Implement and maintain observability solutions for Microsoft Entra ID, Active Directory, and integrated identity services using Azure Monitor, Log Analytics, Splunk, or comparable monitoring and security-event platforms.
Collect, query, correlate, and analyze identity audit, sign-in, provisioning, directory, and infrastructure logs to investigate service failures, identify abnormal trends, validate remediation, and support compliance-oriented operational reporting.
Support incident response for identity-service disruptions by assessing scope and business impact, executing documented runbooks, communicating status, engaging appropriate escalation teams, preserving diagnostic evidence, and validating service recovery.
Perform root-cause analysis for recurring authentication, authorization, provisioning, directory synchronization, monitoring, capacity, and configuration failures; document corrective and preventive actions that improve long-term service reliability.
Develop and support automated remediation workflows that detect known service conditions, validate safeguards, execute approved recovery actions, record results, and escalate when automation does not restore service health.
Use infrastructure-as-code tools, including Terraform, Ansible, ARM templates, or comparable technologies, to define, version, review, deploy, and maintain repeatable cloud and identity-supporting infrastructure configurations.
Apply Git-based version-control practices, including branching, pull requests, peer review, release tagging, change history, configuration rollback, and controlled promotion of infrastructure code across environments.
Support Microsoft Entra ID, Active Directory, or comparable enterprise identity platforms by troubleshooting authentication flows, conditional access effects, application integrations, directory objects, service principals, synchronization dependencies, and access-policy impacts.
Create and maintain operational runbooks, alert-response procedures, dashboards, recovery guides, post-incident records, and reliability-improvement backlogs that enable consistent support and continuous improvement.
Required Qualifications
High school diploma or equivalent and at least 3 years of experience in identity, credential, and access-management engineering, infrastructure operations, cloud operations, DevOps, site reliability engineering, or related enterprise IT roles.
At least 1 year of hands-on experience supporting Microsoft Entra ID, Active Directory, or another enterprise cloud identity or directory platform.
Working knowledge of site reliability engineering concepts, including service-level indicators, service-level objectives, service-level agreements, alert thresholds, availability monitoring, latency measurement, on-call response, reliability reporting, and operational toil reduction.
Experience implementing or supporting monitoring, alerting, log analysis, dashboards, or observability capabilities for cloud, directory, identity, or integrated authentication services.
Experience using infrastructure-as-code tools, such as Terraform, Ansible, ARM templates, or comparable technologies, to define, deploy, and maintain repeatable infrastructure configurations.
Experience using Git-based version control, including branching, pull requests, peer review, change history, tagging, rollback, and controlled promotion of code or configuration across environments.
Demonstrated ability to support incident response, technical triage, root-cause analysis, corrective actions, preventive measures, and recovery validation for service disruptions.
U.S. citizenship and ability to obtain and maintain a Public Trust background investigation.
Preferred Qualifications
Microsoft certification in identity and access administration, cloud administration, security operations, or Azure infrastructure.
HashiCorp Terraform Associate certification or demonstrated experience developing and operating production infrastructure-as-code pipelines.
Experience using Splunk, Azure Monitor, Log Analytics, Microsoft Sentinel, or comparable observability and security-event platforms.
Experience with automated orchestration or scripting using PowerShell, Python, Bash, or comparable technologies.
Compensation Ranges
Compensation ranges for ASM Research positions vary depending on multiple factors; including but not limited to, location, skill set, level of education, certifications, client requirements, contract-specific affordability, government clearance and investigation level, and years of experience. The compensation displayed for this role is a general guideline based on these factors and is unique to each role. Monetary compensation is one component of ASM's overall compensation and benefits package for employees.
EEO Requirements
It is the policy of ASM that an individual's race, color, religion, sex, disability, age, sexual orientation or national origin are not and will not be considered in any personnel or management decisions. We affirm our commitment to these fundamental policies.
All recruiting, hiring, training, and promoting for all job classifications is done without regard to race, color, religion, sex, disability, or age. All decisions on employment are made to abide by the principle of equal employment.
Physical Requirements
The physical requirements described in "Knowledge, Skills and Abilities" above are representative of those which must be met by an employee to successfully perform the primary functions of this job. (For example, "light office duties' or "lifting up to 50 pounds" or "some travel" required.) Reasonable accommodations may be made to enable individuals with qualifying disabilities, who are otherwise qualified, to perform the primary functions.
Disclaimer
The preceding job description has been designed to indicate the general nature and level of work performed by employees within this classification. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications required of employees assigned to this job.
$110,000-$125,000
EEO Requirements
It is the policy of ASM that an individual's race, color, religion, sex, disability, age, gender identity, veteran status, sexual orientation or national origin are not and will not be considered in any personnel or management decisions. We affirm our commitment to these fundamental policies.
All recruiting, hiring, training, and promoting for all job classifications is done without regard to race, color, religion, sex, veteran status, disability, gender identity, or age. All decisions on employment are made to abide by the principle of equal employment.
- ...AnnuallyIndustry Financial ServicesSelling Points Contribute to the reliability of a high-transaction payment platform. Collaborate... ...and SRE principles.Job DescriptionSite Reliability Engineer OverviewThe Site Reliability Engineer ensures the reliability, scalability,...SuggestedRemote work
$87.1k - $157.45k
...throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across... ...to come in and help us build systems that stay reliable when things get complicated. We need a Site Reliability Engineer who has experience building, deploying...SuggestedLocal areaImmediate startWork from homeFlexible hours- ...JD: As a Site Reliability Engineer (SRE) Level II, you will play a key role in maintaining the availability, scalability, and performance of critical infrastructure and services. You will be responsible for building and automating solutions that enhance system...SuggestedWork experience placementWork at office
- ...Senior Site Reliability Engineer We are looking for an adventurous Senior Site Reliability Engineer who loves AWS technologies. You will be a member of an engineering team where collaboration and innovation are a key focus. As part of this team you will design, build...Suggested
$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SuggestedTemporary workImmediate startFlexible hoursShift work$169.3k - $304.7k
...in building and maintaining fast, efficient, scalable, and reliable routing software and infrastructure that is responsible... ...growth and stability of our global platform. As a Principal Site Reliability Engineer - Network, you will be responsible for: Architecting,...Work experience placementWork at office$105.79k - $141.05k
...delivers on-demand networking at scale. As Lead SRE, you'll own the reliability of that platform — partnering with operations teams and... ..., and automation, and you'll coordinate across architecture, engineering, and systems development organizations to measurably improve...Temporary workRemote work- DescriptionJob Description SummaryThe Digital Site Reliability Engineer (SRE) - GCP Cloud Adoption Engineer is responsible for facilitating the migration, adoption, and optimization of Google Cloud Platform (GCP) services within the organization.Job DescriptionSummary:This...Full timeH1bWork at officeRemote workWork from homeFlexible hours
$55k - $187k
...Internal Firm Services - Other Management Level Senior Associate Job Description & Summary The Opportunity As a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...Full timeH1b- ...Responsibilities Build production-grade reliability software, including automation, control... ..., break complex problems into work for engineers, serve as technical lead for medium- to... ...Formal training or certification in site reliability engineering concepts and 5+...Full time
- ...company dedicated to making a positive impact on people's lives. Position Summary Reporting to the Director of Engineering, the Lead Site Reliability Engineer (SRE) is a senior technical contributor responsible for building reliable, scalable software systems and...Full timeRemote workMonday to FridayShift workNight shiftWeekend work
- ...platforms, applying strong experience in Ansible, continuous integration and continuous delivery practices, DevOps and Site Reliability Engineering to design, automate, and optimize geospatial data services.Partner with cross functional teams to ensure reliable map based...
- ...SRE DevOps Engineer Location: Columbus, OH / Jersey City, NJ (LOCALS ONLY) Duration: 6 Months Contract-to-Hire Rate... ...support scalable and secure platforms while improving system reliability and operational stability. Key Responsibilities...Contract workLocal area
- JPMorganChase is seeking a Technology Support III to join the Mainframe and Mid-Range Compute Site Reliability and Engineering team. You will support end-to-end infrastructure services, collaborate with stakeholders, and ensure high availability across a globally distributed...
$51 - $61 per hour
...onsite at the project, significantly reducing and/or eliminating the demands to travel. Key Responsibilities:As a Release Train Engineer, you will be responsible for facilitating Agile Release Train events and processes including communicating with stakeholders...Hourly payLive inWork at officeLocal areaImmediate startFlexible hoursShift work- ...We operate as a hybrid workplace with offices in Boston, MA; Columbus, OH; and College Park, MD. ABOUT THE TEAM The Site Reliability Engineering team helps keep Immuta’s products reliable, deployable, and maintainable in production. The team builds software and automation...Full timeSummer workInternshipLocal areaWorldwide
- Digital - Principal SRE (AI Engineer)Skip to main content#Digital - Principal SRE (AI Engineer... ...intelligence, machine learning, and reliability engineering. This professional is... ...Practice: Visit Huntington's Career Web Site for more details.**Note to Agency Recruiters...Full timeH1bWork at officeRemote workWork from homeFlexible hours
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Enterprise technology, Infrastructure Platforns team, you hold a leadership...
- Distributed Systems Software Engineer, Python / Go Join to apply for the Distributed Systems Software Engineer, Python / Go role at Canonical... ...testing approaches and infrastructure for validating reliability, performance, and resilience of cloud orchestration tools and...Full timeLocal areaRemote workWorldwide
- ...reasoning, and control to deliver field-ready AI that is risk-aware, reliable, and continuously improving through real-world use.Big, hard... ...the impossible possible together.We are seeking an IT Systems Engineer, this role will be responsible for administering and scaling...
$86.4k - $159k
...provide innovative, high-technology deployable equipment and engineered product solutions that not only extend the service life of fielded... ...the evolving needs of national security with precision and reliability.Job SummaryWe are currently seeking a Software Engineer -...Full timeWork experience placementWork at officeLocal areaRemote workFlexible hours$110k - $125k
...organization’s AI platforms, endpoints, infrastructure, networks, SaaS applications, cloud platforms, IAM, and more.ABOUT THE ROLEThis engineering role focuses on building solutions that serve the organization. You will work with stakeholders to discover, design, and...Full timeWork at officeLocal areaWorldwide3 days per week$101.2k - $126.5k
...thinking.Woolpert is an award-winning, global leader in architecture, engineering, and geospatial services. We blend design excellence with... ...for career growth.Position OverviewWoolpert is hiring a Site Civil Engineer (PE) to join our dynamic Land Development Engineering...Local areaFlexible hoursNight shift- ...to support one another. We’re looking for a Senior Software Engineer to be an architect and caretaker of our Perception codebase,... ...registration, seam localization, real-time weld tracking) into reliable production services that ship to our internationally deployed fleet...Full timeFlexible hours
- Description: Software Development Engineer For over two decades, Aspirion has delivered market-leading revenue cycle services.... ...hands-on coding expertise, and a passion for building scalable, reliable, and secure applications that align with the company’s goals....
- ...supervision of the Chief Technology Officer, the Lead Enterprise Systems Engineer works to analyze, plan, implement, maintain, troubleshoot, and... ...tools and gain efficiency• Work as part of a team• Work with on-site equipmentWhat You Bring to the Team• Other duties as assigned•...Full timeFor contractorsWork at officeNight shift
- ...reasoning, and control to deliver field-ready AI that is risk-aware, reliable, and continuously improving through real-world use. Big,... ...possible together. We're seeking a Robotics Software Engineer who thrives on hard systems problems. Welding is where we started...Full timeFlexible hours
- ...winning team.Job Description:Kokosing is seeking a Construction AI Engineer. This individual will serve as a strategic partner between... ...construction sector. You will develop robust systems that function reliably within the high-stakes, risk-management environment of physical...Full timeFor contractorsWork at office
$120k - $140k
...more than 350,000 partnerships that deliver measurable business results. Your Role at impact.com : As a Full Stack Software Engineer on the Tracking team, you will be working within a fast-paced, agile team, implementing the next generation of the core user...Full timeWork at officeHome officeFlexible hoursWeekend work- ...developing, and maintaining software solutions that support Centra’s operational and digital needs. This role focuses on building secure, reliable applications, integrating systems, and supporting technology modernization efforts that enhance internal efficiency and Member...Full timeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Columbus, OH
- site reliability engineer Columbus, OH
- remote website tester Columbus, OH
- IT site lead Columbus, OH
- site safety Columbus, OH
- website content developer Columbus, OH
- site leader Columbus, OH
- on-site clinical research associate (traveling/remote) Columbus, OH
- junior website developer Columbus, OH
- historic site Columbus, OH




