Site Reliability Engineer
TIH
The position is described below. If you want to apply, click the Apply button at the top or bottom of this page. You'll be required to create an account or sign in to an existing one.If you have a disability and need assistance with the application, you can request a reasonable accommodation. Send an email to Accessibility (accommodation requests only; other inquiries won't receive a response).Regular or Temporary:RegularLanguage Fluency: English (Required)Work Shift:1st Shift (United States of America)Please review the following job description:Lead Site Reliability & Environment Monitoring Engineer (Azure / Dynatrace / ServiceNow)We are seeking a Lead Site Reliability & Environment Monitoring Engineer to establish and evolve our enterprise observability and monitoring strategy across cloud and application platforms. This is a full-time leadership role responsible for owning monitoring design, driving platform decisions, and guiding engineering teams toward modern SRE practices.This individual will act as the technical authority for monitoring and alerting, shaping how signals from Dynatrace flow into ServiceNow and enterprise messaging/paging platforms, and enabling a shift toward automated, intelligent, and self-healing operations.Key ResponsibilitiesStrategic Leadership & Decision-MakingDefine and own the enterprise monitoring and SRE observability strategyServe as the subject matter expert for Dynatrace, ServiceNow integration, and alerting architectureEvaluate and recommend tooling, integration patterns, and platform directionDrive decisions on alerting philosophy, noise reduction, and signal quality improvementPlatform Ownership & ArchitectureArchitect and standardize end-to-end monitoring and SRE pipelines:Dynatrace ServiceNow incident lifecycleAlert correlation, deduplication, and prioritizationIntegration with paging systems (PagerDuty, SMS, voice, Teams)Establish best practices for:Event ingestion and enrichmentIncident routing and automated assignmentIntegration with CMDB and service mappingSite Reliability Engineering (SRE) LeadershipLead adoption of SRE principles, including:SLIs, SLOs, and error budgetsReliability engineering practices across servicesProactive monitoring and resilience designChampion a shift from reactive operations to proactive reliability engineeringInfluence application and platform teams to build observable, resilient systems by designAutomation & Self-Healing EnablementDrive development of automated remediation and self-healing capabilitiesLeverage Dynatrace workflows, Azure services, and automation frameworks to:Reduce manual incident handlingEliminate repeatable operational tasksMinimize unnecessary pagingServiceNow & Observability Integration LeadershipOwn integration between Dynatrace and ServiceNow ITSM/ITOM, including:Incident, Event Management, and CMDB alignmentService mapping and dependency visibilityGovernance for application/service taggingDefine standards for:Automated incident creation and resolutionPriority assignment and routing logicMonitoring-to-ITSM data synchronizationTeam Leadership & Cross-Functional InfluenceProvide technical leadership and mentorship across SRE, platform, and application teamsAct as a central point of coordination between engineering, cloud, and ITSM teamsLead workshops and working sessions to:Drive monitoring standardizationAlign teams on reliability practicesInfluence upstream architectural decisionsOperational ExcellenceEstablish KPIs and drive improvement in:Incident response and resolution timesAlert quality and paging effectivenessMonitoring coverage across critical servicesProvide leadership with clear visibility into service health and reliability trendsRequired Qualifications7+ years in Site Reliability Engineering, monitoring, or production engineeringProven experience in a technical leadership or lead engineer roleDeep hands-on experience with:Dynatrace (or equivalent observability platforms)Microsoft Azure (IaaS, PaaS, networking, identity)ServiceNow ITSM / ITOM (incident, event management, CMDB)Demonstrated ability to:Design and lead enterprise monitoring/SRE architecturesDrive platform and tooling decisionsIntegrate observability, ITSM, and paging solutionsPreferred QualificationsExperience leading SRE or observability transformation initiativesStrong expertise with Dynatrace–ServiceNow integrationsExperience modernizing or consolidating paging/on-call toolingFamiliarity with:Azure-based SRE tooling or AI-assisted operationsAutomation frameworks (GitHub Actions, Runbooks, etc.)Infrastructure as Code (Terraform, ARM, Bicep)Success MetricsReduction in alert noise and unnecessary pagingImproved incident routing accuracy and MTTRIncreased adoption of self-healing and automated workflowsStrong alignment between monitoring, CMDB, and service ownershipEnterprise-wide adoption of SRE and monitoring standardsGeneral Description of Available Benefits for Eligible Employees of CRC Group: At CRC Group, we're committed to supporting every aspect of teammates' well-being – physical, emotional, financial, social, and professional. Our best-in-class benefits program is designed to care for the whole you, offering a wide range of coverage and support. Eligible full-time teammates enjoy access to medical, dental, vision, life, disability, and AD&D insurance; tax-advantaged savings accounts; and a 401(k) plan with company match. CRC Group also offers generous paid time off programs, including company holidays, vacation and sick days, new parent leave, and more. Eligible positions may also qualify for restricted stock units and/or a deferred compensation plan.CRC Group supports a diverse workforce and is an Equal Opportunity Employer that does not discriminate against individuals on the basis of race, gender, color, religion, citizenship or national origin, age, sexual orientation, gender identity, disability, veteran status or other classification protected by law. CRC Group is a Drug Free Workplace.EEO is the LawPay Transparency Nondiscrimination ProvisionE-VerifySummaryLocation: Charlotte NC - 600 S Tryon St.Type: Full time
- ...home!Where you’ll be:This position will be based at our Corporate Headquarters located in Charlotte, NC.About the Role:The Site Reliability Engineer plays a critical role in designing, building, and maintaining scalable, secure, and highly available cloud infrastructure...SuggestedFull timeFlexible hours
- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...SuggestedFull time
$152.6k - $191.5k
...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing... ...improvement.Position Summary:The Senior GCP Site Reliability Engineer acts as an advanced senior individual...SuggestedFull timeWork at officeDay shift$50 - $57 per hour
...Infrastructure Engineer 3 – Contingent Client: Financial Services Role: Site Reliability Engineer (SRE) / Production Engineer Location: Zone 2 Work Arrangement: Hybrid Contract Length: 18mo Bill Rate: $50 - $57/hr (OT40) Conversion Opportunity: Yes...SuggestedContract work- ...We are seeking a Senior SRE / DevSecOps Engineer with strong experience in Kubernetes, AWS... .... The role will focus on platform reliability, incident management, SLO/SLI governance... ...Experience with SLO/SLI governance and site reliability practices. ~ Strong understanding...Suggested
- ...Site Reliability Engineer (SRE) – Hiring Fast! Industry: Finance Location: Charlotte, NC Pay Rate: $54-62/HR on W2 Only – NO C2C Setting: Hybrid Required (Remote is NOT an Option) Duration: 18+ months Required Qualifications: ~5+ years of experience...Remote work
- ...Our Financial Services client is seeking a Site Reliability Engineer to join their team. Title: Site Reliability Engineer Location: Charlotte, NC or Chandler, AZ Length: 1 year Pay Rate: 70-75 /hr. w2 Must Have Required Qualifications: : ~3+ years of experience...Weekend work
- ...Job Title: Senior Site Reliability Engineer Duration: 18 months (possibility to extend or convert to FTE) Location: Charlotte, NC - Hybrid Role (3 days onsite in a week) Interview process: 2 rounds #1-hour virtual panel #1 hour on site technical panel...Shift work3 days per week
- ...development of the system. • Collaborates with system analysts, engineers, and programmers to design systems and to determine project... ...Skills: • 3 or more years of experience as DevOps or Site Reliability Engineer • Experience in DevOps/Development • Experience with...
$60 - $65 per hour
...retail industries. Rate Range: $60-$65/Hr Job Description: The Client Document Generation team is seeking a Senior Software Engineer ( IT Onshore Band 4) to participate in the full system development lifecycle (SDLC) of enterprise applications that support high-...Immediate start$84.24k - $142.48k
OverviewJoin us to work collaboratively with our talented team of dynamic and passionate engineers to deliver capabilities that enable our customers to make a difference. You'll deploy and operate ArcGIS Velocity and ArcGIS Workflow Manager SaaS solutions. You will also...WorldwideFlexible hours$55 - $62 per hour
...CREATED BY AI, REVIEW BEFORE POSTING Position: Infrastructure Engineer 3 - Contingent Location: Charlotte, North Carolina... ...automation, and monitoring capabilities. You will bring strong Site Reliability Engineering (SRE) and production support experience while driving...Full timeContract work- ...Site Reliability Engineer Job Location: San Francisco, CA or Charlotte, NC. Job Type: Contract Work with local API development squads, platform teams, product owners, scrum masters, and architects. The SRE ensures that both our internally critical and our externally...Contract workLocal area
- ...Site Reliability Engineer III Our client, a leading organization in the technology and infrastructure sector, is seeking a Site Reliability Engineer III to join their team. As a Site Reliability Engineer III, you will be part of the Infrastructure Support team supporting...Flexible hours
- ...converted to LINK tokens and stored in a strategic [Chainlink Reserve]( Learn more at [chain.link]( About the Role As a Senior Site Reliability Engineer on the CCIP Platform team, you will ensure the reliability, scalability, and operational excellence of the systems...Full timeRemote work
- ...Details Job Description: Mandatory Skills: Observability engineering (metrics/logs/traces, tooling like. Grafana/Splunk/OTel/... ...monitoring). SRE principles (incident response, RCA, automation, reliability). Hybrid/SaaS/on-prem integration monitoring. Vendor-...
- ...English (Required) Work Shift: 1st shift (United States of America) Please review the following job description: The Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- ...PFB the JD must have skills Hands on SRE Engineer with good analytical skills Good Exposure to both incident and Problem Management Must have worked in AWS Must have good knowledge on Middleware components/services (Servers,Load Balancer/Trace Logs...Permanent employmentFull time
- Company DescriptionProSidian is looking for “Great People Who Lead” at all levels in the organization. Are you a talented professional ready to deliver real value to clients in a fast-paced, challenging environment? ProSidian Consulting is looking for professionals who ...For contractorsWork at office
$100k - $160k
..., high-value outcomes every day. Our AI Voice & Telephony Engineering Team builds the systems that power intelligent conversations... ...response flows. Improve system performance, latency, and reliability across a distributed, event-driven stack. Provide technical...Full timeTemporary workFlexible hours- Job Description : Set up CI/CD pipelines and Helm charts for deployments. Implement GitOps for declarative infrastructure. Integrate observability tools for logs, metrics, and traces. Skills: Infrastructure: OpenShift/Kubernetes, Helm/GitOps, Secure pipelines...
$91.2k - $136.8k
Reliability Engineer - IE08GEWe’re determined to make a difference and are proud to be an insurance company that goes well beyond coverages and... ...field.3+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), or DevOps.Hands-on experience...Full timeTemporary workWork at office3 days per week$60 - $87 per hour
...Description Non-negotiables: - Must be a true Release Train Engineer, with experience leading 10 Agile delivery teams. - Must have expert-level experience with JIRA/JQL - Must have experience in PI Planning execution, system demos, inspect & adapt, ART syncs, ART readiness...Contract workTemporary work$110k - $125k
...Release Train Engineer Must Have Technical/Functional Skills • 10+ years of professional experience and Minimum 5+ years of experience as a RTE with RTE certification. • Facilitate Agile Release Train (ART) events and processes. • Lead Program Increment (PI)...- ...Job Title : Release Train Engineer Location: Charlotte, NC (Onsite) Job Description: Wells Fargo is seeking an experienced Release Train Engineer (RTE) to lead scaled Agile planning, execution, and continuous improvement initiatives...
- ...Job Description Insight Global is seeking a Sr Release Engineer for a large insurance service provider. This individual will be responsible for product development and all SREs to define all the steps needed to release software. This team has a new azure application...Remote work
$70 - $85 per hour
...Our customer has several openings for Sr. RTEs for a hybrid, on-site effort here in Charlotte. This is a multi-year engagement with the... ...*Top Skills Required:* * 3+ years of pointed Release Train Engineer tenure * Deep experience supporting a Scaled Agile Framework (...Contract workTemporary work$152.8k - $229.2k
Principal Reliability Engineering - IE06JEWe’re determined to make a difference and are proud to be an insurance company that goes well beyond coverages... ...of the following areas: data, cloud, platform engineering, site/reliability engineering, or large‑scale distributed systems,...Full timeTemporary workWork at office3 days per week$114.08k - $218.03k
...recommendations to business leaders on best solutions.Independently experiments with new patterns and technologies.Helps establish and improve engineering best practices, concepts, and patterns with peers and the business.Understands the customer and proactively identifies innovative...Full timeH1bWork at officeRemote workHome officeRelocation packageFlexible hours- ...Reliability Engineer The Reliability Engineer is responsible for overall evaluation and improvement of product reliability, through collecting and analyzing failure data, recommending design and or process improvement, support MTBF modeling activities, interfacing...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer Charlotte, NC
- site reliability engineer sre Charlotte, NC
- junior website developer Charlotte, NC
- website content developer Charlotte, NC
- on site coordinator Charlotte, NC
- website coordinator Charlotte, NC
- site leader Charlotte, NC
- site recruiter Charlotte, NC
- historic site Charlotte, NC
- remote website tester Charlotte, NC



