Site Reliability Engineer (SRE)
Staffing Spot, Inc.
Key Responsibilities
Design, build, and operate highly available and scalable production systems.
Define and implement SRE practices, standards, and reliability engineering strategies .
Establish and manage SLIs, SLOs, SLAs, and error budgets for critical services.
Own production reliability, availability, performance, and capacity of business-critical applications.
Develop and maintain highly automated infrastructure and operational processes.
Build and manage CI/CD pipelines for reliable and repeatable software delivery.
Design, deploy, and manage containerized workloads using Docker and Kubernetes .
Implement Infrastructure as Code using Terraform, CloudFormation, or similar technologies .
Develop automation using Python, Go, Bash, or similar scripting/programming languages .
Implement comprehensive monitoring, logging, alerting, and observability solutions.
Troubleshoot complex production issues across applications, infrastructure, networking, databases, and cloud services.
Lead incident response, root-cause analysis, and post-incident reviews .
Identify recurring operational problems and eliminate them through automation and engineering solutions.
Perform capacity planning, performance optimization, and scalability assessments.
Implement disaster recovery, backup, business continuity, and high-availability strategies.
Improve deployment reliability through automation, progressive delivery, rollback strategies, and release engineering.
Partner with development teams to improve application reliability, resilience, and observability .
Participate in on-call rotations and provide technical leadership during critical production incidents.
Establish operational runbooks, documentation, and troubleshooting procedures.
Mentor engineers and drive adoption of SRE and DevOps best practices.
Evaluate new technologies and identify opportunities to improve engineering productivity and platform reliability.
Required Qualifications
10+ years of experience in SRE, DevOps, Cloud Engineering, Infrastructure Engineering, or Software Engineering.
Strong experience managing large-scale production environments .
Strong programming/scripting experience with Python, Go, Java, Bash, or similar languages .
Hands-on expertise with AWS, Azure, or Google Cloud Platform (GCP) .
Strong experience with Kubernetes and Docker .
Extensive experience with Infrastructure as Code , particularly Terraform.
Strong understanding of Linux/Unix systems administration .
Experience designing and managing CI/CD pipelines .
Strong understanding of networking concepts including TCP/IP, DNS, TLS, load balancing, and networking in cloud environments.
Experience with microservices and distributed systems .
Strong knowledge of monitoring and observability concepts.
Experience with tools such as Prometheus, Grafana, ELK/Elastic Stack, Datadog, Splunk, or similar platforms .
Strong experience with incident management, troubleshooting, RCA, and problem management.
Understanding of SLOs, SLIs, SLAs, error budgets, and reliability metrics .
Experience with production performance and capacity management.
Strong understanding of security, access management, secrets management, and cloud security best practices.
Excellent problem-solving, communication, and technical leadership skills.
Experience designing highly available distributed systems at scale.
Experience with AWS services such as EC2, EKS, S3,
- ...and make an impact. Join us! Position Summary: The IKCP Site Reliability Engineer Lead is responsible for ensuring the reliability,... ...8+ years of infrastructure, cloud, platform engineering, or SRE experience. ~5+ years managing Kubernetes and/or OpenShift...SuggestedWork at officeFlexible hoursShift workDay shift
- ...Job Summary We are seeking a Senior SRE / DevSecOps Engineer with strong experience in Kubernetes,... ...troubleshooting. The role will focus on platform reliability, incident management, SLO/SLI... ...Experience with SLO/SLI governance and site reliability practices. ~ Strong...Suggested
$83.52k - $125.28k
...are currently seeking a Lead ML Platform Engineer (SRE / FTE / Onsite) to join our team in... ...operate predictive models efficiently and reliably. The successful candidate will lead... ...hire locally to NTT DATA offices or client sites. This ensures we can provide timely and...SuggestedTemporary workWork at officeRemote workFlexible hours$152.6k - $191.5k
...for partnering with leaders across engineering and technology to define objective reliability goals for services. Key... ...improvement.Position Summary: The Senior Site Reliability Engineer acts as an advanced... ...they supportChampions modern SRE, observability, and AIOps...SuggestedFull timeWork at officeDay shift- ...America)Please review the following job description:Lead Site Reliability & Environment Monitoring Engineer (Azure / Dynatrace / ServiceNow)We are seeking a Lead... ..., and guiding engineering teams toward modern SRE practices.This individual will act as the technical authority...SuggestedFull timeTemporary workShift workDay shift
- ...of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence... ...involves standardizing observability practices, mentoring SRE team members, and contributing to enterprise-wide...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- ...home!Where you’ll be:This position will be based at our Corporate Headquarters located in Charlotte, NC.About the Role:The Site Reliability Engineer plays a critical role in designing, building, and maintaining scalable, secure, and highly available cloud infrastructure...Full timeFlexible hours
- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...Full time
$136.75k - $218.8k
...“I can succeed as an AI Platform Engineer at Capital Group.” As an AI Platform... ..., agent execution, and service reliability. Implement logging, tracing, alerting... ...operational excellence: Apply Site Reliability Engineering (SRE) practices to ensure reliability, scalability...Temporary workLocal areaFlexible hours- ...Role Overview: We are seeking a highly skilled Senior Java Backend Engineer to join our Digital Channels API team supporting client profile... ...(Lambda, ECS, RDS, S3, SQS). Collaborate closely with SRE and platform teams to ensure resiliency and observability. Contribute...
- ...specialist Practices; Software Engineering, Data & Analytics, IT... ...Software Delivery Monitoring, Reliability & Production Operations Security... ...technologies Familiarity with SRE concepts such as SLIs, SLOs,... ..., cloud engineering, site reliability engineering, infrastructure...Work at officeRelocationVisa sponsorship3 days per week
- ...Redis, and vector databases, implementing model evaluation, prompt engineering, state management, caching, and high-throughput inference... ...production ML/AI workloads using observability, logging, monitoring, SRE practices, performance tuning, resiliency, disaster recovery,...Temporary work
- ...Job Description Insight Global is seeking a Sr Release Engineer for a large insurance service provider. This individual will be responsible for product development and all SREs to define all the steps needed to release software. This team has a new azure application...Remote work
- Occupation: Computer Systems Engineers/Architects Location: Charlotte... ...: Gravity IT Resources Web Site: gravityitresources.com Onsite... ...infrastructure while ensuring the reliability, availability, and... ...Site Reliability Engineering (SRE) best practices, including monitoring...Full timeRemote workWork from home
$91k - $136.24k
...Release Train Engineer II Work Location: Charlotte, North Carolina, United States of America Hours: 40 Pay Details: $91,000 - $136,240 USD TD is committed to providing fair and equitable compensation opportunities to all colleagues. Growth opportunities and...Immediate start$120k - $200k
...Domain Consultant Senior SAFe-certified SAP S/4HANA transformation delivery leader with proven experience as an Agile Release Train Engineer or large-scale ART execution lead. Direct experience running PI Planning, ART Sync, Scrum Master orchestration, dependency...$110k - $125k
...Release Train Engineer Must Have Technical/Functional Skills 10+ years of professional experience and minimum 5+ years of experience as a Release Train Engineer with Release Train Engineer certification. Facilitate Agile Release Train (ART) events and processes...- ...skilled Senior Platform / DevOps Engineer to design, build, automate, and support... ...teams to improve application reliability, scalability, and deployment efficiency... ...in Platform Engineering, DevOps, Site Reliability Engineering (SRE), or Infrastructure Engineering....
$136.75k - $218.8k
...Senior Infrastructure & Application Resiliency Engineer at Capital Group.” On our Cloud Reliability Engineering team, you’ll architect, build,... ...~ You have at least 10 years in Site Reliability Engineering (SRE), DevOps, or Cloud Architecture roles, including...Full timeTemporary workLocal areaFlexible hours- ...Distributed Systems Software Engineer, Python / Go Join to apply for the Distributed Systems Software Engineer, Python / Go role... ...automated testing approaches and infrastructure for validating reliability, performance, and resilience of cloud orchestration tools and...Full timeLocal areaRemote workWorldwide
- ...opportunities, from hands-on desk support to Cybersecurity, Cloud Engineering, AI, and Modern Application development. We are committed to... ...Retirement Plan Paid Time Off Holiday Time Off (varies by site/state) Associate Shopping Program Health and Wellness...Work at officeLocal areaRemote workFlexible hours
- ...us! LOB Description: The FinOps Engineer will support cloud financial management... ...to implement measures prescribed by the Site Reliability Engineer teams it leads. Key responsibilities... ...the Senior Site Reliability Engineer (SRE) Develops and maintains reliability...Work at officeFlexible hoursShift workDay shift
- ...highly accomplished Agentic AI Platform Engineering & Performance Distinguished level... ...agent runtime platforms to deliver reliable, responsive, and cost-effective AI... ..., orchestration, load testing, and site reliability engineering (SRE). The ideal candidate brings strong...Work at officeFlexible hoursShift workDay shift
$65.05 per hour
...Description Job Title: Mobile Application Release/Deployment Engineer Location: Charlotte, NC Duration: Contract - 12... ...automation and release management tooling. Knowledge of DevOps and Site Reliability Engineering practices. Familiarity with Bank of America...Contract workRelocation3 days per week$91.2k - $136.8k
Reliability Engineer - IE08GEWe’re determined to make a difference and are proud to be an insurance company that goes well beyond... ...+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), or DevOps.Hands-on experience with observability...Full timeTemporary workWork at office3 days per week- ...candidate thrives in complex environments, partners closely with engineering, SRE, and platform teams, and brings a forward-looking mindset to... ...automation while maintaining enterprise-grade security, reliability, and governance.Required Qualifications:Strong experience in...
$110k - $125k
...Functional Skills • 10+ years experience in delivering large scale applications with focus on performance, scalability, security, and reliability. • Experience with WebLogic or JBoss or Mule or tomcat application servers. • Experience in monitoring, triaging and...- ...Platform Engineer 20651 Charlotte, NC 9/16/2025 3:45:00 PM Platform Development FTE - IntraEdge Job Description Role summary: Define and execute DevOps strategy, CI/CD at scale, IaC, observability, and security automation. Key responsibilities...
$110k - $130k
...Java Full Stack Engineer Java Full Stack Engineer with API integration and Camunda Implementation Knowledge is must Backend development (Java / Node.js / APIs) Frontend development (React or equivalent) API development (REST services) Integration adapter development...$54 - $64.62 per hour
...observability platforms, distributed tracing, logging frameworks, and SRE practices. - Experience with event-streaming platforms such... ...domain experience. - Proven track record of driving engineering excellence, technology innovation, and platform standardization...Hourly payContract workTemporary workWork experience placement
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer (SRE). Be the first to apply!
- site reliability engineer sre Charlotte, NC
- site reliability engineer Charlotte, NC
- site recruiter Charlotte, NC
- site services specialist Charlotte, NC
- junior website developer Charlotte, NC
- remote website tester Charlotte, NC
- official site Charlotte, NC
- on site coordinator Charlotte, NC
- site leader Charlotte, NC
- historic site Charlotte, NC



