SRE Engineer
United IT Solutions
Role: SRE Engineer
Location: Atlanta, GA (Hybrid)
Qualifications:
• Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.
• Hands-on experience with incident management and 24/7 production support models.
• Proficiency with monitoring and observability tools such as CloudWatch, Dynatrace, and Quantum Metric.
• Experience building and maintaining monitoring dashboards.
• Strong troubleshooting skills across infrastructure, networking, and application layers.
• Working knowledge of CI/CD pipelines and AWS deployment processes.
• Experience working with databases and Unix/Linux environments.
Key Responsibilities
Incident Management and Production Support
• Provide Level 1 and Level 2 support for production incidents across AWS-hosted applications and infrastructure.
• Triage incidents by identifying root causes, distinguishing infrastructure issues from application defects, and restoring service within defined SLAs.
• Escalate code-level defects to development teams with clear diagnostics, supporting logs, and impact assessments.
• Participate in on-call rotations, major incident bridges, and post-incident reviews.
• Investigate application defects, configuration issues, and infrastructure anomalies reported through monitoring tools or user incidents.
Monitoring and Operational Health
• Perform regular health checks across applications, infrastructure, and AWS services.
• Monitor system health using CloudWatch, Dynatrace, Quantum Metric, and ThousandEyes.
• Respond proactively to alerts related to resource utilization, latency, errors, and availability.
• Maintain and improve monitoring and observability dashboards.
Location: Atlanta, GA (Hybrid)
Qualifications:
• Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.
• Hands-on experience with incident management and 24/7 production support models.
• Proficiency with monitoring and observability tools such as CloudWatch, Dynatrace, and Quantum Metric.
• Experience building and maintaining monitoring dashboards.
• Strong troubleshooting skills across infrastructure, networking, and application layers.
• Working knowledge of CI/CD pipelines and AWS deployment processes.
• Experience working with databases and Unix/Linux environments.
Key Responsibilities
Incident Management and Production Support
• Provide Level 1 and Level 2 support for production incidents across AWS-hosted applications and infrastructure.
• Triage incidents by identifying root causes, distinguishing infrastructure issues from application defects, and restoring service within defined SLAs.
• Escalate code-level defects to development teams with clear diagnostics, supporting logs, and impact assessments.
• Participate in on-call rotations, major incident bridges, and post-incident reviews.
• Investigate application defects, configuration issues, and infrastructure anomalies reported through monitoring tools or user incidents.
Monitoring and Operational Health
• Perform regular health checks across applications, infrastructure, and AWS services.
• Monitor system health using CloudWatch, Dynatrace, Quantum Metric, and ThousandEyes.
• Respond proactively to alerts related to resource utilization, latency, errors, and availability.
• Maintain and improve monitoring and observability dashboards.
Vacancy posted 17 hours ago
Similar jobs that could be interesting for youBased on the SRE Engineer in Atlanta, GA vacancy
- #CareersJC 1483593Qualifications· Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.· Hands-on experience with incident management and 24/7 production support models.· Proficiency with monitoring and observability...Suggested
- ...retain us as their trusted partner. If you're ready to make an impact, you're in the right place. Job Details: Job Title: SRE Engineer Location: Atlanta GA (Hybrid Duration: 1 year Contract Job Summary Mode: A hybrid work schedule will be followed where...SuggestedContract workWork experience placementWork at officeRemote work
- Interview: Face-to-face required Qualifications Strong experience supporting production systems hosted on AWS (EC2, VPC, ALB/NLB, RDS, Lambda, EKS) Hands-on experience with incident management and 24/7 production support models Proficiency with...Suggested
- OneTrust is seeking a Senior Software Engineer in Atlanta, Georgia. The role involves designing and maintaining a reliable application platform, collaborating with engineering teams, and enhancing customer experiences through observability tools. The ideal candidate will...Suggested
- OneTrust is seeking a Senior Software Engineer - SRE in Atlanta to join the Software Engineering team. You will design, implement, and maintain a highly available platform, focusing on observability, automated incident response, and reliable deployments. You will collaborate...Suggested
$35 - $45 per hour
DescriptionKforce has a client that is seeking a remote Site Reliability Engineer to join their team.Summary:The team consists of systems that can... ...but underpinned by a lot of Java/API's hosted in GCP. This SRE will be supporting a full stack React and Java application that...Remote work- ...serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer Unified Observability...Full timeWorldwideFlexible hours
$152.13k - $162.13k
...challenge the status-quo.Unum is changing, and we’re excited about what’s next. Join us.General Summary:Unum Group seeks Site Reliability Engineers in Atlanta, GA.Applicants who are interested in this position may apply at (Ref #66753) for consideration.Design, build, and...Full timeTemporary workWork at officeRemote work$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability... ...impact on our platform.Partners with the larger Cloud Operations, SRE, Engineering teams, and the business-at-large to advance our...Full timeTemporary workWork experience placementFlexible hours$104.9k - $174.7k
...Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory role you will...Full timeWork at officeLocal areaRemote workWork from home- T-Mobile USA, Inc. seeks a Site Reliability Engineer lead to design, build, and operate secure, scalable software platforms using SRE and AI-native practices across cloud-native and distributed systems. You will mentor the SRE team, drive Terraform-based infrastructure,...
- T-Mobile USA, Inc. is seeking an experienced Site Reliability Engineer to lead reliability across cloud platforms and enterprise applications... ..., architecture standardization, and mentorship within a global SRE team. Competitive compensation and a broad benefits program...
$138.1k - $198.2k
...customers, and businesses. We’re making networking easier, faster, and more intuitive with technology that simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments, tools, automation and processes that facilitate developers’...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours- ...Site Reliability Engineering LeadThe Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence... ...role involves standardizing observability practices, mentoring SRE team members, and contributing to enterprise-wide reliability...
- ...Senior Systems Reliability Engineer (SRE)At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees...Full timeTemporary workPart timeWork experience placementFlexible hours
- ...Our employees work at the cutting edge of AI cloud infrastructure alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and listed on Nasdaq, Nebius has a global footprint with R&D hubs across...
$120k - $175k
...Salary: $120,000 - 175,000 per year Requirements: We need at least 5 years of experience as a reliability-focused engineer in a fast-moving, rapidly expanding enterprise setting. We value strong hands-on knowledge of cloud platforms such as AWS, Azure, or Google...Full timeRemote workVisa sponsorshipFlexible hours- ...and highly available software platforms using Site Reliability Engineering and AI-native engineering practices. The engineer collaborates... ...efficiency. This position provides technical leadership for the SRE team, strengthens the India GCC capability, and establishes reusable...Full timeTemporary workPart timeWork experience placementLocal areaFlexible hours
$178.13k - $205.4k
...Salary Range: $178,131 - $205,400 About You Basic Qualification Bachelor's degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5) years of progressive, post-baccalaureate experience in job offered...Work at officeRemote workFlexible hours- ...Description: Qualifications: This position is 60 % SRE and 40% SDE. Also open for candidates to join MTH along with ATL... ...to configure the monitoring and alerting metrics so the support engineers can proactively and timely validate, troubleshoot and resolve...Work experience placement
$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long... ...achieve the automation AWS Pipeline & Infrastructure, DevOps (SRE) activities, Monitoring & Alerting Our client is a leading...Contract workLocal areaImmediate start- ...Job description Snowflake SRE JD Your Role Accountabilities Primarily responsible for administrating Snowflake environments on AWS Identify, tune, and fix the performance issues on priority. Diagnose and troubleshoot Snowflake related errors and work with team to raise...
$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...Temporary workImmediate startFlexible hoursShift work$75.7k - $136.3k
...systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and... ...that support Akamai Cloud's products and services. Our SRE teams solve reliability, security, and usability at scale for our...Work experience placementWork at office- ...Technical Support Specialist In Site Reliability Engineering (Sre) Mandatory skills: Scripting and programming languages like Python, Java, Ruby. Cloud and infrastructure management – AWS, Google cloud and Azure is a plus- CI/CD Automation, Database Management. The...
- ...Site Reliability Engineer (SRE) Location: Atlanta, GA 30303 Duration of the project: 12 Months Strong expertise in Ansible with an SRE background Ability to review and test GitLab Duo generated code CICD pipelines (GitLab, GitHub Actions) Infrastructure automation...
- ...I have an opportunity for "SRE " _ (Atlanta, GA - ONSITE)" and I am looking for a candidate who can join Immediately if you are... ...Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in...Immediate start
- ...Senior Site Reliability Engineer Atlanta, Georgia Who We Are QGenda is redefining healthcare workforce management everywhere care... ...best practices. Actively contribute to fostering an SRE culture within the organization by promoting observability, retrospectives...Permanent employmentFull timeWork at officeRemote workWork from homeWork visa
$167.7k - $245.2k
...supports these customers and their networks. As a Site Reliability Engineer, you will be focused on supporting a specific, highly available,... .... * Keywords: Network Engineering, Production Engineering, SRE, Site Reliability Engineering, DevOps, CCNP, CCIE, JNCP, JNCIE Why...Permanent employmentFull timeTemporary workLocal areaFlexible hours- ...Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and operational... ...role involves standardizing observability practices, mentoring SRE team members, and contributing to enterprise-wide reliability...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE Engineer. Be the first to apply!
Related searches
- site reliability engineer remote Atlanta, GA
- site reliability engineer sre Atlanta, GA
- site reliability engineer Atlanta, GA
- site reliability engineer remote
- site reliability engineer sre
- site reliability engineering manager
- site reliability engineer
- lead site reliability engineer
- junior site reliability engineer


