Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

T Mobile US

Senior Systems Reliability Engineer (SRE)At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees enjoy multiple wealth-building opportunities through our annual stock grant, employee stock purchase plan, 401(k), and access to free, year-round money coaches. That's how we're UNSTOPPABLE for our employees!Are you ready for the next chapter in your Uncarrier journey? The Sr. System Reliability Engineer (SRE) guides and mentors other SREs and improves and protects the software and systems behind all of T-Mobile's IT services, including management of scalability, availability, latency, performance, security, and capacity, and delivering software faster, better, and cheaper. From designing & maintaining CICD Pipelines to building the next generation of T-Mobile applications on cloud native platforms, the SRE's enable great customer experience and product innovation by continuous improvement of operational support. We pride ourselves on encouraging a culture of innovation, advocating for agile methodologies, and promoting transparency in all that we do. Join us in embodying the spirit of the 'Un-carrier' and make a tangible impact! The Compliance Campaign Management (CCM) team within T-Mobile's Identity Security and Access Management (ISAM) organization is leading a major transformation in logical access compliance. What was once a predominantly manual User Access Review process is rapidly evolving into a sophisticated, automation-driven platform powered by cloud technologies, APIs, workflow orchestration, and artificial intelligence. Our team designs and builds solutions using Power Platform, Azure DevOps Pipelines, Microsoft Graph APIs, and AI-powered automation to streamline compliance operations while expanding risk visibility across the enterprise. Every innovation we deliver helps reduce manual effort, strengthen control effectiveness, and improve audit readiness. As a Senior Systems Reliability Engineer (SRE), you will play a critical leadership role in shaping and operating the platform that powers this transformation. CCM is building an environment where human expertise and AI agents work side by side to execute compliance operations with greater speed, consistency, and precision. You will own the reliability, scalability, observability, and continuous improvement of this compliance automation ecosystem, ensuring it remains highly available, performant, secure, and audit ready as it grows across the enterprise. This is more than a traditional operations role. You'll drive engineering excellence across the platform, establish reliability standards, champion automation-first practices, and lead efforts to improve resilience, monitoring, incident response, and operational maturity. Working closely with software engineers, compliance experts, security teams, and AI-enabled solutions, you'll help define the next generation of compliance operations and governance technology. Join a team that's not just supporting compliance but engineering the future of access governance. You'll have the opportunity to solve complex enterprise-scale challenges, influence strategic technology decisions, mentor engineers, and help build intelligent automation capabilities that redefine how access risk is identified, reviewed, and remediated at T-Mobile. If you're energized by the intersection of cloud engineering, automation, AI, cybersecurity, and operational excellence, this is your chance to make a measurable impact on one of the company's most innovative compliance modernization initiatives.Why This Role Is UniqueLead the reliability strategy for a rapidly growing compliance automation platform.Partner with engineers, compliance professionals, and AI solutions to modernize access governance.Build and enhance observability, resilience, and automation across cloud-native services and integrations.Influence the operational foundations of an environment where AI agents and human operators collaborate to execute critical compliance functions.Drive enterprise-scale outcomes that improve security, reduce risk, and strengthen audit readiness across T-Mobile.Job Responsibilities :Utilize DevOps automation tools to enhance continuous integration and continuous delivery pipelines for non-production environmentsManage environments through automated server provisioning and pipeline configuration to support virtual machinesDeliver software improvements that increase availability, scalability, latency, and efficiency of IT servicesCreate and maintain dashboards for continuous monitoring and health checks to improve service quality in non-production environmentsContribute to software delivery process improvements including cloud enablement and microservices containerizationMentor and guide systems reliability engineers and vendor resources to support team development and performanceArchitects scalable, modular automation frameworks that allow compliance workflows to expand across new control types, compliance frameworks, and operational domains without re-engineering — maintaining consistent, audit-defensible output quality at increasing throughput.Designs automated validation and accuracy verification patterns across agent and workflow pipelines, ensuring repeatable, measurable output quality that supports continuous improvement, operational agility, and audit readiness at scale.Education and Work Experience :Bachelor's Degree plus 5 years of related work experience OR Advanced degree with 3 years of related experience (Required)4-7+ years relevant experience. (Required)Experience working in an Agile and DevOps environment. (Required)Experience in one or more of: C, C#, Java, Perl, Python, Go, or scripting experience in Shell and Perl. (Required)Experience in Continuous Integration/Continuous Delivery tools, such as, Jenkins, Cloudbees, etc., and other automation tools. (Required)Experience with DevOps tools, such as, Ansible, Chef, Puppet, etc. Experience in Docker, Kubernetes, etc. is preferable. (Preferred)Experience in APM tools like AppDynamics, logging tools, like Splunk. (Required)Experience working in a cloud environment (public/private). (Required)Experience in migrating to cloud or cloud native environments is preferable. (Preferred)Required Knowledge, Skills and Abilities :Python or comparable programming language experienceRed Hat Ansible or comparable orchestration technologiesAutomation Architecture and Scalable Framework Design experiencePlatform Engineering and Reusable Pipeline DesignPreferred Knowledge, Skills and Abilities :Agentic workflow design, orchestration, and operational governanceAutomated testing and output validation framework designHuman-in-the-loop system architecture for regulated workflowsAPI-first integration architecture for enterprise-scale automation platformsCompliance or regulatory workflow automation in identity, access, or financial controls environmentsAt least 18 years of ageLegally authorized to work in the United StatesTravel:Travel Required (Yes/No): YesDOT Regulated Position (Yes/No): NoSafety Sensitive Position (Yes/No): NoBase Pay Range: $98,500 - $177,700Corporate Bonus Target: 15%The pay range above is the general base pay range for a successful candidate in the role. The successful candidate's actual pay will be based on various factors, such as work location, qualifications, and experience, so the actual starting pay will vary within this range.At T-Mobile, employees in regular, non-temporary roles are eligible for an annual bonus or periodic sales incentive or bonus, based on their role. Most Corporate employees are eligible for a year-end bonus based on company and/or individual performance and which is set at a percentage of the employee's eligible earnings in the prior year. Certain positions in Customer Care are eligible for monthly bonuses based on individual and/or team performance. To find the pay range for this role based on hiring location, T-Mobile, our benefits exemplify the spirit of One Team, Together! A big part of how we care for one another is working to ensure our benefits evolve to meet the needs of our team members. Full and part-time employees have access to the same benefits when eligible. We cover all of the bases, offering medical, dental and vision insurance, a flexible spending account, 401(k), employee stock grants, employee stock purchase plan, paid time off and up to 12 paid holidays - which total about 4 weeks for new full-time employees and about 2.5 weeks for new part-time employees annually - paid parental and family leave, family building benefits, back-up care, enhanced family support, childcare subsidy, tuition assistance, college coaching, short- and long-term disability, voluntary AD&D coverage, voluntary accident coverage, voluntary life insurance, voluntary disability insurance,

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Atlanta, GA vacancy
  •  ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises... 
    Suggested
    Permanent employment
    Full time
    Part time
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Monday to Friday
    Shift work
    Day shift

    Truist

    Atlanta, GA
    4 days ago
  • $35 - $44 per hour

    DescriptionKforce has a client seeking a remote Site Reliability Engineer to join their team. We are seeking a Site Reliability Engineer (SRE) to support a large-scale system modernization and legacy platform retirement initiative. This role will focus on maintaining and... 
    Suggested
    Remote work

    KForce

    Atlanta, GA
    3 days ago
  • $98k - $148.5k

     ...Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization.As a Site Reliability Engineer I on the Core Infrastructure team in our Atlanta office,you'll help build and operate the foundational infrastructure that... 
    Suggested
    Work at office
    Local area
    Flexible hours

    PagerDuty

    Atlanta, GA
    4 days ago
  • Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence... 
    Suggested
    Worldwide

    Inspire Brands

    Atlanta, GA
    4 days ago
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying... 
    Suggested
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Atlanta, GA
    3 days ago
  •  ...Senior Systems Reliability Engineer (SRE)At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees... 
    Full time
    Temporary work
    Part time
    Work experience placement
    Flexible hours

    T Mobile US

    Atlanta, GA
    5 days ago
  •  ...client accounts, mentor staff, and ensure project success while upholding firm standards. Leverage data, algorithms, and software engineering to deliver scalable security solutions and drive business growth. You will guide teams, develop client-ready strategies, and... 

    PwC

    Atlanta, GA
    1 day ago
  • CBRE Group, Inc. seeks an Operations Consultant to provide technical expertise across complex technology systems and support daily operations within the D&T Support function. You will collaborate with IT teams, optimize configurations, and guide stakeholders on technology...

    CBRE Group, Inc.

    Atlanta, GA
    3 days ago
  •  ...resilient software platforms using SRE and AI-native engineering practices. Own production reliability, monitoring, and operational automation while mentoring...  ...), Kubernetes, and production operations. Key Skills Site Reliability Engineering Terraform Python AWS Azure... 
    Temporary work
    Flexible hours

    T-Mobile

    Atlanta, GA
    5 days ago
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability... 

    2T Consulting

    Atlanta, GA
    2 days ago
  • $70 - $85 per hour

     ...redefine what’s possible, give shape to the future—and get there.What You’ll Do* Define and establish enterprise reliability standards, Site Reliability Engineering (SRE) practices, SLIs/SLOs, operational governance models, and engineering guardrails that enable scalable,... 
    Temporary work
    Local area
    Flexible hours
    3 days per week

    Slalom

    Atlanta, GA
    4 days ago
  •  ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview:   We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it... 
    Full time
    Live in
    Work at office

    Incident IQ

    Atlanta, GA
    4 days ago
  •  ...rollout of new infrastructure capabilities to improve platform reliability and scalability. Monitor system health through metrics...  ...Requirements Require 0 to 1+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles. Require hands-... 
    Full time
    Work at office
    Flexible hours

    PagerDuty

    Atlanta, GA
    5 days ago
  • $167.7k - $245.2k

     ...Cisco Meraki, we are responsible for building and growing the cloud that supports these customers and their networks. As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    Atlanta, GA
    8 days ago
  • Direct message the job poster from STAFFWORXS Delivery Manager @ STAFFWORXS | US IT Recruitment Job Opening: AWS Site Reliability Engineer (SRE) We’re hiring a Site Reliability Engineer (SRE) to join our team in Atlanta, GA. This hybrid role offers the opportunity to work... 
    Contract work

    STAFFWORXS

    Atlanta, GA
    2 days ago
  •  ...evolving the foundational systems and practices that ensure the reliability, scalability, performance, and efficiency of our critical...  ...highly resilient systems. Leverage deep expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles... 
    Early shift

    Cloud Analytics Technologies LLC

    Atlanta, GA
    6 days ago
  •  ...our company effectively.    The Lead Systems Engineer is a senior individual contributor responsible for the reliability, scalability, and modernization of Intellum's...  ...infrastructure, DevOps, platform engineering, site reliability engineering, or a related discipline... 
    Remote work
    Flexible hours

    Intellum, Inc.

    Atlanta, GA
    3 days ago
  •  ...Technology Consultant - Site Reliability Engineer (SRE) Location- - Atlanta, GA (hybrid schedule) Client is currently seeking an experienced Technology Consultant Site Reliability Engineer (SRE) with strong hands-on expertise in Kubernetes, Observability... 
    Permanent employment

    Stellent IT LLC

    Atlanta, GA
    3 days ago
  •  ...Job Title: Site Reliability Engineer II (SRE II) Data & Intelligence Location: Atlanta, GA Contract Job Summary The Site Reliability Engineer II (SRE II) is responsible for ensuring the reliability, scalability, performance, security, and operational... 
    Contract work

    VDart Inc

    Atlanta, GA
    2 days ago
  • $151k - $297k

     ...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB's cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Atlanta, GA
    3 days ago
  • $132.23k - $176.31k

     ...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role... 
    Full time
    Temporary work
    Remote work

    Lumen

    Atlanta, GA
    5 hours ago
  • #CareersJC 1483593Qualifications· Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.· Hands-on experience with incident management and 24/7 production support models.· Proficiency with monitoring and observability...

    LTM

    Atlanta, GA
    1 day ago
  •  ...Job Description Job Description We are seeking a highly experienced Site Reliability Engineer (SRE) with deep expertise in Dynatrace, observability engineering, and Azure cloud technologies . This role will be exclusively focused on building, enhancing, and managing... 
    Work at office
    Local area

    RaceTrac

    Atlanta, GA
    2 days ago
  •  ...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's...  ...is founder-led, profitable, and growing. We are hiring a Site Reliability Engineer Our goal is to perfect enterprise infrastructure DevOps... 
    Work at office
    Local area
    Remote work
    Work from home
    Worldwide

    Canonical

    Atlanta, GA
    7 days ago
  •  ...Consultancy and Information Technology Enabled Services.Job DescriptionSCM System EngineerSCM Continuous Integration / Delivery Build Team Engineer with experience in Application Service and Web Application Build, Deployment and Release Management and experience in establishing... 
    Permanent employment
    Full time
    H1b

    Career Guidant

    Atlanta, GA
    1 day ago
  •  ...accurate and real time timesheets, record complete and accurate notes of troubleshooting and communication with clients    Occasional on-site presence to client sites as required Participate in the on-call rotation (1 week every 3-4 months)   Additional duties as... 
    Work at office

    VC3, Inc.

    Atlanta, GA
    4 days ago
  •  ...Job Description Job Description IT Systems Engineer - Level 3 Location: Marietta, GA (near the Battery / Truist Park) Full Time...  ...performance and proactively identify opportunities to improve reliability, security, and operational efficiency. Develop automation... 
    Full time
    Work at office
    Remote work
    3 days per week

    Tier4 Group

    Atlanta, GA
    24 days ago
  • $141.3k - $237.4k

     ...AT&T, you won’t just imagine the future, you’ll build it.We are seeking a highly skilled and hands-on Lead Software Engineer to join Software Reliability Engineering (SRE) Onboarding and automation team. This role will drive innovation through automation, enhancement of... 
    Full time
    Temporary work
    Work at office
    Local area
    Relocation

    AT&T

    Atlanta, GA
    2 days ago
  • $105k - $130k

     ...provide the high-speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage. The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services... 
    Weekly pay
    Permanent employment
    Full time
    Work at office

    Hermeus

    Atlanta, GA
    1 day ago
  • $101.5k - $169.1k

     ...include an incentive program.Job DescriptionThe Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train...  ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures. Maintain... 
    Full time
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours

    Cox Enterprises

    Atlanta, GA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!