Site Reliability Engineer
T Mobile US
Senior Systems Reliability Engineer (SRE)At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees enjoy multiple wealth-building opportunities through our annual stock grant, employee stock purchase plan, 401(k), and access to free, year-round money coaches. That's how we're UNSTOPPABLE for our employees!Are you ready for the next chapter in your Uncarrier journey? The Sr. System Reliability Engineer (SRE) guides and mentors other SREs and improves and protects the software and systems behind all of T-Mobile's IT services, including management of scalability, availability, latency, performance, security, and capacity, and delivering software faster, better, and cheaper. From designing & maintaining CICD Pipelines to building the next generation of T-Mobile applications on cloud native platforms, the SRE's enable great customer experience and product innovation by continuous improvement of operational support. We pride ourselves on encouraging a culture of innovation, advocating for agile methodologies, and promoting transparency in all that we do. Join us in embodying the spirit of the 'Un-carrier' and make a tangible impact! The Compliance Campaign Management (CCM) team within T-Mobile's Identity Security and Access Management (ISAM) organization is leading a major transformation in logical access compliance. What was once a predominantly manual User Access Review process is rapidly evolving into a sophisticated, automation-driven platform powered by cloud technologies, APIs, workflow orchestration, and artificial intelligence. Our team designs and builds solutions using Power Platform, Azure DevOps Pipelines, Microsoft Graph APIs, and AI-powered automation to streamline compliance operations while expanding risk visibility across the enterprise. Every innovation we deliver helps reduce manual effort, strengthen control effectiveness, and improve audit readiness. As a Senior Systems Reliability Engineer (SRE), you will play a critical leadership role in shaping and operating the platform that powers this transformation. CCM is building an environment where human expertise and AI agents work side by side to execute compliance operations with greater speed, consistency, and precision. You will own the reliability, scalability, observability, and continuous improvement of this compliance automation ecosystem, ensuring it remains highly available, performant, secure, and audit ready as it grows across the enterprise. This is more than a traditional operations role. You'll drive engineering excellence across the platform, establish reliability standards, champion automation-first practices, and lead efforts to improve resilience, monitoring, incident response, and operational maturity. Working closely with software engineers, compliance experts, security teams, and AI-enabled solutions, you'll help define the next generation of compliance operations and governance technology. Join a team that's not just supporting compliance but engineering the future of access governance. You'll have the opportunity to solve complex enterprise-scale challenges, influence strategic technology decisions, mentor engineers, and help build intelligent automation capabilities that redefine how access risk is identified, reviewed, and remediated at T-Mobile. If you're energized by the intersection of cloud engineering, automation, AI, cybersecurity, and operational excellence, this is your chance to make a measurable impact on one of the company's most innovative compliance modernization initiatives.Why This Role Is UniqueLead the reliability strategy for a rapidly growing compliance automation platform.Partner with engineers, compliance professionals, and AI solutions to modernize access governance.Build and enhance observability, resilience, and automation across cloud-native services and integrations.Influence the operational foundations of an environment where AI agents and human operators collaborate to execute critical compliance functions.Drive enterprise-scale outcomes that improve security, reduce risk, and strengthen audit readiness across T-Mobile.Job Responsibilities :Utilize DevOps automation tools to enhance continuous integration and continuous delivery pipelines for non-production environmentsManage environments through automated server provisioning and pipeline configuration to support virtual machinesDeliver software improvements that increase availability, scalability, latency, and efficiency of IT servicesCreate and maintain dashboards for continuous monitoring and health checks to improve service quality in non-production environmentsContribute to software delivery process improvements including cloud enablement and microservices containerizationMentor and guide systems reliability engineers and vendor resources to support team development and performanceArchitects scalable, modular automation frameworks that allow compliance workflows to expand across new control types, compliance frameworks, and operational domains without re-engineering — maintaining consistent, audit-defensible output quality at increasing throughput.Designs automated validation and accuracy verification patterns across agent and workflow pipelines, ensuring repeatable, measurable output quality that supports continuous improvement, operational agility, and audit readiness at scale.Education and Work Experience :Bachelor's Degree plus 5 years of related work experience OR Advanced degree with 3 years of related experience (Required)4-7+ years relevant experience. (Required)Experience working in an Agile and DevOps environment. (Required)Experience in one or more of: C, C#, Java, Perl, Python, Go, or scripting experience in Shell and Perl. (Required)Experience in Continuous Integration/Continuous Delivery tools, such as, Jenkins, Cloudbees, etc., and other automation tools. (Required)Experience with DevOps tools, such as, Ansible, Chef, Puppet, etc. Experience in Docker, Kubernetes, etc. is preferable. (Preferred)Experience in APM tools like AppDynamics, logging tools, like Splunk. (Required)Experience working in a cloud environment (public/private). (Required)Experience in migrating to cloud or cloud native environments is preferable. (Preferred)Required Knowledge, Skills and Abilities :Python or comparable programming language experienceRed Hat Ansible or comparable orchestration technologiesAutomation Architecture and Scalable Framework Design experiencePlatform Engineering and Reusable Pipeline DesignPreferred Knowledge, Skills and Abilities :Agentic workflow design, orchestration, and operational governanceAutomated testing and output validation framework designHuman-in-the-loop system architecture for regulated workflowsAPI-first integration architecture for enterprise-scale automation platformsCompliance or regulatory workflow automation in identity, access, or financial controls environmentsAt least 18 years of ageLegally authorized to work in the United StatesTravel:Travel Required (Yes/No): YesDOT Regulated Position (Yes/No): NoSafety Sensitive Position (Yes/No): NoBase Pay Range: $98,500 - $177,700Corporate Bonus Target: 15%The pay range above is the general base pay range for a successful candidate in the role. The successful candidate's actual pay will be based on various factors, such as work location, qualifications, and experience, so the actual starting pay will vary within this range.At T-Mobile, employees in regular, non-temporary roles are eligible for an annual bonus or periodic sales incentive or bonus, based on their role. Most Corporate employees are eligible for a year-end bonus based on company and/or individual performance and which is set at a percentage of the employee's eligible earnings in the prior year. Certain positions in Customer Care are eligible for monthly bonuses based on individual and/or team performance. To find the pay range for this role based on hiring location, T-Mobile, our benefits exemplify the spirit of One Team, Together! A big part of how we care for one another is working to ensure our benefits evolve to meet the needs of our team members. Full and part-time employees have access to the same benefits when eligible. We cover all of the bases, offering medical, dental and vision insurance, a flexible spending account, 401(k), employee stock grants, employee stock purchase plan, paid time off and up to 12 paid holidays - which total about 4 weeks for new full-time employees and about 2.5 weeks for new part-time employees annually - paid parental and family leave, family building benefits, back-up care, enhanced family support, childcare subsidy, tuition assistance, college coaching, short- and long-term disability, voluntary AD&D coverage, voluntary accident coverage, voluntary life insurance, voluntary disability insurance,
- ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises...SuggestedPermanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
$35 - $44 per hour
DescriptionKforce has a client seeking a remote Site Reliability Engineer to join their team. We are seeking a Site Reliability Engineer (SRE) to support a large-scale system modernization and legacy platform retirement initiative. This role will focus on maintaining and...SuggestedRemote work$98k - $148.5k
...Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization.As a Site Reliability Engineer I on the Core Infrastructure team in our Atlanta office,you'll help build and operate the foundational infrastructure that...SuggestedWork at officeLocal areaFlexible hours- Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence...SuggestedWorldwide
$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying...SuggestedFull timeTemporary workWork experience placementFlexible hours- ...Senior Systems Reliability Engineer (SRE)At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees...Full timeTemporary workPart timeWork experience placementFlexible hours
- ...client accounts, mentor staff, and ensure project success while upholding firm standards. Leverage data, algorithms, and software engineering to deliver scalable security solutions and drive business growth. You will guide teams, develop client-ready strategies, and...
- CBRE Group, Inc. seeks an Operations Consultant to provide technical expertise across complex technology systems and support daily operations within the D&T Support function. You will collaborate with IT teams, optimize configurations, and guide stakeholders on technology...
- ...resilient software platforms using SRE and AI-native engineering practices. Own production reliability, monitoring, and operational automation while mentoring... ...), Kubernetes, and production operations. Key Skills Site Reliability Engineering Terraform Python AWS Azure...Temporary workFlexible hours
- ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability...
$70 - $85 per hour
...redefine what’s possible, give shape to the future—and get there.What You’ll Do* Define and establish enterprise reliability standards, Site Reliability Engineering (SRE) practices, SLIs/SLOs, operational governance models, and engineering guardrails that enable scalable,...Temporary workLocal areaFlexible hours3 days per week- ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it...Full timeLive inWork at office
- ...rollout of new infrastructure capabilities to improve platform reliability and scalability. Monitor system health through metrics... ...Requirements Require 0 to 1+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles. Require hands-...Full timeWork at officeFlexible hours
$167.7k - $245.2k
...Cisco Meraki, we are responsible for building and growing the cloud that supports these customers and their networks. As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze...Full timeTemporary workLocal areaFlexible hours- Direct message the job poster from STAFFWORXS Delivery Manager @ STAFFWORXS | US IT Recruitment Job Opening: AWS Site Reliability Engineer (SRE) We’re hiring a Site Reliability Engineer (SRE) to join our team in Atlanta, GA. This hybrid role offers the opportunity to work...Contract work
- ...evolving the foundational systems and practices that ensure the reliability, scalability, performance, and efficiency of our critical... ...highly resilient systems. Leverage deep expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles...Early shift
- ...our company effectively. The Lead Systems Engineer is a senior individual contributor responsible for the reliability, scalability, and modernization of Intellum's... ...infrastructure, DevOps, platform engineering, site reliability engineering, or a related discipline...Remote workFlexible hours
- ...Technology Consultant - Site Reliability Engineer (SRE) Location- - Atlanta, GA (hybrid schedule) Client is currently seeking an experienced Technology Consultant Site Reliability Engineer (SRE) with strong hands-on expertise in Kubernetes, Observability...Permanent employment
- ...Job Title: Site Reliability Engineer II (SRE II) Data & Intelligence Location: Atlanta, GA Contract Job Summary The Site Reliability Engineer II (SRE II) is responsible for ensuring the reliability, scalability, performance, security, and operational...Contract work
$151k - $297k
...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB's cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and...Local areaRemote workWorldwideFlexible hours$132.23k - $176.31k
...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...Full timeTemporary workRemote work- #CareersJC 1483593Qualifications· Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.· Hands-on experience with incident management and 24/7 production support models.· Proficiency with monitoring and observability...
- ...Job Description Job Description We are seeking a highly experienced Site Reliability Engineer (SRE) with deep expertise in Dynatrace, observability engineering, and Azure cloud technologies . This role will be exclusively focused on building, enhancing, and managing...Work at officeLocal area
- ...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's... ...is founder-led, profitable, and growing. We are hiring a Site Reliability Engineer Our goal is to perfect enterprise infrastructure DevOps...Work at officeLocal areaRemote workWork from homeWorldwide
- ...Consultancy and Information Technology Enabled Services.Job DescriptionSCM System EngineerSCM Continuous Integration / Delivery Build Team Engineer with experience in Application Service and Web Application Build, Deployment and Release Management and experience in establishing...Permanent employmentFull timeH1b
- ...accurate and real time timesheets, record complete and accurate notes of troubleshooting and communication with clients Occasional on-site presence to client sites as required Participate in the on-call rotation (1 week every 3-4 months) Additional duties as...Work at office
- ...Job Description Job Description IT Systems Engineer - Level 3 Location: Marietta, GA (near the Battery / Truist Park) Full Time... ...performance and proactively identify opportunities to improve reliability, security, and operational efficiency. Develop automation...Full timeWork at officeRemote work3 days per week
$141.3k - $237.4k
...AT&T, you won’t just imagine the future, you’ll build it.We are seeking a highly skilled and hands-on Lead Software Engineer to join Software Reliability Engineering (SRE) Onboarding and automation team. This role will drive innovation through automation, enhancement of...Full timeTemporary workWork at officeLocal areaRelocation$105k - $130k
...provide the high-speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage. The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services...Weekly payPermanent employmentFull timeWork at office$101.5k - $169.1k
...include an incentive program.Job DescriptionThe Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train... ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures. Maintain...Full timeWork at officeRemote workVisa sponsorshipFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer sre Atlanta, GA
- website coordinator Atlanta, GA
- on-site clinical research associate (traveling/remote) Atlanta, GA
- site safety Atlanta, GA
- junior website developer Atlanta, GA
- construction site safety Atlanta, GA
- IT site lead Atlanta, GA
- website content developer Atlanta, GA
- site recruiter Atlanta, GA





