Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

T-Mobile

At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees enjoy multiple wealth-building opportunities through our annual stock grant, employee stock purchase plan, 401(k), and access to free, year-round money coaches. That’s how we’re UNSTOPPABLE for our employees! Job Overview This role leads the design, development, testing, implementation, and operation of secure, scalable, resilient, and highly available software platforms using Site Reliability Engineering and AI-native engineering practices. The engineer collaborates across technical teams to build and support cloud-native, containerized, and distributed systems while owning production reliability, monitoring, incident response, CI/CD, deployment stability, vulnerability remediation, and operational automation across a growing portfolio of enterprise applications. The role designs and implements Terraform-based infrastructure-as-code and drives enterprise standardization for capabilities such as certificate lifecycle management, secrets management, and Vault solutions. It also applies AI-assisted development, intelligent automation, and automated testing to reduce manual engineering effort, improve software quality, and support the reliable operation of AI-enabled services. Success is measured by improved availability, reduced incidents and recovery time, secure and maintainable solutions, faster delivery, and greater operational efficiency. This position provides technical leadership for the SRE team, strengthens the India GCC capability, and establishes reusable engineering standards that reduce risk, improve engineering productivity, and enable scale. Job Responsibilities Lead SRE architecture, engineering innovation, and standardization across cloud platforms, containers, deployment patterns, certificate lifecycle management, secrets management, and Vault solutions. Integrate AI-native development practices, automation frameworks, modern software architectures, and emerging technologies to improve scalability, efficiency, and software delivery performance, while leading and mentoring the broader SRE team and strengthening the India GCC capability through technical guidance, knowledge transfer, engineering reviews, and consistent operational practices. Maintain clear technical documentation, runbooks, automated knowledge artifacts, architecture decisions, and reusable implementation patterns to improve supportability and collaboration. Own production reliability and engineering for enterprise applications, including monitoring, observability, incident response, root-cause analysis, service recovery, performance, capacity, and operational readiness. Design, develop, automate, test, and optimize software solutions using Terraform-based infrastructure-as-code, CI/CD improvements, AI-driven development workflows, intelligent automation, and modern testing frameworks to reduce manual effort, improve engineering productivity, accelerate delivery, and ensure high-quality releases. Contribute to design innovations that improve systems, processes, or services using new frameworks and industry best practices Collaborate with technical teams to deliver solutions and mentor others through knowledge sharing and training sessions Support technology strategy by evaluating and applying current technologies that align with business goals Create clear documentation for software code, system designs, and business requirements to support knowledge sharing Also responsible for other duties/projects as assigned by business management as needed Required Education and Work Experience Bachelor's Degree plus 5 years of related work experience OR Advanced degree with 3 years of related experience Acceptable areas of study include Computer Science, Software Engineering, Information Management or equivalent experience in field 4-7 years Technical engineering experience Required Knowledge, Skills and Abilities Site Reliability Engineering and production operations: Experience owning reliability, availability, monitoring, incident response, root-cause analysis, and operational readiness for production services. Infrastructure-as-code and automation: Hands-on design and implementation using Terraform and scripting/programming such as Python, PowerShell, and JavaScript/Node.js. Cloud-native platforms and containers: Hands-on experience with AWS and/or Azure, Docker, Kubernetes, platform services, performance, capacity, and cost awareness. CI/CD and deployment engineering: Experience building secure, repeatable pipelines, automated testing, release validation, rollback, and deployment controls. Secure platform operations: Experience with vulnerability and configuration remediation, IAM, certificate lifecycle management, secrets management, and Vault solutions. Observability and reliability practices: Experience with logging, metrics, tracing, alerting, dashboards, SLIs, SLOs, error budgets, and service health reporting. AI-enabled services and engineering workflows: Experience supporting AI-enabled services or applying AI-assisted development, testing, automation, LLM, RAG, embeddings, or anomaly-detection capabilities. Analytical Thinking (Required) Technical leadership and communication: Ability to lead SRE standards, mentor distributed teams, strengthen GCC capabilities, communicate clearly, and collaborate across cybersecurity and engineering teams. Analytical Thinking Analytics Collaboration Communication Customer Service Mentorship Programming Languages Software Design Software Development System Integration Technical Writing Eligibility At least 18 years of age Legally authorized to work in the United States Travel Travel Required (Yes/No): Yes DOT Regulated DOT Regulated Position (Yes/No): No Safety Sensitive Position (Yes/No): No Base Pay Range Base Pay Range: $113,600 - $205,000 Corporate Bonus Target: 15% The pay range above is the general base pay range for a successful candidate in the role. The successful candidate’s actual pay will be based on various factors, such as work location, qualifications, and experience, so the actual starting pay will vary within this range. Benefits At T-Mobile, our benefits exemplify the spirit of One Team, Together! A big part of how we care for one another is working to ensure our benefits evolve to meet the needs of our team members. Full and part-time employees have access to the same benefits when eligible. We cover all of the bases, offering medical, dental and vision insurance, a flexible spending account, 401(k), employee stock grants, employee stock purchase plan, paid time off and up to 12 paid holidays - which total about 4 weeks for new full-time employees and about 2.5 weeks for new part-time employees annually - paid parental and family leave, family building benefits, back-up care, enhanced family support, childcare subsidy, tuition assistance, college coaching, short- and long-term disability, voluntary AD&D coverage, voluntary accident coverage, voluntary life insurance, voluntary disability insurance, and voluntary long-term care insurance. We don't stop there - eligible employees can also receive mobile service & home internet discounts, pet insurance, and access to commuter and transit programs! Career Growth Never stop growing! As part of the T-Mobile team, you know the Un-carrier doesn’t have a corporate ladder- it's more like a jungle gym of possibilities! We love helping our employees grow in their careers, because it's that shared drive to aim high that drives our business and our culture forward. By applying for this career opportunity, you’re living our values while investing in your career growth–and we applaud it. You’re unstoppable! T-Mobile USA, Inc. is an Equal Opportunity Employer. All decisions concerning the employment relationship will be made without regard to age, race, ethnicity, color, religion, creed, sex, sexual orientation, gender identity or expression, national origin, religious affiliation, marital status, citizenship status, veteran status, the presence of any physical or mental disability, or any other status or characteristic protected by federal, state, or local law. Discrimination, retaliation or harassment based upon any of these factors is wholly inconsistent with how we do business and will not be tolerated. Talent comes in all forms at the Un-carrier. If you are an individual with a disability and need reasonable accommodation at any point in the application or interview process, please let us know by emailing View email address on click.appcast.io or calling View phone number on click.appcast.io. Please note, this contact channel is not a means to apply for or inquire about a position and we are unable to respond to non-accommodation related requests. #J-18808-Ljbffr

Vacancy posted 5 hours ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Atlanta, GA vacancy
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying... 
    Suggested
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Atlanta, GA
    2 days ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    RELX Group

    Atlanta, GA
    3 days ago
  •  ...Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer... 
    Suggested
    Full time
    Worldwide
    Flexible hours

    NCR

    Atlanta, GA
    3 days ago
  • $152.13k - $162.13k

     ...challenge the status-quo.Unum is changing, and we’re excited about what’s next. Join us.General Summary:Unum Group seeks Site Reliability Engineers in Atlanta, GA.Applicants who are interested in this position may apply at (Ref #66753) for consideration.Design, build,... 
    Suggested
    Full time
    Temporary work
    Work at office
    Remote work

    Unum Group

    Atlanta, GA
    1 day ago
  • $138.1k - $198.2k

     ...more intuitive with technology that simply works.  The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments...  ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Work experience placement
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Atlanta, GA
    3 days ago
  •  ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and...  ...vision, audio, and emerging multimodal architectures — fast, reliable, and effortless to deploy at massive scale. To deliver on that... 

    GrabJobs

    Atlanta, GA
    4 days ago
  • Job description Snowflake SRE JD Your Role Accountabilities Primarily responsible for administrating Snowflake environments on AWS Identify, tune, and fix the performance issues on priority. Diagnose and troubleshoot Snowflake related errors and work with team to raise...

    Rcinfotech

    Atlanta, GA
    4 days ago
  •  ...Senior Site Reliability Engineer Atlanta, Georgia Who We Are QGenda is redefining healthcare workforce management everywhere care is delivered. We're on a mission to empower the healthcare industry to better onboarding, deploy, and manage their workforce. Over... 
    Permanent employment
    Full time
    Work at office
    Remote work
    Work from home
    Work visa

    QGenda

    Atlanta, GA
    4 days ago
  •  ...provisioning, monitoring, and troubleshooting staging and production cloud environments . Experienced in architectural design for reliability, scalability, and performance. Practical application of SRE principles : SLIs, SLOs, error budgets, automation, incident... 

    Purple Drive

    Atlanta, GA
    13 hours ago
  • $130k - $180k

     ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and...  ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to... 
    Temporary work
    Work at office
    Immediate start
    Remote work
    Flexible hours

    GrabJobs

    Atlanta, GA
    1 day ago
  •  ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San... 
    Internship
    Work at office
    Local area
    Remote work
    Worldwide

    GrabJobs

    Atlanta, GA
    4 days ago
  • $141k

     ...be a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability... 
    Local area
    Remote work
    Home office
    Flexible hours

    GrabJobs

    Atlanta, GA
    1 day ago
  •  ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it... 
    Full time
    Live in
    Work at office

    Incident IQ

    Atlanta, GA
    16 days ago
  •  ...evolving the foundational systems and practices that ensure the reliability, scalability, performance, and efficiency of our critical...  ...highly resilient systems. Leverage deep expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles... 
    Early shift

    Cloud Analytics Technologies LLC

    Atlanta, GA
    1 day ago
  • #CareersJC 1483593Qualifications· Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.· Hands-on experience with incident management and 24/7 production support models.· Proficiency with monitoring and observability...

    LTM

    Atlanta, GA
    13 hours ago
  •  ...Saviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features. As a Principal  Engineer, you will define and drive the reliability strategy for our SaaS platform.... 

    Saviynt

    Atlanta, GA
    a month ago
  •  ...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's...  ...is founder-led, profitable, and growing. We are hiring a Site Reliability Engineer Our goal is to perfect enterprise infrastructure DevOps... 
    Work at office
    Local area
    Remote work
    Work from home
    Worldwide

    Canonical

    Atlanta, GA
    12 days ago
  • $35 per hour

    Kforce has a client that is seeking a remote Site Reliability Engineer to join their team. Summary: The team consists of systems that can track lead management, job management and sales management. It is built on Salesforce but underpinned by a lot of Java/API's hosted... 
    Contract work
    Remote work
    Atlanta, GA
    1 day ago
  •  ...role supports the Subscription Product Engineering organization, including in-house subscription...  ...through automation, monitoring, and reliability-focused practices across production and...  ...support practices (Required)Knowledge of Site Reliability Engineering principles, infrastructure... 
    Full time
    Temporary work
    Part time
    Work experience placement
    Local area
    Flexible hours

    T-Mobile

    Atlanta, GA
    4 days ago
  •  ...Consultancy and Information Technology Enabled Services.Job DescriptionSCM System EngineerSCM Continuous Integration / Delivery Build Team Engineer with experience in Application Service and Web Application Build, Deployment and Release Management and experience in establishing... 
    Permanent employment
    Full time
    H1b

    Career Guidant

    Atlanta, GA
    1 day ago
  • OverviewJob PurposeThe Engineer, Release Engineering will be responsible for ICE’s overall CI strategy. This role is a combination of hands-on and strategic vision around build and deployment working closely with key stakeholders across the company. A successful candidate... 
    Work experience placement

    Black Knight Financial Services

    Atlanta, GA
    1 day ago
  • $101.5k - $169.1k

     ...include an incentive program.Job DescriptionThe Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train...  ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures. Maintain... 
    Full time
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours

    Cox Enterprises

    Atlanta, GA
    4 days ago
  •  ...resource utilization, latency, errors, availability issues Maintain/improve monitoring & observability dashboards Skills Mandatory: CloudWatch, Dynatrace, Git, Observability, Reliability Patterns Good to Have: Chaos Testing, Shell Scripting... 

    1 point system

    Atlanta, GA
    1 day ago
  • $105k - $130k

     ...provide the high-speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage. The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services... 
    Weekly pay
    Permanent employment
    Full time
    Work at office

    Hermeus

    Atlanta, GA
    13 hours ago
  • $160.8k - $214.1k

     ...observability needs of modern infrastructure. The Customer Reliability Engineering team is the deep technical escalation tier for Cisco Hypershield...  .../fix and reliability cases escalated by Cisco TAC, applying Site Reliability Engineering practices across the full stack: the... 
    Full time
    Temporary work
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Atlanta, GA
    1 day ago
  •  ...That’s how we’re UNSTOPPABLE for our employees!Are you ready for the next chapter in your Uncarrier journey? The Sr. System Reliability Engineer (SRE) guides and mentors other SREs and improves and protects the software and systems behind all of T-Mobile's IT services,... 
    Full time
    Temporary work
    Part time
    Work experience placement
    Local area
    Flexible hours

    T-Mobile

    Atlanta, GA
    4 days ago
  • $126k - $167k

     ...autonomy, AI, computer vision, sensor fusion, and networking technology to the military in months, not years.ABOUT THE TEAMThe Reliability Engineering team partners across Anduril's engineering, manufacturing, and operations organizations to ensure our autonomous systems... 
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Atlanta, GA
    13 hours ago
  • $168.5k - $252.7k

     ...secure.About the RoleAs a Senior Software Engineer, you will play a key role in designing...  ...mentor team members to ensure high-velocity, reliable delivery.What You’ll DoDesign, build, and...  ...not Workday Careers. Please be aware of sites that may ask for you to input your data... 
    Full time
    Contract work
    Work at office
    Remote work
    Home office
    Flexible hours

    Workday

    Atlanta, GA
    3 days ago
  • OverviewJob PurposeIntercontinental Exchange presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform. The Engineer will drive enablement and adoption to migrate high performant and critical applications into... 
    Full time

    Intercontinental Exchange

    Atlanta, GA
    1 day ago
  • OverviewJob PurposeIntercontinental Exchange (ICE) presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform. The Engineer will drive enablement and adoption to migrate high performant and critical applications... 
    Full time

    Black Knight Financial Services

    Atlanta, GA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!