Site Reliability Engineer
T-Mobile
At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees enjoy multiple wealth-building opportunities through our annual stock grant, employee stock purchase plan, 401(k), and access to free, year-round money coaches. That’s how we’re UNSTOPPABLE for our employees! Job Overview This role leads the design, development, testing, implementation, and operation of secure, scalable, resilient, and highly available software platforms using Site Reliability Engineering and AI-native engineering practices. The engineer collaborates across technical teams to build and support cloud-native, containerized, and distributed systems while owning production reliability, monitoring, incident response, CI/CD, deployment stability, vulnerability remediation, and operational automation across a growing portfolio of enterprise applications. The role designs and implements Terraform-based infrastructure-as-code and drives enterprise standardization for capabilities such as certificate lifecycle management, secrets management, and Vault solutions. It also applies AI-assisted development, intelligent automation, and automated testing to reduce manual engineering effort, improve software quality, and support the reliable operation of AI-enabled services. Success is measured by improved availability, reduced incidents and recovery time, secure and maintainable solutions, faster delivery, and greater operational efficiency. This position provides technical leadership for the SRE team, strengthens the India GCC capability, and establishes reusable engineering standards that reduce risk, improve engineering productivity, and enable scale. Job Responsibilities Lead SRE architecture, engineering innovation, and standardization across cloud platforms, containers, deployment patterns, certificate lifecycle management, secrets management, and Vault solutions. Integrate AI-native development practices, automation frameworks, modern software architectures, and emerging technologies to improve scalability, efficiency, and software delivery performance, while leading and mentoring the broader SRE team and strengthening the India GCC capability through technical guidance, knowledge transfer, engineering reviews, and consistent operational practices. Maintain clear technical documentation, runbooks, automated knowledge artifacts, architecture decisions, and reusable implementation patterns to improve supportability and collaboration. Own production reliability and engineering for enterprise applications, including monitoring, observability, incident response, root-cause analysis, service recovery, performance, capacity, and operational readiness. Design, develop, automate, test, and optimize software solutions using Terraform-based infrastructure-as-code, CI/CD improvements, AI-driven development workflows, intelligent automation, and modern testing frameworks to reduce manual effort, improve engineering productivity, accelerate delivery, and ensure high-quality releases. Contribute to design innovations that improve systems, processes, or services using new frameworks and industry best practices Collaborate with technical teams to deliver solutions and mentor others through knowledge sharing and training sessions Support technology strategy by evaluating and applying current technologies that align with business goals Create clear documentation for software code, system designs, and business requirements to support knowledge sharing Also responsible for other duties/projects as assigned by business management as needed Required Education and Work Experience Bachelor's Degree plus 5 years of related work experience OR Advanced degree with 3 years of related experience Acceptable areas of study include Computer Science, Software Engineering, Information Management or equivalent experience in field 4-7 years Technical engineering experience Required Knowledge, Skills and Abilities Site Reliability Engineering and production operations: Experience owning reliability, availability, monitoring, incident response, root-cause analysis, and operational readiness for production services. Infrastructure-as-code and automation: Hands-on design and implementation using Terraform and scripting/programming such as Python, PowerShell, and JavaScript/Node.js. Cloud-native platforms and containers: Hands-on experience with AWS and/or Azure, Docker, Kubernetes, platform services, performance, capacity, and cost awareness. CI/CD and deployment engineering: Experience building secure, repeatable pipelines, automated testing, release validation, rollback, and deployment controls. Secure platform operations: Experience with vulnerability and configuration remediation, IAM, certificate lifecycle management, secrets management, and Vault solutions. Observability and reliability practices: Experience with logging, metrics, tracing, alerting, dashboards, SLIs, SLOs, error budgets, and service health reporting. AI-enabled services and engineering workflows: Experience supporting AI-enabled services or applying AI-assisted development, testing, automation, LLM, RAG, embeddings, or anomaly-detection capabilities. Analytical Thinking (Required) Technical leadership and communication: Ability to lead SRE standards, mentor distributed teams, strengthen GCC capabilities, communicate clearly, and collaborate across cybersecurity and engineering teams. Analytical Thinking Analytics Collaboration Communication Customer Service Mentorship Programming Languages Software Design Software Development System Integration Technical Writing Eligibility At least 18 years of age Legally authorized to work in the United States Travel Travel Required (Yes/No): Yes DOT Regulated DOT Regulated Position (Yes/No): No Safety Sensitive Position (Yes/No): No Base Pay Range Base Pay Range: $113,600 - $205,000 Corporate Bonus Target: 15% The pay range above is the general base pay range for a successful candidate in the role. The successful candidate’s actual pay will be based on various factors, such as work location, qualifications, and experience, so the actual starting pay will vary within this range. Benefits At T-Mobile, our benefits exemplify the spirit of One Team, Together! A big part of how we care for one another is working to ensure our benefits evolve to meet the needs of our team members. Full and part-time employees have access to the same benefits when eligible. We cover all of the bases, offering medical, dental and vision insurance, a flexible spending account, 401(k), employee stock grants, employee stock purchase plan, paid time off and up to 12 paid holidays - which total about 4 weeks for new full-time employees and about 2.5 weeks for new part-time employees annually - paid parental and family leave, family building benefits, back-up care, enhanced family support, childcare subsidy, tuition assistance, college coaching, short- and long-term disability, voluntary AD&D coverage, voluntary accident coverage, voluntary life insurance, voluntary disability insurance, and voluntary long-term care insurance. We don't stop there - eligible employees can also receive mobile service & home internet discounts, pet insurance, and access to commuter and transit programs! Career Growth Never stop growing! As part of the T-Mobile team, you know the Un-carrier doesn’t have a corporate ladder- it's more like a jungle gym of possibilities! We love helping our employees grow in their careers, because it's that shared drive to aim high that drives our business and our culture forward. By applying for this career opportunity, you’re living our values while investing in your career growth–and we applaud it. You’re unstoppable! T-Mobile USA, Inc. is an Equal Opportunity Employer. All decisions concerning the employment relationship will be made without regard to age, race, ethnicity, color, religion, creed, sex, sexual orientation, gender identity or expression, national origin, religious affiliation, marital status, citizenship status, veteran status, the presence of any physical or mental disability, or any other status or characteristic protected by federal, state, or local law. Discrimination, retaliation or harassment based upon any of these factors is wholly inconsistent with how we do business and will not be tolerated. Talent comes in all forms at the Un-carrier. If you are an individual with a disability and need reasonable accommodation at any point in the application or interview process, please let us know by emailing View email address on click.appcast.io or calling View phone number on click.appcast.io. Please note, this contact channel is not a means to apply for or inquire about a position and we are unable to respond to non-accommodation related requests. #J-18808-Ljbffr
$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying...SuggestedFull timeTemporary workWork experience placementFlexible hours$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SuggestedFull timeWork at officeLocal areaRemote workWork from home- ...Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer...SuggestedFull timeWorldwideFlexible hours
$152.13k - $162.13k
...challenge the status-quo.Unum is changing, and we’re excited about what’s next. Join us.General Summary:Unum Group seeks Site Reliability Engineers in Atlanta, GA.Applicants who are interested in this position may apply at (Ref #66753) for consideration.Design, build,...SuggestedFull timeTemporary workWork at officeRemote work$138.1k - $198.2k
...more intuitive with technology that simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments... ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering...SuggestedPermanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours- ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...vision, audio, and emerging multimodal architectures — fast, reliable, and effortless to deploy at massive scale. To deliver on that...
- Job description Snowflake SRE JD Your Role Accountabilities Primarily responsible for administrating Snowflake environments on AWS Identify, tune, and fix the performance issues on priority. Diagnose and troubleshoot Snowflake related errors and work with team to raise...
- ...Senior Site Reliability Engineer Atlanta, Georgia Who We Are QGenda is redefining healthcare workforce management everywhere care is delivered. We're on a mission to empower the healthcare industry to better onboarding, deploy, and manage their workforce. Over...Permanent employmentFull timeWork at officeRemote workWork from homeWork visa
- ...provisioning, monitoring, and troubleshooting staging and production cloud environments . Experienced in architectural design for reliability, scalability, and performance. Practical application of SRE principles : SLIs, SLOs, error budgets, automation, incident...
$130k - $180k
...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to...Temporary workWork at officeImmediate startRemote workFlexible hours- ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San...InternshipWork at officeLocal areaRemote workWorldwide
$141k
...be a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability...Local areaRemote workHome officeFlexible hours- ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it...Full timeLive inWork at office
- ...evolving the foundational systems and practices that ensure the reliability, scalability, performance, and efficiency of our critical... ...highly resilient systems. Leverage deep expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles...Early shift
- #CareersJC 1483593Qualifications· Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.· Hands-on experience with incident management and 24/7 production support models.· Proficiency with monitoring and observability...
- ...Saviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features. As a Principal Engineer, you will define and drive the reliability strategy for our SaaS platform....
- ...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's... ...is founder-led, profitable, and growing. We are hiring a Site Reliability Engineer Our goal is to perfect enterprise infrastructure DevOps...Work at officeLocal areaRemote workWork from homeWorldwide
$35 per hour
Kforce has a client that is seeking a remote Site Reliability Engineer to join their team. Summary: The team consists of systems that can track lead management, job management and sales management. It is built on Salesforce but underpinned by a lot of Java/API's hosted...Contract workRemote work- ...role supports the Subscription Product Engineering organization, including in-house subscription... ...through automation, monitoring, and reliability-focused practices across production and... ...support practices (Required)Knowledge of Site Reliability Engineering principles, infrastructure...Full timeTemporary workPart timeWork experience placementLocal areaFlexible hours
- ...Consultancy and Information Technology Enabled Services.Job DescriptionSCM System EngineerSCM Continuous Integration / Delivery Build Team Engineer with experience in Application Service and Web Application Build, Deployment and Release Management and experience in establishing...Permanent employmentFull timeH1b
- OverviewJob PurposeThe Engineer, Release Engineering will be responsible for ICE’s overall CI strategy. This role is a combination of hands-on and strategic vision around build and deployment working closely with key stakeholders across the company. A successful candidate...Work experience placement
$101.5k - $169.1k
...include an incentive program.Job DescriptionThe Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train... ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures. Maintain...Full timeWork at officeRemote workVisa sponsorshipFlexible hours- ...resource utilization, latency, errors, availability issues Maintain/improve monitoring & observability dashboards Skills Mandatory: CloudWatch, Dynatrace, Git, Observability, Reliability Patterns Good to Have: Chaos Testing, Shell Scripting...
$105k - $130k
...provide the high-speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage. The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services...Weekly payPermanent employmentFull timeWork at office$160.8k - $214.1k
...observability needs of modern infrastructure. The Customer Reliability Engineering team is the deep technical escalation tier for Cisco Hypershield... .../fix and reliability cases escalated by Cisco TAC, applying Site Reliability Engineering practices across the full stack: the...Full timeTemporary workLocal areaRemote workFlexible hours- ...That’s how we’re UNSTOPPABLE for our employees!Are you ready for the next chapter in your Uncarrier journey? The Sr. System Reliability Engineer (SRE) guides and mentors other SREs and improves and protects the software and systems behind all of T-Mobile's IT services,...Full timeTemporary workPart timeWork experience placementLocal areaFlexible hours
$126k - $167k
...autonomy, AI, computer vision, sensor fusion, and networking technology to the military in months, not years.ABOUT THE TEAMThe Reliability Engineering team partners across Anduril's engineering, manufacturing, and operations organizations to ensure our autonomous systems...Full timeWork experience placementImmediate start$168.5k - $252.7k
...secure.About the RoleAs a Senior Software Engineer, you will play a key role in designing... ...mentor team members to ensure high-velocity, reliable delivery.What You’ll DoDesign, build, and... ...not Workday Careers. Please be aware of sites that may ask for you to input your data...Full timeContract workWork at officeRemote workHome officeFlexible hours- OverviewJob PurposeIntercontinental Exchange presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform. The Engineer will drive enablement and adoption to migrate high performant and critical applications into...Full time
- OverviewJob PurposeIntercontinental Exchange (ICE) presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform. The Engineer will drive enablement and adoption to migrate high performant and critical applications...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre Atlanta, GA
- site reliability engineer Atlanta, GA
- on-site clinical research associate (traveling/remote) Atlanta, GA
- website coordinator Atlanta, GA
- junior website developer Atlanta, GA
- site leader Atlanta, GA
- historic site Atlanta, GA
- website content developer Atlanta, GA
- construction site safety Atlanta, GA
- official site Atlanta, GA


