Lead Site Reliability Engineer
$105.79k - $141.05kLumen
Lumen is the trusted network for the AI‑powered world, connecting people, data, and applications through our expansive fiber network and connected ecosystem. We enable secure, high‑performance connectivity across cloud, edge, and AI workloads for enterprises, governments, and communities.
At Lumen, you’ll work on infrastructure customers rely on today and build for what’s next, where performance, security, and resilience matter.
This is a high accountability environment where bold ideas drive real innovation for our customers, partners, and industry. The work is challenging, expectations are clear, and trust is built into how we operate. If you’re ready to take ownership, deliver meaningful impact, and help shape the future of AI‑ready connectivity, join us today.
The Role
We are seeking a highly skilled and proactive Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role is critical to ensuring the reliability, scalability, and efficiency of our systems, with a strong emphasis on AWS infrastructure, observability, automation, and AI-assisted engineering practices.
The Lead SRE requires an AI-native mindset, understands the software development lifecycle (from coding to support) and applies modern AI tools to enhance productivity, quality, and operational excellence. This role will shape how Lumen combines the latest technologies, including AI-driven automation, to modernize software delivery and application lifecycle management.
This role will collaborate with key stakeholders across the engineering organization — including product owners, developers, and testers — to design, optimize, and automate business and technical processes, while effectively navigating multiple teams within a large and complex organization.
Location
This role is designated as a fully remote position within the United States.
The Main Responsibilities
Production Support & Incident Management
- Implement AI systems and automations to assist during ongoing outages and triage potential ones. You will work with the development teams to ensure that they have all the data normally needed during an outage at their fingertips including preliminary analysis by AI.
- Provide Tier 3 support for issues across portal services by troubleshooting and resolving technical issues in test and production environments.
- Lead root cause analysis and post-mortem processes, incorporating AI-assisted analysis and pattern detection to ensure continuous improvement.
Performance Optimization
- Monitor system performance and proactively identify bottlenecks or degradation using AI-driven observability and anomaly detection tools.
- Implement tuning strategies across application layers, databases, and infrastructure.
- Drive initiatives to improve latency, throughput, and resource utilization.
Monitoring & Observability
- Deploy improved alerting for Lumen Connect in depth, focusing on outside in but including early indicators for fulfilment and other areas. Combining traditional monitoring with AI-based anomaly detection and noise reduction.
- Proactively monitor the errors and performance on Lumen Connect. Implement rules to detect deviations, implement improvements together with the teams.
- Design and maintain dashboards, alerts, and metrics using tools like Datadog, AppInsights, CloudWatch, or similar.
Automation & Infrastructure as Code
- Develop and maintain automation scripts and tools for deployment, scaling, and recovery.
- Develop and maintain automation scripts and tools for deployment, scaling, and recovery, leveraging AI-assisted code generation and validation tools
- Use Terraform, or similar IaC tools to manage AWS resources.
Reliability Engineering
- Perform an in-depth analysis of the overall system and its dependencies, implementing techniques to increase the global availability, reduce the reliance on unstable dependencies and guide ecosystem improvements.
- Champion SRE principles such as SLIs, SLOs, and error budgets.
- Advocate for resilient architecture and fault-tolerant design patterns, incorporating AI-assisted design reviews and architecture evaluation.
Collaboration & Communication
- Work closely with software engineers, DevOps, and product teams to align reliability goals.
- Document processes, runbooks, and best practices for knowledge sharing.
- Provide mentorship and guidance on reliability and operational excellence.
What We Look For in a Candidate
Required Qualifications:
- 5 years overall professional experience in SRE, DevOps, or infrastructure engineering roles.
- Experience with Terraform, or similar IaC tools to manage Cloud resources.
- Proficiency in scripting languages (Python, Bash, etc.) and automation frameworks.
- Experience with CI/CD pipelines and tools like GitHub Actions, Jenkins or GitLab CI.
- Solid understanding of monitoring and logging tools (e.g., CloudWatch, ELK, Datadog).
- Familiarity with containerization and orchestration (Docker, Kubernetes).
- Excellent AI and problem-solving skills, and a proactive mindset.
Preferred Qualifications:
- Experience in AWS services (EC2, CloudFront, EKS, RDS, S3, etc.).
- Certifications in AWS or related technologies are a plus.
- Experience of application development using Java Microservices and Spring Boot framework
- Experience with Agile/SCRUM Methodologies and development practices
Compensation
This information reflects the anticipated base salary range for this position based on current national data. Minimums and maximums may vary based on location. Individual pay is based on skills, experience and other relevant factors.
Location Based Pay Ranges
$105,786 - $141,047 in these states: AL AR AZ FL GA IA ID IN KS KY LA ME MO MS MT ND NE NM OH OK PA SC SD TN UT VT WI WV WY $111,074 - $148,099 in these states: CO HI MI MN NC NH NV OR RI $116,364 - $155,152 in these states: AK CA CT DC DE IL MA MD NJ NY TX VA WA
Lumen offers a comprehensive package featuring a broad range of Health, Life, Voluntary Lifestyle benefits and other perks that enhance your physical, mental, emotional and financial wellbeing. We're able to answer any additional questions you may have about our bonus structure (short-term incentives, long-term incentives and/or sales compensation) as you move through the selection process.
Learn more about Lumen's:
LI-Remote
LI-VK1
Requisition #: 342698
Life at Lumen
Life at Lumen is human and connected, even in a fast moving, AI‑focused organization. We set clear expectations and trust people to meet them. With real support and shared accountability, teams collaborate better, move faster, and deliver meaningful outcomes.
Our Lumen 8 behaviors guide how we interact, make decisions, and work together, shaping a culture built to perform and win.
To learn more about Life at Lumen and how we live the Lumen 8, please visit:
Background Screening
If you are selected for a position, there will be a background screen, which may include checks for criminal records and/or motor vehicle reports and/or drug screening, depending on the position requirements. For more information on these checks, please refer to the Post Offer section of our FAQ page . Job-related concerns identified during the background screening may disqualify you from the new position or your current role. Background results will be evaluated on a case-by-case basis.
Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.
Equal Employment Opportunities
We are committed to providing equal employment opportunities to all persons regardless of race, color, ancestry, citizenship, national origin, religion, veteran status, disability, genetic characteristic or information, age, gender, sexual orientation, gender identity, gender expression, marital status, family status, pregnancy, or other legally protected status (collectively, “protected statuses”). We do not tolerate unlawful discrimination in any employment decisions, including recruiting, hiring, compensation, promotion, benefits, discipline, termination, job assignments or training.
Privacy Notice
Lumen is committed to protecting the privacy and security of personal information collected during the recruitment and hiring process. Our Applicant Privacy Notice explains how we collect, use, disclose, and protect applicant information, as well as how individuals may request access to or deletion of their personal data.
To review Lumen’s Global Employment Applicant and Talent Community Privacy Notice, please visit:
Disclaimer
The job responsibilities described above indicate the general nature and level of work performed by employees within this classification. It is not intended to include a comprehensive inventory of all duties and responsibilities for this job. Job duties and responsibilities are subject to change based on evolving business needs and conditions.
In any materials you submit, you may redact or remove age-identifying information such as age, date of birth, or dates of school attendance or graduation. You will not be penalized for redacting or removing this information.
Please be advised that Lumen does not require any form of payment from job applicants during the recruitment process. All legitimate job openings will be posted on our official website or communicated through official company email addresses. If you encounter any job offers that request payment in exchange for employment at Lumen, they are not for employment with us, but may relate to another company with a similar name.
- ...works. You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for... ...better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives...SuggestedFull timeMonday to FridayFlexible hoursShift workNight shift
$110.7k - $171.8k
...operational toil and clear failure modes Lead the design and implementation of... ...Participation in oncall rotation as a platform reliability escalation point Incident response,... ...requirements. Collaborate with engineering teams across the organization to...SuggestedWork experience placementWork at officeLocal area$127k - $249k
...MongoDB, Inc. is seeking an experienced Senior or Staff Engineer for their SRE, InfraSec team, responsible for guiding the security of cloud-based infrastructure. The role involves hands-on technical work and mentorship of a small team while collaborating with engineering...SuggestedRemote workFlexible hours- ...METRIX IT SOLUTIONS INC is seeking a senior database engineer to design, deploy, and manage multi-region CockroachDB clusters in production. The role focuses on high availability, data consistency, and scalable capacity planning for global deployments. You will monitor...Suggested
- ...A leading entertainment company in Austin, Texas, is looking for a Manager of the Technical Operations Center (TOC) to oversee... ...and cloud services. With a strong technical background in site reliability engineering, the ideal applicant will have excellent communication...Suggested
$127k - $249k
...A leading technology company is seeking an experienced Senior or Staff Engineer for their SRE, InfraSec team in Austin. This role focuses on leading the design and implementation of security solutions for cloud platforms while mentoring a team. Candidates should have over...$174.35k - $210k
...Site Reliability Engineer, IBM Corporation, Austin, TX (Up to 80% telecommuting permitted): Analyze business needs, determine problems, and advise... .... Contribute with technical and operational guidance to lead professional teams. Manage special projects or departmental...Remote work- ...commercialization, and mass production to change the world for the better. JOB SUMMARY We are seeking an experienced Site Reliability Engineer to own and maintain the deployment of our cloud-based infrastructure to customer sites. In this role, you will work...Full timeLocal area
$121.4k - $218.6k
...and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages... ...Employee Stock Purchase Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings plan, company...Work experience placementWork at office- ...Site Reliability Engineer Location: Austin, Texas Schedule: Full-Time Pay Range: Competitive pay, based on experience and qualifications. Make a Difference as a Site Reliability Engineer in Austin! Are you a problem-solving Site Reliability Engineer...Full timeFlexible hours
$152k - $195k
...Senior Site Reliability Engineer Austin, TX (Hybrid) SecurityScorecard is the global leader in cybersecurity ratings, with over 12 million... ...system observability — define SLOs, alerts, and dashboards. Lead incident response and postmortems, focusing on root cause...$98.58k - $138.02k
...Site Reliability Engineer II Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back-office operations for restaurants. Restaurant365's culture is focused on empowering...Work at office- ...Title: Site Reliability Engineer (SRE) Location: Austin, TX Description: We're searching for a driven Site Reliability Engineer (SRE) to join our innovative team. As an SRE, you'll be a cornerstone of our production software, ensuring our systems...Work experience placement
$131.6k - $210.3k
...Progress starts with you. Job Description The Staff Site Reliability Engineer (Azure)is responsible for designing, building, and... ...our platform maturity by supporting cross-functional squads, leading complex technical initiatives, and ensuring the availability...Work experience placementWork at officeLocal areaRemote work$102k - $234.6k
...architecting infrastructure and service for reliability and functionality. Provides day-to-day... ...technology, execute improvements, build site reliability knowledge, and provide clear... ...inclusive culture. Problem Solving: Leads team to identify and address moderately...Temporary workImmediate startFlexible hours$51.9 per hour
...This job is responsible for the reliability, availability, and... ...efficiency. This role blends software engineering, clinical engineering, and... ...cross-functionally with AHN site leaders and teams to navigate... ...drills and exercises, as needed. Leads or participates in post-...For contractorsLocal area$167.18k - $203.61k
...Cox Automotive - USA Job Family Group Engineering / Product Development Job Profile Lead Software Engineer Management Level Manager... ...Cox Automotive Corporate Services, LLC LEAD SITE RELIABILITY ENGINEER Job Description : Lead...Work at officeRemote workFlexible hoursShift work$132.23k - $176.31k
...the future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This...Full timeTemporary workRemote work$94.82k - $130.41k
...To Work®, and named one of BuiltIn's Best Places to Work for the seventh year in a row! We are is seeking a Senior Solutions Engineer to lead the delivery of our most advanced technical solutions. As a vital member of the Professional Services team, you’ll work...Permanent employmentFull timeH1bWork at officeLocal areaWork visa- ...SRE Engineer Experienced SRE Engineer with experience in designing, managing and supporting distributed systems across multi-cloud... ...environment monitoring and task automation Identify application reliability and availability improvements and build solutions to drive an...Local areaRelocation
- ***Must be submitted with LinkedIn Profile*** SRE Engineer Must-haves: Candidate should have sound knowledge in Observability... ...Role Descriptions: We are looking for a highly skilled Site Reliability Engineer (SRE) to design| build| and maintain reliable| scalable...
- ...Hello, Hope you are doing good. Position: SRE Engineer Location: Austin TX (Day 1 on-site) Duration: Long Term Client: Cigniti/Apple Job Description: • Expert-level experience in Kubernetes administration, including building...
$90k - $145k
...and your family. World-class facilities and the technology you need to thrive - in our offices or yours. Job Summary The DevOps Engineer is responsible for designing and managing CI/CD pipelines, developing Infrastructure as Code with Terraform, and ensuring system...Full timeWork experience placementWorldwideFlexible hours- ...everywhere. More about our mission and what we offer . Job Description The role of a Wise Platform Enterprise Support Team Lead is a blend of team management, strategic oversight, and hands-on operational support, all focused on ensuring Wise's business-to-...Work at officeVisa sponsorshipFlexible hours
- ...environment. OCI’s mission is to provide secure, reliable, and performant cloud infrastructure... ...looking for a strong Senior software engineer who can design, build, and operate... .... Discover your potential at a company leading the way in AI and cloud solutions that impact...Flexible hours
$24.1 - $41.23 per hour
...are is what we need. Job Category Retail Position Summary The Lead Technician position requires your experience and technical... ...understanding of any of the following: Electrical/electronic systems Engine repair Engine performance Automatic transmission/transaxle...Daily paidMinimum wagePrice workFull timeTemporary workPart timeLocal areaFlexible hours- ...Job Description Job Description Sr. Software Engineer - Site Reliability About ShipperHQ: ShipperHQ is a trusted leader in the e-commerce shipping space, with over 15 years of experience helping merchants deliver better checkout experiences. Founded in 2009, we...Full timeWork at office
$24 - $30 per hour
..., Kerry, and ex-MBB consultants. Position Overview: Nulixir is looking for a highly qualified and experienced professional to lead Nulixir’s day-to-day operations of our in-house beverage canning/bottling line operations. Position Location: This position will...Full timeRelocationShift work$180k - $230k
...offer. Job Description We’re looking for a Senior Solutions Engineer to join our growing team in Austin, Texas. The Wise Platform... ...happen by working with partners to build integrations to our class-leading international finance products, while improving Wise Platform...Full timeWork experience placementWork at office$184k - $287.5k
...teams with the most inquisitive people in the world. Join us at the forefront of technological advancement. We are hiring software engineers to work on the CUDA driver for Windows. CUDA is NVIDIA’s platform for accelerating general purpose computation on the GPU. Our...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead Site Reliability Engineer. Be the first to apply!
- lead engineer Austin, TX
- lead operating engineer Austin, TX
- site reliability engineer Austin, TX
- site reliability engineer remote Austin, TX
- site reliability engineer sre Austin, TX
- construction site safety Austin, TX
- site recruiter Austin, TX
- on site coordinator Austin, TX
- website content developer Austin, TX
- website coordinator Austin, TX




