Sr Lead Site Reliability Engineer
$132.23k - $176.31kLumen
Lumen is the trusted network for the AI‑powered world, connecting people, data, and applications through our expansive fiber network and connected ecosystem. We enable secure, high‑performance connectivity across cloud, edge, and AI workloads for enterprises, governments, and communities.
At Lumen, you’ll work on infrastructure customers rely on today and build for what’s next, where performance, security, and resilience matter.
This is a high accountability environment where bold ideas drive real innovation for our customers, partners, and industry. The work is challenging, expectations are clear, and trust is built into how we operate. If you’re ready to take ownership, deliver meaningful impact, and help shape the future of AI‑ready connectivity, join us today.
The Role
We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role is critical to ensuring the reliability, scalability, and efficiency of our systems, with a strong emphasis on AWS infrastructure, observability, automation and AI-assisted engineering practices.
The Senior Lead SRE requires an AI-native mindset, understands the software development lifecycle (from coding to support) and applies modern AI tools to enhance productivity, quality, and operational excellence. This role will shape how Lumen combines the latest technologies, including AI-driven automation, to modernize software delivery and application lifecycle management.
This role will collaborate with key stakeholders across the engineering organization — including product owners, developers, and testers — to design, optimize, and automate business and technical processes, while effectively navigating multiple teams within a large and complex organization.
Location
This role is designated as a fully remote position within the United States.
The Main Responsibilities
Production Support & Incident Management
- Help define and improve the processes and best practices for incident management within the customer facing applications.
- Implement AI systems and automations to assist during ongoing outages and triage potential ones. You will work with the development teams to ensure that they have all the data normally needed during an outage at their fingertips including preliminary analysis by AI.
- Design and implement improved SRE processes to handle incidents. From proactive analysis, faster response, automatic remediation, AI guided analysis, partially automatic root cause analysis.
Performance Optimization
- Monitor system performance and proactively identify bottlenecks or degradation using AI-driven observability and anomaly detection tools.
- Implement tuning strategies across application layers, databases, and infrastructure.
- Drive initiatives to improve latency, throughput, and resource utilization.
Monitoring & Observability
- Deploy improved alerting for Lumen Connect in depth, focusing on outside in but including early indicators for fulfilment and other areas. Combining traditional monitoring with AI-based anomaly detection and noise reduction.
- Proactively monitor the errors and performance on Lumen Connect. Implement rules to detect deviations, implement improvements together with the teams.
- Design and maintain dashboards, alerts, and metrics using tools like Datadog, AppInsights, CloudWatch, or similar.
Automation & Infrastructure as Code
- Develop and maintain automation scripts and tools for deployment, scaling, and recovery, leveraging AI-assisted code generation and validation tools
- Use Terraform, or similar IaC tools to manage AWS resources.
Reliability Engineering
- Perform an in-depth analysis of the overall system and its dependencies, implementing techniques to increase the global availability, reduce the reliance on unstable dependencies and guide ecosystem improvements.
- Champion SRE principles such as SLIs, SLOs, and error budgets.
- Advocate for resilient architecture and fault-tolerant design patterns, incorporating AI-assisted design reviews and architecture evaluation.
- Help define and improve better SRE processes and lead significant improvements in reliability for the Lumen Connect platform.
Collaboration & Communication
- Work closely with software engineers, DevOps, and product teams to align reliability goals.
- Document processes, runbooks, and best practices for knowledge sharing.
- Provide mentorship and guidance on reliability and operational excellence.
What We Look For in a Candidate
Required Qualifications:
- 10+ years overall professional experience in SRE, DevOps, or infrastructure engineering roles.
- Experience with Terraform, or similar IaC tools to manage Cloud resources.
- Proficiency in scripting languages (Python, Bash, etc.) and automation frameworks.
- Experience with CI/CD pipelines and tools like GitHub Actions, Jenkins or GitLab CI.
- Solid understanding of monitoring and logging tools (e.g., CloudWatch, ELK, Datadog).
- Familiarity with containerization and orchestration (Docker, Kubernetes).
- Excellent AI and problem-solving skills, and a proactive mindset.
Preferred Qualifications:
- Experience in AWS services (EC2, CloudFront, EKS, RDS, S3, etc.).
- Certifications in AWS or related technologies are a plus.
- Experience of application development using Java Microservices and Spring Boot framework
- Experience with Agile/SCRUM Methodologies and development practices
Compensation
This information reflects the anticipated base salary range for this position based on current national data. Minimums and maximums may vary based on location. Individual pay is based on skills, experience and other relevant factors.
Location Based Pay Ranges
$132,232 - $176,310 in these states: AL AR AZ FL GA IA ID IN KS KY LA ME MO MS MT ND NE NM OH OK PA SC SD TN UT VT WI WV WY
$138,844 - $185,124 in these states: CO HI MI MN NC NH NV OR RI
$145,456 - $193,940 in these states: AK CA CT DC DE IL MA MD NJ NY TX VA WA
Lumen offers a comprehensive package featuring a broad range of Health, Life, Voluntary Lifestyle benefits and other perks that enhance your physical, mental, emotional and financial wellbeing. We're able to answer any additional questions you may have about our bonus structure (short-term incentives, long-term incentives and/or sales compensation) as you move through the selection process.
Learn more about Lumen's:
#LI-Remote
#LI-VK1
Requisition #: 342707
Life at Lumen
Life at Lumen is human and connected, even in a fast moving, AI‑focused organization. We set clear expectations and trust people to meet them. With real support and shared accountability, teams collaborate better, move faster, and deliver meaningful outcomes.
Our Lumen 8 behaviors guide how we interact, make decisions, and work together, shaping a culture built to perform and win.
To learn more about Life at Lumen and how we live the Lumen 8, please visit:
Background Screening
If you are selected for a position, there will be a background screen, which may include checks for criminal records and/or motor vehicle reports and/or drug screening, depending on the position requirements. For more information on these checks, please refer to the Post Offer section of our FAQ page . Job-related concerns identified during the background screening may disqualify you from the new position or your current role. Background results will be evaluated on a case-by-case basis.
Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.
Equal Employment Opportunities
We are committed to providing equal employment opportunities to all persons regardless of race, color, ancestry, citizenship, national origin, religion, veteran status, disability, genetic characteristic or information, age, gender, sexual orientation, gender identity, gender expression, marital status, family status, pregnancy, or other legally protected status (collectively, “protected statuses”). We do not tolerate unlawful discrimination in any employment decisions, including recruiting, hiring, compensation, promotion, benefits, discipline, termination, job assignments or training.
Privacy Notice
Lumen is committed to protecting the privacy and security of personal information collected during the recruitment and hiring process. Our Applicant Privacy Notice explains how we collect, use, disclose, and protect applicant information, as well as how individuals may request access to or deletion of their personal data.
To review Lumen’s Global Employment Applicant and Talent Community Privacy Notice, please visit:
Disclaimer
The job responsibilities described above indicate the general nature and level of work performed by employees within this classification. It is not intended to include a comprehensive inventory of all duties and responsibilities for this job. Job duties and responsibilities are subject to change based on evolving business needs and conditions.
In any materials you submit, you may redact or remove age-identifying information such as age, date of birth, or dates of school attendance or graduation. You will not be penalized for redacting or removing this information.
Please be advised that Lumen does not require any form of payment from job applicants during the recruitment process. All legitimate job openings will be posted on our official website or communicated through official company email addresses. If you encounter any job offers that request payment in exchange for employment at Lumen, they are not for employment with us, but may relate to another company with a similar name.
- ...Sr. Site Reliability Engineer Comtech is a woman-owned small business founded in 1998 and headquartered in Reston, VA. We offer IT solutions across the disciplines of program/project management, applications development, infrastructure, Cyber security, and enterprise...SeniorLocal area
$134.25k - $214.8k
...change. Constantly grow as you work hard for a mission that matters at a company where you matter. Your Impact As a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and...SeniorWork experience placementWork at officeRemote workFlexible hours$174k - $253k
...Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent... ...distributed systems. 2 years of experience leading projects and providing technical leadership... ...or Engineering. About The Job Site Reliability Engineering (SRE) is what you get when...SeniorTemporary work- ...Guidacent, Inc. is seeking a Principal DevOps Engineer in the Greater Seattle area, onsite in Bellevue. You will own the architecture of... ...engineers while driving AI tooling adoption across teams. You will lead design of IaC practices, cloud deployments, and observability...SuggestedFull time
- Technical/Functional Skills Windows Servers, Digital: Microsoft Azure Windows Powershell, Digital: DevOps Roles & Responsibilities Windows Server 2012 -2019 Administration Microsoft Azure Azure AAD DFSR, DHCP DNS, KMS, WSUS TCP/IP Hyper-V High Availability Clusters ...Suggested
$17 - $27.75 per hour
...deliver an exceptional customer experience * Serves as a Brand Ambassador embodying of Coach values and increasing brand awareness * Leads implementation of Company initiatives and support full operation of the business * Maintain a growth mindset for business and...Minimum wageFull timeShift work$194k - $267k
...something more than once, automate it" and who can rapidly self-educate on new concepts and tools. Position Overview: The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and...Permanent employmentWork at officeLocal areaWorldwideFlexible hours- ...seeking a Managing Principal for their Bellevue, WA office. This leader will drive the Building Structures practice, manage a team of engineers, and expand into new markets. Responsibilities include oversight of project delivery, resource management, and client engagement....Full timeWork at office
$160k - $250k
...public clouds when the right fit. As we continue to commercialize our machine learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS offering for our customers. Our ideal candidate is someone who is able...Senior$119.8k - $234.7k
...thinking in a cloud-enabled world. Microsoft's Azure Data engineering team is leading the transformation of analytics in the world of data... ...and We are looking for a self-driven Senior Site Reliability Engineer (SRE) who likes taking a data driven and systems...SeniorOngoing contractLocal area$152.2k - $243.7k
...We are looking for dedicated, curious, and energetic Software Engineers who embrace solving complex challenges at scale. As a Visa Software... ...meaningfully to planning and collaborate with PMs and team leads to shape team vision. Experience in one or more of the high-level...SeniorFull timeWork at officeLocal areaWork visa$125.4k - $181.88k
...Empower Yourself. The Senior Private Wealth Advisor, Practice Lead plays a vital role in helping high net worth clients achieve long... ...For remote and hybrid positions you will be required to provide reliable high-speed internet with a wired connection as well as a place in...Senior16 hoursTemporary workWork experience placementCasual workPrivate practiceWork at officeLocal areaRemote workWork from homeFlexible hours$225k - $235k
...A global engineering firm is seeking a Principal Project Manager in Seattle to lead rolling stock engineering consulting teams. The role requires 10+ years of relevant experience, strong project management skills, and a Bachelor's degree in Mechanical or Electrical Engineering...SeniorFull time$150k - $175k
...exploit the full potential of the cloud. As an Amazon Connect Lead / Senior Delivery Consultant, you will collaborate with Storm Reply... ...environments Demonstrated Amazon Connect solution design and engineering expertise Demonstrated DevSecOps experience on AWS with AWS...SeniorFull time- ...Principal Site Reliability Engineer (IC4) As a Principal Site Reliability Engineer (IC4), you will be responsible for designing, building, and... ...across large-scale distributed systems. You will lead complex reliability initiatives, drive automation-first operational...
$133.2k - $219.6k
...Site Reliability Engineer of Container Service Direct message the job poster from Alibaba Cloud Global Talent Acquisition Talent Sourcer Job Description: As a leader in Kubernetes management, the Alibaba Cloud Container Service team is dedicated to delivering...Full time$147k - $202k
...and we are looking for an Observability Engineer to help ensure that our Product and... ...stability. If you have experience within the Site Reliability Engineering (SRE) field or working as a... ...junior engineers. Proven ability to lead cross-functional technical initiatives...SeniorFull timeLocal areaWorldwideFlexible hours$250k - $315k
...customers, and are trusted by the world's leading companies. We build innovative solutions... ...the Role As part of the Infrastructure Engineering team and a people leader, you will ensure the health, stability, and reliability of Rokt's critical cloud infrastructure...Full timeWork at office$173.5k - $234.7k
...and real-time sync mechanisms that deliver the right data to each consumer within strict latency and reliability targets. The role requires applying seasoned engineering judgment to solve complex systems challenges and to make recommendations on architecture, technology...Full timeTemporary workPart timeWork experience placementLocal areaFlexible hours$248.83k - $348.36k
...is seeking a Senior Software Development Engineer to join the Software Engineering team supporting... ...networking performance. You will shape reliable, high-performance systems that operate... ...is $248,830.00 - $348,361.65 Other site ranges may differ Culture Statement Don...SeniorPermanent employmentFull timeTemporary workLocal areaWorldwide$126.08k - $189.12k
...timeliness Education Qualification For roles outside USA: Bachelor's Degree in Computer Science or “STEM” Majors (Science, Technology, Engineering and Math) with advanced experience. For roles in USA:Bachelor's Degree in Computer Science or “STEM” Majors (Science, Technology...SeniorOdd jobFull timeContract workImmediate startVisa sponsorshipWork visaRelocation package$190k - $210k
...channels. Our flagship brand, Micro Ingredients, is a leading supplement brand on Amazon, supported by a broad... ...more scalable, and more profitable marketplace growth engine. About the Role We are seeking a Director / Sr. Director, Amazon & Marketplace Growth to lead the...SeniorFull timeFreelanceWork at officeLocal areaRemote workShift work$172.5k - $313.7k
...duplicating efforts. Job Category Software Engineering Job Details About Salesforce... ...to level-up your career at the company leading workforce transformation in the agentic... ...enables AI to operate accurately and reliably. * Critically evaluate code (human or...Full time$139k - $257.55k
The Opportunity We’re looking for an iOS Engineer with strong programming fundamentals and keen attention to detail to help shape our flagship... ...while contributing to overall software architecture. • Lead code reviews and provide constructive feedback to team members....SeniorFull timeTemporary workLocal areaWorldwide- ...Site Reliability Engineer Comtech is a woman-owned small business founded in 1998 and headquartered in Reston, VA. We offer IT solutions across... ...ISMS), and CMMI-DEV Level 3. Job Description Job Title: Sr. Site Reliability Engineer Duration: Long Term These...
- ...in every community we are in. About this team Site Reliability Engineering We are looking for a motivated engineer to join the Foundations... ..., and creates the space for others to do the same. Leads with courage, knowing the possibility of greatness is...
$100k - $170k
...who wants to own systems, not just watch them. You'll take real surface area: the automation and tooling other engineers depend on, and the reliability of production services running AI and GPU workloads at scale. You'll sit in the incident rotation, and you'll...Flexible hoursShift work$264.1k - $369.74k
Sr Principal Software Engineering - Enterprise Technology page is loaded## Sr Principal Software Engineering... ...experienced Sr Principal Software Engineer to lead the technical strategy for Blue... ...is $264,103.00 - $369,743.85**Other site ranges may differ****Culture...SeniorPermanent employmentTemporary workLocal areaRelocation- ...Position Overview SingleStore is seeking a Site Reliability Engineer to help optimize and scale our managed service offering across all three... .... In this role, you will be at the intersection of leading technology trends - A highly performant distributed database...Worldwide
$65 per hour
...SRE Engineer in Redmond,WA Lateral - 120-130K per annum Subcon- 65$ per hr Microsoft Responsibilities... ...solutions Monitor and ensure service uptime, availability, reliability, and latency Track and integrate SRE metrics with...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr Lead Site Reliability Engineer. Be the first to apply!
- lead engineer Bellevue, WA
- lead operating engineer Bellevue, WA
- senior hvac project manager Bellevue, WA
- senior technical product manager Bellevue, WA
- senior medical science liaison Bellevue, WA
- senior accountant remote Bellevue, WA
- senior marketing account manager Bellevue, WA
- senior robotics software engineer Bellevue, WA
- sr project manager Bellevue, WA
- senior dynamics crm developer Bellevue, WA








