Sr Lead Site Reliability Engineer
$132.23k - $176.31kLumen
Lumen is the trusted network for the AI‑powered world, connecting people, data, and applications through our expansive fiber network and connected ecosystem. We enable secure, high‑performance connectivity across cloud, edge, and AI workloads for enterprises, governments, and communities.
At Lumen, you’ll work on infrastructure customers rely on today and build for what’s next, where performance, security, and resilience matter.
This is a high accountability environment where bold ideas drive real innovation for our customers, partners, and industry. The work is challenging, expectations are clear, and trust is built into how we operate. If you’re ready to take ownership, deliver meaningful impact, and help shape the future of AI‑ready connectivity, join us today.
The Role
We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role is critical to ensuring the reliability, scalability, and efficiency of our systems, with a strong emphasis on AWS infrastructure, observability, automation and AI-assisted engineering practices.
The Senior Lead SRE requires an AI-native mindset, understands the software development lifecycle (from coding to support) and applies modern AI tools to enhance productivity, quality, and operational excellence. This role will shape how Lumen combines the latest technologies, including AI-driven automation, to modernize software delivery and application lifecycle management.
This role will collaborate with key stakeholders across the engineering organization — including product owners, developers, and testers — to design, optimize, and automate business and technical processes, while effectively navigating multiple teams within a large and complex organization.
Location
This role is designated as a fully remote position within the United States.
The Main Responsibilities
Production Support & Incident Management
- Help define and improve the processes and best practices for incident management within the customer facing applications.
- Implement AI systems and automations to assist during ongoing outages and triage potential ones. You will work with the development teams to ensure that they have all the data normally needed during an outage at their fingertips including preliminary analysis by AI.
- Design and implement improved SRE processes to handle incidents. From proactive analysis, faster response, automatic remediation, AI guided analysis, partially automatic root cause analysis.
Performance Optimization
- Monitor system performance and proactively identify bottlenecks or degradation using AI-driven observability and anomaly detection tools.
- Implement tuning strategies across application layers, databases, and infrastructure.
- Drive initiatives to improve latency, throughput, and resource utilization.
Monitoring & Observability
- Deploy improved alerting for Lumen Connect in depth, focusing on outside in but including early indicators for fulfilment and other areas. Combining traditional monitoring with AI-based anomaly detection and noise reduction.
- Proactively monitor the errors and performance on Lumen Connect. Implement rules to detect deviations, implement improvements together with the teams.
- Design and maintain dashboards, alerts, and metrics using tools like Datadog, AppInsights, CloudWatch, or similar.
Automation & Infrastructure as Code
- Develop and maintain automation scripts and tools for deployment, scaling, and recovery, leveraging AI-assisted code generation and validation tools
- Use Terraform, or similar IaC tools to manage AWS resources.
Reliability Engineering
- Perform an in-depth analysis of the overall system and its dependencies, implementing techniques to increase the global availability, reduce the reliance on unstable dependencies and guide ecosystem improvements.
- Champion SRE principles such as SLIs, SLOs, and error budgets.
- Advocate for resilient architecture and fault-tolerant design patterns, incorporating AI-assisted design reviews and architecture evaluation.
- Help define and improve better SRE processes and lead significant improvements in reliability for the Lumen Connect platform.
Collaboration & Communication
- Work closely with software engineers, DevOps, and product teams to align reliability goals.
- Document processes, runbooks, and best practices for knowledge sharing.
- Provide mentorship and guidance on reliability and operational excellence.
What We Look For in a Candidate
Required Qualifications:
- 10+ years overall professional experience in SRE, DevOps, or infrastructure engineering roles.
- Experience with Terraform, or similar IaC tools to manage Cloud resources.
- Proficiency in scripting languages (Python, Bash, etc.) and automation frameworks.
- Experience with CI/CD pipelines and tools like GitHub Actions, Jenkins or GitLab CI.
- Solid understanding of monitoring and logging tools (e.g., CloudWatch, ELK, Datadog).
- Familiarity with containerization and orchestration (Docker, Kubernetes).
- Excellent AI and problem-solving skills, and a proactive mindset.
Preferred Qualifications:
- Experience in AWS services (EC2, CloudFront, EKS, RDS, S3, etc.).
- Certifications in AWS or related technologies are a plus.
- Experience of application development using Java Microservices and Spring Boot framework
- Experience with Agile/SCRUM Methodologies and development practices
Compensation
This information reflects the anticipated base salary range for this position based on current national data. Minimums and maximums may vary based on location. Individual pay is based on skills, experience and other relevant factors.
Location Based Pay Ranges
$132,232 - $176,310 in these states: AL AR AZ FL GA IA ID IN KS KY LA ME MO MS MT ND NE NM OH OK PA SC SD TN UT VT WI WV WY $138,844 - $185,124 in these states: CO HI MI MN NC NH NV OR RI $145,456 - $193,940 in these states: AK CA CT DC DE IL MA MD NJ NY TX VA WA
Lumen offers a comprehensive package featuring a broad range of Health, Life, Voluntary Lifestyle benefits and other perks that enhance your physical, mental, emotional and financial wellbeing. We're able to answer any additional questions you may have about our bonus structure (short-term incentives, long-term incentives and/or sales compensation) as you move through the selection process.
Learn more about Lumen's:
LI-Remote
LI-VK1
Requisition #: 342707
Life at Lumen
Life at Lumen is human and connected, even in a fast moving, AI‑focused organization. We set clear expectations and trust people to meet them. With real support and shared accountability, teams collaborate better, move faster, and deliver meaningful outcomes.
Our Lumen 8 behaviors guide how we interact, make decisions, and work together, shaping a culture built to perform and win.
To learn more about Life at Lumen and how we live the Lumen 8, please visit:
Background Screening
If you are selected for a position, there will be a background screen, which may include checks for criminal records and/or motor vehicle reports and/or drug screening, depending on the position requirements. For more information on these checks, please refer to the Post Offer section of our FAQ page . Job-related concerns identified during the background screening may disqualify you from the new position or your current role. Background results will be evaluated on a case-by-case basis.
Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.
Equal Employment Opportunities
We are committed to providing equal employment opportunities to all persons regardless of race, color, ancestry, citizenship, national origin, religion, veteran status, disability, genetic characteristic or information, age, gender, sexual orientation, gender identity, gender expression, marital status, family status, pregnancy, or other legally protected status (collectively, “protected statuses”). We do not tolerate unlawful discrimination in any employment decisions, including recruiting, hiring, compensation, promotion, benefits, discipline, termination, job assignments or training.
Privacy Notice
Lumen is committed to protecting the privacy and security of personal information collected during the recruitment and hiring process. Our Applicant Privacy Notice explains how we collect, use, disclose, and protect applicant information, as well as how individuals may request access to or deletion of their personal data.
To review Lumen’s Global Employment Applicant and Talent Community Privacy Notice, please visit:
Disclaimer
The job responsibilities described above indicate the general nature and level of work performed by employees within this classification. It is not intended to include a comprehensive inventory of all duties and responsibilities for this job. Job duties and responsibilities are subject to change based on evolving business needs and conditions.
In any materials you submit, you may redact or remove age-identifying information such as age, date of birth, or dates of school attendance or graduation. You will not be penalized for redacting or removing this information.
Please be advised that Lumen does not require any form of payment from job applicants during the recruitment process. All legitimate job openings will be posted on our official website or communicated through official company email addresses. If you encounter any job offers that request payment in exchange for employment at Lumen, they are not for employment with us, but may relate to another company with a similar name.
$134.25k - $214.8k
...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance...SeniorWork experience placementWork at officeRemote workFlexible hours- ...Sr. Site Reliability Engineer Comtech is a woman-owned small business founded in 1998 and headquartered in Reston, VA. We offer IT solutions across the disciplines of program/project management, applications development, infrastructure, Cyber security, and enterprise...SeniorLocal area
$120k - $170k
Sr. Manager/Manager Site Reliability Engineering Join to apply for the Sr. Manager/Manager Site Reliability Engineering role at Aritzia Sr. Manager/Manager Site... ...be part of the Quality and Service Delivery team and lead the SRE team responsible for continuously improving...SeniorFull timeWork at officeRemote workFlexible hours- ...AECOM office, remote location or at a client site, you will be working in a dynamic environment... ...global network of experts - planners, designers, engineers, scientists, consultants, program and construction managers - leading the change toward a more sustainable and...SeniorFor contractorsWork at officeLocal areaRemote workFlexible hours
$118.2k - $160k
The ESS Enablement Lead is a newly created, centralized role within the AWS Events, Sports, and Sponsorship org responsible for owning all pre- and post-event enablement plans across the full ESS portfolio. This role serves as the marketing integrator between ESS, Sales...SeniorFlexible hours$170k - $220k
...We're Looking ForWe’re looking for a hands-on, high-agency Site Reliability Engineer to help shape and scale the reliability layer of our stack.... ...shipping process.You’ll work closely with engineers, product leads, and company leadership to ensure uptime, speed, and...Senior$166k - $258k
...considered for this position.Nordstrom is looking for a Senior Engineer 2 to join our Site Reliability Engineering (SRE) team — and we think that person could... .... You'll set and maintain high operational standards, lead key projects from concept to execution, and bring fresh...SeniorFull timeWork at office$147k - $202.4k
...let's talk.Our company is seeking a highly skilled Senior Site Reliability Engineer to join our team. We are a SaaS company specializing in securing... ...and be a primary responder for critical incidents, leading root cause analysis and implementing preventative measures...SeniorWork at officeLocal areaWorldwideFlexible hoursShift work$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$134.25k - $214.8k
...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed... ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,...SeniorWork experience placementWork at officeRemote work$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build,...SeniorPermanent employmentLocal areaWorldwideFlexible hours$167.7k - $245.2k
...also delivering AI-powered assurance insights within Cisco’s leading Networking, Security, Collaboration, and Observability... ...maintaining our FedRAMP offering. Your ImpactAs a FedRAMP Site Reliability Engineer(SRE), you will lead the operations and architecture of our...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours2 days per week$151.2k - $204.6k
Amazon Prime Air is seeking a Sr. Systems Development Engineer to serve as our internal DO-178C Designated Engineering Representative (DER) and certification... ...processes that satisfy SOI-1 through executing and leading SOI-2, SOI-3, SOI-4, and Conformity Audits.Your impact...SeniorFlexible hours$127k - $249k
We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure... ...Responsibilities:Cloud Security Design and Implementation: Help lead the design and deployment of security solutions for cloud...SeniorLocal areaRemote workWorldwideFlexible hours- AECOM is seeking a Senior Drainage Lead to be based in Seattle, WA . Job Summary/Responsibilities... ...during project implementation. Conduct site visits and field investigations to assess... ...- from advisory, planning, design and engineering to program and construction management....SeniorFor contractorsLocal area
$160k - $250k
...DevOps And Systems Engineer Hive is the leading provider of cloud-based AI solutions to understand, search, and generate content, and is trusted... ...learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS...Senior- ...team is responsible for the reliability, scalability, and efficiency... ...building features, but about engineering the resilience and performance... ...maintain system stability.As a Site Reliability Engineer, you... ...Capacity and cost optimization: Lead initiatives in capacity planning...Senior
$150k - $175k
...them exploit the full potential of the cloud.As an Amazon Connect Lead / Senior Delivery Consultant, you will collaborate with Storm... ...environments-Demonstrated Amazon Connect solution design and engineering expertise-Demonstrated DevSecOps experience on AWS with AWS and...SeniorFull time$232k - $319k
...service with great people and reliable, cost-effective, and efficient... ...processes, and tooling. As the Sr. Manager of Infrastructure... ...tooling. What you’ll be doing Lead the Infra platform and shared... ...the velocity of SRE and product engineering by developing robust platforms...SeniorPermanent employmentLocal areaWorldwideFlexible hours- ...Participate in on-call rotations, respond to critical incidents, lead root cause analysis, and implement preventative measures.... ...AI/ML experience or a strong interest in applying AI/ML to reliability, security, or operational efficiency is a plus. Benefits...SeniorFull timeWork at officeFlexible hours
- ...of the cloud footprint Evaluate AWS compute, networking, and security architecture for scalability and growth Collaborate across engineering and product teams to scope projects to core business requirements Drive engineering-wide improvements to service management for deployments...Senior
$160k - $210k
...change and achieving remarkable growth in a rapidly evolving industry. Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management across Cognitiv. Our immediate challenge is to scale...SeniorWork at officeImmediate startRemote workWork from home$210k - $256.67k
...team, and you’ll be able to see your own ideas transform into breakthrough results.The RoleWe are looking for Program Director, to lead a large & complex multi pillar Oracle fusion cloud transformation program for a North America based customer with proven experience...SeniorFull timeTemporary work$143k - $191k
...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental... ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and...Full timeTemporary workWork experience placementImmediate start$117.2k - $313.7k
...up your career at the company leading workforce transformation in... ...Distributed Systems Software Engineer - Public Cloud (Senior/Lead/Principal... ...on our platform to be highly reliable, lightning fast, supremely... ...experience balancing live-site management, feature delivery,...SeniorFull time- ## Sr. Client Lead Recruiter, Space VehiclesApplylocations: Greater Seattle Area: Space Coast, FL: Huntsville, AL: Denver, COtime type: Full... ...$173,054.70WA applicants is $134,434.00 - $188,207.25**Other site ranges may differ****Culture Statement**Don’t meet all desired...SeniorPermanent employmentTemporary workLocal areaRelocation
$158k - $217k
Snowflake Partner Solution Lead - West RegionAs a Snowflake Partner Solution Lead, you will take the helm of one of Slalom’s most critical technology alliances. Representing a Snowflake Elite Services Partner and 7-time Snowflake Partner of the Year, you will act as a...SeniorTemporary workLocal area- ...will build the product team, partner with engineering and sales, and shape how customers... ...next-gen cloud services at scale. You’ll lead a cross-functional, data-driven effort to... ...solutions, balance ambitious innovation with reliability, and drive business outcomes through...Senior
$172.5k - $260.1k
...the heart of it all.Ready to level-up your career at the company leading workforce transformation in the agentic era? You’re in the... ...behind the Headless 360 Data Foundation Platform — a predictive P&L engine replacing static spreadsheet analysis with customer-level P&L...SeniorFull time$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ...servicesIncident Response: Participate in on-call rotations and lead post-incident reviews to drive systemic improvements and "...Permanent employmentWork at officeLocal areaWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr Lead Site Reliability Engineer. Be the first to apply!
- lead network engineer Seattle, WA
- lead infrastructure engineer Seattle, WA
- lead operating engineer Seattle, WA
- lead engineer Seattle, WA
- site reliability engineer Seattle, WA
- site reliability engineer remote Seattle, WA
- site reliability engineer sre Seattle, WA
- senior technical analyst Seattle, WA
- senior associate attorney Seattle, WA
- senior developer Seattle, WA


