Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead Site Reliability Engineer

$105.79k - $141.05k

Lumen Inc

Lumen is the trusted network for the AI-powered world, connecting people, data, and applications through our expansive fiber network and connected ecosystem. We enable secure, high-performance connectivity across cloud, edge, and AI workloads for enterprises, governments, and communities.

At Lumen, you'll work on infrastructure customers rely on today and build for what's next, where performance, security, and resilience matter.

This is a high accountability environment where bold ideas drive real innovation for our customers, partners, and industry. The work is challenging, expectations are clear, and trust is built into how we operate. If you're ready to take ownership, deliver meaningful impact, and help shape the future of AI-ready connectivity, join us today.

The Role

Lumen's Network as a Service (NaaS) platform delivers on-demand networking at scale. As Lead SRE, you'll own the reliability of that platform - partnering with operations teams and development counterparts to drive technical direction and resolve systemic issues across a broad range of network topologies and applications.

You'll be accountable for platform observability, incident management, and automation, and you'll coordinate across architecture, engineering, and systems development organizations to measurably improve reliability. You'll also use AI and agentic tooling to build utilities that accelerate deployment automation, platform administration, and incident investigation.

Success in this role draws on networking fundamentals, cloud platforms, software development and troubleshooting methodology, and a bias toward automating what you'd otherwise do twice. We're looking for a change maker - someone who sees where the platform should go next and drives meaningful impact for the customers who rely on it

Location

This role is designated as a fully remote position within the United States.

The Main Responsibilities

  • Reliability & Observability

  • Serve as subject matter expert for network automation platform applications, services, and hosting environments

  • Build and maintain the observability stack: instrument services, collect and curate metrics, and create dashboards and visualizations that make system health obvious at a glance

  • Define and tune proactive alerting so issues surface before customers feel them

  • Champion core SRE principles - SLIs, SLOs, and error budgets - and advocate for resilient, fault tolerant architecture

  • Incident Management

  • Participate in an on-call rotation and lead incident response for service outages and unplanned downtime

  • Drive blameless postmortems and root cause analysis; own follow-up actions through to completion

  • Prevent recurrence through process improvements, tooling, and knowledge sharing across teams

  • Automation & Infrastructure

  • Automate deployment pipelines (CI/CD) and cloud infrastructure provisioning, scaling, and configuration using infrastructure as code

  • Develop tools and utilities that reduce toil and empower operations and development teams to manage services independently

  • Apply AI-assisted and agentic workflows to development, support, and investigation work

  • Collaboration & Leadership

  • Collaborate with cross-functional development teams to support, enhance, and scale NaaS applications

  • Provide guidance and mentorship to junior engineers

  • Maintain clear documentation for processes and architecture

What We Look For in a Candidate

Required Qualifications:

  • Bachelor's degree or equivalent in engineering, computer science, or related field.

  • 8+ years in software development, systems engineering, and/or networking

  • 5+ years of related experience required.

  • Hands-on experience with at least one major cloud platform (AWS, Azure, or GCP), including compute, networking, and identity services

  • Strong automation and infrastructure-as-code skills: Terraform, Ansible, and Python

  • Experience running containerized workloads on Kubernetes

  • Working knowledge of modern observability and monitoring tooling (e.g., Datadog, CloudWatch, Grafana, Prometheus), including building dashboards and defining alerts

  • Demonstrated experience with incident management and blameless postmortems

  • Comfort using AI-assisted development and agentic tools as part of daily engineering practice

  • Understanding of network technologies including Internet, Ethernet, IPVPN, Edge Compute, and Optical transport

  • Strong listening and communication skills; able to operate with autonomy while knowing when to escalate

Preferred Qualifications:

  • Multi-cloud experience across AWS, Azure, and GCP

  • Asynchronous programming concepts and distributed systems design

  • Zero-downtime deployment strategies

  • High availability and multi-region architectures

  • Source control and CI/CD practices at scale

  • Experience applying agentic workflows to operational support and investigation

Compensation

This information reflects the anticipated base salary range for this position based on current national data. Minimums and maximums may vary based on location. Individual pay is based on skills, experience and other relevant factors.

Location Based Pay Ranges

$105,786 - $141,047 in these states: AL AR AZ FL GA IA ID IN KS KY LA ME MO MS MT ND NE NM OH OK PA SC SD TN UT VT WI WV WY

$111,074 - $148,099 in these states: CO HI MI MN NC NH NV OR RI

$116,364 - $155,152 in these states: AK CA CT DC DE IL MA MD NJ NY TX VA WA

Lumen offers a comprehensive package featuring a broad range of Health, Life, Voluntary Lifestyle benefits and other perks that enhance your physical, mental, emotional and financial wellbeing. We're able to answer any additional questions you may have about our bonus structure (short-term incentives, long-term incentives and/or sales compensation) as you move through the selection process.

Learn more about Lumen's:

Benefits (

Bonus Structure

#LI-Remote

#LI-VK1

Requisition #: 343429

Life at Lumen

Life at Lumen is human and connected, even in a fast moving, AI-focused organization. We set clear expectations and trust people to meet them. With real support and shared accountability, teams collaborate better, move faster, and deliver meaningful outcomes.

Our Lumen 8 behaviors guide how we interact, make decisions, and work together, shaping a culture built to perform and win.

To learn more about Life at Lumen and how we live the Lumen 8, please visit:

Background Screening

If you are selected for a position, there will be a background screen, which may include checks for criminal records and/or motor vehicle reports and/or drug screening, depending on the position requirements. For more information on these checks, please refer to the Post Offer section of our FAQ page ( . Job-related concerns identified during the background screening may disqualify you from the new position or your current role. Background results will be evaluated on a case-by-case basis.

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Equal Employment Opportunities

We are committed to providing equal employment opportunities to all persons regardless of race, color, ancestry, citizenship, national origin, religion, veteran status, disability, genetic characteristic or information, age, gender, sexual orientation, gender identity, gender expression, marital status, family status, pregnancy, or other legally protected status (collectively, "protected statuses"). We do not tolerate unlawful discrimination in any employment decisions, including recruiting, hiring, compensation, promotion, benefits, discipline, termination, job assignments or training.

Privacy Notice

Lumen is committed to protecting the privacy and security of personal information collected during the recruitment and hiring process. Our Applicant Privacy Notice explains how we collect, use, disclose, and protect applicant information, as well as how individuals may request access to or deletion of their personal data.

To review Lumen's Global Employment Applicant and Talent Community Privacy Notice, please visit:

Disclaimer

The job responsibilities described above indicate the general nature and level of work performed by employees within this classification. It is not intended to include a comprehensive inventory of all duties and responsibilities for this job. Job duties and responsibilities are subject to change based on evolving business needs and conditions.

In any materials you submit, you may redact or remove age-identifying information such as age, date of birth, or dates of school attendance or graduation. You will not be penalized for redacting or removing this information.

Please be advised that Lumen does not require any form of payment from job applicants during the recruitment process. All legitimate job openings will be posted on our official website or communicated through official company email addresses. If you encounter any job offers that request payment in exchange for employment at Lumen, they are not for employment with us, but may relate to another company with a similar name.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Lead Site Reliability Engineer in Boston, MA vacancy
  • $130k - $150k

     ...industry experts, and academics. At CRA you will be exposed to leading minds who use economic, financial, and business analysis to...  ...is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are reliable... 
    Suggested
    Work at office
    Work from home
    3 days per week

    CRA International

    Boston, MA
    1 day ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Suggested
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    4 days ago
  • $134.25k - $214.8k

     ...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed...  ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,... 
    Suggested
    Work experience placement
    Work at office
    Remote work

    Axon

    Boston, MA
    9 hours ago
  • $160k - $200k

     ...equip their workforce with composable, connected apps, leading to higher quality work, improved efficiency, and end-to...  ...evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage &... 
    Suggested
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interface

    Somerville, MA
    3 days ago
  • $90k - $110k

    SS&C is a leading provider of mission-critical, AI-powered technology and services empowering...  ..., and technology.Job DescriptionSite Reliability EngineerLocations: Boston/Waltham, MA |...  ...for the position ofSite Reliability Engineer. This role is based out of one of our Boston... 
    Suggested
    Ongoing contract
    Full time
    Casual work
    Work at office
    Worldwide
    Flexible hours

    SS&C Technologies

    Boston, MA
    2 days ago
  • $134.25k - $214.8k

     ...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance... 
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Axon

    Boston, MA
    2 days ago
  •  ...and best in class outcomesVisionary in future focused problem-solvingExceptional in execution and impactThe RoleAs a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to deliver... 
    Full time
    Flexible hours

    Proofpoint

    Boston, MA
    2 days ago
  • $55k - $151.47k

     ...LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in...  ...architecture to support data integrity and accessibility- Leading incident management and resolution efforts to maintain operational... 
    Full time
    H1b

    PwC

    Boston, MA
    22 hours ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that...  ...maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering... 
    Local area
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    4 days ago
  • $185.5k - $232k

     ...Senior Site Reliability Engineer New York, NY; Boston, MA; San Francisco, CA About Formation Bio Formation Bio is a tech and AI driven pharma company differentiated by radically more efficient drug development. Advancements in AI and drug discovery are creating... 
    Work experience placement
    Work at office
    Local area
    Relocation
    3 days per week

    Formation Bio (Formerly TrailSpark)

    Boston, MA
    20 hours ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role...  ...future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives... 
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Boston, MA
    3 days ago
  • $168k - $200k

     ...is passionate about creating transformative change in healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the forefront of building and operating a resilient, observable, and... 
    Remote work

    Datavant

    Boston, MA
    1 day ago
  • $160k - $200k

     ...equip their workforce with composable, connected apps, leading to higher quality work, improved efficiency, and end-to...  ...evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's... 
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interfaces

    Somerville, MA
    2 days ago
  • $160k - $200k

    Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware... 
    Local area
    Remote work

    QuEra Computing

    Boston, MA
    3 days ago
  •  ...), and Check Services. We are currently leading a strategic effort to transform FRFS to...  ...our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team,...  ...who loves building and maintaining reliable and scalable systems, CI/CD tooling, and... 
    Full time

    Federal Reserve Bank of Boston

    Boston, MA
    5 days ago
  •  ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed... 

    Watershed Informatics

    Cambridge, MA
    4 days ago
  • $95k - $171k

     ...infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for:...  ...Akamai powers and protects life online. Leading companies worldwide choose Akamai to... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Cambridge, MA
    5 days ago
  • $140k - $210.9k

     ...Check Services. We are currently leading a strategic effort to...  ...position will be primarily on-site with residency commutable to...  ...DevOps backgrounds or software engineering backgrounds (e.g., Java Python...  ...interest in operating and improving reliability of distributed production... 
    Full time
    Temporary work
    Part time
    Work at office
    Shift work

    Federal Reserve Bank

    Boston, MA
    3 days ago
  • $121.4k - $218.6k

     ...for ensuring best-in-class uptime and reliability of our AI hardware infrastructure...  ...when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing...  ...building technical runbooks, leading complex incident response bridges, and... 
    Work experience placement
    Work at office

    Akamai

    Boston, MA
    5 days ago
  • $160k - $225k

     ...Staff Site Reliability Engineer Cambridge, MA Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients...  ...Manifold to operate faster and more effectively. Backed by leading investors including Reach Capital, TQ Ventures, Calibrate... 

    Manifold

    Cambridge, MA
    1 day ago
  • $200k - $250k

     ...The Crown Is Yours As a Principal Site Reliability Engie r , you'll shape the long-term...  ...cloud and on-premise platforms, helping engineering teams build, deploy, and operate highly...  ...capacity planning, and cost optimization. Lead large-scale platform initiatives across... 
    Full time
    Immediate start

    DraftKings

    Boston, MA
    4 days ago
  • $169.3k - $304.7k

     ...maintaining fast, efficient, scalable, and reliable routing software and infrastructure...  ...global platform. As a Principal Site Reliability Engineer - Network, you will be responsible...  ...Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings... 
    Work experience placement
    Work at office

    Akamai

    Boston, MA
    3 days ago
  • $148k - $185k

     ...at a time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure Engineering organization. You'll define how we measure... 
    Full time
    Immediate start

    DraftKings

    Boston, MA
    5 days ago
  •  ...Technology group delivers secure, reliable technology solutions that...  ...a Senior Application Support Engineer, you will help power DTCC's global...  ...and settlement.Leveraging Site Reliability Engineering (SRE)...  ...mission-critical applications.Lead incident response, troubleshooting... 
    Remote work
    Flexible hours

    DTCC- The Depository Trust & Clearing Corporation

    Boston, MA
    22 hours ago
  • $119k - $170k

     ...impact at the company pioneering security transformation in the AI era? Join us at Zscaler.RoleWe are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler... 
    Full time
    Work at office
    Local area
    Remote work
    Shift work
    3 days per week

    Zscaler

    Boston, MA
    22 hours ago
  • $118.3k - $147.9k

     ...CMT is looking for a Senior Site Reliability Engineer, SecOps to help us change the world. CMT has helped protect over 65 million drivers and...  ...Fusion®, proactively identifies and reduces driving risk, leading to fewer crashes and injuries. To date, CMT’s technology has... 
    Full time
    Temporary work
    Summer work
    Work from home
    Worldwide
    Flexible hours

    Cambridge Mobile Telematics

    Cambridge, MA
    1 day ago
  • $146.25k - $225k

     ...to solving complex problems and making a huge impact. We are expanding our team and recruiting for a skilled Senior/Staff Site Reliability Engineer focused on designing, building, and operating our on-prem/cloud environment.The OpportunityYou will advance the state of our... 
    Local area
    Remote work
    Relocation package

    PathAI

    Boston, MA
    9 hours ago
  •  ...reviews across the platform. Support platform architecture decisions and technology standardization under the guidance of senior engineers. Help design and build a Kubernetes-based platform that automates core infrastructure lifecycle management. Collaborate with... 
    Full time
    Work at office
    Remote work

    Axon

    Boston, MA
    2 days ago
  • $220k - $290k

     ...Backed by some of the world's leading investors prior to its public...  .... You know what it takes to engineer the data pipeline that makes...  ...and delivery — ensuring the reliability and scale that training and simulation...  ....This position is based on-site at Merlin HQ in Boston, MA.... 
    Full time

    Merlin Labs

    Boston, MA
    2 days ago
  •  ...will build on it.The roleWe are seeking a Developer Relations Lead to help QuEra’s QPU and software ecosystem become the platform...  ...growing a developer community of advanced quantum scientists and engineers, while helping users move from first contact with QuEra’s tools... 

    QuEra Computing

    Boston, MA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead Site Reliability Engineer. Be the first to apply!