Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead Site Reliability Engineer

$198.05k - $267.95k

Job Description

At Boeing, we innovate and collaborate to make the world a better place. We're committed to fostering an environment for every teammate that's welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us.

The Boeing Company is looking for a Lead Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO.

We are seeking a highly talented, motivated, and creative technical leader responsible for the reliability strategy, architecture, operational maturity, and long-term technical direction of mission-critical developer platforms used by Air Dominance engineering teams.

This role will provide technical leadership across GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, related software delivery tools such as Artifactory and SonarQube, and the supporting infrastructure, automation, monitoring, backup, recovery, and security controls required to operate these services. The selected candidate will define standards, guide architecture decisions, mentor engineers, lead complex technical investigations, and partner with program leadership and stakeholders to ensure developer tooling remains secure, reliable, scalable, and supportable.

Position Responsibilities:

  • Define and lead the Site Reliability Engineering technical strategy for GitLab, CI/CD runners, Jira, Confluence, PostgreSQL, Artifactory, SonarQube, and related developer tooling infrastructure
  • Establish platform reliability architecture, operational standards, SLIs, SLOs, SLAs, KPIs, error budgets, observability patterns, capacity models, backup strategies, and disaster recovery approaches
  • Serve as the senior technical authority for complex reliability, performance, scalability, integration, database, automation, and security-related platform decisions
  • Lead architecture and design reviews for developer tooling infrastructure, CI/CD runner topology, PostgreSQL operations, cloud-based and on-premises infrastructure, monitoring, alerting, access controls, and platform integrations
  • Drive automation, Infrastructure as Code, Ansible, configuration management, and repeatable operational patterns that reduce toil and improve reliability
  • Guide major upgrades, migrations, lifecycle planning, patch strategies, recovery planning, and technical roadmaps for supported platforms
  • Lead the most complex incidents and technical investigations, including root cause analysis, corrective action planning, and systemic reliability improvements
  • Mentor and technically guide SREs in operational excellence, troubleshooting, automation, secure administration, and architectural thinking
  • Partner with program leadership, cybersecurity, infrastructure, software engineering, database, networking, suppliers, customers, and other stakeholders
  • Identify platform risks, technical debt, capacity constraints, single points of failure, compliance concerns, and operational gaps, then drive remediation plans
  • Define, collect, analyze, and refine software delivery and platform reliability metrics for team execution, management visibility, and continuous improvement
  • Lead standards for provisioning, platform scaling, configuration management, monitoring, troubleshooting, and software delivery tool integration
  • Develop and maintain architecture documentation, design patterns, standards, operational readiness criteria, and executive technical briefings
  • Influence support models, maintenance strategies, escalation paths, and investment priorities based on mission impact and operational risk
  • Lead efforts to operationally field higher-quality end-to-end system software more frequently
  • Participate in after-hours support and escalation for urgent or mission-impacting issues as required

This position is expected to be 100% onsite. The selected candidate will be required to work onsite at one of the listed location options.

Basic Qualifications (Required Skills/ Experience):

  • Active Secret U.S. Security Clearance. (A U.S. Security Clearance that has been active in the past 24 months is considered active.)
  • Ability to obtain access to Special Access Programs (SAP)
  • Bachelor's Degree
  • 14+ years of experience with DevOps, Site Reliability Engineering, software engineering, and/or cloud engineering
  • Experience with GitLab, Azure DevOps and CI/CD (Continuous Integration and Continuous Delivery (CI/CD)
  • Experience with technical leadership
  • Experience with designing and implementing scalable computing infrastructure for data solutions, including cloud architectures (AWS, Azure, Google Cloud)
  • 5 years of experience conducting root cause analysis of design failures

Preferred Qualifications (Desired Skills/Experience):

  • Bachelor of Science degree from an accredited course of study in engineering, engineering technology (includes manufacturing engineering technology), chemistry, physics, mathematics, data science, or computer science and 14+ years of related work experience or Bachelor's Degree and 18+ years of directly related work experience or 22+ years of related, relevant experience
  • Experience communicating technical strategy, risk, tradeoffs, and recommendations to senior technical and program leadership
  • Deep experience administering or architecting GitLab, GitLab CI/CD, GitLab runners, or comparable enterprise source control and CI/CD platforms
  • Deep experience administering or architecting Jira, Confluence, or other Atlassian products in an enterprise environment
  • Deep experience with PostgreSQL architecture and operations, including backup and recovery, replication, performance tuning, storage planning, maintenance, and high-availability patterns
  • Experience with AWS, Microsoft Azure, Infrastructure as Code, Ansible, configuration management, containers, Docker, Kubernetes, virtualization, artifact management, secrets management, and secure software delivery practices
  • Experience administering or architecting Artifactory, SonarQube, Jenkins, or similar software delivery tools
  • Experience designing observability platforms, alerting strategies, SLO frameworks, service health dashboards, and operational reporting
  • Experience supporting Air Dominance, classified, air-gapped, or highly regulated engineering environments
  • Experience developing disaster recovery strategy, continuity of operations plans, recovery time objectives, recovery point objectives, and restore validation programs
  • Experience guiding cybersecurity hardening, vulnerability remediation, audit readiness, privileged access controls, and compliance-driven operations
  • Ability to obtain Security+ certification
  • Demonstrated ability to lead through influence across engineering teams, customers, suppliers, cybersecurity, infrastructure, and program stakeholders
  • Strong written and verbal communication skills with the ability to produce architecture documentation, executive briefings, technical roadmaps, and decision records

Conflict of Interest:

Successful candidates for this job must satisfy the Company's Conflict of Interest (COI) assessment process.

Drug Free Workplace:

Boeing is a Drug Free Workplace where post offer applicants and employees are subject to testing for marijuana, cocaine, opioids, amphetamines, PCP, and alcohol when criteria is met as outlined in our policies.

CodeVue Coding Challenge:

To be considered for this position you will be required to complete a technical assessment as part of the selection process. Failure to complete the assessment will remove you from consideration.

Pay & Benefits:

At Boeing, we strive to deliver a Total Rewards package that will attract, engage and retain the top talent. Elements of the Total Rewards package include competitive base pay and variable compensation opportunities.

The Boeing Company also provides eligible employees with an opportunity to enroll in a variety of benefit programs, generally including health insurance, flexible spending accounts, health savings accounts, retirement savings plans, life and disability insurance programs, and a number of programs that provide for both paid and unpaid time away from work.

The specific programs and options available to any given employee may vary depending on eligibility factors such as geographic location, date of hire, and the applicability of collective bargaining agreements.

Pay is based upon candidate experience and qualifications, as well as market and business considerations.

Summary Pay Range: $198,050 - $267,950

Applications for this position will be accepted until Jul. 31, 2026

Export Control Requirements:

This position must meet U.S. export control compliance requirements. To meet U.S. export control compliance requirements, a "U.S. Person" as defined by 22 C.F.R. 120.62 is required. "U.S. Person" includes U.S. Citizen, U.S. National, lawful permanent resident, refugee, or asylee.

Export Control Details:

US based job, US Person required

Education

Bachelor's Degree or Equivalent Required

Relocation

This position offers relocation based on candidate eligibility.

Security Clearance

This position requires an active U.S. Secret Security Clearance (U.S. Citizenship Required). (A U.S . click apply for full job details

Vacancy posted 6 days ago
Similar jobs that could be interesting for youBased on the Lead Site Reliability Engineer in Berkeley, CA vacancy
  •  ...ambitious goals and attract incredibly creative scientists and engineers from leading academic institutions and from frontier AI labs. Neuro...  ...brain. Position Summary We are looking for a Site Reliability Engineer to own the digital infrastructure that powers... 
    Suggested
    Visa sponsorship

    Astera

    Emeryville, CA
    4 days ago
  • $174.92k - $209.91k

     ...access to data as simple and reliable as electricity. With Fivetran...  ...and ready to query, with no engineering or maintenance required. We’re...  ...together two industry-leading companies with a shared mission...  ...our teams, systems, and career sites. About the Role Fivetran... 
    Suggested
    Full time
    Work at office
    Remote work

    Fivetran

    Oakland, CA
    5 days ago
  • $27 - $30 per hour

     ...We are hiring immediately for a full‑time BOH Lead Supervisor. This position is based at 5000 MacArthur Boulevard, Oakland, CA 94613. Location & Schedule Location: Mills College. Full time schedule; days and hours may vary, primarily PM shifts. Open availability. Details... 
    Suggested
    Hourly pay
    Full time
    Temporary work
    Seasonal work
    Local area
    Immediate start
    Relocation
    Flexible hours
    Shift work

    Chartwells Higher Education Dining Services

    Oakland, CA
    2 days ago
  •  ...practices comply with company policies and procedures. Essential Duties and Responsibilities: Assists in ordering and keeping inventory of products. Maintains product cost and labor cost according to budge Lead Supervisor, Supervisor, Benefits, Retail, Insurance, Associate... 
    Suggested
    Hourly pay
    Full time

    Compass Group, PLC

    Oakland, CA
    3 days ago
  •  ...small team of former Google and Stripe engineers, including the founding team of Google...  ...looking for a skilled and passionate Site Reliability Engineer to join our team. As a SRE, you...  ...toil and build self-healing systems Lead incident response, conduct root-cause analysis... 
    Suggested
    Remote work
    1 day per week

    Runloop AI, Inc

    San Francisco, CA
    2 days ago
  • $160k - $250k

     ...public clouds when the right fit. As we continue to commercialize our machine learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS offering for our customers. Our ideal candidate is someone who is able... 

    Hive

    San Francisco, CA
    2 days ago
  • $170k - $230k

     ...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make...  ...GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community, including... 
    Work at office
    Local area
    1 day per week

    Mithril

    San Francisco, CA
    2 days ago
  • $86k - $105k

     ...generation of application infrastructure and to be responsible for reliability, automation and scalability using and the latest best...  ...certifications. Minimum of 2 years prior DevOps, software engineering or related experience. Must be able to work different schedules... 
    Hourly pay
    Work at office
    Immediate start
    Visa sponsorship
    Work visa
    Flexible hours

    Early Warning Services

    San Francisco, CA
    4 days ago
  • $98.58k - $138.02k

     ...Site Reliability Engineer II Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back-office operations for restaurants. Restaurant365's culture is focused on empowering... 
    Work at office

    Restaurant365

    San Francisco, CA
    2 days ago
  • $100k - $170k

     ...Site Reliability Engineer Houston; San Francisco; Seattle About Nscale Nscale is the GPU cloud built for AI. We run high-performance, cost-efficient infrastructure for AI-native startups and global enterprises, from bare metal up through the platform services... 
    Flexible hours
    Shift work

    Nscale

    San Francisco, CA
    2 days ago
  • $150k

     ...Site Reliability Engineer San Francisco, CA About The Role We are seeking an experienced Site Reliability Engineer (SRE) with a strong...  ...observability tooling (e.g., CloudWatch, Datadog, Grafana). Lead periodic infrastructure and dependency audits; produce... 

    VantageScore®

    San Francisco, CA
    3 days ago
  • $163.71k - $306k

     ...infrastructure, behind their own controls, with the reliability and operational clarity they would...  ...TAMs to trust. Partner with product engineers on infrastructure requirements for new...  ...when they introduce new dependencies Lead through ambiguity, make careful risk... 

    Retool

    San Francisco, CA
    3 days ago
  • $200k - $300k

     ...Site Reliability Engineer Title of Role: Site Reliability Engineer Location: San Francisco, onsite Company Stage of Funding: Venture Round - Healthcare, AI Office Type: Onsite Salary: $200K-$300K Company Description We're representing a dynamic... 
    Work at office

    Recruiting from Scratch

    San Francisco, CA
    2 days ago
  •  ...Site Reliability Engineer (SRE) FLUIX is building the AI operating system that plans, designs, and optimizes AI infrastructure. We are based in Silicon Valley. We specialize in providing AI-driven solutions for data centers and power providers, leveraging cutting-edge... 
    Work at office
    Weekend work

    Fluix AI

    San Francisco, CA
    2 days ago
  •  ...About the job Senior Site Reliability Engineer About the Company Stellar is a decentralized, public blockchain that gives developers the tools to create experiences that are more like cash than crypto. The network is faster, cheaper, and far more energy-efficient... 

    TechChain Talent

    San Francisco, CA
    7 days ago
  • $170k - $250k

     ...Site Reliability Engineer (SRE) Location: San Francisco, CA / Palo Alto, CA Company Stage of Funding: Growth-Stage AI Infrastructure Company ($80M Raised) Office Type: Onsite (4 Days Per Week) Salary: $170,000-$250,000 + Competitive Equity Company Description... 
    Work at office
    Visa sponsorship
    Flexible hours

    Recruiting from Scratch

    San Francisco, CA
    4 days ago
  •  ...Udaip Cloud-Based Data And Ai Platform Engineer At U.S. Bank, we're on a journey to do our best. Helping the customers and businesses...  ...the company. At high level, UDAIP success includes industry leading hybrid cloud infrastructure, innovative data capabilities with... 
    Temporary work
    Work experience placement

    Phenom People

    San Francisco, CA
    2 days ago
  • $204k - $281k

     ...all in on this mission. If you are too, let’s talk. Manager, Site Reliability Engineering San Francisco, California Okta authenticates, authorizes...  ...self‑service capabilities, and robust self‑healing patterns. Lead, mentor, and grow a high‑performing team of engineers and... 
    Permanent employment
    Worldwide
    Flexible hours

    Okta, Inc.

    San Francisco, CA
    2 days ago
  •  ...The Boeing Company is seeking a Systems Engineers to join one of our St. Louis, MO (Berkeley...  ...disciplines including, Experienced and Lead Levels:Systems Architecture,...  ...systemsPerform analyses in affordability, safety, reliability, maintainability, testability, human factors... 
    Work experience placement
    Interim role
    Currently hiring
    Immediate start
    Remote work
    Relocation
    Visa sponsorship
    Work visa
    Flexible hours

    Boeing

    Berkeley, CA
    1 day ago
  • $205k - $305k

     ...Director Of Site Reliability Engineering Interested in working on cutting-edge blockchain technology and creating equitable access to the global...  ...looking for a Director of Site Reliability Engineering to lead a small, high-leverage SRE team and help shape how engineering... 
    Temporary work
    Work at office
    Local area
    Worldwide
    Flexible hours

    Stellar

    San Francisco, CA
    12 days ago
  •  ...Staff Site Reliability Engineer (SRE) Location: San Francisco, CA Job Responsibilities As our Staff SRE, you'll be the primary expert...  ...across organizational boundaries. Design, implement, and lead large-scale, cross-functional projects to improve the... 

    United IT

    San Francisco, CA
    2 days ago
  • $175k - $250k

     ...Senior Cloud Infrastructure Engineer Location: San Francisco, CA....  ...Remote unavailable. Modality: On-Site only. Must live within...  ...this role, you will take the lead on designing, deploying, and...  ...scalability, performance, and reliability across environments. What You... 
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    San Francisco, CA
    4 days ago
  •  ...intersection of labor markets and AI research. We partner with leading AI labs and enterprises to provide the human intelligence...  ...our new San Francisco headquarters. About the Role As a Site Reliability Engineer (SRE) at Mercor, you’ll own production reliability across our... 

    Mercor

    San Francisco, CA
    4 days ago
  • $210k - $240k

    Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $210,000.00/yr - $2... 
    Full time

    Alembic Technologies

    San Francisco, CA
    4 days ago
  • $210.8k - $272.8k

    About Thumbtack Thumbtack helps millions of people confidently care for their homes. About the Site Reliability Engineering Team The Site Reliability Engineering team focuses on creating and maintaining a reliable, secure, and scalable platform vital for a seamless user... 
    Local area

    Thumbtack

    San Francisco, CA
    3 days ago
  • $205k - $235k

     ...have ambitious goals for the future.  As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to...  ...DevOps roles, preferably working as a senior engineer or tech lead. ~ Knowledge of HPC/batch compute frameworks (Slurm,... 
    Remote job
    Local area

    The Voleon Group

    Berkeley, CA
    more than 2 months ago
  • $160.65k - $217.35k

     ...future with us. The Boeing Company is looking for a Senior Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team...  ...operational workflows, troubleshoot complex incidents, lead planned maintenance activities, and help establish mature Site... 
    Permanent employment
    Work experience placement
    Relocation
    Visa sponsorship
    Work visa
    Flexible hours
    Shift work
    Day shift
    Berkeley, CA
    6 days ago
  • $27 per hour

     ...Job Description Job Description   Location: Mills College We are hiring immediately for a full time BOH LEAD SUPERVISOR position. Address : 5000 MacArthur Boulevard, Oakland, CA 94613.  Note: online applications accepted only. Schedule : Full time schedule... 
    Hourly pay
    Full time
    Temporary work
    Part time
    Summer holiday
    Local area
    Immediate start
    Remote work
    Flexible hours
    Shift work

    Chartwells HE

    Oakland, CA
    29 days ago
  •  ...scale the service with great people and reliable, cost-effective, and efficient infrastructure...  ...& tooling. What you’ll be doing Lead the Infra platform and shared services org...  ...partnership with architects and product engineering Build a world-class observability platform... 

    Gravity Engineering Services Pvt Ltd.

    San Francisco, CA
    2 days ago
  • $99.45k - $134.55k

     ...inclusive, with great opportunity for professional growth. Find your future with us. 1The Boeing Company is looking for an Site Reliability Engineer (Associate or Experienced) to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO. We are... 
    Permanent employment
    Work experience placement
    Interim role
    Relocation
    Visa sponsorship
    Work visa
    Flexible hours
    Shift work
    Day shift
    Berkeley, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead Site Reliability Engineer. Be the first to apply!