Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer — HPC & Automation (Silicon Engineering)

$125k - $175k
Full-time

SpaceX

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.

SITE RELIABILITY ENGINEER — HPC & AUTOMATION (SILICON ENGINEERING)

At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most advanced broadband internet system. Starlink is the world’s largest satellite constellation and is providing fast, reliable internet to millions of users worldwide. We design, build, test, and operate all parts of the system – thousands of satellites, consumer receivers that allow users to connect within minutes of unboxing, and the software that brings it all together. We’ve only begun to scratch the surface of Starlink’s potential global impact and are looking for best-in-class engineers to help maximize Starlink’s utility for communities and businesses around the globe.

We are seeking a motivated, proactive, and intellectually curious engineer who will work alongside world-class cross-disciplinary teams (systems, firmware, architecture, design, validation, product engineering, ASIC implementation). As a Site Reliability Engineer on the Silicon Engineering team you will get the opportunity to design, operate, scale, and automate the high performance computing infrastructure we use to develop the chips powering the world's largest satellite constellation and a global internet service. This position will have a meaningful impact on Starlink silicon by enabling faster design-iterations, simulations, and regression turnaround times that gate how fast our chip teams can ship.

RESPONSIBILITIES:

  • Deploy, upgrade, operate, maintain, and scale our suite of clusters and services
  • Collaborate with engineers to develop automated, full turnkey solutions for silicon simulation workflows to speed up project timelines
  • Manage our underlying infrastructure as code and use modern observability tools to provide a complete picture of cluster and infrastructure health
  • Operate the continuous integration pipeline, build and release systems, and version control across the environment
  • Identify and eliminate performance bottlenecks using measurement and creative engineering

BASIC QUALIFICATIONS:

  • Bachelor’s degree in computer science, information systems, or an engineering discipline; OR 2+ years of professional experience in system administration, high performance computing, or site reliability engineering
  • 1+ years of development experience with Bash, Python, and/or other programming languages
  • 1+ years of experience with Linux operating systems

PREFERRED SKILLS AND EXPERIENCE:

  • Familiarity with containerization technologies (i.e. Docker, Kubernetes)
  • Knowledge in computer system concepts (computer architecture, computer organization, operating systems and concurrency)
  • Experience with databases and data modeling (e.g., MySQL, PostgreSQL, SQLite)
  • Networking knowledge of TCP/IP
  • Experience with high performance computing and workload managers (e.g., Slurm, LSF)
  • Experience with Terraform, Ansible, Puppet, or similar automation frameworks
  • Experience building monitoring and alerting as code (e.g., Grafana, Prometheus, custom exporters)
  • Experience with CI/CD automation at scale (e.g., Jenkins, Bamboo, build systems)
  • Experience with infrastructure as code (IaC) tools for managing fleets of servers
  • Experience with using & building REST API clients/servers
  • Experience with enterprise/networked storage automation (e.g., NetApp ONTAP REST API/CLI, NFS)
  • Experience with ASIC design flows and tools (e.g., Cadence, Synopsys, Ansys, Keysight, Siemens)
  • Strong desire to find performance bottlenecks and performance improvement techniques
  • Excellent communication skills with the ability to communicate with customers, peers, management, etc. in both formal and informal situations
  • Ability to quickly learn new tools and frameworks
  • Interest in or experience with AI/LLM-assisted tooling (e.g., Grok, Claude Code)

ADDITIONAL REQUIREMENTS:

  • Ability to work extended hours and weekends as needed to meet critical milestones

COMPENSATION AND BENEFITS:

Pay Range:
Level 1: $125,000.00 - $150,000.00
Level 2: $145,000.00 - $175,000.00

Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.

Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees in Washington State accrue paid sick time in compliance with state and federal law. Company shuttles are offered to employees for roundtrip travel from select Seattle locations to the SpaceX Redmond office Monday to Friday.

ITAR REQUIREMENTS:

  • To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here.

SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status.

Applicants wishing to view a copy of SpaceX’s Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to View email address on aiapply.co .

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer — HPC & Automation (Silicon Engineering) in Redmond, WA vacancy
  • $165k - $270k

     ...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD) At SpaceX we’re leveraging our experience in...  ...Developer Operations, and GPU platforms. You will develop automation to deploy and manage on-premise compute resources, create... 
    Suggested
    Permanent employment
    Temporary work
    Work at office
    Immediate start
    Monday to Friday
    Weekend work

    SpaceX

    Redmond, WA
    1 day ago
  • $119.8k - $234.7k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole...  ...EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...world.Microsoft’s Azure Data engineering team is leading the transformation...  ...and shaping the Livesite Automation and AI Ops stack in Cosmos... 
    Suggested
    3 days per week

    Microsoft Corporation

    Redmond, WA
    3 days ago
  •  ...manage AI resources on Microsoft Azure, including AI Foundry and RAG solutions Monitor and ensure service uptime, availability, reliability, and latency Track and integrate SRE metrics with enterprise monitoring systems Support CI/CD and DevOps workflows using... 
    Suggested

    Tech M USAAvance Consulting

    Redmond, WA
    22 hours ago
  • $142.8k - $274.8k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole...  ...EngineeringCompany: MicrosoftOverviewMicrosoft Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team behind...  ...performance, efficiency, reliability, and operational scalability.... 
    Suggested
    Ongoing contract
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Redmond, WA
    2 days ago
  • $165k - $230k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK) At SpaceX we’re leveraging our experience in building rockets and spacecraft to... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago
  •  ...Responsibilities Bring up new RF silicon and power up chips for the first time. Design...  ...hardware-in-the-loop test platforms and automated integration and regression tests....  ...Requirements Bachelor’s degree in engineering, mathematics, or physics. At least 1... 
    Full time
    Work at office
    Monday to Friday
    Weekend work

    SpaceX

    Redmond, WA
    13 hours ago
  • $165k - $260k

     ...of enabling human life on Mars. SR. TECHNICAL PROJECT LEAD (SILICON ENGINEERING) At SpaceX we’re leveraging our experience in building...  ...world’s largest satellite constellation and is providing fast, reliable internet to millions of users worldwide. We design, build,... 
    Permanent employment
    Full time
    Temporary work
    Work at office
    Worldwide
    Monday to Friday
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago
  • $119.8k - $234.7k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole...  ...EngineeringCompany: MicrosoftOverviewMicrosoft Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team behind...  ...performance, efficiency, reliability, and operational scalability.... 
    Ongoing contract
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Redmond, WA
    1 day ago
  • $165k - $230k

     ...enabling human life on Mars. SR. HARDWARE / INFRASTRUCTURE SITE RELIABILITY ENGINEER (STARLINK) At SpaceX we’re leveraging our experience...  ...Operations, to our internal Kubernetes platforms. You will develop automation to deploy and manage on-premise compute resources, create... 
    Permanent employment
    Full time
    Temporary work
    Work at office
    Worldwide
    Monday to Friday
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago
  • $125k - $145k

     ...ultimate goal ofenabling human life on Mars.DESIGN VERIFICATION ENGINEER (SILICON ENGINEERING)At SpaceX we’re leveraging our experience in...  ...world’s largest satellite constellation and is providing fast, reliable internet to millions of users worldwide. We design, build,... 
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago
  • $119.8k - $234.7k

     ...for a Quantum Error Correction Software Engineer. This position offers an opportunity to...  ...analysis, research gathering, day to day task automation). Additional or Preferred...  ...equivalent experience. Experience with HPC, scientific programming, and/or computational... 
    Ongoing contract
    Permanent employment
    Local area

    Microsoft Corporation

    Redmond, WA
    3 days ago
  • $142.8k - $274.8k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole...  ...EngineeringDiscipline: Silicon EngineeringCompany: MicrosoftOverviewMicrosoft...  ...Hardware, and Infrastructure Engineering (SCHIE) is the team behind...  ...are looking for a Principal Reliability Engineer to join the team.... 
    Ongoing contract
    Permanent employment
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Redmond, WA
    18 hours ago
  • $160k - $210k

     ...change and achieving remarkable growth in a rapidly evolving industry. Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management across Cognitiv. Our immediate challenge is to scale... 
    Work at office
    Immediate start
    Remote work
    Work from home

    Cognitiv

    Bellevue, WA
    28 days ago
  •  ...Test Automation Engineer - RemoteBright Vision Technologies is a technology consulting and software development company delivering cloud,...  ...tests.Drive continuous improvement of test coverage, test reliability, and time-to-feedback.Mentor junior QA engineers and uphold... 
    Contract work
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship
    Shift work

    Bright Vision Technologies

    Kirkland, WA
    22 hours ago
  •  ...Site Reliability Engineer Join the innovators connecting just about anything—from families to cars to now things—on T-Mobile's biggest and best network yet. The SyncUP Things platform team has an immediate need for a Site Reliability Engineer. Responsibilities:... 
    Contract work
    Immediate start
    Remote work

    Software Technology Inc

    Bellevue, WA
    22 hours ago
  • $119.8k - $234.7k

     ...type: Full-TimeWork site: 3 days / week in-...  ...Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team...  ...teams, quality and reliability teams, business managers...  ...serviceability, packaging, automation, tolerance, and...  ...engineering for AI, GPU, HPC, or rack-scale... 
    Ongoing contract
    Contract work
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Redmond, WA
    2 days ago
  • Technical/Functional Skills Windows Servers, Digital: Microsoft Azure Windows Powershell, Digital: DevOps Roles & Responsibilities Windows Server 2012 -2019 Administration Microsoft Azure Azure AAD DFSR, DHCP DNS, KMS, WSUS TCP/IP Hyper-V High Availability Clusters ...

    The Dignify Solutions, LLC

    Bellevue, WA
    1 day ago
  • $143.7k - $194.4k

     ...network. Our mission is to deliver fast, reliable internet connectivity to customers...  ...with every device we design, from custom silicon to secure software, to enable innovative...  ...reality.As a Device Software Development Engineer on the Amazon Leo Satellite Gateway team... 
    Permanent employment
    Internship
    Local area
    Flexible hours

    Amazon

    Redmond, WA
    2 days ago
  • $151.2k - $204.6k

     ...network. Its mission is to deliver fast, reliable internet to customers and communities...  ...connectivity.We're looking for a hands-on Systems Engineer with a background in the requirements...  ...- Experience leading the design, automation, deployment, and support of large-scale... 
    Permanent employment
    Flexible hours

    Amazon

    Redmond, WA
    18 hours ago
  •  ...portfolio of networking, security, automation, and AI-powered operations...  ..., migration plans, and engineering bills of materials (BOMs) for...  ...high-performance computing (HPC), cloud, and large-scale distributed...  ...disaggregation, and merchant silicon‑based networking platforms.... 
    Work experience placement
    Remote work
    Work from home

    Hewlett Packard Enterprise

    Redmond, WA
    22 hours ago
  • $194k - $267k

     ...be challenged and have a passion for solving large‑scale automation, testing, and tuning problems, we would love to hear from...  ...educate on new concepts and tools. Position Overview The Site Reliability Engineer (SRE) will play a key role in building and managing... 
    Permanent employment
    Work at office
    Local area
    Flexible hours

    Segment (Twilio)

    Bellevue, WA
    1 day ago
  • $116k - $189.75k

     ...workloads. We are looking for a Software Engineer focused on bring-up, triage, benchmarking...  ...implementation of the benchmarking tooling, automation, and debugging workflows that support...  ...of experience developing software for AI, HPC, or systems-level applications.Hands-on experience... 
    Full time
    Remote work

    Nvidia

    Redmond, WA
    18 hours ago
  • $184k - $287.5k

     ...looking for a Senior Software Engineer to lead the bring-up, triage,...  ...run efficiently and reliably at scale. You will lead deep...  ...repeatable benchmark suites, automation, acceptance criteria, and qualification...  ...for large-scale AI or HPC systems, including a track record... 
    Full time
    Remote work

    Nvidia

    Redmond, WA
    3 days ago
  • $125k - $200k

     ...enabling human life on Mars. NEW GRADUATE ENGINEER, SOFTWARE - '26/'27 (STARLINK) At...  ...satellite constellation and is providing fast, reliable internet to millions of users worldwide....  ...hardware, supporting our in-house RF Silicon designs. Full life cycle... 
    Permanent employment
    Full time
    Temporary work
    Internship
    Work at office
    Worldwide
    Monday to Friday
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago
  •  ...Senior Automation Engineer – Global MSAT Location: Redmond, WA At Just Evotec Biologics, we...  ...validation, and rollout across international sites. Architect, standardize, and sustain...  ...integration, maintaining secure and reliable connectivity between PI systems and... 
    Flexible hours

    Just – Evotec Biologics

    Redmond, WA
    1 day ago
  •  ...Employment Type: Full TimeIndustry: Computer SoftwareClient: WiproContact: Bharath Gampa, Ram MaddulaCompany: SRI Tech SolutionsRole : Automation DeveloperJob Type : Full-TimeRedmond, WA (Day one Onsite) locals preferred.Automation framework: TAEF, and WTT.Tools/Tech Skills:... 
    Local area

    Sri Tech

    Redmond, WA
    1 day ago
  • $60 - $65 per hour

     ...Consulting is a design-led, software development and hardware engineering company, offering end-to-end digital services to help companies...  ...software for the above Assist in the development of automated tests and/or CI for the above Work closely with design engineers... 
    Temporary work

    Fresh Consulting

    Redmond, WA
    26 days ago
  • $119.8k - $234.7k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-...  ...Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team behind...  ...validation methodologies, automation infrastructure, and deployment...  ...qualification coverage, reliability requirements, usage models... 
    Ongoing contract
    Permanent employment
    Local area
    Remote work
    3 days per week

    Microsoft

    Redmond, WA
    1 day ago
  • $125k - $200k

     ...enabling human life on Mars. SOFTWARE ENGINEER, C++ SIMULATIONS (STARLINK) As a C++...  ...parallelization, dispersion handling, and reliability for large-scale runs. Drive...  ...simulators). High-performance computing (HPC) background: parallel computing, multi-threading... 
    Permanent employment
    Full time
    Temporary work
    Work at office
    Monday to Friday
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago
  • $125k - $175k

     ...ultimate goal of enabling human life on Mars. SOFTWARE ENGINEER, HARDWARE TEST & AUTOMATION (STARLINK) As a Software Engineer on the Starlink...  ...components and entire satellites for maximum performance and reliability in extreme environments, with mission success being... 
    Permanent employment
    Full time
    Temporary work
    Work at office
    Monday to Friday
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer — HPC & Automation (Silicon Engineering). Be the first to apply!