Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Application Software Engineer, Inference

$135k - $160k
Full-time

Spacex

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.

APPLICATION SOFTWARE ENGINEER, INFERENCE

The application software team is the central nervous system of SpaceX – we create mission critical applications that are used throughout SpaceX to accelerate launch vehicle production and flight as well as systems that allow Starlink to grow into a worldwide fast, reliable Internet service. We are looking for engineers who treat fellow teammates with fairness, respect, and support.

Our team maintains a high-performance AI inference platform that serves the best models internally at SpaceX to accelerate our most ambitious engineering goals. As part of this effort in Palo Alto, you will design and optimize large-scale model serving systems end-to-end, owning everything from distributed infrastructure to deep low-level optimizations. You will work on systems that deliver reliable, high-throughput inference to power SpaceX’s mission-critical applications while maintaining the highest standards of performance and availability.

Aerospace experience is not required to be successful here - rather we look for smart, motivated, respectful, collaborative engineers who love solving problems and want to make an impact on a super inspiring mission. You will have full ownership of challenging problems, working with a team of enthusiastic engineers with diverse perspectives to design and produce solutions that enable SpaceX to achieve its loftiest engineering goals at a rapid pace. The success of the missions at SpaceX depends on the software that you and your team produce.

This role will report through SpaceX Application Software while also working closely with xAI engineering teams. 

RESPONSIBILITIES:


  • Develop highly reliable, high-throughput inference systems that serve the best AI models internally across SpaceX

  • Architect and implement scalable distributed infrastructure for model serving, including load balancing, auto-scaling, batch scheduling, global KV cache, and continuous batching  

  • Optimize latency and throughput of model inference under real production workloads, including low-level GPU kernel work, quantization, speculative decoding, and other acceleration techniques  

  • Build reliable, high-concurrency serving systems with 100% uptime, low tail latency, and excellent observability  

  • Own end-to-end components such as request routing, SDK development, rate limiting, and efficient scaling for internal SpaceX AI inference platforms  

  • Benchmark, fine-tune, and accelerate inference engines (e.g., SGLang, vLLM, TensorRT-LLM) 

  • Develop custom tools for tracing, replaying, and resolving issues across the full stack — from orchestration down to GPU kernels  

  • Create robust CI/CD infrastructure for seamless endpoint deployment, image publishing, and inference engine updates  

  • Collaborate across SpaceXAI   teams to integrate inference capabilities into broader systems and workflows  

BASIC QUALIFICATIONS:


  • Bachelor's degree in computer science, engineering, math, or scientific discipline; OR 2+ years of professional experience building software in lieu of a degree

  • Experience in designing, implementing, and maintaining reliable and horizontally scalable distributed systems

  • 1+ years of experience in full stack development or backend development with production systems

  • 1+ years of experience with Rust or C++

PREFERRED SKILLS AND EXPERIENCE:


  • Experience with LLM inference engines and serving frameworks (e.g., SGLang, vLLM, Triton, TensorRT-LLM) 

  • Deep low-level systems programming and optimizations: GPU kernels, code generation, batching, caching, parallelism, quantization, and speculative decoding  

  • Experience with large-scale, high-concurrency production serving systems  

  • Knowledge of service observability and reliability best practices  

  • Experience operating commonly used databases such as PostgreSQL, ClickHouse, or MongoDB  

  • Experience designing or building with agent SDKs and agent orchestration frameworks  

  • Experience with Docker, Kubernetes, and containerized applications  

  • Expert knowledge of gRPC (unary, response streaming, bi-directional streaming, REST mapping) 

  • Programming experience in Python, Go, or similar languages  

  • Experience with version control, continuous integration, continuous delivery, build systems, and monitoring  

  • Expertise in profiling and improving application performance  

ADDITIONAL REQUIREMENTS:


  • You may be asked to work extended hours/weekends dependent on launch cadence and platform demands  

  • This role requires you to be onsite in Palo Alto. Remote and/or hybrid work will not be considered  

COMPENSATION AND BENEFITS:

 
Pay Range:
Software Engineer/Level I: $135,000.00 - $160,000.00/per year
Software Engineer/Level II: $155,000.00 - $185,000.00/per year

Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.

Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock, stock options, or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law.

ITAR REQUIREMENTS:


  • To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here .

SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status.

Applicants wishing to view a copy of SpaceX’s Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to  View email address on jobs.jobcopilot.com

Vacancy posted 9 hours ago
Similar jobs that could be interesting for youBased on the Application Software Engineer, Inference in Palo Alto, CA vacancy
  •  ...frameworks. Strong track record of working with machine learning systems and/or platforms. Experience in serving LLMs using inference engines like vLLM, TensorRT-LLM, TEI, SGLang, and knowing tradeoffs between them. Experience serving fine-tuned LLMs (PEFT, DPO, RL... 
    Suggested
    Full time

    Snowflake

    Menlo Park, CA
    9 hours ago
  • $207k - $300k

    Analyze and optimize AI inference workloads across the application, model, and distributed fleet infrastructure...  ...degree in Computer Science, Computer Engineering, Electrical Engineering, Applied...  ....8 years of experience in software development.Experience in Python and... 
    Suggested

    Google

    Mountain View, CA
    14 days ago
  • $2,000 per month

     ...architecture and design of the Sohu host software stack Implement high-performance,...  ...handling continuous batching and real time inference Implement inference-time...  ...person team in Cupertino, and greatly value engineering skills. We do not have boundaries between... 
    Suggested
    Full time
    Work at office
    Relocation package

    Etched

    Cupertino, CA
    9 hours ago
  •  ...the transformation of technology. We are at the forefront of software and hardware innovation, pushing the boundaries of what is...  ...3 days per week. The role: Principal System Software Engineer, AI Inference Execution What you will do: The role requires you to... 
    Suggested
    Work experience placement
    3 days per week

    Jobleads-US

    Santa Clara, CA
    5 days ago
  • $188k - $275k

     ...Learn more at  . What You’ll Do: Inference Platform Team The Inference team...  ...systems. About the role: As a Staff Software Engineer (IC5) on the Inference team, you will...  ...Consumer Privacy Act - California applicants only CoreWeave is an equal opportunity... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    Core Weave

    Sunnyvale, CA
    9 hours ago
  •  ...Luma Model Serving Engineer You'll own how Luma's models get served — integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping expensive GPU fleets busy while meeting internal SLOs. This is large-scale... 

    Luma AI

    Redwood City, CA
    9 hours ago
  • $224k - $356.5k

    We are now looking for a Senior System Software Engineer to work on Dynamo. NVIDIA is hiring...  ...fast-paced team building Generative AI inference platform to make design and deployment...  ...also be eligible for equity and benefits.Applications for this job will be accepted at least... 
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $92k - $135k

     ...more at  What You'll Do: Join the Inference team to ship production features that improve...  ...with mentorship from experienced engineers. About the role: Implement well-scoped...  ...innovative disruption California Applicants California Consumer Privacy Act... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Internship
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    18 days ago
  •  ...deliver industry-leading training and inference speeds; over 10 times faster than GPU-...  ...transforming the user experience of AI applications, unlocking real-time iteration and...  ...inference. About the Role We're hiring a Software Engineer to help contribute to projects on our... 
    Full time

    Cerebras Systems

    Sunnyvale, CA
    9 hours ago
  •  ...industry-leading training and inference speeds; over 10 times faster...  ...the user experience of AI applications, unlocking real-time iteration...  ...We're hiring a Staff Engineer to own major areas of the architecture...  ...~8+ years of experience in software engineering, with... 
    Full time

    Cerebras Systems

    Sunnyvale, CA
    9 hours ago
  • $193.3k - $261.5k

    We develop AWS Neuron, the complete software stack for Trainium, Amazon's custom cloud...  ....As a Sr. Software Development Engineer on the Inference Model Enablement team, you will onboard...  ...protected status.Los Angeles County applicants: Job duties for this position include... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    a month ago
  • $165.2k - $223.6k

     ...AWS) builds AWS Neuron, the software development kit used to accelerate...  ...ML compiler, runtime, and application framework that seamlessly...  ...enabling unparalleled ML inference and training performance.The...  ...hardware-software boundary, our engineers build systematic... 
    Work experience placement
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    29 days ago
  •  ...deliver industry-leading training and inference speeds; over 10 times faster than...  ...the user experience of AI applications, unlocking real-time iteration and...  ...decode on the Cerebras Wafer-Scale Engine.We are hiring a Software Engineer to productionize and optimize... 

    Cerebras Systems

    Sunnyvale, CA
    7 days ago
  • $184k - $287.5k

    We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency...  ....You will also be eligible for equity and benefits.Applications for this job will be accepted at least until May 2,... 
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  •  ...Role We're looking for early career Software Engineers to join our engineering team. You'll...  ...features across our web, mobile, and backend applications, including offline-first experiences...  ...Building and improving AI systems: inference pipelines, orchestration (fallbacks,... 
    Full time
    Internship
    Work at office
    Immediate start

    Commure

    Mountain View, CA
    9 hours ago
  • $193.3k - $261.5k

     ...team builds.We're looking for an engineer to work at the frontier of disaggregated inference: splitting LLM serving into...  ...optimize the low-level data-movement software across accelerators, servers,...  ...status.Los Angeles County applicants: Job duties for this position include... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    a month ago
  •  ...seeking experienced product-focused engineers to join our team in building the...  ...-facing surfaces of our inference and training infrastructure. As a Software Engineer, Product, you will partner...  ...-end, including APIs, SDKs, web applications, or developer tools. ~ Experience... 
    Full time

    Radixark

    Palo Alto, CA
    9 hours ago
  • $193.93k - $291.15k

     ...Nuro Driver™, to support a wide range of applications, from robotaxis and commercial fleets...  ...autonomy. We are looking for strong software engineers to research, develop, and implement technologies...  ...systems, distributed training, or inference optimization.   At Nuro, we... 
    Full time
    Immediate start

    Nuro

    Mountain View, CA
    9 hours ago
  • $187.74k - $198k

     ...its commercial self-driving software to develop, test and deploy...  ...hands-on Perception Software Engineer to join the Perception team....  ...systems for autonomous trucking applications. Implement and evaluate...  ...optimization for real-time inference using TensorRT and ONNX and... 
    Full time
    Temporary work
    Work at office
    Visa sponsorship
    Flexible hours

    Kodiak

    Mountain View, CA
    9 hours ago
  • $200k - $420k

     ...rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructure, next-generation UIs, and...  ...deep learning research. Who we are We are scientists, engineers, and builders from the industry's top tech companies and AI... 
    Full time
    Local area
    Visa sponsorship
    Relocation package

    River Ai

    Palo Alto, CA
    9 hours ago
  • $238k - $302k

     ...that are core to our autonomous driving software. We help our partners by offering the best...  ...driving.  We are looking for engineers with ML system expertise to help us train...  ...and implement optimizations to improve inference speed and resource utilization. Collaborate... 
    Full time
    Remote work

    Waymo

    Mountain View, CA
    9 hours ago
  • $135k - $200k

     ...and Okta, our team includes engineers from premier tech companies...  ...frameworks, including training and inference of large language models....  .... Good understanding of software development principles, data...  ...to all employees and applicants for employment without regard... 
    Full time

    Ema

    Mountain View, CA
    9 hours ago
  • $135k - $160k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. APPLICATION SOFTWARE ENGINEER The application software team is the central nervous system of SpaceX - we create mission critical applications that are... 
    Permanent employment
    Temporary work
    Remote work
    Worldwide
    Weekend work

    SpaceX

    Palo Alto, CA
    1 day ago
  • $230k - $350k

     ...Member of Technical Staff, Inference Systems, to build and optimize...  ...systems and strong systems engineering skills, with Rust experience...  ...systems and production software engineering fundamentals....  ...opportunity employer. Qualified applicants are considered based on their... 
    Work at office

    Premier Global Links

    Palo Alto, CA
    2 days ago
  •  ...a collective of visionary scientists, engineers, and entrepreneurs are dedicated to transforming...  ...modern, responsive full-stack web applications Translate user flows and designs...  ...language models or working with AI/ML model inference infrastructure Familiarity with AI... 
    Full time
    Work at office

    GenBio AI

    Palo Alto, CA
    9 hours ago
  • $274k - $304k

     ...human and non-human access to all of an organization's applications, data, and business processes. Customers trust Saviynt to...  ...institutions. For more information, please visit AI Platform Engineer - Training & Inference Saviynt's AI-powered identity platform manages and... 

    Saviynt

    Milpitas, CA
    2 days ago
  •  ...Institutional Page About the role We are hiring a Staff Software Engineer to be the foundational technical lead on the newly established...  ..., and knowing when a deterministic component beats an inference call. Ship the first-generation workflow tools: Build and... 
    Full time

    Nubank

    Palo Alto, CA
    9 hours ago
  •  ...Applied Performance Group (APG) Engineering team provides a highly...  ...warehouses, data processing, and applications Coding/programming...  ...Kafka, Apache Flink, etc.) Software development experience with...  ...like PyTorch for training and inference). University degree in computer... 
    Full time

    Snowflake

    Menlo Park, CA
    9 hours ago
  •  ...moving data out to a separate system to run inference or build an agent, our customers do it...  ...and process events at scale. As an engineer, you'll own delivery of significant...  ...organization. By proceeding with this application, you understand that Confluent will share... 
    Full time
    Live in

    Confluent

    Mountain View, CA
    9 hours ago
  • $215k - $260k

     ...production. That means owning the inference stack end to end: profiling...  ...work directly with customer engineering teams to tailor deployments...  .... Build and support the software and product features around...  ...to be determined by the applicant's knowledge, education, and... 
    Temporary work

    Crusoe

    Sunnyvale, CA
    26 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Application Software Engineer, Inference. Be the first to apply!