Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, Inference (AI Data Engineering)

$135k - $210k
Full-time

SpaceX

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.

SOFTWARE ENGINEER, INFERENCE (AI DATA ENGINEERING)

The application software team is the central nervous system of SpaceX – we create mission critical applications that are used throughout SpaceX to accelerate launch vehicle production and flight as well as systems that allow Starlink to grow into a worldwide fast, reliable Internet service. We are looking for engineers who treat fellow teammates with fairness, respect, and support.

Our team maintains a high-performance AI inference platform that serves the best models internally at SpaceX to accelerate our most ambitious engineering goals. As part of this effort in Palo Alto, you will design and optimize large-scale model serving systems end-to-end, owning everything from distributed infrastructure to deep low-level optimizations. You will work on systems that deliver reliable, high-throughput inference to power SpaceX’s mission-critical applications while maintaining the highest standards of performance and availability.

Aerospace experience is not required to be successful here - rather we look for smart, motivated, respectful, collaborative engineers who love solving problems and want to make an impact on a super inspiring mission. You will have full ownership of challenging problems, working with a team of enthusiastic engineers with diverse perspectives to design and produce solutions that enable SpaceX to achieve its loftiest engineering goals at a rapid pace. The success of the missions at SpaceX depends on the software that you and your team produce.

This role will report through SpaceX Internal AI Infrastructure, as we'll also be providing support for training workloads.

RESPONSIBILITIES:

  • Develop highly reliable, high-throughput inference systems that serve the best AI models internally across SpaceX
  • Architect and implement scalable distributed infrastructure for model serving, including load balancing, auto-scaling, batch scheduling, global KV cache, and continuous batching
  • Optimize latency and throughput of model inference under real production workloads, including low-level GPU kernel work, quantization, speculative decoding, and other acceleration techniques
  • Build reliable, high-concurrency serving systems with 100% uptime, low tail latency, and excellent observability
  • Own end-to-end components such as request routing, SDK development, rate limiting, and efficient scaling for internal SpaceX AI inference platforms
  • Benchmark, fine-tune, and accelerate inference engines (e.g., SGLang, vLLM, TensorRT-LLM)
  • Develop custom tools for tracing, replaying, and resolving issues across the full stack — from orchestration down to GPU kernels
  • Create robust CI/CD infrastructure for seamless endpoint deployment, image publishing, and inference engine updates
  • Collaborate across SpaceXAI teams to integrate inference capabilities into broader systems and workflows

BASIC QUALIFICATIONS:

  • Bachelor's degree in computer science, engineering, math, or scientific discipline; OR 2+ years of professional experience building software in lieu of a degree
  • Experience in designing, implementing, and maintaining reliable and horizontally scalable distributed systems
  • 1+ years of experience in full stack development or backend development with production systems
  • 1+ years of experience with Rust or C++

PREFERRED SKILLS AND EXPERIENCE:

  • Experience with LLM inference engines and serving frameworks (e.g., SGLang, vLLM, Triton, TensorRT-LLM)
  • Deep low-level systems programming and optimizations: GPU kernels, code generation, batching, caching, parallelism, quantization, and speculative decoding
  • Experience with large-scale, high-concurrency production serving systems
  • Knowledge of service observability and reliability best practices
  • Experience operating commonly used databases such as PostgreSQL, ClickHouse, or MongoDB
  • Experience designing or building with agent SDKs and agent orchestration frameworks
  • Experience with Docker, Kubernetes, and containerized applications
  • Expert knowledge of gRPC (unary, response streaming, bi-directional streaming, REST mapping)
  • Programming experience in Python, Go, or similar languages
  • Experience with version control, continuous integration, continuous delivery, build systems, and monitoring
  • Expertise in profiling and improving application performance

ADDITIONAL REQUIREMENTS:

  • You may be asked to work extended hours/weekends dependent on launch cadence and platform demands
  • This role requires you to be onsite in Palo Alto. Remote and/or hybrid work will not be considered

COMPENSATION AND BENEFITS:

Pay Range:
Level 1: $135,000.00 - $175,000.00
Level 2: $155,000.00 - $210,000.00

Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.

Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law.

ITAR REQUIREMENTS:

  • To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here.

SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status.

Applicants wishing to view a copy of SpaceX’s Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to View email address on aiapply.co .

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Software Engineer, Inference (AI Data Engineering) in Palo Alto, CA vacancy
  • $190k - $250k

    A leading financial technology firm in California seeks an AI Inference engineer to join its team. The role involves developing APIs for AI inference, improving system reliability, and optimizing LLM performance. Required qualifications include experience with ML systems... 
    Suggested

    Pantera Capital

    Palo Alto, CA
    5 days ago
  • $190k - $250k

    Location San Francisco Employment Type Full time Location Type Hybrid Department AI We are looking for an AI Inference engineer to join our growing team. Our current stack is Python, Rust, C++, PyTorch, Triton, CUDA, Kubernetes. You will have the opportunity to work on... 
    Suggested
    Full time

    Kindredventures

    Palo Alto, CA
    6 days ago
  • $117.7k - $221.4k

     ...understand, and curate high-value data from large-scale real-world...  ...cost efficient for embodied AI systems. We believe the next generation...  ...model reflects how Cola engineers think: build durable...  ...processing, featurization, and inference foundations that power scalable... 
    Suggested
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  •  ...AI/ML Software EngineerAt Gallatin, we are rebuilding logistics infrastructure...  ...operate at the layer where data becomes decisions, and...  ...looking for an AI/ML Software Engineer to play a foundational role...  ...ML pipelines and real-time inference systems—while collaborating... 
    Suggested
    Local area

    Gallatin AI, Inc.

    Palo Alto, CA
    3 days ago
  •  ...reputation with the clients.Currently, we are looking for entry-level software programmers, Java full stack developers, Python/Java developers, data analysts/data scientists, machine learning engineers for full time positions with clients. Who should apply? Recent... 
    Suggested
    Full time
    H1b
    Remote work

    SynergisticIT

    Mountain View, CA
    3 days ago
  •  ...NVIDIA Corporation is seeking a Senior Software Engineer for the TensorRT Edge-LLM team in the US. You will develop a high-performance inference framework in modern C++ that extends TensorRT for autoregressive model serving, including speculative decoding and KV cache... 

    NVIDIA Corporation

    Santa Clara, CA
    5 hours ago
  • An innovative company is seeking a talented data analyst to harness the power of Google Cloud's Generative AI technology. In this role, you'll observe, organize, and operationalize diverse datasets, transforming them into actionable business insights. Collaborate with cutting... 

    humanfirst.ai

    Mountain View, CA
    4 days ago
  • SpaceXAI is seeking a hands-on Engineer to ensure engineering teams get high-quality data for training and evaluation. You will craft data projects, partner with model teams, and drive data strategy at the intersection of engineering and data operations. This role focuses... 

    Socket.dev

    Palo Alto, CA
    5 days ago
  • AI-Native Data Engineer @ TrueMeter SF Bay Area | Hybrid (3 days onsite, 2 remote) About Us We’re building the AI Energy Agent that’s becoming the default way any business pays for power and saves on energy. The grid is breaking under the weight of AI and electrification... 
    Immediate start
    Remote work

    Pear VC

    Palo Alto, CA
    2 days ago
  • $250k - $300k

     ...vertically integrated AI infrastructure company...  ...energy, manufacturing, data center construction, and...  ...That means owning the inference stack end to end: profiling...  ...directly with customer engineering teams to tailor...  ...Build and support the software and product features around... 
    Temporary work

    Crusoe

    Sunnyvale, CA
    15 days ago
  •  ...Unity is seeking a Senior Software Engineer to own and improve analytics datasets and ETL pipelines supporting an internal AI product analytics agent. You will work with data scientists and analysts to enhance retrieval, knowledge, and evaluation systems, while building... 

    Unity Enterprise

    Mountain View, CA
    5 hours ago
  • Cerebras Systems, Inc. is seeking engineers for its Inference Core Platform group in Sunnyvale, California. This role involves building foundational software and hardware infrastructure to enhance AI inference performance on the Cerebras Wafer-Scale Engine. Ideal candidates... 

    Cerebras Systems, Inc.

    Sunnyvale, CA
    5 days ago
  • Intel in Santa Clara, CA, seeks an experienced software engineer to optimize local inference for edge devices. You will work on llama.cpp, vLLM, and quantization...  ...engines while advancing hardware-aware optimizations for efficient AI on local devices. #J-18808-Ljbffr Intel
    Local area

    Intel

    Santa Clara, CA
    6 days ago
  • Intel is seeking a seasoned software engineer to accelerate AI inference on edge hardware. You will optimize llama.cpp/vLLM, tune KV cache, batching and scheduling, and push quantization strategies to balance speed and quality. This role focuses on low-latency, privacy... 

    PVH (Tommy Hilfiger/Calvin Klein)

    Santa Clara, CA
    6 days ago
  •  ...Systems, Inc. is looking for a Senior Performance Engineer to enhance the performance benchmarking and competitive pricing models for their AI chip. The ideal candidate will have extensive experience with open-source inference frameworks and an understanding of ML systems.... 

    Cerebras Systems, Inc.

    Sunnyvale, CA
    5 days ago
  • Rhombus Power, Inc. is seeking a Data Engineer to design and implement data engineering activities on mission-driven projects. Key responsibilities include developing code for data ingestion, architecting data repositories, and collaborating with analysts and developers... 

    Rhombus Power, Inc.

    Palo Alto, CA
    2 days ago
  • $116.2k - $269.1k

     ...Sr. AI Software Engineer – Coding Agent Tencent Overseas IT supports rapid global growth with future‑ready IT platforms and leads strategy...  ..., and iterate on the tool’s features. Optimize model inference speed and resource usage to ensure the agent is fast and responsive... 
    Overseas
    Relocation package

    Tencent

    Palo Alto, CA
    5 hours ago
  • A pioneering tech company in Mountain View seeks an early Software Engineer to drive the development of innovative data platforms for AI agents. You will design and maintain ETL processes, manage core data models, and partner with engineering teams to ensure data quality... 

    MAI Agents

    Mountain View, CA
    2 days ago
  • $160k - $225k

    A cutting-edge advertising technology firm in California is looking for a Software Engineer to own and enhance its data platform. You will design robust ETL/ELT processes and optimize data handling for analytics. The ideal candidate has a strong background in data engineering... 

    MAI

    Mountain View, CA
    3 days ago
  • A pioneering energy management company in the SF Bay Area is seeking an AI-Native Data Engineer to own the data and AI infrastructure. The candidate will build high-reliability data pipelines, design GCP infrastructure for scalability, and have a strong background in production... 

    Pear VC

    Palo Alto, CA
    2 days ago
  •  ...NVIDIA is seeking a Senior Agentic AI Software Engineer (Finance) to build agentic systems and advance AI inference workloads in real-world finance contexts. You will design and implement scalable software that drives experimental agents, optimize performance, and contribute... 

    Nvidia Corporation in

    Santa Clara, CA
    5 hours ago
  •  ...NVIDIA seeks a Senior Systems Software Engineer to tackle client-side AI challenges on Windows and Linux PCs with limited resources. You will...  ...and DGX ecosystems, while optimizing AI models, data pipelines, and inference runtimes for performance on next-generation GPUs.... 
    Local area

    NVIDIA

    Santa Clara, CA
    5 hours ago
  •  ...NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models... 

    NVIDIA

    Santa Clara, CA
    5 hours ago
  •  ...We're an Industrial AI start-up founded by a...  ...recognized leaders in Data Intelligence for Aerospace...  ...proprietary Explainable AI engines for SME-explainable...  ...for the models, ML, and inferences Automation of data wrangling...  ...with the cloud software team and occasionally with... 
    Remote job
    Full time
    Contract work
    Part time
    For contractors
    Work at office

    Brain Gain Recruiting

    Palo Alto, CA
    1 day ago
  • $180k

     ...About the Position As a Reasoning engineer, you will build frameworks to improve the reasoning capability, build distributed reinforcement learning systems, techniques for inference time compute (e.g. tree search and planning), and develop environments for agents.... 
    Full time
    Relocation

    Xai

    Palo Alto, CA
    1 day ago
  • SpaceXAI seeks an engineer to craft high-value data tasks and evaluations that power training and model development. You will operate at the intersection of engineering and data operations, partnering with model teams to design projects that capture meaningful signals and... 

    SpaceXAI

    Palo Alto, CA
    2 days ago
  • NVIDIA Corporation in Santa Clara, CA is seeking outstanding AI systems engineers to develop groundbreaking inference technologies for the hardware-accelerated stack. You will create libraries, code generators, and GPU kernel innovations for LLM workloads. Join a team... 

    NVIDIA Corporation

    Santa Clara, CA
    4 days ago
  • A leader in AI technology in Palo Alto is seeking a Senior AI Systems Performance Engineer to optimize the latest foundation models on their innovative platform. This role involves collaborating with cross-functional teams to push the performance limits of AI systems.... 

    SambaNova

    Palo Alto, CA
    5 days ago
  •  ...NVIDIA is seeking a Senior Agentic AI Software Engineer to advance agentic AI systems and workloads from scalable research to production-grade solutions. You will build agentic components, analyze inference dynamics, and collaborate with teams owning evaluation pipelines... 

    NVIDIA

    Santa Clara, CA
    5 hours ago
  • $206k - $250k

     ...a highly motivated, hands-on AI Engineer to spearhead the end-to-end development...  ...full stack framework from data ingestion and vector database...  ...+ years in Data Engineering, Software Engineering, or Data Science,...  ...model performance and reduce inference costs.Previous experience in... 
    Full time
    Contract work
    Temporary work
    Relocation package

    Zoox

    Foster, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, Inference (AI Data Engineering). Be the first to apply!