Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, Inference (AI Data Engineering)

SpaceX

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SOFTWARE ENGINEER, INFERENCE (AI DATA ENGINEERING)The application software team is the central nervous system of SpaceX – we create mission critical applications that are used throughout SpaceX to accelerate launch vehicle production and flight as well as systems that allow Starlink to grow into a worldwide fast, reliable Internet service. We are looking for engineers who treat fellow teammates with fairness, respect, and support.Our team maintains a high-performance AI inference platform that serves the best models internally at SpaceX to accelerate our most ambitious engineering goals. As part of this effort in Palo Alto, you will design and optimize large-scale model serving systems end-to-end, owning everything from distributed infrastructure to deep low-level optimizations. You will work on systems that deliver reliable, high-throughput inference to power SpaceX’s mission-critical applications while maintaining the highest standards of performance and availability.Aerospace experience is not required to be successful here - rather we look for smart, motivated, respectful, collaborative engineers who love solving problems and want to make an impact on a super inspiring mission. You will have full ownership of challenging problems, working with a team of enthusiastic engineers with diverse perspectives to design and produce solutions that enable SpaceX to achieve its loftiest engineering goals at a rapid pace. The success of the missions at SpaceX depends on the software that you and your team produce.This role will report through SpaceX Internal AI Infrastructure, as we'll also be providing support for training workloads.RESPONSIBILITIES:Develop highly reliable, high-throughput inference systems that serve the best AI models internally across SpaceXArchitect and implement scalable distributed infrastructure for model serving, including load balancing, auto-scaling, batch scheduling, global KV cache, and continuous batching Optimize latency and throughput of model inference under real production workloads, including low-level GPU kernel work, quantization, speculative decoding, and other acceleration techniques Build reliable, high-concurrency serving systems with 100% uptime, low tail latency, and excellent observability Own end-to-end components such as request routing, SDK development, rate limiting, and efficient scaling for internal SpaceX AI inference platforms Benchmark, fine-tune, and accelerate inference engines (e.g., SGLang, vLLM, TensorRT-LLM) Develop custom tools for tracing, replaying, and resolving issues across the full stack — from orchestration down to GPU kernels Create robust CI/CD infrastructure for seamless endpoint deployment, image publishing, and inference engine updates Collaborate across SpaceXAI teams to integrate inference capabilities into broader systems and workflows BASIC QUALIFICATIONS:Bachelor's degree in computer science, engineering, math, or scientific discipline; OR 2+ years of professional experience building software in lieu of a degreeExperience in designing, implementing, and maintaining reliable and horizontally scalable distributed systems1+ years of experience in full stack development or backend development with production systems1+ years of experience with Rust or C++PREFERRED SKILLS AND EXPERIENCE:Experience with LLM inference engines and serving frameworks (e.g., SGLang, vLLM, Triton, TensorRT-LLM) Deep low-level systems programming and optimizations: GPU kernels, code generation, batching, caching, parallelism, quantization, and speculative decoding Experience with large-scale, high-concurrency production serving systems Knowledge of service observability and reliability best practices Experience operating commonly used databases such as PostgreSQL, ClickHouse, or MongoDB Experience designing or building with agent SDKs and agent orchestration frameworks Experience with Docker, Kubernetes, and containerized applications Expert knowledge of gRPC (unary, response streaming, bi-directional streaming, REST mapping) Programming experience in Python, Go, or similar languages Experience with version control, continuous integration, continuous delivery, build systems, and monitoring Expertise in profiling and improving application performance ADDITIONAL REQUIREMENTS:You may be asked to work extended hours/weekends dependent on launch cadence and platform demands This role requires you to be onsite in Palo Alto. Remote and/or hybrid work will not be considered COMPENSATION AND BENEFITS:Pay Range:Level 1: $135,000.00 - $175,000.00Level 2: $155,000.00 - $210,000.00Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law.ITAR REQUIREMENTS:To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here. SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status.Applicants wishing to view a copy of SpaceX’s Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to View email address on click.appcast.io.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Software Engineer, Inference (AI Data Engineering) in Palo Alto, CA vacancy
  • Data ScientistPosition OverviewWe are seeking an Data Scientist / Applied AI Engineer t to turn data into production-ready ML solutions that...  ...containerize, deploy and automate inference and monitoring workflows in...  ...experiments; follow software engineering best practices... 
    Suggested
    Remote work

    CyberCoders

    Redwood City, CA
    2 days ago
  •  ...opportunity for you to take your software engineering career to the next level. As...  ...enterprise-authorized AI coding assist tools within the...  ...from large, diverse data sets in service of continuous...  ...TensorFlow Serving, Triton Inference Server)Familiarity with distributed... 
    Suggested

    JP Morgan Chase

    Palo Alto, CA
    3 days ago
  •  ...AI/ML Software Engineer At Gallatin, we are rebuilding logistics infrastructure for the national...  ...foxhole, we operate at the layer where data becomes decisions, and decisions make...  ...large scale ML pipelines and real-time inference systems—while collaborating with cross... 
    Suggested
    Local area

    Gallatin AI, Inc.

    Palo Alto, CA
    4 days ago
  • $120k - $220k

     ...information powered by advanced AI, recommendation systems, and...  ...every cycle.We're hiring the engineer who owns this agent end-to-end...  ...them, trained a LoRA, optimized inference, built a ComfyUI workflow you’...  ...first 100 ads shipped CPI data back tighter loop. Not “I’ll spec... 
    Suggested
    Full time
    Local area
    Work from home

    News Break

    Mountain View, CA
    13 hours ago
  • $165.2k - $223.6k

     ...purchase regretSearch Science Data Infrastructure (SSDI)...  ...an ML Engineer you will:Lead development...  ...and operations using AWS AI services, DL compute resources...  ...deployment to our large scale inference services. You will...  ...professional software development experience-... 
    Suggested
    Internship
    Local area
    Worldwide
    Flexible hours

    Amazon

    Palo Alto, CA
    1 day ago
  • $144k - $236k

     ...of the team.Responsibilities: AI is at the core of how...  ...trust platforms. As a Senior AI Software Engineer you will own end-to-end machine...  ...or quality improvement (i.e. inference/training efficiency, engineer...  ...tech-debt removal) backed by data.You operate systems reliably... 
    For contractors
    Work at office
    Immediate start
    Flexible hours

    Linkedin

    Mountain View, CA
    2 days ago
  • $153k - $179k

     ...everything we do. Expectations are high, and so are the rewards.The Data Engineering team builds and maintains the foundational datasets that...  ...experience building end-to-end data pipelinesHands-on software engineering experience, with the ability to write production-... 
    Work at office
    Flexible hours
    Shift work
    3 days per week

    Robinhood Financial

    Menlo Park, CA
    4 days ago
  • $129k - $212k

     ...of the team.LinkedIn's Data Science team leverages...  ...for an ambitious data engineer to have an impact and transform...  ...talented and driven Sr Software Engineer, Data Science...  ...with LinkedIn’s AI-powered data stack. Successful...  ..., causal inference, measurement, and diagnostics... 
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    4 days ago
  • $175k - $287k

     ...of the team.Responsibilities: AI is at the core of how...  ...trust platforms. As a Staff AI Software Engineer you will own end-to-end machine...  ...or quality improvement (i.e. inference/training efficiency, engineer...  ...members of the team backed by data.You operate systems reliably... 
    For contractors
    Work at office
    Immediate start
    Flexible hours

    Linkedin

    Mountain View, CA
    2 days ago
  • $165.2k - $223.6k

     ...(AWS) builds AWS Neuron, the software development kit used to accelerate...  ...JAX enabling unparalleled ML inference and training performance....  ...-software boundary, our engineers build systematic infrastructure...  ...boundaries of what's possible in AI acceleration. As part of... 
    Full time
    Work experience placement
    Internship
    Local area
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    6 hours ago
  •  ...Cornerstone OnDemand is looking for a AI Engineer who will lead the design and...  ...document processing and data extraction. Build and...  ...model training and real-time inference. Define AI solution standards...  ...degree in computer science, Software Engineering, Artificial Intelligence... 

    Namely

    Mountain View, CA
    3 days ago
  •  ...Job Overview: LiveX AI is building the next generation...  ...Machine Learning Engineer to help us train, fine‑...  ...full model lifecycle—from data curation and training...  ...models, to low‑latency inference optimization for live,...  ...streaming pipelines. ~ Solid software engineering... 

    LiveX AI Inc.

    Palo Alto, CA
    3 days ago
  • $2,000 per month

     ...Elastic, the Search AI Company, enables everyone to find the answers...  ...in real time, using all their data, at scale - unleashing the...  ...for an innovative Agentic AI Engineer to join our team. You will build...  ...-concurrency demands of LLM inferences and multi-agent coordination.... 
    Local area
    Flexible hours

    Elastic

    Mountain View, CA
    3 days ago
  •  ...Join our AI team to build cutting-edge machine learning systems that power personalized...  ...use Quantize models for optimal inference on edge devices and cloud infrastructure...  ...years experience in machine learning or AI engineering ~ Strong proficiency in Python and ML... 

    Pintours

    Menlo Park, CA
    3 days ago
  • $120k - $195k

     ...of the team.Responsibilities: AI is at the core of how...  ...and trust platforms. As an AI Software Engineer you will own end-to-end machine...  ...or quality improvement (i.e. inference/training efficiency, engineer...  ...tech-debt removal) backed by data.You operate systems reliably... 
    For contractors
    Work at office
    Immediate start
    Flexible hours

    LinkedIn

    Mountain View, CA
    20 hours ago
  • $286.4k - $358k

    Uniphore is the Business AI company. Our sovereign,...  ...connects enterprise data, fine-tunes AI models and...  ...are seeking a VP of AI Engineering to lead the...  ...support SLLM fine-tuning, inference, prompt engineering, and...  ...organizations. • Own end-to-end software delivery: planning,... 
    Full time

    Uniphore

    Palo Alto, CA
    3 days ago
  •  ...We are seeking a Principal AI Engineer to design, build, and scale AI...  ...capabilities across products; implement data pipelines supporting feature...  ...(RAG), vector search, and inference; apply best practices for...  ...~8–10 years of hands‑on software engineering experience building... 

    Namely

    Mountain View, CA
    3 days ago
  • $184k - $287.5k

    NVIDIA is the platform upon which every new AI-powered application is built. We are seeking a Senior Software Engineer - AI Inference Performance to advance innovative LLM and...  ...who turns performance models and profiler data into working code. You will collaborate with... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. As a key contributor, you will help design, build, and optimize...  ...software that powers today’s most sophisticated AI applications. Our team is responsible for developing... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $275.8k - $340.5k

     ...to meet the unique demands of AI and ML innovation, supporting...  ...such as Embodied AI, Simulation, Data Science, and more. We enable...  ...enhance the productivity of ML engineers, and drive the adoption of...  ...includes: AI Validation & Inference: Ensures robust model performance... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Mountain View, CA
    18 hours ago
  •  ...Job Description Palona’s AI agents operate in real restaurant...  ...evaluation, high-quality data, modeling judgment,...  ...for an applied AI Modeling Engineer to improve the intelligence...  ...services, or training and inference pipelines. ~ Strong software engineering judgment; your... 
    Temporary work
    Immediate start

    Palona AI

    Los Altos, CA
    a month ago
  • $200k - $270k

     ...Description Samsung SDS America AI Team is researching the...  ...AI systems, spanning data collection through...  ...a Senior Physical AI Engineer to join the team...  ...learning, and production software engineering. You will connect...  ...Optimize low-latency inference and control systems for... 
    Worldwide
    Flexible hours

    Samsung SDS America

    Mountain View, CA
    a month ago
  •  ...Labs builds Industrial AI for the world's leading...  ...volumes of real production data. A core focus of this...  .... As a Senior/Staff AI Engineer, you will turn ML...  ...work with AI Scientists, Software Engineers, and Program...  ...production: data, training, and inference pipelines, CI/CD,... 
    Full time
    Shift work

    Gauss Labs

    Palo Alto, CA
    a month ago
  •  ...builds the world's largest AI chip, 56 times larger...  ...industry-leading training and inference speeds; over 10 times...  ...a growing number of data centers. As the fleet expands, the software used to monitor and manage...  ...a senior software engineer to build the platforms and... 

    Cerebras Systems

    Sunnyvale, CA
    4 days ago
  •  ...You'll own how Luma's models get served - integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping expensive GPU fleets busy while meeting internal SLOs. This is large-scale inference systems work: scheduling... 

    Luma

    Redwood City, CA
    13 hours ago
  •  ...Primary Function of Position Advancing embodied AI in robotic surgery requires high-quality data across the robotic platform, the surgical field, and...  ...operating room environment. The Senior AI/Data Science Engineer will advance the data strategy and engineering that turns... 

    Jobleads-US

    Sunnyvale, CA
    3 days ago
  •  ...and high-impact talent who see AI as a teammate – leveraging it...  ...hands-on AI / Machine Learning Engineer who can frame business...  ...machine learning initiatives from data preparation and model development...  ...workflows for batch or real-time inference, evaluation, monitoring,... 
    Full time
    Flexible hours

    MoneyLion

    Mountain View, CA
    10 days ago
  • $70 per hour

     ...the hardware, compiler, and software frameworks for the ML accelerators...  ...! You will: Work with AI agents to create high...  ...in Computer Science, Computer Engineering, Computational Science, or a...  ...key kernels for transformer inference including performance analysis... 
    Hourly pay
    Full time
    Internship
    Summer internship

    Waymo

    Mountain View, CA
    1 day ago
  • $193.3k - $261.5k

     ...model generates depends on data reaching the right...  ...at the right moment. As AI models outgrow any single...  ...builds.We're looking for an engineer to work at the frontier of disaggregated inference: splitting LLM serving...  ...low-level data-movement software across accelerators, servers... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $228k - $285k

     ...generations. Role Summary As a Staff Software Engineer, ML training and inference infrastructure, you will be a member...  ...amount of labeled and unlabeled data Qualifications PhD in CS/CE/EE...  ...such jurisdictions. How We Use AI in Our Hiring Process: To ensure... 
    Full time
    Contract work
    Local area

    Rivian

    Palo Alto, CA
    18 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, Inference (AI Data Engineering). Be the first to apply!