Software Engineer, Inference (AI Data Engineering)
$135k - $175kSpaceX
SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. SOFTWARE ENGINEER, INFERENCE (AI DATA ENGINEERING) The application software team is the central nervous system of SpaceX - we create mission critical applications that are used throughout SpaceX to accelerate launch vehicle production and flight as well as systems that allow Starlink to grow into a worldwide fast, reliable Internet service. We are looking for engineers who treat fellow teammates with fairness, respect, and support. Our team maintains a high-performance AI inference platform that serves the best models internally at SpaceX to accelerate our most ambitious engineering goals. As part of this effort in Palo Alto, you will design and optimize large-scale model serving systems end-to-end, owning everything from distributed infrastructure to deep low-level optimizations. You will work on systems that deliver reliable, high-throughput inference to power SpaceX's mission-critical applications while maintaining the highest standards of performance and availability. Aerospace experience is not required to be successful here - rather we look for smart, motivated, respectful, collaborative engineers who love solving problems and want to make an impact on a super inspiring mission. You will have full ownership of challenging problems, working with a team of enthusiastic engineers with diverse perspectives to design and produce solutions that enable SpaceX to achieve its loftiest engineering goals at a rapid pace. The success of the missions at SpaceX depends on the software that you and your team produce. This role will report through SpaceX Internal AI Infrastructure, as we'll also be providing support for training workloads. RESPONSIBILITIES:
Level 1: $135,000.00 - $175,000.00
Level 2: $155,000.00 - $210,000.00 Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience. Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law. ITAR REQUIREMENTS:
Learn more about the ITAR here. SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status. Applicants wishing to view a copy of SpaceX's Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to
- Develop highly reliable, high-throughput inference systems that serve the best AI models internally across SpaceX
- Architect and implement scalable distributed infrastructure for model serving, including load balancing, auto-scaling, batch scheduling, global KV cache, and continuous batching
- Optimize latency and throughput of model inference under real production workloads, including low-level GPU kernel work, quantization, speculative decoding, and other acceleration techniques
- Build reliable, high-concurrency serving systems with 100% uptime, low tail latency, and excellent observability
- Own end-to-end components such as request routing, SDK development, rate limiting, and efficient scaling for internal SpaceX AI inference platforms
- Benchmark, fine-tune, and accelerate inference engines (e.g., SGLang, vLLM, TensorRT-LLM)
- Develop custom tools for tracing, replaying, and resolving issues across the full stack - from orchestration down to GPU kernels
- Create robust CI/CD infrastructure for seamless endpoint deployment, image publishing, and inference engine updates
- Collaborate across SpaceXAI teams to integrate inference capabilities into broader systems and workflows
- Bachelor's degree in computer science, engineering, math, or scientific discipline; OR 2+ years of professional experience building software in lieu of a degree
- Experience in designing, implementing, and maintaining reliable and horizontally scalable distributed systems
- 1+ years of experience in full stack development or backend development with production systems
- 1+ years of experience with Rust or C++
- Experience with LLM inference engines and serving frameworks (e.g., SGLang, vLLM, Triton, TensorRT-LLM)
- Deep low-level systems programming and optimizations: GPU kernels, code generation, batching, caching, parallelism, quantization, and speculative decoding
- Experience with large-scale, high-concurrency production serving systems
- Knowledge of service observability and reliability best practices
- Experience operating commonly used databases such as PostgreSQL, ClickHouse, or MongoDB
- Experience designing or building with agent SDKs and agent orchestration frameworks
- Experience with Docker, Kubernetes, and containerized applications
- Expert knowledge of gRPC (unary, response streaming, bi-directional streaming, REST mapping)
- Programming experience in Python, Go, or similar languages
- Experience with version control, continuous integration, continuous delivery, build systems, and monitoring
- Expertise in profiling and improving application performance
- You may be asked to work extended hours/weekends dependent on launch cadence and platform demands
- This role requires you to be onsite in Palo Alto. Remote and/or hybrid work will not be considered
Level 1: $135,000.00 - $175,000.00
Level 2: $155,000.00 - $210,000.00 Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience. Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law. ITAR REQUIREMENTS:
Learn more about the ITAR here. SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status. Applicants wishing to view a copy of SpaceX's Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Software Engineer, Inference (AI Data Engineering) in Palo Alto, CA vacancy
- ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SOFTWARE ENGINEER, INFERENCE (AI DATA ENGINEERING)The application software team is the central nervous system of SpaceX - we create mission critical...SuggestedPermanent employmentTemporary workRemote workWorldwideWeekend work
$230k - $350k
...experienced Member of Technical Staff, Inference Systems, to build and optimize a high-performance AI inference platform from the... ...systems and strong systems engineering skills, with Rust experience... ...systems and production software engineering fundamentals. ~...SuggestedWork at office- ...opportunity for you to take your software engineering career to the next level. As... ...enterprise-authorized AI coding assist tools within the... ...from large, diverse data sets in service of continuous... ...TensorFlow Serving, Triton Inference Server)Familiarity with distributed...Suggested
- ...AI/ML Software Engineer At Gallatin, we are rebuilding logistics infrastructure for the national... ...foxhole, we operate at the layer where data becomes decisions, and decisions make... ...large scale ML pipelines and real-time inference systems—while collaborating with cross...SuggestedLocal area
$117.7k - $221.4k
...understand, and curate high-value data from large-scale real-world... ...cost efficient for embodied AI systems. We believe the next generation... ...model reflects how Cola engineers think: build durable... ...processing, featurization, and inference foundations that power scalable...SuggestedFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours- ...reputation with the clients. Currently, we are looking for entry-level software programmers, Java full stack developers, Python/Java developers, data analysts/data scientists, machine learning engineers for full time positions with clients. Who should apply? Recent...Remote jobFull timeH1b
- ...for millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one purpose: to... ...DescriptionPrimary Function of PositionAdvancing embodied AI in robotic surgery requires high-quality data across the robotic platform, the surgical field, and...Local areaWorldwideFlexible hours
$274k - $304k
Saviynt's AI-powered identity platform manages and governs human and non-human... ...all of an organization's applications, data, and business processes. Customers trust... ..., please visit AI Platform Engineer - Training & Inference Saviynt's AI-powered identity platform...$120k - $220k
...information powered by advanced AI, recommendation systems, and... ...every cycle.We're hiring the engineer who owns this agent end-to-end... ...them, trained a LoRA, optimized inference, built a ComfyUI workflow you’... ...first 100 ads shipped CPI data back tighter loop. Not “I’ll spec...Full timeLocal areaWork from home$144k - $236k
...of the team.Responsibilities: AI is at the core of how... ...trust platforms. As a Senior AI Software Engineer you will own end-to-end machine... ...or quality improvement (i.e. inference/training efficiency, engineer... ...tech-debt removal) backed by data.You operate systems reliably...For contractorsWork at officeImmediate startFlexible hours$160.36k - $240.54k
...most immediate and profound opportunity for AI to drive positive change in the physical... ...from designing robust and scalable model & data pipelines to building to deploying the... ...on-road validation.Maintain an in-house ML inference platform to serve large language models efficiently...Immediate startFlexible hours$165.2k - $223.6k
...purchase regretSearch Science Data Infrastructure (SSDI)... ...an ML Engineer you will:Lead development... ...and operations using AWS AI services, DL compute resources... ...deployment to our large scale inference services. You will... ...professional software development experience-...InternshipLocal areaWorldwideFlexible hours$153k - $179k
...everything we do. Expectations are high, and so are the rewards.The Data Engineering team builds and maintains the foundational datasets that... ...experience building end-to-end data pipelinesHands-on software engineering experience, with the ability to write production-...Work at officeFlexible hoursShift work3 days per week$129k - $212k
...of the team.LinkedIn's Data Science team leverages... ...for an ambitious data engineer to have an impact and transform... ...talented and driven Sr Software Engineer, Data Science... ...with LinkedIn’s AI-powered data stack. Successful... ..., causal inference, measurement, and diagnostics...For contractorsWork at officeFlexible hours- ...is a research lab of top researchers and engineers, building the world’s top-ranked realtime... ...used to power the largest consumer-facing AI applications available, across categories... ...-of-the-art models, optimizing realtime inference, and creating best-in-class APIs and products...Full timeContract workWork at officeRelocation
- ...forefront of a new era in enterprise AI — one defined not by model... ...viable at scale. Our Data & AI practice brings together... ...frontier AI research and production engineering — investigating the foundational... ..., model selection and inference routing strategies, autonomy and...Full timeWork experience placementLive inWork at officeLocal areaRelocation
- ...Job Description We are seeking a hands-on Senior Data & AI Full Stack Engineer with marketing technology experience to deliver an outcome-based program across lead scoring and CRM prioritization, Marketo and GenStudio integration, Google PLA enablement, affiliate attribution...
$171k
Sr. Staff ll, AI Engineer - Search (L7-2)We exist to wow our customers. We know we’re... ...vector databases, high-throughput inference, and real-time data pipelines for context enrichment and... ...chains to build intelligent, stable software services.Enforce AI reliability, monitoring...Temporary workFlexible hours$175k - $287k
...of the team.Responsibilities: AI is at the core of how... ...trust platforms. As a Staff AI Software Engineer you will own end-to-end machine... ...or quality improvement (i.e. inference/training efficiency, engineer... ...members of the team backed by data.You operate systems reliably...For contractorsWork at officeImmediate startFlexible hours$144k - $236k
...team. Responsibilities: AI is at the core of how... ...trust platforms. As a Senior AI Software Engineer you will own end-to-end machine... ...or quality improvement (i.e. inference/training efficiency, engineer... ...tech-debt removal) backed by data. You operate systems reliably...For contractorsWork at officeImmediate startFlexible hours$207k - $300k
...month Forward Deployed Engineer (FDE) embeds and 2-4-week... ...years of experience in software development.5 years of... ..., model evaluation, data processing, debugging,... ...integrating generative AI tools or LLM interfaces... ...efficiency pipelines, optimize inference serving engines, and...$286.4k - $358k
Uniphore is the Business AI company. Our sovereign,... ...connects enterprise data, fine-tunes AI models and... ...are seeking a VP of AI Engineering to lead the... ...support SLLM fine-tuning, inference, prompt engineering, and... ...organizations. • Own end-to-end software delivery: planning,...Full time$275.8k - $340.5k
...to meet the unique demands of AI and ML innovation, supporting... ...such as Embodied AI, Simulation, Data Science, and more. We enable... ...enhance the productivity of ML engineers, and drive the adoption of... ...includes: AI Validation & Inference: Ensures robust model performance...Full timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours$110k - $150k
The AI & Data Engineer at EVERFORCE LLC will help build and operate the data and AI foundation behind enterprise initiatives. This onsite role in Santa Clara, CA focuses on scalable pipeline engineering, end-to-end AI/ML model development, and practical integration with...$165.2k - $223.6k
...builds the AWS Neuron SDK, the software development kit used to... ...learning accelerators.The Neuron Inference organization is at the forefront... ...-software boundary, our engineers build systematic infrastructure... ...boundaries of what'spossible in AI acceleration.As part of the...InternshipLocal areaFlexible hours$184k - $287.5k
NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. As a key contributor, you will help design, build, and optimize... ...software that powers today’s most sophisticated AI applications. Our team is responsible for developing...Full time$151.3k - $283.8k
...technological advancements such as cloud, AI, and network security. While driving the... ...context of Large Language Model (LLM) inference and training.2.Operator & Performance Optimization... ...: Master’s or Ph.D. degree in Computer Engineering, Electronic Engineering,...Full timeRelocation package- ...Labs builds Industrial AI for the world's leading... ...volumes of real production data. A core focus of this... .... As a Senior/Staff AI Engineer, you will turn ML... ...work with AI Scientists, Software Engineers, and Program... ...production: data, training, and inference pipelines, CI/CD,...Full timeShift work
- ...Description Job Description We are looking for a product-focused AI Software Engineer to turn advances in AI into restaurant products that work... .... Develop customer-facing interfaces, backend APIs, data models, agent tools, and integrations as one coherent product...Temporary work
$200k - $270k
...Description Samsung SDS America AI Team is researching the... ...AI systems, spanning data collection through... ...a Senior Physical AI Engineer to join the team... ...learning, and production software engineering. You will connect... ...Optimize low-latency inference and control systems for...WorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Software Engineer, Inference (AI Data Engineering). Be the first to apply!
Related searches
- software developer fintech Palo Alto, CA
- startup software engineer Palo Alto, CA
- financial software developer Palo Alto, CA
- junior software developer remote Palo Alto, CA
- software engineer Palo Alto, CA
- software data engineer Palo Alto, CA
- software developer internship no experience Palo Alto, CA
- part time software developer Palo Alto, CA
- software engineer entry level Palo Alto, CA
- software engineer amazon Palo Alto, CA


