Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Application Software Engineer, Inference

$135k - $160k

SpaceX

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.APPLICATION SOFTWARE ENGINEER, INFERENCEThe application software team is the central nervous system of SpaceX – we create mission critical applications that are used throughout SpaceX to accelerate launch vehicle production and flight as well as systems that allow Starlink to grow into a worldwide fast, reliable Internet service. We are looking for engineers who treat fellow teammates with fairness, respect, and support.Our team maintains a high-performance AI inference platform that serves the best models internally at SpaceX to accelerate our most ambitious engineering goals. As part of this effort in Palo Alto, you will design and optimize large-scale model serving systems end-to-end, owning everything from distributed infrastructure to deep low-level optimizations. You will work on systems that deliver reliable, high-throughput inference to power SpaceX’s mission-critical applications while maintaining the highest standards of performance and availability.Aerospace experience is not required to be successful here - rather we look for smart, motivated, respectful, collaborative engineers who love solving problems and want to make an impact on a super inspiring mission. You will have full ownership of challenging problems, working with a team of enthusiastic engineers with diverse perspectives to design and produce solutions that enable SpaceX to achieve its loftiest engineering goals at a rapid pace. The success of the missions at SpaceX depends on the software that you and your team produce.This role will report through SpaceX Application Software while also working closely with xAI engineering teams. RESPONSIBILITIES:Develop highly reliable, high-throughput inference systems that serve the best AI models internally across SpaceXArchitect and implement scalable distributed infrastructure for model serving, including load balancing, auto-scaling, batch scheduling, global KV cache, and continuous batching Optimize latency and throughput of model inference under real production workloads, including low-level GPU kernel work, quantization, speculative decoding, and other acceleration techniques Build reliable, high-concurrency serving systems with 100% uptime, low tail latency, and excellent observability Own end-to-end components such as request routing, SDK development, rate limiting, and efficient scaling for internal SpaceX AI inference platforms Benchmark, fine-tune, and accelerate inference engines (e.g., SGLang, vLLM, TensorRT-LLM) Develop custom tools for tracing, replaying, and resolving issues across the full stack — from orchestration down to GPU kernels Create robust CI/CD infrastructure for seamless endpoint deployment, image publishing, and inference engine updates Collaborate across SpaceXAI teams to integrate inference capabilities into broader systems and workflows BASIC QUALIFICATIONS:Bachelor's degree in computer science, engineering, math, or scientific discipline; OR 2+ years of professional experience building software in lieu of a degreeExperience in designing, implementing, and maintaining reliable and horizontally scalable distributed systems1+ years of experience in full stack development or backend development with production systems1+ years of experience with Rust or C++PREFERRED SKILLS AND EXPERIENCE:Experience with LLM inference engines and serving frameworks (e.g., SGLang, vLLM, Triton, TensorRT-LLM) Deep low-level systems programming and optimizations: GPU kernels, code generation, batching, caching, parallelism, quantization, and speculative decoding Experience with large-scale, high-concurrency production serving systems Knowledge of service observability and reliability best practices Experience operating commonly used databases such as PostgreSQL, ClickHouse, or MongoDB Experience designing or building with agent SDKs and agent orchestration frameworks Experience with Docker, Kubernetes, and containerized applications Expert knowledge of gRPC (unary, response streaming, bi-directional streaming, REST mapping) Programming experience in Python, Go, or similar languages Experience with version control, continuous integration, continuous delivery, build systems, and monitoring Expertise in profiling and improving application performance ADDITIONAL REQUIREMENTS:You may be asked to work extended hours/weekends dependent on launch cadence and platform demands This role requires you to be onsite in Palo Alto. Remote and/or hybrid work will not be considered COMPENSATION AND BENEFITS:Pay Range:Software Engineer/Level I: $135,000.00 - $160,000.00/per yearSoftware Engineer/Level II: $155,000.00 - $185,000.00/per yearYour actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock, stock options, or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law.ITAR REQUIREMENTS:To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here. SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status.Applicants wishing to view a copy of SpaceX’s Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to View email address on click.appcast.io.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Application Software Engineer, Inference in Palo Alto, CA vacancy
  • $228k - $285k

     ...protect it for future generations. Role SummaryAs a Staff Software Engineer, ML training and inference infrastructure, you will be a member of the...  ...teamsPay DisclosureSalary Range for California Based Applicants: $228,000.00 - $285,000.00 (actual compensation will... 
    Suggested
    Full time
    Contract work
    Local area

    Rivian Automotive

    Palo Alto, CA
    1 day ago
  •  ...technology. We are at the forefront of software and hardware innovation, pushing the...  ....The role: Principal System Software Engineer, AI Inference ExecutionWhat you will do:The role...  ...consistent and fair evaluation of al applicants. Thank you for your understanding and... 
    Suggested
    3 days per week

    d-Matrix

    Santa Clara, CA
    2 days ago
  • $207k - $301k

    Lead the software architecture and development for on-device inference, ensuring system readiness for external deployments.Coordinate and drive engineering efforts across multiple cross-functional teams to deliver critical system components.Ensure consistency, performance... 
    Suggested

    Google

    Mountain View, CA
    2 days ago
  • $224k - $356.5k

    We are now looking for a Senior System Software Engineer to work on Dynamo. NVIDIA is hiring...  ...fast-paced team building Generative AI inference platform to make design and deployment...  ...also be eligible for equity and benefits.Applications for this job will be accepted at least... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...and SOTA LLM and Multimodal inference at scale across multi-GPU and...  ...across internal GPU software teams and engage with open-source...  ...ecosystem. THE PERSON:   Skilled engineer with strong technical and...  ...employers and will consider all applicants without regard to age,... 
    Suggested

    AMD

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

    We are now looking for a Senior Software Engineer for Quantized Inference! NVIDIA is seeking software engineers to accelerate the discovery and deployment...  ...4.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until July 26,... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

    We are now looking for a Senior Software Engineer for Deep Learning Inference! Would you like to make a big impact in Deep Learning by helping build a...  ...4.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until June 13... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

    NVIDIA is the platform upon which every new AI‑powered application is built. We are seeking a Senior Software Engineer - AI Inference to advance open‑source LLM serving by contributing directly to upstream inference engines like vLLM and SGLang-ensuring they run best‑in... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ...About The Role The Principal Compiler Engineer - ML Systems position will be...  ...fundamentals. Experience building and deploying software products. Experience with one or more...  ...note that in order to be considered an applicant for any position at SambaNova Systems,... 
    Full time
    Temporary work
    Local area
    Flexible hours

    Sambanova Systems

    Palo Alto, CA
    15 hours ago
  • $193.3k - $261.5k

    We develop AWS Neuron, the complete software stack for Trainium, Amazon's custom cloud...  ...hardware.As a Software Development Engineer on the Inference Model Enablement team, you will...  ...protected status.Los Angeles County applicants: Job duties for this position include... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $207k - $300k

    Exercise judgment to guide sustainable engineering choices for ML systems at scale.Innovate next directions for infrastructure over a 12...  ...degree or equivalent practical experience.8 years of experience in software development.5 years of experience testing, and launching... 

    Google

    Mountain View, CA
    4 days ago
  • $193.3k - $261.5k

     ...AWS) builds AWS Neuron, the software development kit used to accelerate...  ...ML compiler, runtime, and application framework that seamlessly...  ...enabling unparalleled ML inference and training performance.The...  ...hardware-software boundary, our engineers build systematic... 
    Work experience placement
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $135k - $160k

     ...possible, with the ultimate goal of enabling human life on Mars.APPLICATION SOFTWARE ENGINEERThe application software team is the central...  ...worldwide fast, reliable Internet service. We are looking for engineers who treat fellow teammates with fairness, respect, and... 
    Permanent employment
    Temporary work
    Remote work
    Worldwide
    Weekend work

    SpaceX

    Palo Alto, CA
    4 days ago
  •  ...Luma Model Serving Engineer You'll own how Luma's models get served — integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping expensive GPU fleets busy while meeting internal SLOs. This is large-scale... 

    Luma AI

    Redwood City, CA
    4 days ago
  • $184k - $287.5k

    We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency...  ....You will also be eligible for equity and benefits.Applications for this job will be accepted at least until May 2,... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...to work on cutting-edge AI technology for safety-critical applications? Join NVIDIA's TensorRT team as a Senior Software Engineer, and be at the forefront of technology, enabling high-performance AI inference solutions for automotive safety and other specialized platforms... 
    Full time

    Nvidia

    Santa Clara, CA
    11 hours ago
  • $152k - $241.5k

     ...for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for its Deep Learning & AI...  ...has been the backbone of NVIDIA’s inference engine, spanning across data centers...  ...eligible for equity and benefits.Applications for this job will be accepted at least... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...industry-leading training and inference speeds; over 10 times faster...  ...the user experience of AI applications, unlocking real-time iteration...  ...'re hiring a Principal Engineer for our Inference Cloud Platform...  ...10+ years of experience in software engineering, with substantial... 

    Cerebras Systems

    Sunnyvale, CA
    2 days ago
  •  ...industry-leading training and inference speeds; over 10 times faster...  ...the user experience of AI applications, unlocking real-time iteration...  ...RoleWe're hiring a Staff Engineer to own major areas of the architecture...  ...8+ years of experience in software engineering, with... 

    Cerebras Systems

    Sunnyvale, CA
    2 days ago
  • $194k

     ...RoleWe're looking for a Manager - Mobile Enginerring, who will lead the Search API Product &...  ...integration of ML models (on-device inference, server-side ranking signal consumption...  ...compensation that would be offered to the hired applicant in addition to their established salary... 
    Temporary work

    Coupang

    Mountain View, CA
    1 day ago
  • $117.7k - $221.4k

     ...This operating model reflects how Cola engineers think: build durable intermediate artifacts...  ...data processing, featurization, and inference foundations that power scalable world...  ...that match their skills and capabilities. Applicants in the recruitment process may be... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  •  ...a research lab of top researchers and engineers, building the world’s top-ranked realtime...  ...power the largest consumer-facing AI applications available, across categories like health...  ...of-the-art models, optimizing realtime inference, and creating best-in-class APIs and products... 
    Full time
    Contract work
    Work at office
    Relocation

    Inworld AI

    Mountain View, CA
    4 days ago
  • $224k - $356.5k

     ...platform upon which every new AI-powered application is built. We are seeking a deeply technical software manager to lead production AI inference for NVIDIA Inference Microservices (NIM...  ...stack, combining optimized inference engines, model profiles/recipes, validated runtime... 

    Socket.dev

    Santa Clara, CA
    3 days ago
  • $157k - $235k

     ...Development, and Employee Relations.We’re looking for a People Application Developer to join the People Tech team at Snap Inc!What you’...  ...the People Team solutionsWork closely with the product and engineering teams to integrate AI solutions into our systemsExperience integrating... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    2 days ago
  •  ...synthesis and timing closure Support the test program development, chip validation and chip life until production maturity Work with FPGA engineers to perform early prototyping Support hand-off and integration of blocks into larger SOC environments Assist with Algorithm... 

    Intelliswift

    Menlo Park, CA
    2 days ago
  • $119.8k - $234.7k

     ...ContributorTravel: Less than 25%Profession: Software EngineeringDiscipline:...  ...for a Principal Software Engineer - Responsible AI who is...  ...for AI models and agents Inference, routing, orchestration, and...  ...There is a different range applicable to specific work locations,... 
    Ongoing contract
    Work at office
    Local area
    3 days per week

    Microsoft

    Mountain View, CA
    3 days ago
  • $193.93k - $291.15k

     ...investors.About the RoleWe’re looking for an Autonomy Engineer focused on onboard autonomy—the software that runs on the robot/vehicle/embedded computer and...  ...management, and field telemetry.Experience in inference optimizationAt Nuro, we celebrate differences and are... 
    Local area
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    2 days ago
  • $193.3k - $261.5k

     ...scale ML training and real-time inference systems that deliver highly...  .... If you're energized by ML engineering at massive scale with direct...  ...non-internship professional software development experience- 5+...  ...status.Los Angeles County applicants: Job duties for this position... 
    Internship
    Local area
    Worldwide
    Flexible hours

    Amazon

    Palo Alto, CA
    2 days ago
  • $213.51k - $245k

     ...'re looking for an exceptional Senior Software Engineer to help shape the future of our core platforms...  ...systems, ensuring low-latency inference and reliable service operation. Define...  ...benefits.Base pay for the successful applicant will depend on a variety of job-related... 
    Work at office
    Remote work
    Flexible hours
    Shift work
    3 days per week

    Robinhood Financial

    Menlo Park, CA
    4 days ago
  • $193.93k - $291.15k

     ...driverless autonomy. We are looking for strong software engineers to research, develop, and implement...  ...can power L4 driving for robotaxi applications as well as serve as an L2 solution for...  ...systems, distributed training, or inference optimization.At Nuro, we celebrate differences... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Application Software Engineer, Inference. Be the first to apply!