Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, ML Inference Performance

Full-time

Sambanova Systems


The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their data, accelerate processes, reduce costs, drive efficiency and innovation to fundamentally transform their businesses and operations at scale.

About The Role


The Principal Compiler Engineer - ML Systems position will be responsible for working with the different layers of the compiler stack and coordinating with other development teams here at SambaNova. It is a critical role responsible for driving innovation in compiler infrastructure and optimization algorithms that enable state-of-the-art ML model performance on the SambaNova platform. This can involve anything from digging through PyTorch and machine learning models to determining how to map operations on to our underlying hardware.

Responsibilities



  • Lead compiler engineering through ensuring standard methodologies, enterprise product insertion and process evolution.

  • Work with peers, domain experts, developers, customers, and work across the enterprise seeking optimal solutions.

  • Develop, integrate, and implement products.

  • Provide support for proposals in key areas aligned with core team competencies.

Basic Qualifications



  • Bachelor’s or Master’s Degree in Computer Science, Computer Engineering, or equivalent with 5-10 years of industry experience.

Additional Qualifications



  • Deep theoretical understanding of compiler fundamentals.

  • Experience building and deploying software products.

  • Experience with one or more deep learning frameworks (i.e. TensorFlow, PyTorch) is a plus.

  • Experience with common compiler development practices and methodologies.

  • Excitement about high-performance systems engineering and performance debugging.

  • An appreciation for process and developing cross-disciplinary collaboration.

Preferred Qualifications



  • Experience with MLIR.

  • Familiarity with machine learning models and frameworks.

  • Familiarity with accelerated computing.

  • Exposure to dataflow architectures.

 


Submission Guidelines
Please note that in order to be considered an applicant for any position at SambaNova Systems, you must submit an application form for each position for which you believe you are qualified. 

EEO Policy
SambaNova Systems is an Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard basis of age (40 and over), color, disability, gender identity, genetic information, marital status, military or veteran status, national origin/ancestry, race, religion, creed, sex (including pregnancy, childbirth, breastfeeding), sexual orientation, and any other applicable status protected by federal, state, or local laws.

Benefits Summary for US-Based, Full-Time Employment Positions
SambaNova offers a competitive total rewards package, including the base salary, plus equity and benefits. We cover 95% premium coverage for employee medical insurance, and 77% premium coverage for dependents and offer a Health Savings Account (HSA) with employer contribution. We also offer Dental, Vision, Short/Long term Disability, Basic Life, Voluntary Life, and AD&D insurance plans in addition to Flexible Spending Account (FSA) options like Health Care, Limited Purpose, and Dependent Care. Our library of well-being benefits available to you and your dependents includes a full subscription to Headspace, Gympass+ membership with access to physical gyms, One Medical membership, counseling services with an Employee Assistance Program, and much more.

Vacancy posted 23 hours ago
Similar jobs that could be interesting for youBased on the Software Engineer, ML Inference Performance in Palo Alto, CA vacancy
  • $126k - $248k

     ...re looking for a Senior Engineer to help build the next-generation inference platform that supports embedding...  ...and collaborate with ML researchers and...  ...designed for reliability, performance, and ease of use. We're...  ...systems at scale Strong software engineering skills in languages... 
    Performance
    Local area
    Flexible hours

    United States Digital Space LLC

    Palo Alto, CA
    4 days ago
  • $153k - $222k

     ...defense. As a full-stack engineer, you'll be involved in all...  ...bringing the latest and greatest software advancements to the...  ...learning model training, and inference Work on ML-adjacent tooling and infrastructure...  ...as it grows, ensuring performance, reliability, and security... 
    Performance
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    Applied Intuition

    Mountain View, CA
    23 hours ago
  • $160k - $240k

     ...looking for self-motivated engineers to build the next-generation...  ...mission is to provide a high-performance, highly reliable foundation...  ...modules Collaborate with other software teams to build foundational...  ...Robotics experience, ML inference optimization experience, computer... 
    Performance

    Nuro

    Mountain View, CA
    4 days ago
  •  ...proprietary providers. Our inference platform serves trillions of...  .... We are looking for strong Software Engineers to join our team. Why this role...  ...If you’re excited about AI/ML, have built and shipped...  ...new features, improve system performance, and contribute to overall system... 
    Performance
    Full time

    Deepinfra

    Palo Alto, CA
    1 day ago
  • $169k - $192k

     ...and IPO readiness. As a Senior Software Engineer at Obsidian, you’ll Own...  ...shipped software Improve the performance, reliability, scalability, and...  ...production Understand core AI/ML concepts such as LLMs, embeddings, vector databases, inference, and evaluation Experience integrating... 
    Performance
    Temporary work
    Fixed term contract
    Worldwide
    Flexible hours

    Obsidian Security

    Palo Alto, CA
    4 days ago
  • $255k - $405k

     ...user experience. This role focuses on performance, reliability, and thoughtful UX for AI-...  .... Collaborate closely with backend and ML engineers on API contracts and system behavior. Ensure...  ...with MLKit or light on-device inference. Published production apps on the Google... 
    Performance

    A1 Services

    Palo Alto, CA
    3 days ago
  • $213.51k - $245k

     ...their careers. We’re a high-performing, fast-moving team with ethics...  ...for an exceptional Senior Software Engineer to help shape the future of...  ...implementation of end‑to‑end ML pipelines, including data ingestion...  ..., ensuring low‑latency inference and reliable service... 
    Performance
    Work at office
    Remote work
    Flexible hours
    Shift work
    3 days per week

    Robinhood

    Menlo Park, CA
    3 days ago
  • $180k - $258.75k

     ...Models. We are looking for a Senior Software Engineer to join our end-to-end automated...  ...in C++ and Python, that supports ML training, evaluation, and inference workflows. Build and maintain ML...  ..., runtime integration, and performance validation on embedded compute platforms... 
    Performance
    Local area
    Shift work

    Toyota Research Institute

    Los Altos, CA
    2 days ago
  • $152k - $204k

     ...combines superior infrastructure performance with deep technical expertise...  ...What You'll Do: Senior engineers are area owners who lead designs...  ...evolve our Kubernetes-native inference platform and meet strict P99...  .... ~ Optimize end-to-end ML system performance by developing... 
    Performance
    Permanent employment
    Temporary work
    Casual work
    Work at office
    Flexible hours
    Shift work

    CoreWeave

    Sunnyvale, CA
    4 days ago
  •  ...About the Role As a software engineering intern, you will work closely...  ...Platform, Onboard Systems, ML Infrastructure, Simulation,...  ...provide a reliable and high-performance platform that allows our autonomy...  ...-cloud training and onboard inference. Our solutions include a distributed... 
    Performance
    Internship

    Nuro

    Mountain View, CA
    23 hours ago
  •  ...experience, how AI feels, responds, and performs in users’ hands. This is not a thin...  .... Collaborate closely with backend and ML engineers on API design and system behavior. Maintain...  ...SQL / noSQL TensorFlow Lite (on-device inference) How We Work The best products today... 
    Performance
    Remote work

    A1

    Palo Alto, CA
    1 day ago
  • $92k - $135k

     ...combines superior infrastructure performance with deep technical expertise...  ...What You'll Do: Join the Inference team to ship production...  ...mentorship from experienced engineers. About the role: Implement...  ...that deployed a microservice or ML inference demo. Coursework... 
    Performance
    Permanent employment
    Temporary work
    Casual work
    Internship
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    4 days ago
  • # Senior Software Engineer, AI PlatformMeta## Job DescriptionMeta AI is looking for experienced Senior Software Engineers...  ...systems for AI development.- Optimize the performance and efficiency of ML training and inference infrastructure.- Collaborate with research scientists... 
    Performance

    Aibreakingwire

    Menlo Park, CA
    2 days ago
  •  ...and APIs that support model inference, orchestration, and tool execution...  ...features Collaborate with ML engineers to integrate, evaluate, and...  ...Help improve system performance, reliability, and scalability...  ...1-3+ years of professional software engineering experience Strong... 
    Performance
    Full time

    Eightfold

    Santa Clara, CA
    23 hours ago
  • $120k - $400k

     ...to systems including hardware and software to train and run the largest ML workloads for AGI. We primarily use...  ...and maintain functional and performance models of our hardware Develop...  ...equivalent degree Excellent software engineering skills, with a focus on maintainable... 
    Performance
    Full time
    Work experience placement
    Local area

    Matx

    Mountain View, CA
    23 hours ago
  • $204k - $259k

     ...5+ U.S. states. The Waymo ML Frameworks & Efficiency team...  ...core to our autonomous driving software. We help our partners by...  ...driving. We are looking for engineers with ML system expertise to help...  ...quickly to improve model performance and training workflows. Stay... 
    Performance
    Full time
    Remote work

    Waymo

    Mountain View, CA
    23 hours ago
  • $175k - $220k

     ...models and the fastest, most scalable inference. We’ve been independently...  ...The Role:  We're looking for a Software Engineer focused on Performance Optimization to help push the boundaries...  ...and video models) Collaborate with ML researchers to co-design and tune model... 
    Performance
    Full time

    Fireworks Ai

    Redwood City, CA
    23 hours ago
  • $170k - $216k

     ...strategic insights into driver performance. We design automated analytic...  ...to directly feed the ML training flywheel and ensure...  ...role, you will report to an engineering manager.   You will:...  ...related field. ~2+ years of software development experience (3+ years... 
    Performance
    Full time
    Remote work

    Waymo

    Mountain View, CA
    23 hours ago
  • $120k - $400k

     ...silicon to systems including hardware and software to train and run the largest ML workloads for AGI. We primarily use the...  ...equivalent degree  Possess outstanding software engineering skills with a focus on efficiency and performance Possess experience in compiler... 
    Performance
    Full time
    Work experience placement
    Local area

    Matx

    Mountain View, CA
    23 hours ago
  • $180k

     ...small, highly motivated, and focused on engineering excellence. This organization is for...  ...that streamline workflows for training, inference, evaluation, and deployment. If you're...  ...cutting-edge GPU resources for maximum performance. Innovate new approaches that bring... 
    Performance
    Full time
    Temporary work

    Xai

    Palo Alto, CA
    23 hours ago
  •  ...the AI Data Cloud—delivering industry-leading performance and price at scale across data engineering, analytics, and AI/ML workloads. Every day, new customers begin their...  ...to prototypes, early products. OUR IDEAL SOFTWARE ENGINEER WILL HAVE: Strong software... 
    Performance
    Full time

    Snowflake

    Menlo Park, CA
    23 hours ago
  •  ...We are looking for a Senior Software Engineer to build scalable AI...  ...environments to create a high-performance platform for Physical AI....  ...training artifacts into on-robot inference stacks. Requirements...  ...proficiency with C++, Python, and ML frameworks (e.g., PyTorch, JAX... 
    Performance
    Full time
    Work at office
    Visa sponsorship

    RoboForce

    Milpitas, CA
    23 hours ago
  • $180k - $300k

     ...dramatically increase model performance as if you had trained...  ...far less compute at inference time, substantially...  ...research and data engineering necessary to solve this...  ...role, you will build software from the ground up to...  ...Have prior experience in ML/AI (preferred but not... 
    Performance
    Full time
    Work at office
    Relocation package

    Datologyai

    Redwood City, CA
    23 hours ago
  • $166k - $244k

     ...we’re a team of scientists, engineers, machine learning experts and...  ...up for success as an Senior Software Engineer in the Gemini App team...  ...to enhance and refine model performance About You In order to...  ...analysis skills Experience with ML, especially LLMs Ability... 
    Performance
    Full time

    Deepmind

    Mountain View, CA
    23 hours ago
  • $145k - $219k

     ...efficiently. We continuously refine our routing engine to calculate more efficient routes,...  ...and know how to balance correctness and performance. You are proficient in C++ programming...  ...Experience in training and inferencing ML models. You have extensive experience... 
    Performance
    Full time

    Nuro

    Mountain View, CA
    23 hours ago
  • $175k - $220k

     ...and the fastest, most scalable inference. We’ve been independently...  ...As a Training Infrastructure Engineer, you'll design, build, and optimize...  ...essential to developing high-performance AI training infrastructure....  ...with distributed systems and ML infrastructure ~ Experience... 
    Performance
    Full time

    Fireworks Ai

    Redwood City, CA
    23 hours ago
  • $152k - $241.5k

    We are now looking for a Senior Software Engineer for Quantized Inference! NVIDIA is seeking software engineers...  ...specifications into functionally correct, performant code, e.g., writing Triton kernels,...  ...‑assisted tooling Experience with ML accelerators with a basic... 

    NVIDIA

    Santa Clara, CA
    5 days ago
  • $204k - $259k

     ...Driver. Our team is a diverse, and collaborative group of software engineers, machine learning (ML) engineers, and data scientists. We develop industry-...  ...advanced ML algorithms that measure and enhance the performance of the Waymo Driver. We develop industry-leading... 
    Performance
    Full time
    Work experience placement
    Remote work

    Waymo

    Mountain View, CA
    4 days ago
  • $152k - $228k

     ...regression test different aspects of the software and hardware integration layer. This performance simulation platform includes...  ...every autonomy code change, from ML model updates to radius of map...  ...profiling tools, and much much more. Engineers across the company rely on this... 
    Performance
    Full time
    Temporary work

    Who We Are Nuro

    Mountain View, CA
    23 hours ago
  • $200k - $287.5k

     ...In Snowflake Notebooks, you can perform exploratory data analysis, develop...  ...perform other data science and data engineering tasks all in one place. As a Senior Software Engineer for Snowflake Notebooks,...  ...tools, data infrastructure, or ML is a plus Demonstrated technical... 
    Performance
    Flexible hours

    Snowflake Computing

    Menlo Park, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, ML Inference Performance. Be the first to apply!