Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Embedded ML Inference Optimization Engineer

Decisive Point

Decisive Point is seeking a Software Engineer in Sunnyvale, California, with expertise in optimizing machine learning models for embedded systems. This role involves performance optimization for embedded compute platforms, collaborating with ML engineers, and requires strong software development skills. The ideal candidate has a Bachelor’s degree and at least 3 years of experience in ML accelerators and deep learning frameworks such as PyTorch and JAX. The position offers a competitive salary and comprehensive benefits. #J-18808-Ljbffr Decisive Point

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Embedded ML Inference Optimization Engineer in Sunnyvale, CA vacancy
  • Intel Corporation in Santa Clara seeks an Inference Optimization Engineer to optimize AI models for local and edge environments. Candidates should possess over 5 years of experience in software development, proficient in C++ and Python, and comfortable with performance... 
    Suggested
    Local area

    Intel Corporation

    Santa Clara, CA
    5 days ago
  • $170.5k - $315.49k

     ...future of AI should belong to the people it serves. Role Summary Make models fast on the hardware people actually own. You optimize inference engines (llama.cpp, vLLM) for constrained local and edge environments — GPU/iGPUs, Vulkan backends — not datacenter H100... 
    Suggested
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    4 days ago
  • Intel in Santa Clara, California is seeking a talented individual to optimize inference engines for local environments, impacting the future of AI. Applicants should have a strong background in C++ and software development, with experience in profiling performance issues... 
    Suggested
    Local area

    Intel

    Santa Clara, CA
    4 days ago
  • ScOp Venture Capital is looking for an ML Systems Engineer to optimize LLM inference systems crucial for their AI platform. The role focuses on enhancing performance and efficiency via low-level systems optimization, directly impacting industry leader processes in semiconductor... 
    Suggested

    ScOp Venture Capital

    Santa Clara, CA
    5 days ago
  • $178.7k - $268.1k

     ...We’re hiring a Senior Backend Engineer to build and operate the...  ...reliability, and scalability of inference systems. What You’ll Be Doing...  ...infrastructure. Partner with ML engineers to ensure online...  ...Prometheus and Grafana. Manage and optimize cloud infrastructure on GCP,... 
    Suggested
    Work at office
    Relocation package

    Unity

    Mountain View, CA
    5 days ago
  • NVIDIA Gruppe is looking for a skilled engineer to join their TensorRT Edge-LLM team in...  ...involves developing a state-of-the-art inference framework for large language models and optimizing it for real-time performance on embedded platforms. Candidates should have a strong... 

    NVIDIA Gruppe

    Santa Clara, CA
    5 days ago
  • Hippocratic AI in Palo Alto is seeking an experienced LLM Inference Engineer to optimize large language model serving infrastructure. You will design multi-node architectures, implement quantization, and push latency reductions in production environments. The role requires... 

    Hippocratic-Ai

    Palo Alto, CA
    3 days ago
  • GMI Cloud, based in Mountain View, California, is seeking talented Machine Learning Engineers to advance their leading inference optimization solutions and token platform. You will focus on cutting-edge research and production methodologies that set industry benchmarks... 

    GMI Cloud

    Mountain View, CA
    5 days ago
  • $184k - $287.5k

    NVIDIA is seeking AI systems engineers to innovate in the inference systems software stack. The role involves designing and optimizing libraries and kernel technologies, significantly impacting...  ...) and over 6 years of experience in ML/DL systems development, with strong... 

    NVIDIA

    Santa Clara, CA
    5 days ago
  •  ...automotive software development team part of the Volkswagen Group, seeks a Sr Software Engineer for Embedded Machine Learning in Mountain View, CA. You will design, train, and optimize ML models for embedded accelerators, ensuring real-time performance and production-... 

    Cariad, Inc.

    Mountain View, CA
    5 days ago
  •  ...Technologies in Mountain View, CA, seeks a Senior Machine Learning Engineer to join the Vector Bidding Science team. Define the technical...  ...real-time bidding systems powered by cutting‑edge AI and optimization frameworks. As a key member of the Vector AI org, you will develop... 

    Unity Technologies

    Mountain View, CA
    4 days ago
  •  ...inventory management will optimize the global physical...  ...Learning Software Engineers to build compute-constrained...  ...teams to deploy CV/ML models into production...  ...networks (both training and inference) Knowledge of model...  ...techniques for embedded systems, including knowledge... 
    Visa sponsorship

    Corvus Robotics, Inc.

    Mountain View, CA
    1 day ago
  • $184.7k - $324.8k

    Senior Software Engineer, Middleware - Special Projects...  ...and distributed inference platforms to large‑scale...  ...with hardware, robotics, ML, design, and platform...  ...Collaborate with hardware, embedded systems, and AI...  ...Commitment to measuring and optimizing performance; skilled... 
    Relocation

    Apple Inc.

    Cupertino, CA
    2 days ago
  • Unity in Mountain View is hiring a Senior Backend Engineer to design and operate distributed systems that power large-scale online model inference. You will work on performance, reliability, and scalability of these critical systems. The position requires expertise in... 
    Worldwide

    Unity

    Mountain View, CA
    5 days ago
  • $152k - $241.5k

     ...TensorRT team as a Senior Software Engineer, and be at the forefront of...  ...high‑performance AI inference solutions for automotive safety...  .... Contribute to performance optimization and benchmarking efforts for...  ...with systems programming, embedded systems, and/or compiler development... 

    Nvidia Corporation

    Santa Clara, CA
    5 days ago
  • $174k - $253k

    Senior Software Engineer, Sensor AI/ML, Watch Software Google — Mountain View...  ...production delivery on embedded targets, and experience using...  ..., model evaluation, optimization, data processing, debugging...  ...efficient, micro‑watt edge inference. Develop and maintain high... 

    Google Inc.

    Mountain View, CA
    1 day ago
  • $184k - $287.5k

    NVIDIA is recruiting a Senior Inference Engineer to advance AIConfigurator ( a system that automatically...  ...models on NVIDIA platforms by optimizing efficiency, latency, parallelism, and resource...  ...GPU computing, distributed systems, ML infrastructure, or high-performance... 

    NVIDIA

    Santa Clara, CA
    5 days ago
  • $190k - $230k

     ...industry‑leading training and inference speeds and empowers machine...  ...effortlessly run large‑scale ML applications, without the hassle...  ...Role As a Senior Mechanical Engineer at Cerebras, you will lead...  ...manufacturing‑production, diagnostic and embedded software engineering teams,... 
    For contractors

    CEREBRAS SYSTEMS INC.

    Sunnyvale, CA
    2 days ago
  • $152k - $222k

     ...or leveraging AI solutions, ML APIs, prompting, agent tooling...  ...and modern AI frameworks, and embedding into demos. Experience...  ...product marketing management and engineering teams to stay on top of industry...  ...advice to customers to optimize Google Cloud effectiveness.... 
    Temporary work
    Local area

    Google

    Sunnyvale, CA
    5 days ago
  • $166k - $244k

    Google is looking for a Senior Software Engineer in Sunnyvale, CA to lead GPU performance optimizations for cutting-edge AI and machine learning technologies. This role offers the opportunity to work on innovative projects that impact billions of users around the globe.... 

    Google

    Sunnyvale, CA
    1 day ago
  • $174k - $253k

     ...experience delivering AI/ML algorithms for...  ...production delivery on embedded targets, and experience...  ...deployment, model evaluation, optimization, data processing,...  ...Job Google's software engineers develop the next-generation...  ..., micro-watt edge inference. Develop and maintain... 
    Immediate start

    Google

    Mountain View, CA
    5 days ago
  • $152k - $287.5k

    NVIDIA is seeking a Senior Software Engineer for Quantized Inference in Santa Clara, CA. This role involves implementing quantized and sparse recipes...  ...engines, ensuring efficient model export pipelines, and optimizing throughput for large language models. The ideal... 

    NVIDIA

    Santa Clara, CA
    5 days ago
  •  ...Senior Software & Machine Learning Engineer to join our Energy Optimization team. This role focuses on building...  ...integration applications Build automated ML pipelines for model training,...  ...managing containerized applications in embedded or firmware environments. Expertise... 

    Pentangle Tech Services | P5 Group

    Palo Alto, CA
    5 days ago
  • NVIDIA is seeking a System Software Engineer to work on Dynamo‑Triton Inference Server in Santa Clara, CA. You will contribute to a high-performance, GPU-...  ...Python skills, plus experience with large distributed/ml systems and excellent communication in an agile team.... 

    2100 NVIDIA USA

    Santa Clara, CA
    3 days ago
  •  ...Overview Job Description As the Engineering Manager for the Machine...  ...talented engineers in building, optimizing, and scaling the end-to-end systems for the entire ML/LLM lifecycle. This includes our...  ...for distributed training and inference, model evaluation frameworks,... 
    Work at office
    Remote work
    Flexible hours

    ServiceNow

    Mountain View, CA
    3 days ago
  • $180k - $250k

     ...infrastructure firm is seeking a TPU Systems Engineer to develop high-performance systems using...  ...-model workloads on TPU hardware and optimizing performance across the stack. Candidates...  ...least 3 years of experience in production ML systems and a strong foundation in Python... 

    RadixArk

    Palo Alto, CA
    3 days ago
  • Google is seeking a Software Engineer III for Pixel Performance, focusing on developing cutting...  ..., you'll work on system resources and optimize performance for Pixel devices. The ideal...  ...software development, particularly with embedded systems. This position offers... 

    Google

    Mountain View, CA
    5 days ago
  • $100k

    Staff Field Application Engineer, Customer Success Tenstorrent...  ...who’s wired for AI/ML, fluent in real‑world problem...  .... Experience with embedded systems, FPGA Architecture...  ...hardware/software co‑optimization is often critical for efficient inference at the edge. Good understanding... 
    Permanent employment

    Tenstorrent Inc.

    Santa Clara, CA
    5 days ago
  • $165k - $242k

    Dormont Manufacturing Co is seeking a Senior Engineer to lead designs and improve engineering standards. The role focuses on evolving our Kubernetes-native inference platform and ensuring reliability across multiple services. Qualified candidates should have 5-8 years... 

    Dormont Manufacturing Company

    Sunnyvale, CA
    2 days ago
  •  ...inc. is seeking a Staff Runtime Systems Engineer to join our team in Santa Clara, CA. This...  ...for multiprocessor systems, ensuring optimal runtime performance. Ideal candidates have...  ...Electrical Engineering with at least 5 years of embedded software experience. #J-18808-Ljbffr d-... 
    3 days per week

    d-Matrix inc.

    Santa Clara, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Embedded ML Inference Optimization Engineer. Be the first to apply!