Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

On-Device ML Compiler Engineer

Full-time

Apple Oakbrook

Responsibilities

  • Develop and improve MLIR-based compiler infrastructure for machine learning model compilation.
  • Target the Neural Engine, GPU, and CPU to maximize hardware capabilities for ML execution.
  • Support model optimization, transformation, execution, debugging, profiling, and analysis across Apple devices.
  • Collaborate with model authoring, runtime, and performance teams.
  • Write kernels and compiler optimizations for efficient ML model execution.

Requirements

  • 3–5 years of experience working on MLIR-based compilers.
  • Familiarity with common machine learning model architectures, execution schemes, and operations.
  • Familiarity with C++.
  • Familiarity with PyTorch or related training frameworks.
  • Preferred: familiarity with Swift.
  • Preferred: familiarity with programming paradigms for the GPU, CPU, and Neural Engine.
  • Preferred: familiarity with writing kernels for ML model execution.
Vacancy posted 8 days ago
Similar jobs that could be interesting for youBased on the On-Device ML Compiler Engineer in Cupertino, CA vacancy
  •  ...learning model execution across Apple devices. Build model orchestration...  ...Collaborate with model authoring, compiler, and runtime teams. Support model...  ...analysis workflows. Work on ML execution across CPU, GPU, Neural Engine, embedded systems, and larger compute... 
    Suggested
    Full time

    Apple

    Cupertino, CA
    8 days ago
  • $165.2k - $223.6k

    The AWS Neuron Compiler team is actively seeking skilled compiler engineers to join our efforts in developing a state-of-the-art deep learning compiler stack. This...  ...represent the forefront of AWS innovation for advanced ML capabilities, powering solutions like Generative AI... 
    Suggested
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $193.3k - $261.5k

     ...The Inferentia chip delivers best-in-class ML inference performance at the lowest cost...  ...Kit (SDK), which includes an ML compiler, runtime and natively integrates into popular...  ...covers multiple disciplines including silicon engineering, hardware design and verification,... 
    Suggested
    Internship
    Local area
    Work from home
    Relocation
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $193.3k - $261.5k

     ...Neuron is the SDK that optimizes the performance of complex ML models executed on AWS Inferentia and Trainium, our custom chips...  ...deep-learning workloadsThis role is for a senior software engineer in the Compiler team for AWS Neuron. As part of this role, you will be... 
    Suggested
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $193.5k

     ...least 2+ years of experience developing compiler features and optimizations. We need proficiency...  .... Knowledge of compilers or building ML models on accelerators is a plus....  ...closely with chip architects, runtime and OS engineers, scientists, and ML application teams to... 
    Suggested
    Full time

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $165.6k

     ...-grade development experience in object-oriented languages such as C++ or Java. We prefer experience with compiler design for CPU, GPU, vector engines, or ML accelerators. We prefer experience with open-source compiler toolchains such as LLVM or MLIR. We prefer... 
    Full time
    Internship

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $152k - $241.5k

     ...are now looking for a Senior Machine Learning Applications and Compiler Engineer!NVIDIA is seeking engineers to develop algorithms and...  ...approaches for inference and related spatial accelerators at top tier ML, compiler, and computer architecture venues.What we need to see... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $193.3k - $261.5k

     ...forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML...  ...accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that... 
    Internship
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $173k - $253k

    Matterport - Senior ML Ops Engineer Job Description CoStar Group is a leading global provider of commercial and residential real estate...  ...on various hardware platforms (e.g., GPUs, CPUs, edge devices).Collaborate with ML R&D Engineers to understand model architectures... 
    Full time
    Work at office
    Work from home

    Matterport

    Sunnyvale, CA
    1 day ago
  • $152k - $241.5k

     ...lasting impact on the worldWe are looking for a Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for its Deep Learning...  ...NVIDIA inference engine, spanning across data centers, personal devices, automotive, and robotics. The compiler must deliver leading... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $152k - $241.5k

     ...make a lasting impact on the world.NVIDIA is hiring a Senior Compiler Engineer to join our team driving the next generation of GPU systems programming...  ...frameworks, and JIT compilation systems that bridge host and device execution—allowing developers to write memory-safe, high-... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $130.7k - $261.3k

     ...with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 122,000 colleagues...  ...executives, and scientists.THE OPPORTUNITYThis Senior Staff ML Ops Engineer position can work out of our Santa Clara, CA location.Senior... 
    Shift work

    Abbott

    Santa Clara, CA
    2 days ago
  • $100k

     ...to unify innovations in software models, compilers, platforms, networking, and semiconductors...  ...is seeking an Physical Design Engineer to lead cross-functional efforts to solve...  ...will architect, integrate, and deploy AI/ML-driven solutions into production physical... 
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    5 days ago
  • $157.3k - $212.8k

     ...Scientist II with strong science application skills to join our Device Economics team. This role will focus primarily on Amazon's...  ...planning initiatives across Amazon Devices. The DSO team of 300+ engineers, scientists, and PMs applies quantitative methods and data-driven... 
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    3 days ago
  •  ...to identify and resolve performance bottlenecks. Implement compiler optimizations including fusion, sharding, tiling, and scheduling...  ...decisions. Publish research and mentor experienced engineers. Requirements At least 3 years of non-internship professional... 
    Full time
    Internship
    Flexible hours

    Amazon

    Cupertino, CA
    8 days ago
  • $165.6k

     ...perform detailed performance analysis and remove bottlenecks Apply compiler optimizations such as fusion, sharding, tiling, and scheduling...  ...-focused kernels and help customers get the most out of AWS ML accelerators. We operate across the broader Neuron Compiler organization... 
    Full time
    Internship

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  • $193.5k

     ...Experience serving as a mentor, tech lead, or engineering team lead ~ Bachelors degree in...  ...Expertise in accelerator architectures for ML or HPC, such as GPUs, CPUs, FPGAs, or custom...  ...and identify bottlenecks Develop compiler optimizations including fusion, sharding,... 
    Full time
    Internship
    Flexible hours

    Annapurna Labs (U.S.) Inc.

    Cupertino, CA
    1 day ago
  •  ...CAEmployment Type: 1099, C2C, W-2Industry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: ML Engineer with LLMLocation: Sunnyvale, CA(onsite)Job Description:6-8 years of experience in machine learning and LLM, with a proven track... 

    SRI Tech

    Sunnyvale, CA
    3 days ago
  • $161k - $221k

     ...Materials is a global leader in materials engineering solutions used to produce virtually every...  ...and semiconductor chips - the brains of devices we use every day. As the foundation of the...  ...and feasibility for classical, and ML/DL based computer vision algorithms, including... 
    Full time
    Work experience placement

    Applied Materials

    Santa Clara, CA
    4 days ago
  • $250k - $350k

    About the RoleWe are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this...  ...quantization, attention acceleration, and deep learning compiler stacks.GPU & Parallelism: Deep knowledge of GPU programming (CUDA... 
    Work at office
    3 days per week

    Pika

    Palo Alto, CA
    4 days ago
  • $105k - $115k

     ...Quest Global delivers world-class end-to-end engineering solutions by leveraging our deep industry...  ..., energy, hi-tech, healthcare, medical devices, rail and semiconductor industries.We are...  ...in cloud platformsAI Engineer/ ML Engineer with python or typescript. They... 
    Temporary work

    Quest Global Services

    Sunnyvale, CA
    1 day ago
  • $138k - $197k

     ...infrastructure and architecture for on-device AI/ML accelerators.Implement, model, analyze,...  ...architects, design verification, firmware, and compiler teams to build cycle-approximate models...  ...:Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science... 
    Worldwide

    Google

    Mountain View, CA
    3 days ago
  •  ...delivered for millions of patients worldwide. We’re a team of engineers, clinicians, and innovators united by one purpose: to make...  ...other image insights toolkits • Project experience in medical devices or another regulated industry Additional Information Due... 
    Full time
    Local area
    Worldwide
    Flexible hours

    Intuitive & Co

    Sunnyvale, CA
    13 hours ago
  • $159.05k - $199.3k

     ...or leaving earlier when needed to accommodate family commitments. About the role We are looking for a software engineer with deep experience in optimizing ML models and deploying them on production-grade embedded runtime environments. You’ll work across the entire ML framework... 
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    Applied Intuition

    Sunnyvale, CA
    3 days ago
  •  ...Title: ML Ops Engineer Location: Sunnyvale, CA Job Description: Strong Software Engineering fundamentals Experience deploying machine learning models into production environment Strong DevOps, Data Engineering and ML background with... 

    Fisec Global

    Sunnyvale, CA
    4 days ago
  •  ...ML Engineer Sunnyvale, California, United States About the Job Our client is a rapidly growing Tier 1 VC backed startup based in New York with $60 million in funding revolutionizing how outside sales and service teams work. Their AI technology captures and analyzes... 
    Full time

    Catalyst Labs, LLC

    Sunnyvale, CA
    13 hours ago
  • $150.4k - $277.6k

     ...AI/ML Software EngineerThe Video Computer Vision organization is working on exciting...  ...closely with Multimodal GenAI researchers and engineers to develop world-class audio-video...  ...image and video modalities for efficient on-device inferenceBuild tooling for visualization,... 
    Relocation

    Apple

    Sunnyvale, CA
    2 days ago
  • $150.4k - $277.6k

     ...Machine Learning Engineer The Applied Sensing & Health team has built innovative ways for...  .... When you exercise and move with your devices, it's the sensor fusion algorithms from the...  ...other Apple products. We are looking for ML engineers who care deeply about their... 
    Relocation

    Apple

    Cupertino, CA
    13 hours ago
  •  ...Position: Client Engineer Location: Cupertino, CA (Onsite) Duration: C2C Contract Experience: 12+ Years Job Description: • 12+ years of experience in Client Engineering with experince in NLP • Expereince in deploying Client models • Strong understanding... 
    Contract work
    Immediate start

    Syntricate Technologies

    Cupertino, CA
    4 days ago
  •  ...AWS Neuron compiler team is seeking a senior software engineer to build the next generation compiler that transforms ML models into executable code for AWS Inferentia and Trainium. You will collaborate with chip architects, runtime/OS engineers, and ML Apps teams to deliver... 
    Worldwide

    Amazon

    Cupertino, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to On-Device ML Compiler Engineer. Be the first to apply!