Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

On-Device ML Infra Engineer — API & Model Optimization

Socket.dev

Apple is seeking an ML Infrastructure Engineer focused on ML user experience APIs and integration for on-device AI. You will develop new model authoring and conversion APIs that serve as entry points into Apple’s ML infrastructure and drive onboarding of popular models into Apple’s stack with strong, competitive performance on Apple devices.

You will collaborate across teams to integrate these APIs into internal and external model repositories, optimize pipelines from authoring to runtime, and

#J-18808-Ljbffr
Vacancy posted 15 hours ago
Similar jobs that could be interesting for youBased on the On-Device ML Infra Engineer — API & Model Optimization in Cupertino, CA vacancy
  •  ...Apple silicon. The On-Device Machine Learning...  ...run powerful AI models locally, privately...  ...research, software engineering, hardware...  ...systems, developing optimization toolkits for model...  ...acceleration, building ML compilers and runtimes...  ...user experience APIs and integration.... 
    Suggested

    Socket.dev

    Cupertino, CA
    14 hours ago
  • $160.5k - $240.7k

     ...Inc. Job Area Engineering Group, Engineering...  ...tools to help developers optimize and deploy machine learning models on edge and mobile...  ...optimizing and deploying ML models – especially for edge devices. What You'll Do...  ...and ONNX Develop APIs and developer-facing... 
    Suggested
    Work experience placement
    Immediate start
    Work from home

    Qualcomm

    Santa Clara, CA
    14 hours ago
  •  ...seeking a Senior Tech Lead for ML Infrastructure to drive...  ...of large-scale models across diverse hardware....  ...You will work across data engineering, model development, and on-device deployment to enable scalable...  ...solutions. You will guide optimization efforts for billions-parameter... 
    Suggested

    Neura Market

    Mountain View, CA
    2 days ago
  • $170.6k - $261.3k

     ...About the team: The AV ML Infra team at GM builds end-to...  ...the productivity of ML engineers, and drive the adoption...  ...: Ensures robust model performance by running...  ...Compute: Streamlines and optimizes large-scale ML training...  ...scalable backend services and APIs that power frontend... 
    Suggested
    Full time
    Local area
    Work from home
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  •  ...We are looking for a senior ML infrastructure engineer to build and evolve the systems that support model training, deployment, and production...  ..., inference deployment, APIs, queues, and observability...  ...architecture, or low-level performance optimization. Experience with open-... 
    Suggested

    Maxinsights

    Santa Clara, CA
    28 days ago
  • $174.72k - $295.68k

     ...for a full-time Machine Learning Engineer / Research Scientist to drive the modeling and algorithmic development of XPENG...  ....Contribute to model deployment optimization, including quantization, export,...  ...cross-functionally with infra, perception, and planning teams to... 
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  •  ...worldwide.We’re a team of engineers, clinicians, and...  ...platforms. As a Senior AI/ML Research Engineer, you...  ...fine-tune the foundation models—VFMs, VLMs, and VLA models...  ...on, knowing when to optimize versus when to move fast...  ...level: AssociateIndustry: Medical Device
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    10 hours ago
  •  ...world.Role OverviewAs our Staff Software Engineer, ML infra Engineer for Search & Discovery...  ...unstructured data needed to train complex ML models and efficiently serve them online.The...  ...organization is responsible for optimizing customers' navigation experience and driving... 
    Temporary work

    Coupang

    Mountain View, CA
    2 days ago
  •  ...Principal Machine Learning Engineer to join our Models and Applications team. If you...  ...training pipeline performance.Optimize the distributed training...  ...EXPERIENCE:Experience with ML/DL frameworks such as PyTorch...  ...at scale.Experience with ML infra at kernel, framework, or system... 

    AMD

    San Jose, CA
    4 days ago
  • $153.2k - $234.1k

     ...world scenarios. As a Senior ML Infra Engineer, you will work on the core...  ...advanced Autonomous Driving models. From enabling large foundational...  ...with durable, well-designed APIs. Solid understanding of...  ...performance profiling and training optimization techniques and their impact... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    10 hours ago
  • $153.2k - $234.1k

     ...Join the Embodied AI Infra Foundation team at...  ...machine learning engineer working on our...  ...Autonomous Driving models. From foundational...  ...state-of-the-art optimization, our work is at the...  ...vehicles.As a Senior ML Infra Engineer,...  ...quality, long-lasting APIs.Deep understanding... 
    Full time
    Work at office
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  • $159.05k - $199.3k

     ...leaving earlier when needed to accommodate family commitments. About the role We are looking for a software engineer with deep experience in optimizing ML models and deploying them on production-grade embedded runtime environments. You’ll work across the entire ML framework... 
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    Applied PC

    Sunnyvale, CA
    4 days ago
  • $117.7k - $221.4k

     ...and robotics depends not only on stronger models, but also on better infrastructure for...  ...This operating model reflects how Cola engineers think: build durable intermediate artifacts...  ...quality, speed, and cost instead of optimizing any one of them in isolation.The RoleWe... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • Applied Intuition Inc. is seeking a software engineer with deep expertise in optimizing ML models for production-grade embedded runtime environments. You will...  ..., and driving efficient architectures for memory-constrained devices. #J-18808-Ljbffr Applied Intuition Inc.

    Applied Intuition Inc.

    Sunnyvale, CA
    3 days ago
  • Apple Inc. is seeking an Algorithm Engineer for Home and Audio Devices in Cupertino, CA. You will drive sensing algorithms, ML and data fusion, and lead integration into embedded...  ...teams. You will prototype, validate, and optimize solutions to meet performance and latency... 

    Apple Inc.

    Cupertino, CA
    21 hours ago
  • A leading technology company in Cupertino is seeking a Senior ML Software Engineer to develop innovative machine learning features for Apple Watch. You will work on optimizing ML algorithms using multimodal data and collaborate with cross-functional teams to enhance user... 

    Jobleads-US

    Cupertino, CA
    3 days ago
  •  ...An innovative AI startup is seeking a Founding ML Infrastructure Engineer to take charge of deploying and optimizing production-grade LLM systems. In this core role, you will be responsible for building and managing a full ML serving stack, working closely with product... 

    Realmlabs

    Sunnyvale, CA
    22 hours ago
  •  ...learning and GPU programming engineers to deliver robust compute...  ...Silicon. You will contributes to optimized kernels, ML workflows, and high-...  ...collaboration with architecture teams, API work in Metal, and kernel-...  ...performance across Apple devices. A strong foundation in... 

    Apple Inc.

    Cupertino, CA
    21 hours ago
  • $174.72k - $295.68k

     ...smart connectivity. We are seeking Machine Learning Engineers with strong expertise in generative modeling and large-scale deep learning systems, along with...  ...of data structures, algorithms, code optimization and large-scale data processing. Excellent problem... 
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  •  ...Mountain View, CA seeks an experienced ML engineer to advance growth analytics by building scalable models for lead scoring, conversion prediction, and campaign optimization. You will partner with marketing...  ...hands-on work with marketing APIs and ML frameworks. Authorization... 

    Jobtailor

    Mountain View, CA
    21 hours ago
  •  ...your career. THE ROLE:The AI Models and Applications team at AMD...  ...Sr. Staff or Principal level engineer who is passionate about enabling...  ...inference at scale on AMD devices.Why Join Us?Exciting Opportunities...  ...model training and inference optimizations across a variety of... 

    AMD

    San Jose, CA
    2 days ago
  • $161k - $221k

     ...global leader in materials engineering solutions used to...  ...chips - the brains of devices we use every day. As the...  ...feasibility for classical, and ML/DL based computer...  ...teams such as compute infra, SW, systems and applications...  ..., profile, and optimize algorithms to reduce computational... 
    Full time
    Work experience placement

    Applied Materials

    Santa Clara, CA
    3 days ago
  • $195.2k - $262.2k

     ...enterprises from data and model training through to...  ...building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large...  ...orchestration to inference optimization, we own the hard problems...  ...structured outputs, streaming APIs, high concurrency, and... 
    Full time
    Temporary work
    Immediate start
    Remote work

    Nebius

    Palo Alto, CA
    1 day ago
  • $250k - $350k

     ...are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's...  ...acceleration, GPU parallelism, advanced model deployment, and video generation technologies...  ...at scale.You will design and optimize inference pipelines, implement state-of-... 
    Work at office
    3 days per week

    Pika

    Palo Alto, CA
    3 days ago
  •  ...Cupertino, CA is seeking a Machine Learning Engineering Manager to lead a team building the next...  ...and monetization. Role requires deep ML expertise, experience leading ML teams,...  ...drive innovation in GenAI, campaign optimization, and related areas while ensuring #J-1... 

    Apple Inc.

    Cupertino, CA
    15 hours ago
  •  ...Apple seeks a senior engineer to design the core auction system powering its ads platform at global scale. You will drive optimal outcomes for advertisers, users, and the platform, applying advanced auction theory and machine learning in production environments. You... 

    Socket.dev

    Cupertino, CA
    14 hours ago
  •  ...the Institute of Foundation Models  We are a dedicated research...  ...researchers, data scientists, and engineers, tackling the most...  ...Learning Engineer focused on ML infrastructure and MLOps to design...  ...systems.   ~ Knowledge of cost optimization, security, and networking in... 
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    more than 2 months ago
  • $170k - $216k

     ...advanced machine learning models to deliver training and...  ...and software engineers who are passionate about...  ...to the production and optimization of machine learning models...  ...distributed systems covering the ML lifecycle, supporting...  ...both the user facing API and the internal... 
    Full time

    Waymo

    Mountain View, CA
    2 days ago
  • $189.4k - $300.6k

     ...vehicle development. We engineer high-performance tools that...  ...identify top-performing models and partner with data-intensive ML teams to drive rapid innovation...  ...experiment path toward optimal models.Develop and...  ...Consumption/Mining/Quality), Infra Foundations, and... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    10 hours ago
  • $161k - $239k

     ...Experience in training generative models (e.g., LLMs, GANs, diffusion...  ...PyTorch. ~ Publications in CV/ML venues. About the job Google's software engineers develop the next-generation technologies...  ...We develop ML models for on-device ML solutions, bridging cutting-... 
    Full time

    Google

    Sunnyvale, CA
    14 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to On-Device ML Infra Engineer — API & Model Optimization. Be the first to apply!