Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

On-Device ML Infra Engineer — API & Model Optimization

Socket.dev

Apple is seeking an ML Infrastructure Engineer focused on ML user experience APIs and integration for on-device AI. You will develop new model authoring and conversion APIs that serve as entry points into Apple’s ML infrastructure and drive onboarding of popular models into Apple’s stack with strong, competitive performance on Apple devices. You will collaborate across teams to integrate these APIs into internal and external model repositories, optimize pipelines from authoring to runtime, and #J-18808-Ljbffr Socket.dev

Vacancy posted 13 hours ago
Similar jobs that could be interesting for youBased on the On-Device ML Infra Engineer — API & Model Optimization in Cupertino, CA vacancy
  •  ...Apple silicon. The On-Device Machine Learning...  ...run powerful AI models locally, privately...  ...research, software engineering, hardware...  ...systems, developing optimization toolkits for model...  ...acceleration, building ML compilers and runtimes...  ...user experience APIs and integration.... 
    Suggested

    Socket.dev

    Cupertino, CA
    13 hours ago
  • $160.5k - $240.7k

     ..., Inc. Job Area: Engineering Group, Engineering...  ...machine learning models on edge and mobile...  ...supportcutting-edgemodel optimization workflows —...  ...especially for edge devices. What You'll Do...  ...workflows with popular ML frameworks —...  ...PyTorchand ONNX Develop APIs and developer-... 
    Suggested
    Work experience placement
    Immediate start
    Work from home

    Socket.dev

    Santa Clara, CA
    2 days ago
  •  ...seeking a Senior Tech Lead for ML Infrastructure to drive...  ...of large-scale models across diverse hardware....  ...You will work across data engineering, model development, and on-device deployment to enable scalable...  ...solutions. You will guide optimization efforts for billions-parameter... 
    Suggested

    Neura Market

    Mountain View, CA
    1 day ago
  • $170.6k - $261.3k

     ...About the team: The AV ML Infra team at GM builds end-to...  ...the productivity of ML engineers, and drive the adoption...  ...: Ensures robust model performance by running...  ...Compute: Streamlines and optimizes large-scale ML training...  ...scalable backend services and APIs that power frontend... 
    Suggested
    Full time
    Local area
    Work from home
    Flexible hours

    General Motors

    Sunnyvale, CA
    23 hours ago
  •  ...We are looking for a senior ML infrastructure engineer to build and evolve the systems that support model training, deployment, and production...  ..., inference deployment, APIs, queues, and observability...  ...architecture, or low-level performance optimization. Experience with open-... 
    Suggested

    Maxinsights Corporation

    Santa Clara, CA
    22 days ago
  • $174.72k - $295.68k

     ...for a full-time Machine Learning Engineer / Research Scientist to drive the modeling and algorithmic development of XPENG...  ....Contribute to model deployment optimization, including quantization, export,...  ...cross-functionally with infra, perception, and planning teams to... 
    Full time

    XPENG Motors

    Santa Clara, CA
    1 day ago
  •  ...worldwide.We’re a team of engineers, clinicians, and...  ...platforms. As a Senior AI/ML Research Engineer, you...  ...fine-tune the foundation models—VFMs, VLMs, and VLA models...  ...on, knowing when to optimize versus when to move fast...  ...level: AssociateIndustry: Medical Device
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    4 days ago
  •  ...world.Role OverviewAs our Staff Software Engineer, ML infra Engineer for Search & Discovery...  ...unstructured data needed to train complex ML models and efficiently serve them online.The...  ...organization is responsible for optimizing customers' navigation experience and driving... 
    Temporary work

    Coupang

    Mountain View, CA
    1 day ago
  •  ...Principal Machine Learning Engineer to join our Models and Applications team. If you...  ...training pipeline performance.Optimize the distributed training...  ...EXPERIENCE:Experience with ML/DL frameworks such as PyTorch...  ...at scale.Experience with ML infra at kernel, framework, or system... 

    AMD

    San Jose, CA
    3 days ago
  • $153.2k - $234.1k

     ...world scenarios. As a Senior ML Infra Engineer, you will work on the core...  ...advanced Autonomous Driving models. From enabling large foundational...  ...with durable, well-designed APIs. Solid understanding of...  ...performance profiling and training optimization techniques and their impact... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago
  • $153.2k - $234.1k

     ...Join the Embodied AI Infra Foundation team at...  ...machine learning engineer working on our...  ...Autonomous Driving models. From foundational...  ...state-of-the-art optimization, our work is at the...  ...vehicles.As a Senior ML Infra Engineer,...  ...quality, long-lasting APIs.Deep understanding... 
    Full time
    Work at office
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    23 hours ago
  • $117.7k - $221.4k

     ...and robotics depends not only on stronger models, but also on better infrastructure for...  ...This operating model reflects how Cola engineers think: build durable intermediate artifacts...  ...quality, speed, and cost instead of optimizing any one of them in isolation.The RoleWe... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  • Applied Intuition Inc. is seeking a software engineer with deep expertise in optimizing ML models for production-grade embedded runtime environments. You will...  ..., and driving efficient architectures for memory-constrained devices. #J-18808-Ljbffr Applied Intuition

    Applied Intuition

    Sunnyvale, CA
    2 days ago
  • Apple Inc. is seeking an Algorithm Engineer for Home and Audio Devices in Cupertino, CA. You will drive sensing algorithms, ML and data fusion, and lead integration into embedded...  ...teams. You will prototype, validate, and optimize solutions to meet performance and latency... 

    Apple Inc.

    Cupertino, CA
    13 hours ago
  • A leading technology company in Cupertino is seeking a Senior ML Software Engineer to develop innovative machine learning features for Apple Watch. You will work on optimizing ML algorithms using multimodal data and collaborate with cross-functional teams to enhance user... 

    Jobleads-US

    Cupertino, CA
    2 days ago
  • $159.05k - $199.3k

     ...leaving earlier when needed to accommodate family commitments. About the role We are looking for a software engineer with deep experience in optimizing ML models and deploying them on production‑grade embedded runtime environments. You’ll work across the entire ML framework... 
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    Applied Intuition

    Sunnyvale, CA
    2 days ago
  •  ...learning and GPU programming engineers to deliver robust compute...  ...Silicon. You will contributes to optimized kernels, ML workflows, and high-...  ...collaboration with architecture teams, API work in Metal, and kernel-...  ...performance across Apple devices. A strong foundation in... 

    Apple

    Cupertino, CA
    13 hours ago
  • An innovative AI startup is seeking a Founding ML Infrastructure Engineer to take charge of deploying and optimizing production-grade LLM systems. In this core role, you will be responsible for building and managing a full ML serving stack, working closely with product... 

    Realm Labs LLC

    Sunnyvale, CA
    1 day ago
  • $174.72k - $295.68k

     ...smart connectivity. We are seeking Machine Learning Engineers with strong expertise in generative modeling and large-scale deep learning systems, along with...  ...of data structures, algorithms, code optimization and large-scale data processing. Excellent problem... 
    Full time

    XPENG Motors

    Santa Clara, CA
    1 day ago
  •  ...Mountain View, CA seeks an experienced ML engineer to advance growth analytics by building scalable models for lead scoring, conversion prediction, and campaign optimization. You will partner with marketing...  ...hands-on work with marketing APIs and ML frameworks. Authorization... 

    Jobtailor

    Mountain View, CA
    13 hours ago
  •  ...the Institute of Foundation Models  We are a dedicated research...  ...researchers, data scientists, and engineers, tackling the most...  ...Learning Engineer focused on ML infrastructure and MLOps to design...  ...systems.   ~ Knowledge of cost optimization, security, and networking in... 
    Visa sponsorship

    Institute of Foundation Models

    Sunnyvale, CA
    more than 2 months ago
  •  ...your career. THE ROLE:The AI Models and Applications team at AMD...  ...Sr. Staff or Principal level engineer who is passionate about enabling...  ...inference at scale on AMD devices.Why Join Us?Exciting Opportunities...  ...model training and inference optimizations across a variety of... 

    AMD

    San Jose, CA
    1 day ago
  • $250k - $350k

     ...are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's...  ...acceleration, GPU parallelism, advanced model deployment, and video generation technologies...  ...at scale.You will design and optimize inference pipelines, implement state-of-... 
    Work at office
    3 days per week

    Pika

    Palo Alto, CA
    2 days ago
  • $161k - $221k

     ...global leader in materials engineering solutions used to...  ...chips - the brains of devices we use every day. As the...  ...feasibility for classical, and ML/DL based computer...  ...teams such as compute infra, SW, systems and applications...  ..., profile, and optimize algorithms to reduce computational... 
    Full time
    Work experience placement

    Applied Materials

    Santa Clara, CA
    2 days ago
  • $189.4k - $300.6k

     ...vehicle development. We engineer high-performance tools that...  ...identify top-performing models and partner with data-intensive ML teams to drive rapid innovation...  ...experiment path toward optimal models.Develop and...  ...Consumption/Mining/Quality), Infra Foundations, and... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago
  •  ...Mountain View is seeking an experienced professional to enhance ML infrastructure for training workloads. Responsibilities include...  ...distributed input data pipelines and collaborating with research teams to optimize systems. Candidates should have a strong background in Python... 

    Waymo

    Mountain View, CA
    1 day ago
  • $161k - $239k

     ...Experience in training generative models (e.g., LLMs, GANs, diffusion...  ...or PyTorch. Publications in CV/ML venues. About the job Google's software engineers develop the next-generation technologies...  .... We develop ML models for on-device ML solutions, bridging cutting-edge... 
    Full time

    Google

    Sunnyvale, CA
    1 day ago
  •  ...ML Engineer Sunnyvale, California, United States About the Job...  ...how businesses learn from and optimize their in-person customer experiences...  ...of applied intelligence from model optimization to productized...  ...with FastAPI, OpenAI APIs, Baseten, LiteLLM, LiveKit, PostgreSQL... 
    Full time

    Catalyst Labs, LLC

    Sunnyvale, CA
    3 days ago
  • $157.3k - $212.8k

     ...application skills to join our Device Economics team. This role...  ...the intersection of economic modeling, forecasting science, and...  ...Devices. The DSO team of 300+ engineers, scientists, and PMs applies...  ...engineering, operations research, optimization, data mining, analytics, or... 
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    1 day ago
  • Decisive Point is seeking a Software Engineer in Sunnyvale, California, with expertise in optimizing machine learning models for embedded systems. This role involves performance...  ...compute platforms, collaborating with ML engineers, and requires strong software development... 

    Decisive Point

    Sunnyvale, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to On-Device ML Infra Engineer — API & Model Optimization. Be the first to apply!