Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Compiler Engineer, Edge-AI & Low-Latency

femtoAI

A technology company based in San Bruno is seeking a Compiler Engineer to work on a custom ML compiler for their AI accelerator. The role requires a minimum of 2 years experience in compilers or edge-AI, with strong proficiency in Python and/or C++. You will be responsible for building model ingestion pipelines, implementing graph transformations, and debugging. The position offers benefits including medical insurance, 401(k), and paid leave for parents. #J-18808-Ljbffr femtoAI

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the ML Compiler Engineer, Edge-AI & Low-Latency in San Bruno, CA vacancy
  • A tech company specializing in AI solutions, based in San Bruno, is looking for a Compiler Developer. The role involves building...  ...execution for a custom ML compiler. Candidates should have...  ...years of experience in compilers or edge AI, with proficiency in Python and... 
    Suggested

    Femtosense

    San Bruno, CA
    2 days ago
  •  ...pioneered a high-performance AI accelerator integrated with an...  ...AI platform, enabling low-latency operation with less energy at...  ...Description You will work on a custom ML compiler that transforms modern ML and...  ...in compilers and/or edge-AI Proficiency in Python and... 
    Suggested

    Femtosense

    San Bruno, CA
    2 days ago
  •  ...Francisco seeks candidates with expertise in AI simulation development. The role...  ...enhancing GPU performance, and ensuring low-latency inference. Applicants should be proficient...  ...opportunities for innovation and cutting-edge technology implementation. #J-18808-Ljbffr... 
    Suggested

    Embedding VC

    San Francisco, CA
    2 days ago
  • $128.7k - $261.3k

     ...accessible mobility. For the AI Kernels & Compilers team, that mission...  ...: turning cutting‑edge perception,...  ...development, and performance engineering so that every cycle on...  ..., and effortless for ML engineers across the...  ...fidelity, and on‑vehicle latency. Along the way, you’... 
    Suggested
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    San Francisco, CA
    1 day ago
  • A leading AI evaluation platform in San Francisco is looking for a Senior Software Engineer specializing in ML infrastructure. The successful candidate will design and develop robust...  ...compensation and a chance to work with cutting-edge AI technologies in a collaborative... 
    Suggested

    LMArena

    San Francisco, CA
    1 day ago
  • $170.1k - $258.3k

     ...mobility. For the AI Kernels & Compilers team, that...  ...turning cutting‑edge perception, prediction...  ...and performance engineering so that every cycle...  ...of our on‑vehicle ML inference for ADAS...  ...meeting strict latency, throughput, and...  ...basedExperience with low latency or real... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    San Francisco, CA
    3 days ago
  • A media technology company in San Francisco is seeking a Founding Engineer specializing in ML Inference. This highly technical role requires expertise in the ML infrastructure stack and aims to optimize generative media performance. The ideal candidate will drive innovations... 
    Relocation package

    Reactor.am

    San Francisco, CA
    1 day ago
  •  ...next generation of AI-driven game...  ...Machine Learning Engineer for On-Device &...  ...directly shapes the latency, quality, memory...  ...quality bars.Do low-level...  ...integration between the ML runtime and the...  ...on on-device / edge inference or real...  ...Familiarity with compiler stacks (MLIR, TVM... 
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    4 days ago
  •  ...for the physical world. Our AI platform provides real-time...  ...at scale, we rely on cutting-edge optimization to ensure our vision...  ...About the Role As an ML / DevOps Engineer, you will play a pivotal role...  ...will ensure high throughput, low latency, and rock-solid reliability... 
    Work at office

    Zensors

    San Francisco, CA
    more than 2 months ago
  • $273k - $345k

     ...Atoms builds Physical AI- real-world robots for...  ...We are roboticists, engineers, operators, and builders...  ...Research and develop cutting edge RL and distillation...  ...models to run with low latency on vehicle edge...  ...PyTorch or JAX) and runtime compilation. ~ Robust programming... 
    Full time
    Internship
    Work at office
    Flexible hours

    Atoms

    San Francisco, CA
    2 days ago
  • Google is seeking a Senior Software Engineer to advance ML compiler technology for TPUs and YouTube-scale recommendations. You will contribute to the Accelerated Linear Algebra (XLA) compiler, author custom kernels in JAX, and optimize models for TPU hardware across the... 

    Google

    San Bruno, CA
    4 days ago
  • Position: Senior ML Performance Engineer Location: SF Bay Area (US) or Toronto...  ...: Full-Time Industry: AI Infrastructure / Compiler Systems Overview A...  ...anywhere.” This includes cloud, edge, and hybrid environments...  ...metrics, and test suites (latency, throughput, memory... 
    Full time

    Amadeus Search

    San Francisco, CA
    1 day ago
  • $300k - $400k

     ...Zeta Global (NYSE: ZETA) is the AI-Powered Marketing Cloud that...  ...As a Principal AI/ML Engineer in our AdTech team, you will...  ...operating at large scale and low latency to handle billions of ad events...  ...capabilities on the cutting edge .   ~ AI & Agentic Applications... 
    Full time

    Zeta Global

    San Francisco, CA
    1 day ago
  • $180k - $270k

     ...building the world's most trusted AI work companion for...  ...Gain exposure to cutting-edge AI for Pro tools and play a...  ...deploying high-throughput, ultra-low-latency inference engines for large language models or...  ...intersection between the core ML training team and the backend... 
    Full time
    Work at office
    Worldwide

    Plaud

    San Francisco, CA
    1 day ago
  • $160k - $230k

     ...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference...  ...engineers to create cutting-edge AI solutions. Join us in...  ...Excellent understanding of low-level operating systems concepts...  ...of Rust, Cython and compilers. About Together AI... 
    Full time

    Together Ai

    San Francisco, CA
    1 day ago
  •  ...founding Machine Learning Engineers (MLEs) to own and...  ...browser understanding, and low-latency systems, shipping...  ..., or consumer-focused "AI browsers," we run AI directly...  ...creates unique ML challenges. This is...  ...app, Cloudflare Workers edge proxy, and inference providers... 
    Full time
    Sleeping nights

    Composite

    San Francisco, CA
    1 day ago
  •  ...at the intersection of AI, robotics, and healthcare...  ...Senior Machine Learning Engineer, you will build the intelligence...  ...-world data, ambiguous edge cases, and high-leverage...  ..., error handling, and low-quality input recovery....  .... Optimize cost, latency, and reliability across... 
    Full time
    Work at office

    Hike Medical

    San Francisco, CA
    1 day ago
  • $137.1k - $201.6k

     ...learning, optimization, and systems engineering to the core decisions behind...  ..., parcel, and catering. ML models and optimization...  ...that require high reliability, low latency, and strong operational discipline...  ...and lead DoorDash’s cutting-edge AI vision for logistics: an LLM-... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    1 day ago
  • $200k - $280k

     ...industries get work done. Our Physical AI and fleet services turn heavy...  ...-knit group of bold thinkers—engineers, innovators, and industry...  ...We're looking for a skilled ML engineer to build the perception...  .... Optimize models for low-latency inference on resource-constrained... 
    Full time
    Temporary work
    Flexible hours
    Shift work

    Agtonomy

    South San Francisco, CA
    1 day ago
  • $345.04k - $399.42k

     ...Principal Machine Learning Engineer within the Creator...  ...development of Embodied AI and Behavioral Agents that...  ...gap between cutting-edge research and massive-scale...  ...to ensure quality, to ML Players with human-like...  ...experience working on low-latency motor control (human-... 
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    3 days ago
  •  ...seeking a Machine Learning Infrastructure Engineer to help architect the compute, training,...  ...experimentation, flexibility, and ultra-low latency serving.Establish comprehensive...  ...optimize resource efficiency.Evaluate cutting-edge developments in model optimization and integrate... 
    Full time
    Work at office
    Flexible hours

    Objective Paradigm

    San Francisco, CA
    4 days ago
  •  ...learning, optimization, and systems engineering to the core decisions behind...  ...retail, parcel, and catering.ML models and optimization...  ...that require high reliability, low latency, and strong operational...  ...and lead DoorDash’s cutting-edge AI vision for logistics: an LLM-... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    1 day ago
  • $197.3k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans with...  ...Staff Machine Learning Engineer with deep expertise in...  ...finetuning to join our ML team. You'll design,...  ...-on: you'll work at a low level with training...  ...speculative decoding, compilation via TorchScript/TensorRT... 
    Full time

    Salesforce

    San Francisco, CA
    2 days ago
  • $218.4k - $273k

     ...bottleneck in Physical AI. This position will be a...  ...Robotics and developing ML pipelines for training and...  ...Senior Machine Learning Engineer, Computer Vision to drive cutting-edge research and development...  ...on edge devices (low power, low latency) and/or large-scale cloud... 
    Full time

    Scale Ai

    San Francisco, CA
    1 day ago
  •  ...Job Description The AI Infrastructure team...  ...Zensors builds the engine that powers our...  ...Learning Engineer in ML Runtime & Optimization...  ...current server and edge compute...  ...throughput and minimize latency. Building Efficient...  ...entire ML framework/compiler stack (e.g., PyTorch... 

    Zensors

    San Francisco, CA
    more than 2 months ago
  • $190k - $250k

     ...harnessed the power of cutting-edge deep learning...  ...be at the forefront of AI-driven advertising solutions...  ..., machine learning engineering, and platform architecture...  ..., build, and maintain low-latency, high-throughput...  ...maintaining, and scaling robust ML and data pipelines. ~... 
    Work at office
    Remote work
    Work from home

    Cognitiv

    San Mateo, CA
    20 days ago
  • $180k - $240k

     ...client, a venture-backed AI Startup, is hiring a talented ML/AI Research Engineer to join their team in...  ...working at the cutting edge of LLMs, vector search,...  ...systems. ~Optimize inference latency and GPU utilization...  ...strategies, reranking, and low-latency deployment (e.g.... 
    Full time

    Alldus International Consulting Ltd

    San Francisco, CA
    more than 2 months ago
  • $295.25k - $345.04k

     ...Learning Infrastructure Engineer, you’ll build scalable,...  ...that powers ML systems across our organization...  ...core modelers, data and AI infrastructure engineers...  ...distributed training to low-latency inference and...  ...the adoption of leading-edge technologies and practices... 
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    4 days ago
  •  ...we’re not just building AI models—we’re redefining...  ...’t: on-device, at the edge, under real-time constraints...  ...of world-class engineers, researchers, and builders...  ...debug models in popular ML frameworks, and...  ...others stall: on CPUs, with low latency, minimal memory, and maximum... 

    Liquid AI

    San Francisco, CA
    3 days ago
  •  ...build, train, and serve AI models tailored to...  ...Applied Machine Learning Engineer, you will serve as a vital...  ...bridge between cutting-edge AI research and...  ...successful deployment of ML solutions. Demo /...  ...AI infrastructure, from low-latency inference to scalable model... 

    Fireworks AI

    San Mateo, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Compiler Engineer, Edge-AI & Low-Latency. Be the first to apply!