Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Machine Learning Engineer

Unity Technologies

Remote, California, USA

San Francisco, CA, USA

Mountain View, CA, USA

Full time

JOBREQ-2616041

The opportunity

We are building the next generation of AI-driven game experiences, running generative models on-device, right where the players are — on phones, tablets, laptops, and desktops. Our games run inside a modern, browser-native runtime (built on technologies such as WebGPU and WebNN), so the models that power these experiences must be deployed and accelerated entirely within that runtime. As a Senior Machine Learning Engineer for On-Device & Mobile AI, you will take state-of-the-art multi-modal models — transformers, diffusion networks, and vision-language models (VLMs) — and make them run fast, small, and reliably on mobile and constrained hardware.

This is a deeply hands-on role. You will own the optimization and deployment of significant parts of the inference stack — from a trained checkpoint leaving research, through export, quantization, and kernel-level tuning, to a shipped feature running inside the engine at interactive frame rates within a fixed memory and power budget. Your work directly shapes the latency, quality, memory footprint, and battery profile of AI features experienced by billions of players.

This role is for an engineer who is energized by the gap between a research model and a shipping, on-device product. If you enjoy profilers, frame captures, op-fusion, and shaving milliseconds and megabytes, this is your role.

What you'll be doing

  • Inference & On-Device Optimization

  • Own the optimization pipeline for the models you ship: model export, graph transformation, operator fusion, memory-layout planning, and hardware-specific tuning across NPU, mobile GPU, and desktop/laptop GPU.

  • Apply quantization (INT4/INT8/FP16), weight sharing, structured/unstructured pruning, and knowledge distillation to hit hard latency, memory, and power budgets — and validate them against quality bars.

  • Do low-level performance work: write and tune WebGPU compute shaders (WGSL) and, where relevant, native kernels (Metal, Vulkan/SPIR-V compute, CUDA); profile with browser and platform tools (Chrome/Dawn GPU traces, PIX, Instruments/Metal System Trace,

  • Snapdragon Profiler, Nsight, RenderDoc), and eliminate bottlenecks at the op and memory-bandwidth level.

  • Apply efficiency techniques — dynamic resolution, token reduction, cross-frame caching/reuse, reduced-step diffusion samplers — as engineering levers to meet budgets on target SKUs.

  • Runtime & Systems Integration

  • Work with WebGPU-targeted inference runtimes (ONNX Runtime Web, Transformers.js, WebLLM, TensorFlow.js) alongside native options (CoreML, ONNX Runtime, TFLite, ExecuTorch), and extend or build glue code where off-the-shelf options fall short of our diffusion and VLM workloads.

  • Build parts of the integration between the ML runtime and the game engine: real-time scheduling, memory pooling, zero-copy buffer sharing between the inference and render paths, and frame-budget management alongside the renderer.

  • Build supporting engineering for your components: model packaging and asset pipelines, on-device fallbacks and SKU-aware capability tiers, crash/quality telemetry, and automated on-device benchmarking in CI.

  • Research Productionization

  • Partner with research scientists to turn novel CV and multi-modal architectures into implementations that are deployable, debuggable, and fast on device.

  • Provide a feedback loop into research: surface hardware constraints, op-support gaps, and cost models early so model design and deployment converge.

  • Track breakthroughs in efficient inference (efficient attention, distillation, reduced-step diffusion) and assess them pragmatically: what actually moves latency/memory/power on our target devices.

  • Collaboration & Engineering Quality

  • Contribute to engineering best practices, code-review standards, performance-regression gates, and on-device benchmarking methodology.

  • Support a culture of measurement: track KPIs for latency, quality, memory, and power for the systems you work on, across the device matrix.

  • Partner with platform engineers, product managers, and runtime teams to align your work with device-SKU constraints and product roadmaps.

  • Share knowledge and mentor junior and mid-level engineers through code review, pairing, and design discussion.

What we're looking for

  • 5+ years in software/ML engineering, with meaningful time focused on on-device / edge inference or real-time, performance-critical systems.

  • Production deployment of transformer- and/or diffusion-based models (e.g., ViT, Stable Diffusion, CLIP/SigLIP-style encoders) on mobile, desktop, or embedded hardware — shipped, not just prototyped.

  • Hands-on experience with at least one major inference runtime (ONNX Runtime / ORT Web, CoreML, TFLite, ExecuTorch) and a working understanding of operator fusion, memory layout, and runtime scheduling.

  • Low-level performance engineering: solid command of at least one GPU/compute API — WebGPU/WGSL, Metal, Vulkan, D3D12, or CUDA — and the profiling tools to go with it. You can read a frame capture and a kernel trace and reason about where the time and memory go.

  • Working knowledge of model-optimization techniques — quantization (INT4/INT8/FP16), weight sharing, pruning, and distillation — and the judgment to apply them to hit latency and memory budgets. You use them effectively as engineering tools.

  • Understanding of target hardware: mobile SoCs (Apple Neural Engine, Qualcomm Hexagon/Adreno, ARM Mali) and/or desktop/laptop GPUs (Apple Silicon, NVIDIA, AMD, Intel).

  • Strong Python for export pipelines and training-side tooling; familiarity with the core languages of a browser-native runtime (TypeScript/JavaScript, WGSL) is a plus.

  • Working fluency with the models you deploy — enough to read an architecture, modify it for deployment, and reason about accuracy trade-offs.

  • A collaborative working style: clear communication, reliable delivery, and a willingness to support and learn from teammates.

You might also have

  • Experience shipping world-model, neural-rendering, or real-time generative pipelines NeRF, 3DGS, real-time diffusion, or similar) on device.

  • Hands-on experience deploying models through WebGPU — e.g., ONNX Runtime Web WebGPU EP), Transformers.js, WebLLM, or TensorFlow.js — including writing/tuning WGSL compute shaders.

  • Game-engine or real-time-graphics background (Unity, Unreal, or a custom engine; Metal/Vulkan/D3D/OpenGL ES render pipelines) — especially integrating compute workloads alongside a renderer.

  • Contributions to open-source ML inference frameworks, runtimes, or GPU/compute libraries especially in the WebGPU ecosystem (Dawn, wgpu, ORT Web, Transformers.js, WebLLM).

  • Familiarity with compiler stacks (MLIR, TVM, IREE, XLA) for custom kernel generation and graph optimization.

  • Experience with on-device benchmarking infrastructure, performance-regression CI, and device-farm matrices.

  • Proficiency in C++/Objective-C/Swift for runtime integration.

Additional information

  • Relocation support is not available for this position

  • Work visa/immigration sponsorship is not available for this position

  • ​Salary range : 218 400.00 USD - 283 900. 00 USD

This range reflects the anticipated base salary for this position. Beyond base salary, this role may be eligible for equity awards and participation in our company incentive plans (such as annual discretionary bonuses or sales commissions). The final offer amount will depend on several factors, including geographic location and the candidate’s relevant experience, professional background, and skill set.

Benefits

At Unity, we want our team members to thrive. We offer a wide range of benefits designed to support well-being and work-life balance.

Please note: Benefits eligibility, specific offerings, and coverage vary based on the country and employment status.

While specific benefits vary, here are some of the ways we strive to take care of our eligible team members globally: Comprehensive health, life, and disability insurance | Commute subsidy | Employee stock ownership | Competitive retirement/pension plans | Generous vacation and personal days | Support for new parents through leave and family-care programs | Office food snacks | Mental Health and Wellbeing programs and support | Employee Resource Groups | Global Employee Assistance Program | Training and development programs | Volunteering and donation matching program

Life at Unity

Unity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D — closing the gap between ideas and reality. For more information, please visit .

Unity is a proud equal opportunity employer. We are committed to fostering an inclusive, innovative environment and celebrate our employees across age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable law. Our differences are strengths that enable us to support the growing and evolving needs of our customers, partners, and collaborators. If you have a disability that means there are preparations or accommodations we can make to help ensure you have a comfortable and positive interview experience, please fill out this form ( to let us know.

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

This position requires the incumbent to have a sufficient knowledge of English to have professional verbal and written exchanges in this language since the performance of the duties related to this position requires frequent and regular communication with colleagues and partners located worldwide and whose common language is English.

Headhunters and recruitment agencies may not submit resumes/CVs through this Web site or directly to managers. Unity does not accept unsolicited headhunter and agency resumes. Unity will not pay fees to any third-party agency or company that does not have a signed agreement with Unity.

Your privacy is important to us. Please take a moment to review our Prospect ( and Applicant ( Privacy Policies. Should you have any concerns about your privacy, please contact us at View email address on click.appcast.io.

Unity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D — closing the gap between ideas and reality.

At Unity, we're committed to creating a workplace that fosters collaboration and teamwork that allows you to harness your unique skills. Join us in creating industry tools to help creators of all levels bring their projects to reality.

Learn more about our culture and values ( , and get a head start by reading about how our hiring process works ( .

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Staff Machine Learning Engineer in San Francisco, CA vacancy
  • $250k - $350k

     ...getting started — this team will define what dependable, production-grade agentic AI looks like.About the RoleAs a Staff Machine Learning Research Engineer, you will operate across the full breadth of AIS’s technical needs — wherever the hardest ML problem in agentic AI... 
    Suggested
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  • $298k - $368k

     ...team builds the system which learns the spatial-temporal representation...  ...set of sensors, enabling engineers like you to (1) develop...  ...role, you'll report to a Senior Staff Technical Lead Manager. You...  ...Computer Science, Robotics, Machine Learning, or equivalent practical... 
    Suggested
    Full time
    Remote work
    Shift work

    Waymo

    San Francisco, CA
    22 days ago
  •  ...published research papers and won research competitions at top conferences like ICLR, ICML, AAAI. Role Description As a Machine Learning Engineer at Advex, you will play a pivotal role in shaping the company's technical direction. As an MLE you are well versed with... 
    Suggested
    Full time

    Openreq

    San Francisco, CA
    a month ago
  •  ...About us Our mission is to reinvent the way people learn, starting with language. Learning a language can change a life...  ...About this role We are looking for an experienced Machine Learning Engineer to join our team and help develop cutting-edge speech recognition... 
    Suggested
    Full time
    Live in
    Work at office
    Worldwide

    Speak

    San Francisco, CA
    more than 2 months ago
  • $225k - $300k

     ...Machine Learning Engineer About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access,...  ...direction along the way. We are primarily hiring for senior and staff-level engineers who are comfortable owning critical systems... 
    Suggested
    Full time
    Work at office
    Immediate start

    Latent

    San Francisco, CA
    more than 2 months ago
  • $244k - $292k

     ...relationships. Yes, you can build an exciting business AND have real-life real-customer impact. We are seeking a Senior Machine Learning Engineer to join our team. This role will focus on developing and maintaining machine learning infrastructure and operations,... 
    Full time
    Local area

    Kikoff

    San Francisco, CA
    more than 2 months ago
  •  ...and Y Combinator and angels from Google DeepMind, OpenAI, Anthropic, Meta Superintelligence Labs, and Microsoft AI. Machine Learning Engineer, Quality Intelligence Overview AfterQuery builds the data and evaluation systems that power frontier AI models.... 
    Full time

    AfterQuery

    San Francisco, CA
    more than 2 months ago
  • $180k - $270k

     ...highest standards of data security and privacy protection. To learn more about Plaud, please visit and follow along on Instagram...  ...building and deploying high-throughput, ultra-low-latency inference engines for large language models or foundational speech models.... 
    Full time
    Work at office
    Worldwide

    Plaud

    San Francisco, CA
    more than 2 months ago
  •  ...tools consistently fail. We are a small, fast-growing team of engineers in San Francisco powering Fortune 100 enterprises, YC startups,...  ...~5 days in-office at our San Francisco office ~ Eager to learn and adapt quickly ~ Prior startup or founding experience is a... 
    Full time
    Work at office
    Visa sponsorship
    Relocation package

    The Pulse

    San Francisco, CA
    a month ago
  •  ...Our mission is to recover human embodied intelligence as a learned model. To achieve this, we build custom hardware products,...  ...civilizational scale, join us. The Opportunity As a Machine Learning Engineer, you’ll work on multimodal perception, VLA training, robotics... 
    Full time
    Shift work

    Human Archive

    San Francisco, CA
    a month ago
  •  ...forward can earn meaningful scope. About the Role You will build machine-learning systems that remove real bottlenecks from drug discovery and development. The role spans research and engineering: identifying valuable problems, adapting modern biomolecular models,... 
    Full time
    H1b
    Visa sponsorship

    Capable Labs

    San Francisco, CA
    20 days ago
  •  ...enterprise that runs the real economy. Learn more about our vision in our manifesto....  ...production. Collaborate with product and engineering teams to integrate and deploy models...  ...curve. Must Have Strong experience in machine learning, deep learning, and NLP.... 
    Full time
    Worldwide
    Shift work

    HappyRobot

    San Francisco, CA
    a month ago
  •  ...future of healthcare, we’d love to meet you. Apply now to join our growing team. About the Role We're looking for a Machine Learning Engineer to design, build, and deploy production-grade ML systems that power the next generation of Plenful's AI platform. You'll own... 
    Full time
    Work at office
    Remote work
    Flexible hours
    2 days per week

    Plenful

    San Francisco, CA
    a month ago
  • $175k - $215k

     ...driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states. As a Perception Machine Learning Engineer, you will build the intelligent systems that "see" the world, directly shaping the future of autonomous travel. Within... 
    Full time
    Remote work

    Waymo

    San Francisco, CA
    a month ago
  • $160k - $230k

     ...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and... 
    Full time

    Together Ai

    San Francisco, CA
    a month ago
  •  ...industry expertise with advanced software, automation, and data-driven decision-making. Role Overview We are hiring Machine Learning Engineers (Autonomy) to build the autonomy and sensor-integration software that lets our mining vehicles perceive, decide, and... 
    Full time

    Mariana Minerals

    San Francisco, CA
    a month ago
  • $150k - $190k

     ...We are building an AI-driven simulation software stack for engineering and manufacturing across advanced industries. By enabling high...  ...and career goals. Who We're Looking For As a Machine Learning Engineer in Delivery, you are a problem solver who stays anchored... 
    Remote job
    Full time
    Flexible hours

    Physicsx

    San Francisco, CA
    more than 2 months ago
  • $244k - $320k

     ...email, and push notifications, our AI-powered personalization engine delivers bespoke experiences that drive performance, revenue,...  ...'s Corporate Equality Index ! About the Role Our Machine Learning Engineering team powers personalized experiences for hundreds... 
    Full time

    Attentive

    San Francisco, CA
    more than 2 months ago
  •  ...at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices. About the Role As a Machine Learning Engineer on the Marketplace team, you will build the models and decision systems that power Mercor’s hiring engine. This includes... 
    Full time
    Work at office
    Relocation package

    Mercor

    San Francisco, CA
    more than 2 months ago
  • $165k - $230k

     ...that isn't just moving fast, but rewriting what fast looks like. About the role We're looking for exceptional Machine Learning Engineers focused on Ads to help take Higgsfield's advertising platform to the next level. You'll work at the intersection of large... 
    Full time
    Work at office
    Remote work
    Worldwide
    3 days per week

    Higgsfield

    San Francisco, CA
    a month ago
  •  ...their best active lives. We believe in the power of movement to connect and drive people forward. We are looking for a Machine Learning Engineer to join the growing AI and Machine Learning team at Strava. This team is responsible for sophisticated machine learning... 
    Full time
    Work at office
    Worldwide
    Flexible hours
    3 days per week

    Strava

    San Francisco, CA
    a month ago
  • $137.1k - $201.6k

     ...We leverage artificial intelligence and advanced ML, deep learning techniques to power decision-making in real time — from optimizing...  ...marketplace.  About the Role We’re looking for a Machine Learning Engineer to help design, build, optimize and scale large-scale ML... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash Usa

    San Francisco, CA
    more than 2 months ago
  • $200k - $400k

     ...company using interpretability to understand, learn from, and design AI systems. Our mission...  ..., or shape what models learn. Every engineering discipline has been gated by fundamental...  ...About the role We’re looking for Machine Learning Engineers to help build our platform... 
    Full time
    Work at office
    Remote work

    Заявка На Вакансию «machine Learning Engineer» В Компании «g...

    San Francisco, CA
    more than 2 months ago
  • $140k - $265k

     ...About the Role: Glean is looking for engineers to help build the world’s best search and...  ...test Mentor more junior engineers, or learn from battle-tested ones About you:...  ...processing, or other large systems involving machine learning ~ Strong analytical skills and... 
    Full time
    Home office
    Flexible hours
    3 days per week

    Glean

    San Francisco, CA
    more than 2 months ago
  • $151.3k - $178k

     ...digital advertising? The Modeling team is responsible for Machine Learning (ML) systems at Quantcast. We build and maintain multiple ML...  ...interested in relevant content.   As a Machine Learning Engineer you care about the health and maintainability of our systems... 
    Full time
    Work at office
    Relocation

    Quantcast

    San Francisco, CA
    more than 2 months ago
  •  ...Our mission is to reinvent the way people learn, starting with language. Learning a...  ...About this role We’re hiring an ML Engineer, Assessments to help build best-in-class...  ...Design Lead (Content/Learning Design) , Machine Learning, Product, and Engineering to turn... 
    Full time
    Live in
    Immediate start

    Speak

    San Francisco, CA
    more than 2 months ago
  • $213k - $263k

     ...the ultimate data challenge: How do you mathematically prove that a virtual world is "real"? We are seeking visionary machine learning engineers and researchers to architect the scalable deep learning systems, novel data workflows, and eval tools that power our research... 
    Full time
    Remote work

    Waymo

    San Francisco, CA
    more than 2 months ago
  •  ...London, and Amsterdam. The Fraud Data team at Plaid builds the machine learning systems that power Plaid’s fraud detection products,...  ...from evolving fraud threats. As a Senior Machine Learning Engineer on Plaid's Fraud Data team, you will develop models that improve... 
    Full time
    Work experience placement
    Local area

    Plaid Inc.

    San Francisco, CA
    12 days ago
  •  ...About the company We’re a team of engineers, neuroscientists, and designers solving the most difficult...  ...of autonomy and demand intense, fast-paced learning. You will: Critically evaluate and implement the best machine learning approaches for our unique design... 
    Full time

    Orbit Neuro Co.

    San Francisco, CA
    more than 2 months ago
  •  ...compounds with every signal. About the Role At Poesis, machine learning and artificial intelligence open the door to improved alpha...  ...management. We're looking for an exceptional Machine Learning Engineer to help build the systems that make this possible. In this... 
    Full time
    Work at office
    Visa sponsorship
    Work visa
    Relocation package
    3 days per week

    Poesis Inc

    San Francisco, CA
    more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Machine Learning Engineer. Be the first to apply!