Staff Machine Learning Engineer
Unity Technologies
Remote, California, USA
San Francisco, CA, USA
Mountain View, CA, USA
Full time
JOBREQ-2616041
The opportunity
We are building the next generation of AI-driven game experiences, running generative models on-device, right where the players are — on phones, tablets, laptops, and desktops. Our games run inside a modern, browser-native runtime (built on technologies such as WebGPU and WebNN), so the models that power these experiences must be deployed and accelerated entirely within that runtime. As a Senior Machine Learning Engineer for On-Device & Mobile AI, you will take state-of-the-art multi-modal models — transformers, diffusion networks, and vision-language models (VLMs) — and make them run fast, small, and reliably on mobile and constrained hardware.
This is a deeply hands-on role. You will own the optimization and deployment of significant parts of the inference stack — from a trained checkpoint leaving research, through export, quantization, and kernel-level tuning, to a shipped feature running inside the engine at interactive frame rates within a fixed memory and power budget. Your work directly shapes the latency, quality, memory footprint, and battery profile of AI features experienced by billions of players.
This role is for an engineer who is energized by the gap between a research model and a shipping, on-device product. If you enjoy profilers, frame captures, op-fusion, and shaving milliseconds and megabytes, this is your role.
What you'll be doing
Inference & On-Device Optimization
Own the optimization pipeline for the models you ship: model export, graph transformation, operator fusion, memory-layout planning, and hardware-specific tuning across NPU, mobile GPU, and desktop/laptop GPU.
Apply quantization (INT4/INT8/FP16), weight sharing, structured/unstructured pruning, and knowledge distillation to hit hard latency, memory, and power budgets — and validate them against quality bars.
Do low-level performance work: write and tune WebGPU compute shaders (WGSL) and, where relevant, native kernels (Metal, Vulkan/SPIR-V compute, CUDA); profile with browser and platform tools (Chrome/Dawn GPU traces, PIX, Instruments/Metal System Trace,
Snapdragon Profiler, Nsight, RenderDoc), and eliminate bottlenecks at the op and memory-bandwidth level.
Apply efficiency techniques — dynamic resolution, token reduction, cross-frame caching/reuse, reduced-step diffusion samplers — as engineering levers to meet budgets on target SKUs.
Runtime & Systems Integration
Work with WebGPU-targeted inference runtimes (ONNX Runtime Web, Transformers.js, WebLLM, TensorFlow.js) alongside native options (CoreML, ONNX Runtime, TFLite, ExecuTorch), and extend or build glue code where off-the-shelf options fall short of our diffusion and VLM workloads.
Build parts of the integration between the ML runtime and the game engine: real-time scheduling, memory pooling, zero-copy buffer sharing between the inference and render paths, and frame-budget management alongside the renderer.
Build supporting engineering for your components: model packaging and asset pipelines, on-device fallbacks and SKU-aware capability tiers, crash/quality telemetry, and automated on-device benchmarking in CI.
Research Productionization
Partner with research scientists to turn novel CV and multi-modal architectures into implementations that are deployable, debuggable, and fast on device.
Provide a feedback loop into research: surface hardware constraints, op-support gaps, and cost models early so model design and deployment converge.
Track breakthroughs in efficient inference (efficient attention, distillation, reduced-step diffusion) and assess them pragmatically: what actually moves latency/memory/power on our target devices.
Collaboration & Engineering Quality
Contribute to engineering best practices, code-review standards, performance-regression gates, and on-device benchmarking methodology.
Support a culture of measurement: track KPIs for latency, quality, memory, and power for the systems you work on, across the device matrix.
Partner with platform engineers, product managers, and runtime teams to align your work with device-SKU constraints and product roadmaps.
Share knowledge and mentor junior and mid-level engineers through code review, pairing, and design discussion.
What we're looking for
5+ years in software/ML engineering, with meaningful time focused on on-device / edge inference or real-time, performance-critical systems.
Production deployment of transformer- and/or diffusion-based models (e.g., ViT, Stable Diffusion, CLIP/SigLIP-style encoders) on mobile, desktop, or embedded hardware — shipped, not just prototyped.
Hands-on experience with at least one major inference runtime (ONNX Runtime / ORT Web, CoreML, TFLite, ExecuTorch) and a working understanding of operator fusion, memory layout, and runtime scheduling.
Low-level performance engineering: solid command of at least one GPU/compute API — WebGPU/WGSL, Metal, Vulkan, D3D12, or CUDA — and the profiling tools to go with it. You can read a frame capture and a kernel trace and reason about where the time and memory go.
Working knowledge of model-optimization techniques — quantization (INT4/INT8/FP16), weight sharing, pruning, and distillation — and the judgment to apply them to hit latency and memory budgets. You use them effectively as engineering tools.
Understanding of target hardware: mobile SoCs (Apple Neural Engine, Qualcomm Hexagon/Adreno, ARM Mali) and/or desktop/laptop GPUs (Apple Silicon, NVIDIA, AMD, Intel).
Strong Python for export pipelines and training-side tooling; familiarity with the core languages of a browser-native runtime (TypeScript/JavaScript, WGSL) is a plus.
Working fluency with the models you deploy — enough to read an architecture, modify it for deployment, and reason about accuracy trade-offs.
A collaborative working style: clear communication, reliable delivery, and a willingness to support and learn from teammates.
You might also have
Experience shipping world-model, neural-rendering, or real-time generative pipelines NeRF, 3DGS, real-time diffusion, or similar) on device.
Hands-on experience deploying models through WebGPU — e.g., ONNX Runtime Web WebGPU EP), Transformers.js, WebLLM, or TensorFlow.js — including writing/tuning WGSL compute shaders.
Game-engine or real-time-graphics background (Unity, Unreal, or a custom engine; Metal/Vulkan/D3D/OpenGL ES render pipelines) — especially integrating compute workloads alongside a renderer.
Contributions to open-source ML inference frameworks, runtimes, or GPU/compute libraries especially in the WebGPU ecosystem (Dawn, wgpu, ORT Web, Transformers.js, WebLLM).
Familiarity with compiler stacks (MLIR, TVM, IREE, XLA) for custom kernel generation and graph optimization.
Experience with on-device benchmarking infrastructure, performance-regression CI, and device-farm matrices.
Proficiency in C++/Objective-C/Swift for runtime integration.
Additional information
Relocation support is not available for this position
Work visa/immigration sponsorship is not available for this position
Salary range : 218 400.00 USD - 283 900. 00 USD
This range reflects the anticipated base salary for this position. Beyond base salary, this role may be eligible for equity awards and participation in our company incentive plans (such as annual discretionary bonuses or sales commissions). The final offer amount will depend on several factors, including geographic location and the candidate’s relevant experience, professional background, and skill set.
Benefits
At Unity, we want our team members to thrive. We offer a wide range of benefits designed to support well-being and work-life balance.
Please note: Benefits eligibility, specific offerings, and coverage vary based on the country and employment status.
While specific benefits vary, here are some of the ways we strive to take care of our eligible team members globally: Comprehensive health, life, and disability insurance | Commute subsidy | Employee stock ownership | Competitive retirement/pension plans | Generous vacation and personal days | Support for new parents through leave and family-care programs | Office food snacks | Mental Health and Wellbeing programs and support | Employee Resource Groups | Global Employee Assistance Program | Training and development programs | Volunteering and donation matching program
Life at Unity
Unity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D — closing the gap between ideas and reality. For more information, please visit .
Unity is a proud equal opportunity employer. We are committed to fostering an inclusive, innovative environment and celebrate our employees across age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable law. Our differences are strengths that enable us to support the growing and evolving needs of our customers, partners, and collaborators. If you have a disability that means there are preparations or accommodations we can make to help ensure you have a comfortable and positive interview experience, please fill out this form ( to let us know.
Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.
This position requires the incumbent to have a sufficient knowledge of English to have professional verbal and written exchanges in this language since the performance of the duties related to this position requires frequent and regular communication with colleagues and partners located worldwide and whose common language is English.
Headhunters and recruitment agencies may not submit resumes/CVs through this Web site or directly to managers. Unity does not accept unsolicited headhunter and agency resumes. Unity will not pay fees to any third-party agency or company that does not have a signed agreement with Unity.
Your privacy is important to us. Please take a moment to review our Prospect ( and Applicant ( Privacy Policies. Should you have any concerns about your privacy, please contact us at View email address on click.appcast.io.
Unity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D — closing the gap between ideas and reality.
At Unity, we're committed to creating a workplace that fosters collaboration and teamwork that allows you to harness your unique skills. Join us in creating industry tools to help creators of all levels bring their projects to reality.
Learn more about our culture and values ( , and get a head start by reading about how our hiring process works ( .
$250k - $350k
...getting started — this team will define what dependable, production-grade agentic AI looks like.About the RoleAs a Staff Machine Learning Research Engineer, you will operate across the full breadth of AIS’s technical needs — wherever the hardest ML problem in agentic AI...SuggestedFull time$298k - $368k
...team builds the system which learns the spatial-temporal representation... ...set of sensors, enabling engineers like you to (1) develop... ...role, you'll report to a Senior Staff Technical Lead Manager. You... ...Computer Science, Robotics, Machine Learning, or equivalent practical...SuggestedFull timeRemote workShift work- ...published research papers and won research competitions at top conferences like ICLR, ICML, AAAI. Role Description As a Machine Learning Engineer at Advex, you will play a pivotal role in shaping the company's technical direction. As an MLE you are well versed with...SuggestedFull time
- ...About us Our mission is to reinvent the way people learn, starting with language. Learning a language can change a life... ...About this role We are looking for an experienced Machine Learning Engineer to join our team and help develop cutting-edge speech recognition...SuggestedFull timeLive inWork at officeWorldwide
$225k - $300k
...Machine Learning Engineer About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access,... ...direction along the way. We are primarily hiring for senior and staff-level engineers who are comfortable owning critical systems...SuggestedFull timeWork at officeImmediate start$244k - $292k
...relationships. Yes, you can build an exciting business AND have real-life real-customer impact. We are seeking a Senior Machine Learning Engineer to join our team. This role will focus on developing and maintaining machine learning infrastructure and operations,...Full timeLocal area- ...and Y Combinator and angels from Google DeepMind, OpenAI, Anthropic, Meta Superintelligence Labs, and Microsoft AI. Machine Learning Engineer, Quality Intelligence Overview AfterQuery builds the data and evaluation systems that power frontier AI models....Full time
$180k - $270k
...highest standards of data security and privacy protection. To learn more about Plaud, please visit and follow along on Instagram... ...building and deploying high-throughput, ultra-low-latency inference engines for large language models or foundational speech models....Full timeWork at officeWorldwide- ...tools consistently fail. We are a small, fast-growing team of engineers in San Francisco powering Fortune 100 enterprises, YC startups,... ...~5 days in-office at our San Francisco office ~ Eager to learn and adapt quickly ~ Prior startup or founding experience is a...Full timeWork at officeVisa sponsorshipRelocation package
- ...Our mission is to recover human embodied intelligence as a learned model. To achieve this, we build custom hardware products,... ...civilizational scale, join us. The Opportunity As a Machine Learning Engineer, you’ll work on multimodal perception, VLA training, robotics...Full timeShift work
- ...forward can earn meaningful scope. About the Role You will build machine-learning systems that remove real bottlenecks from drug discovery and development. The role spans research and engineering: identifying valuable problems, adapting modern biomolecular models,...Full timeH1bVisa sponsorship
- ...enterprise that runs the real economy. Learn more about our vision in our manifesto.... ...production. Collaborate with product and engineering teams to integrate and deploy models... ...curve. Must Have Strong experience in machine learning, deep learning, and NLP....Full timeWorldwideShift work
- ...future of healthcare, we’d love to meet you. Apply now to join our growing team. About the Role We're looking for a Machine Learning Engineer to design, build, and deploy production-grade ML systems that power the next generation of Plenful's AI platform. You'll own...Full timeWork at officeRemote workFlexible hours2 days per week
$175k - $215k
...driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states. As a Perception Machine Learning Engineer, you will build the intelligent systems that "see" the world, directly shaping the future of autonomous travel. Within...Full timeRemote work$160k - $230k
...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and...Full time- ...industry expertise with advanced software, automation, and data-driven decision-making. Role Overview We are hiring Machine Learning Engineers (Autonomy) to build the autonomy and sensor-integration software that lets our mining vehicles perceive, decide, and...Full time
$150k - $190k
...We are building an AI-driven simulation software stack for engineering and manufacturing across advanced industries. By enabling high... ...and career goals. Who We're Looking For As a Machine Learning Engineer in Delivery, you are a problem solver who stays anchored...Remote jobFull timeFlexible hours$244k - $320k
...email, and push notifications, our AI-powered personalization engine delivers bespoke experiences that drive performance, revenue,... ...'s Corporate Equality Index ! About the Role Our Machine Learning Engineering team powers personalized experiences for hundreds...Full time- ...at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices. About the Role As a Machine Learning Engineer on the Marketplace team, you will build the models and decision systems that power Mercor’s hiring engine. This includes...Full timeWork at officeRelocation package
$165k - $230k
...that isn't just moving fast, but rewriting what fast looks like. About the role We're looking for exceptional Machine Learning Engineers focused on Ads to help take Higgsfield's advertising platform to the next level. You'll work at the intersection of large...Full timeWork at officeRemote workWorldwide3 days per week- ...their best active lives. We believe in the power of movement to connect and drive people forward. We are looking for a Machine Learning Engineer to join the growing AI and Machine Learning team at Strava. This team is responsible for sophisticated machine learning...Full timeWork at officeWorldwideFlexible hours3 days per week
$137.1k - $201.6k
...We leverage artificial intelligence and advanced ML, deep learning techniques to power decision-making in real time — from optimizing... ...marketplace. About the Role We’re looking for a Machine Learning Engineer to help design, build, optimize and scale large-scale ML...Hourly payWork at officeLocal areaRemote workFlexible hours$200k - $400k
...company using interpretability to understand, learn from, and design AI systems. Our mission... ..., or shape what models learn. Every engineering discipline has been gated by fundamental... ...About the role We’re looking for Machine Learning Engineers to help build our platform...Full timeWork at officeRemote work$140k - $265k
...About the Role: Glean is looking for engineers to help build the world’s best search and... ...test Mentor more junior engineers, or learn from battle-tested ones About you:... ...processing, or other large systems involving machine learning ~ Strong analytical skills and...Full timeHome officeFlexible hours3 days per week$151.3k - $178k
...digital advertising? The Modeling team is responsible for Machine Learning (ML) systems at Quantcast. We build and maintain multiple ML... ...interested in relevant content. As a Machine Learning Engineer you care about the health and maintainability of our systems...Full timeWork at officeRelocation- ...Our mission is to reinvent the way people learn, starting with language. Learning a... ...About this role We’re hiring an ML Engineer, Assessments to help build best-in-class... ...Design Lead (Content/Learning Design) , Machine Learning, Product, and Engineering to turn...Full timeLive inImmediate start
$213k - $263k
...the ultimate data challenge: How do you mathematically prove that a virtual world is "real"? We are seeking visionary machine learning engineers and researchers to architect the scalable deep learning systems, novel data workflows, and eval tools that power our research...Full timeRemote work- ...London, and Amsterdam. The Fraud Data team at Plaid builds the machine learning systems that power Plaid’s fraud detection products,... ...from evolving fraud threats. As a Senior Machine Learning Engineer on Plaid's Fraud Data team, you will develop models that improve...Full timeWork experience placementLocal area
- ...About the company We’re a team of engineers, neuroscientists, and designers solving the most difficult... ...of autonomy and demand intense, fast-paced learning. You will: Critically evaluate and implement the best machine learning approaches for our unique design...Full time
- ...compounds with every signal. About the Role At Poesis, machine learning and artificial intelligence open the door to improved alpha... ...management. We're looking for an exceptional Machine Learning Engineer to help build the systems that make this possible. In this...Full timeWork at officeVisa sponsorshipWork visaRelocation package3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Machine Learning Engineer. Be the first to apply!
- staff data engineer San Francisco, CA
- senior staff engineer San Francisco, CA
- senior staff systems engineer San Francisco, CA
- engineering aide San Francisco, CA
- software engineer staff San Francisco, CA
- staff design engineer San Francisco, CA
- assistant engineer San Francisco, CA
- staff security engineer San Francisco, CA
- technology administrator San Francisco, CA
- staff engineer San Francisco, CA

