Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Machine Learning Engineer

Unity Technologies

The opportunityWe are building the next generation of AI-driven game experiences, running generative models on-device, right where the players are — on phones, tablets, laptops, and desktops. Our games run inside a modern, browser-native runtime (built on technologies such as WebGPU and WebNN), so the models that power these experiences must be deployed and accelerated entirely within that runtime. As a Senior Machine Learning Engineer for On-Device & Mobile AI, you will take state-of-the-art multi-modal models — transformers, diffusion networks, and vision-language models (VLMs) — and make them run fast, small, and reliably on mobile and constrained hardware.This is a deeply hands-on role. You will own the optimization and deployment of significant parts of the inference stack — from a trained checkpoint leaving research, through export, quantization, and kernel-level tuning, to a shipped feature running inside the engine at interactive frame rates within a fixed memory and power budget. Your work directly shapes the latency, quality, memory footprint, and battery profile of AI features experienced by billions of players.This role is for an engineer who is energized by the gap between a research model and a shipping, on-device product. If you enjoy profilers, frame captures, op-fusion, and shaving milliseconds and megabytes, this is your role.What you'll be doingInference & On-Device OptimizationOwn the optimization pipeline for the models you ship: model export, graph transformation, operator fusion, memory-layout planning, and hardware-specific tuning across NPU, mobile GPU, and desktop/laptop GPU.Apply quantization (INT4/INT8/FP16), weight sharing, structured/unstructured pruning, and knowledge distillation to hit hard latency, memory, and power budgets — and validate them against quality bars.Do low-level performance work: write and tune WebGPU compute shaders (WGSL) and, where relevant, native kernels (Metal, Vulkan/SPIR-V compute, CUDA); profile with browser and platform tools (Chrome/Dawn GPU traces, PIX, Instruments/Metal System Trace,Snapdragon Profiler, Nsight, RenderDoc), and eliminate bottlenecks at the op and memory-bandwidth level.Apply efficiency techniques — dynamic resolution, token reduction, cross-frame caching/reuse, reduced-step diffusion samplers — as engineering levers to meet budgets on target SKUs.Runtime & Systems IntegrationWork with WebGPU-targeted inference runtimes (ONNX Runtime Web, Transformers.js, WebLLM, TensorFlow.js) alongside native options (CoreML, ONNX Runtime, TFLite, ExecuTorch), and extend or build glue code where off-the-shelf options fall short of our diffusion and VLM workloads.Build parts of the integration between the ML runtime and the game engine: real-time scheduling, memory pooling, zero-copy buffer sharing between the inference and render paths, and frame-budget management alongside the renderer.Build supporting engineering for your components: model packaging and asset pipelines, on-device fallbacks and SKU-aware capability tiers, crash/quality telemetry, and automated on-device benchmarking in CI.Research ProductionizationPartner with research scientists to turn novel CV and multi-modal architectures into implementations that are deployable, debuggable, and fast on device.Provide a feedback loop into research: surface hardware constraints, op-support gaps, and cost models early so model design and deployment converge.Track breakthroughs in efficient inference (efficient attention, distillation, reduced-step diffusion) and assess them pragmatically: what actually moves latency/memory/power on our target devices.Collaboration & Engineering QualityContribute to engineering best practices, code-review standards, performance-regression gates, and on-device benchmarking methodology.Support a culture of measurement: track KPIs for latency, quality, memory, and power for the systems you work on, across the device matrix.Partner with platform engineers, product managers, and runtime teams to align your work with device-SKU constraints and product roadmaps.Share knowledge and mentor junior and mid-level engineers through code review, pairing, and design discussion.What we're looking for5+ years in software/ML engineering, with meaningful time focused on on-device / edge inference or real-time, performance-critical systems.Production deployment of transformer- and/or diffusion-based models (e.g., ViT, Stable Diffusion, CLIP/SigLIP-style encoders) on mobile, desktop, or embedded hardware — shipped, not just prototyped.Hands-on experience with at least one major inference runtime (ONNX Runtime / ORT Web, CoreML, TFLite, ExecuTorch) and a working understanding of operator fusion, memory layout, and runtime scheduling.Low-level performance engineering: solid command of at least one GPU/compute API — WebGPU/WGSL, Metal, Vulkan, D3D12, or CUDA — and the profiling tools to go with it. You can read a frame capture and a kernel trace and reason about where the time and memory go.Working knowledge of model-optimization techniques — quantization (INT4/INT8/FP16), weight sharing, pruning, and distillation — and the judgment to apply them to hit latency and memory budgets. You use them effectively as engineering tools.Understanding of target hardware: mobile SoCs (Apple Neural Engine, Qualcomm Hexagon/Adreno, ARM Mali) and/or desktop/laptop GPUs (Apple Silicon, NVIDIA, AMD, Intel).Strong Python for export pipelines and training-side tooling; familiarity with the core languages of a browser-native runtime (TypeScript/JavaScript, WGSL) is a plus.Working fluency with the models you deploy — enough to read an architecture, modify it for deployment, and reason about accuracy trade-offs.A collaborative working style: clear communication, reliable delivery, and a willingness to support and learn from teammates.You might also haveExperience shipping world-model, neural-rendering, or real-time generative pipelines NeRF, 3DGS, real-time diffusion, or similar) on device.Hands-on experience deploying models through WebGPU — e.g., ONNX Runtime Web WebGPU EP), Transformers.js, WebLLM, or TensorFlow.js — including writing/tuning WGSL compute shaders.Game-engine or real-time-graphics background (Unity, Unreal, or a custom engine; Metal/Vulkan/D3D/OpenGL ES render pipelines) — especially integrating compute workloads alongside a renderer.Contributions to open-source ML inference frameworks, runtimes, or GPU/compute libraries especially in the WebGPU ecosystem (Dawn, wgpu, ORT Web, Transformers.js, WebLLM).Familiarity with compiler stacks (MLIR, TVM, IREE, XLA) for custom kernel generation and graph optimization.Experience with on-device benchmarking infrastructure, performance-regression CI, and device-farm matrices.Proficiency in C++/Objective-C/Swift for runtime integration.Additional informationRelocation support is not available for this positionWork visa/immigration sponsorship is not available for this position​Salary range : 218 400.00 USD - 283 900. 00 USDThis range reflects the anticipated base salary for this position. Beyond base salary, this role may be eligible for equity awards and participation in our company incentive plans (such as annual discretionary bonuses or sales commissions). The final offer amount will depend on several factors, including geographic location and the candidate’s relevant experience, professional background, and skill set.BenefitsAt Unity, we want our team members to thrive. We offer a wide range of benefits designed to support well-being and work-life balance.Please note: Benefits eligibility, specific offerings, and coverage vary based on the country and employment status.While specific benefits vary, here are some of the ways we strive to take care of our eligible team members globally: Comprehensive health, life, and disability insurance | Commute subsidy | Employee stock ownership | Competitive retirement/pension plans | Generous vacation and personal days | Support for new parents through leave and family-care programs | Office food snacks | Mental Health and Wellbeing programs and support | Employee Resource Groups | Global Employee Assistance Program | Training and development programs | Volunteering and donation matching programLife at UnityUnity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D — closing the gap between ideas and reality. For more information, please visit .Unity is a proud equal opportunity employer. We are committed to fostering an inclusive, innovative environment and celebrate our employees across age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable law. Our differences are strengths that enable us to support the growing and evolving needs of our customers, partners, and collaborators. If you have a disability that means there are preparations or accommodations we can make to help ensure you have a comfortable and positive interview experience, please fill out this form to let us know.Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.This position requires the incumbent to have a sufficient knowledge of English to have professional verbal and written exchanges in this language since the performance of the duties related to this position requires frequent and regular communication with colleagues and partners located worldwide and whose common language is English.Headhunters and recruitment agencies may not submit resumes/CVs through this Web site or directly to managers. Unity does not accept unsolicited headhunter and agency resumes. Unity will not pay fees to any third-party agency or company that does not have a signed agreement with Unity.Your privacy is important to us. Please take a moment to review ourProspect andApplicant Privacy Policies. Should you have any concerns about your privacy, please contact us at View email address on click.appcast.io: Remote, California, USA; San Francisco, CA, USA; Mountain View, CA, USAType: Full time

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Staff Machine Learning Engineer in San Francisco, CA vacancy
  •  ...and the world's top educational institutions Work alongside engineers, scientists, operators, and more from Palantir, Meta, Scale...  ...the largest scale. About the Role Handshake is hiring a Staff Machine Learning Engineer for the Network and Handshake AI Marketplace... 
    Suggested
    Full time
    Work at office
    Remote work
    Flexible hours

    Handshake

    San Francisco, CA
    2 days ago
  • $209k - $313k

     ...express themselves, live in the moment, learn about the world, and have fun together.The...  ...Saturn, and other digital services.Snap Engineering teams build fun and technically...  ...privacy at the forefront.We’re looking for a Machine Learning Engineer to join Snap Inc!What... 
    Suggested
    Full time
    Live in
    Work at office
    Local area

    Snap

    San Francisco, CA
    11 hours ago
  • $163.42k - $285.98k

     ...in our recruiting process here.With more than 600 million users around the world and 300 billion ideas saved, Pinterest Machine Learning engineers build personalized experiences to help Pinners create a life they love. With just over 4,000 global employees, our teams... 
    Suggested
    Local area
    Relocation package

    Pinterest

    San Francisco, CA
    2 days ago
  • $140k - $200k

     ...GV, and Accel and enjoy multi-year runway.About the RoleWe’re looking for an ML Engineer to build the production systems that train, deploy, monitor, retrain, and serve our machine-learning models reliably. You sit between software engineering, data engineering, and modeling... 
    Suggested
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    4 days ago
  • $151.8k - $265.35k

     ...& Entertainment, marketing, and consumer retail, and is expanding rapidly into adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services that turn Firefly Foundry’s models into reliable, enterprise-grade products. You will... 
    Suggested
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    1 day ago
  • $138.91k - $285.98k

     ...group is dedicated to the development and research of applied machine learning. Our initiatives span a diverse range of AI/ML fields,...  ...core visual pod, a collaborative group of approximately six engineers and a product prototyping team, to create specialized evaluation... 
    Work at office
    Local area
    Remote work
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    11 hours ago
  • $117k - $152k

     ...streaming data, supporting analytics, product intelligence, machine learning pipelines, and business operations. As data volume and complexity...  ...production ML systems.We’re looking for a Machine Learning Engineer to join our Offline Infrastructure team. This is an ideal... 
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    11 hours ago
  • $148.7k - $199.4k

    Job Posting Title:Senior Machine Learning Engineer - ESPNReq ID:10150610Job Description:Disney Entertainment & ESPN TechnologyOn any given day at Disney Entertainment & ESPN Technology, we’re reimagining ways to create magical viewing experiences for the world’s most beloved... 
    Full time
    Worldwide

    Hulu

    San Francisco, CA
    1 day ago
  • $180k - $220k

    At Ouster, we build sensors and tools for engineers, roboticists, and researchers, so they can make the world safer and more efficient...  ...and need your help!We are looking for a highly technical Machine Learning Engineer to lead our efforts in Object Detection and... 
    Work experience placement
    Local area

    Ouster

    San Francisco, CA
    2 days ago
  • $173k - $259k

     ...express themselves, live in the moment, learn about the world, and have fun together.The...  ...Saturn, and other digital services.Snap Engineering teams build fun and technically...  ...privacy at the forefront.We’re looking for a Machine Learning Engineer to join Snap Inc!What... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    San Francisco, CA
    11 hours ago
  •  ...We leverage artificial intelligence and advanced ML, deep learning techniques to power decision-making in real time — from optimizing...  ...loop marketplace. About the RoleWe’re looking for a Machine Learning Engineer to help design, build, optimize and scale large-scale ML... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    2 days ago
  • $172.5k - $306.63k

     ...center of Adobe’s creative ecosystem. Our mission is to employ machine learning to enhance our comprehension of the creative content that...  ...intelligence to new heights.What You'll DoAs a Senior Machine Learning Engineer on the Content Intelligence team, you will lead the... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    2 days ago
  • $133.5k - $212k

     ...candidates from all backgrounds and encourage you to apply. Learn more about our story and mission on our Culture and About...  ...engagement together!Position Overview:We are looking for a Senior Machine Learning Engineer to build the core Machine Learning foundations that power... 
    Contract work
    Local area
    Immediate start
    Remote work
    Worldwide
    Home office

    Iterable

    San Francisco, CA
    2 days ago
  • $189.72k - $332.01k

     ...in our recruiting process here.With more than 500 million users around the world and 300 billion ideas saved, Pinterest Machine Learning engineers build personalized experiences to help Pinners create a life they love. With just over 4,000 global employees, our teams... 
    Local area
    Relocation package

    Pinterest

    San Francisco, CA
    2 days ago
  • $290k - $359.6k

     ...DescriptionDUTIES: Research, design, develop and test robotic and machine learning applications for an AV company. Duties may include:...  ...telecommute.REQUIREMENTS: Master's degree in Data Science, Aerospace Engineering, Automotive Engineering, Computer Engineering, Robotics or a... 
    Full time
    Temporary work
    Local area
    Remote work
    Work from home

    General Motors

    San Francisco, CA
    2 days ago
  • $112k - $269k

    SummaryYelp engineering culture is driven by our values: we’re a cooperative team that values...  ...requires the use of cutting-edge Machine Learning (ML) and Artificial Intelligence (AI) to...  ...spanning various geographical locations. As a Staff-level ML Engineer on the Content... 
    Work experience placement
    Local area
    Remote work

    Yelp

    San Francisco, CA
    11 hours ago
  • $189.72k - $332.01k

     ...ofthe Advanced Technologies Group (ATG), Pinterest’s advanced machine learning team. ATG’s goal is to keep Pinterest at the forefront of...  ...that technology to the product in collaboration with product engineering teams. The team also publishes its work in applied research... 
    Work experience placement
    Work at office
    Local area
    Remote work
    Relocation package

    Pinterest

    San Francisco, CA
    4 days ago
  • $244k - $320k

     ...email, and push notifications, our AI-powered personalization engine delivers bespoke experiences that drive performance, revenue,...  ...Foundation's Corporate Equality Index!About the RoleOur Machine Learning Engineering team powers personalized experiences for hundreds... 
    Full time

    Attentive

    San Francisco, CA
    2 days ago
  • $162.8k - $203.5k

     ...immersive transportation to improve people's lives warrants modern ML utilizing peta-byte scale data. Our highly motivated Machine Learning Engineers work on these challenging problems and define solutions to directly impact various aspects of our core business.If you... 
    Hourly pay
    Work at office
    Local area
    3 days per week

    Lyft

    San Francisco, CA
    2 days ago
  •  ...Senior Machine Learning Engineer Location: SF or Waterloo, with ability to travel Start Date: Flexible, ideally Q3 2025 About Hum.ai Hum.ai is building planetary superintelligence. Backed by top funds, we’ve raised $10M+ and are now heads down building... 
    Remote job
    Full time
    Work experience placement
    Flexible hours

    hum.ai

    San Francisco, CA
    15 hours ago
  • $244k - $292k

     ...relationships. Yes, you can build an exciting business AND have real-life real-customer impact. We are seeking a Senior Machine Learning Engineer to join our team. This role will focus on developing and maintaining machine learning infrastructure and operations, particularly... 
    Full time
    Local area

    Kikoff

    San Francisco, CA
    15 hours ago
  • $225k - $300k

     ...Machine Learning Engineer About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access,...  ...direction along the way. We are primarily hiring for senior and staff-level engineers who are comfortable owning critical systems... 
    Full time
    Work at office
    Immediate start

    Latent

    San Francisco, CA
    15 hours ago
  •  ...and Y Combinator and angels from Google DeepMind, OpenAI, Anthropic, Meta Superintelligence Labs, and Microsoft AI. Machine Learning Engineer, Quality Intelligence Overview AfterQuery builds the data and evaluation systems that power frontier AI models.... 
    Full time

    AfterQuery

    San Francisco, CA
    15 hours ago
  • $180k - $270k

     ...highest standards of data security and privacy protection. To learn more about Plaud, please visit and follow along on Instagram...  ...building and deploying high-throughput, ultra-low-latency inference engines for large language models or foundational speech models.... 
    Full time
    Work at office
    Worldwide

    Plaud

    San Francisco, CA
    15 hours ago
  •  ...published research papers and won research competitions at top conferences like ICLR, ICML, AAAI. Role Description As a Machine Learning Engineer at Advex, you will play a pivotal role in shaping the company's technical direction. As an MLE you are well versed with... 
    Full time

    Openreq

    San Francisco, CA
    15 hours ago
  •  ...on problems that sit at the frontier of AI, manufacturing, and healthcare, this is the place. The Role As a Senior Machine Learning Engineer, you will build the intelligence layer that automates complex healthcare compliance and document procurement workflows.... 
    Full time
    Work at office

    Hike Medical

    San Francisco, CA
    15 hours ago
  •  ...have to know a.) precisely where they are, and b.) everything about the fruit they are seeing. We are looking for a Machine Learning Engineer to build creative, practical, and robust solutions to ML/CV software and infrastructure problems, relating to training edge... 
    Full time
    Work at office
    Flexible hours
    Weekend work

    Orchard Robotics

    San Francisco, CA
    15 hours ago
  • $150k - $200k

     ...practice—real customer use-cases with clear success criteria. Required Qualifications ● PhD or MS degree in Computer Science, Machine Learning, Robotics, or equivalent technical discipline ● Deep expertise in machine learning fundamentals, reinforcement learning, and... 
    Full time

    Deft Ai, Inc.

    San Francisco, CA
    15 hours ago
  • $200k - $400k

     ...a team of the world’s top interpretability researchers and engineers from organizations like OpenAI and DeepMind. We’ve raised $5...  ...Clinic, and Rakuten. About the role We’re looking for Machine Learning Engineers to help build our platform for training,... 
    Full time

    Goodfire

    San Francisco, CA
    15 hours ago
  •  ...to explore new frontiers, create unforgettable experiences, and build a legacy that inspires future generations. Senior Machine Learning Engineer Primary: Bay Area (San Francisco / Peninsula) | Secondary: NYC   The Opportunity We're doing an AI-first... 
    Full time
    Work at office
    Relocation package
    Flexible hours
    2 days per week

    Mrbeast

    San Francisco, CA
    15 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Machine Learning Engineer. Be the first to apply!