Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Machine Learning Engineer

Unity Technologies

The opportunityWe are building the next generation of AI-driven game experiences, running generative models on-device, right where the players are — on phones, tablets, laptops, and desktops. Our games run inside a modern, browser-native runtime (built on technologies such as WebGPU and WebNN), so the models that power these experiences must be deployed and accelerated entirely within that runtime. As a Senior Machine Learning Engineer for On-Device & Mobile AI, you will take state-of-the-art multi-modal models — transformers, diffusion networks, and vision-language models (VLMs) — and make them run fast, small, and reliably on mobile and constrained hardware.This is a deeply hands-on role. You will own the optimization and deployment of significant parts of the inference stack — from a trained checkpoint leaving research, through export, quantization, and kernel-level tuning, to a shipped feature running inside the engine at interactive frame rates within a fixed memory and power budget. Your work directly shapes the latency, quality, memory footprint, and battery profile of AI features experienced by billions of players.This role is for an engineer who is energized by the gap between a research model and a shipping, on-device product. If you enjoy profilers, frame captures, op-fusion, and shaving milliseconds and megabytes, this is your role.What you'll be doingInference & On-Device OptimizationOwn the optimization pipeline for the models you ship: model export, graph transformation, operator fusion, memory-layout planning, and hardware-specific tuning across NPU, mobile GPU, and desktop/laptop GPU.Apply quantization (INT4/INT8/FP16), weight sharing, structured/unstructured pruning, and knowledge distillation to hit hard latency, memory, and power budgets — and validate them against quality bars.Do low-level performance work: write and tune WebGPU compute shaders (WGSL) and, where relevant, native kernels (Metal, Vulkan/SPIR-V compute, CUDA); profile with browser and platform tools (Chrome/Dawn GPU traces, PIX, Instruments/Metal System Trace,Snapdragon Profiler, Nsight, RenderDoc), and eliminate bottlenecks at the op and memory-bandwidth level.Apply efficiency techniques — dynamic resolution, token reduction, cross-frame caching/reuse, reduced-step diffusion samplers — as engineering levers to meet budgets on target SKUs.Runtime & Systems IntegrationWork with WebGPU-targeted inference runtimes (ONNX Runtime Web, Transformers.js, WebLLM, TensorFlow.js) alongside native options (CoreML, ONNX Runtime, TFLite, ExecuTorch), and extend or build glue code where off-the-shelf options fall short of our diffusion and VLM workloads.Build parts of the integration between the ML runtime and the game engine: real-time scheduling, memory pooling, zero-copy buffer sharing between the inference and render paths, and frame-budget management alongside the renderer.Build supporting engineering for your components: model packaging and asset pipelines, on-device fallbacks and SKU-aware capability tiers, crash/quality telemetry, and automated on-device benchmarking in CI.Research ProductionizationPartner with research scientists to turn novel CV and multi-modal architectures into implementations that are deployable, debuggable, and fast on device.Provide a feedback loop into research: surface hardware constraints, op-support gaps, and cost models early so model design and deployment converge.Track breakthroughs in efficient inference (efficient attention, distillation, reduced-step diffusion) and assess them pragmatically: what actually moves latency/memory/power on our target devices.Collaboration & Engineering QualityContribute to engineering best practices, code-review standards, performance-regression gates, and on-device benchmarking methodology.Support a culture of measurement: track KPIs for latency, quality, memory, and power for the systems you work on, across the device matrix.Partner with platform engineers, product managers, and runtime teams to align your work with device-SKU constraints and product roadmaps.Share knowledge and mentor junior and mid-level engineers through code review, pairing, and design discussion.What we're looking for5+ years in software/ML engineering, with meaningful time focused on on-device / edge inference or real-time, performance-critical systems.Production deployment of transformer- and/or diffusion-based models (e.g., ViT, Stable Diffusion, CLIP/SigLIP-style encoders) on mobile, desktop, or embedded hardware — shipped, not just prototyped.Hands-on experience with at least one major inference runtime (ONNX Runtime / ORT Web, CoreML, TFLite, ExecuTorch) and a working understanding of operator fusion, memory layout, and runtime scheduling.Low-level performance engineering: solid command of at least one GPU/compute API — WebGPU/WGSL, Metal, Vulkan, D3D12, or CUDA — and the profiling tools to go with it. You can read a frame capture and a kernel trace and reason about where the time and memory go.Working knowledge of model-optimization techniques — quantization (INT4/INT8/FP16), weight sharing, pruning, and distillation — and the judgment to apply them to hit latency and memory budgets. You use them effectively as engineering tools.Understanding of target hardware: mobile SoCs (Apple Neural Engine, Qualcomm Hexagon/Adreno, ARM Mali) and/or desktop/laptop GPUs (Apple Silicon, NVIDIA, AMD, Intel).Strong Python for export pipelines and training-side tooling; familiarity with the core languages of a browser-native runtime (TypeScript/JavaScript, WGSL) is a plus.Working fluency with the models you deploy — enough to read an architecture, modify it for deployment, and reason about accuracy trade-offs.A collaborative working style: clear communication, reliable delivery, and a willingness to support and learn from teammates.You might also haveExperience shipping world-model, neural-rendering, or real-time generative pipelines NeRF, 3DGS, real-time diffusion, or similar) on device.Hands-on experience deploying models through WebGPU — e.g., ONNX Runtime Web WebGPU EP), Transformers.js, WebLLM, or TensorFlow.js — including writing/tuning WGSL compute shaders.Game-engine or real-time-graphics background (Unity, Unreal, or a custom engine; Metal/Vulkan/D3D/OpenGL ES render pipelines) — especially integrating compute workloads alongside a renderer.Contributions to open-source ML inference frameworks, runtimes, or GPU/compute libraries especially in the WebGPU ecosystem (Dawn, wgpu, ORT Web, Transformers.js, WebLLM).Familiarity with compiler stacks (MLIR, TVM, IREE, XLA) for custom kernel generation and graph optimization.Experience with on-device benchmarking infrastructure, performance-regression CI, and device-farm matrices.Proficiency in C++/Objective-C/Swift for runtime integration.Additional informationRelocation support is not available for this positionWork visa/immigration sponsorship is not available for this position​Salary range : 218 400.00 USD - 283 900. 00 USDThis range reflects the anticipated base salary for this position. Beyond base salary, this role may be eligible for equity awards and participation in our company incentive plans (such as annual discretionary bonuses or sales commissions). The final offer amount will depend on several factors, including geographic location and the candidate’s relevant experience, professional background, and skill set.BenefitsAt Unity, we want our team members to thrive. We offer a wide range of benefits designed to support well-being and work-life balance.Please note: Benefits eligibility, specific offerings, and coverage vary based on the country and employment status.While specific benefits vary, here are some of the ways we strive to take care of our eligible team members globally: Comprehensive health, life, and disability insurance | Commute subsidy | Employee stock ownership | Competitive retirement/pension plans | Generous vacation and personal days | Support for new parents through leave and family-care programs | Office food snacks | Mental Health and Wellbeing programs and support | Employee Resource Groups | Global Employee Assistance Program | Training and development programs | Volunteering and donation matching programLife at UnityUnity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D — closing the gap between ideas and reality. For more information, please visit .Unity is a proud equal opportunity employer. We are committed to fostering an inclusive, innovative environment and celebrate our employees across age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable law. Our differences are strengths that enable us to support the growing and evolving needs of our customers, partners, and collaborators. If you have a disability that means there are preparations or accommodations we can make to help ensure you have a comfortable and positive interview experience, please fill out this form to let us know.Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.This position requires the incumbent to have a sufficient knowledge of English to have professional verbal and written exchanges in this language since the performance of the duties related to this position requires frequent and regular communication with colleagues and partners located worldwide and whose common language is English.Headhunters and recruitment agencies may not submit resumes/CVs through this Web site or directly to managers. Unity does not accept unsolicited headhunter and agency resumes. Unity will not pay fees to any third-party agency or company that does not have a signed agreement with Unity.Your privacy is important to us. Please take a moment to review ourProspect andApplicant Privacy Policies. Should you have any concerns about your privacy, please contact us at View email address on click.appcast.io: Remote, California, USA; San Francisco, CA, USA; Mountain View, CA, USAType: Full time

Vacancy posted 20 hours ago
Similar jobs that could be interesting for youBased on the Staff Machine Learning Engineer in San Francisco, CA vacancy
  •  ...worldwide to eliminate busywork and focus on what matters. Learn more at superhuman.com and about our values here. The Opportunity...  ...complex tasks, leveraging Superhuman ubiquitous UI. As a Machine Learning Engineer on this team, you will be at the heart of our company's... 
    Suggested
    Full time
    Worldwide
    Home office
    Flexible hours

    Superhuman

    San Francisco, CA
    2 days ago
  •  ...and discovery experiences. As a Senior Staff Engineer, you will lead the technical direction...  ...techniques such as sequence modeling, deep learning, and large language models (LLMs). Your...  ...the Role Apply state-of-the-art machine learning and LLM techniques to problems... 
    Suggested
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    DoorDash USA

    San Francisco, CA
    1 day ago
  • $172.5k - $306.63k

     ...center of Adobe’s creative ecosystem. Our mission is to employ machine learning to enhance our comprehension of the creative content that...  ...intelligence to new heights.What You'll DoAs a Senior Machine Learning Engineer on the Content Intelligence team, you will lead the... 
    Suggested
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    5 days ago
  • $209k - $313k

     ...express themselves, live in the moment, learn about the world, and have fun together.The...  ...Saturn, and other digital services.Snap Engineering teams build fun and technically...  ...privacy at the forefront.We’re looking for a Machine Learning Engineer to join Snap Inc!What... 
    Suggested
    Full time
    Live in
    Work at office
    Local area

    Snap

    San Francisco, CA
    3 days ago
  • $173k - $259k

     ...express themselves, live in the moment, learn about the world, and have fun together.The...  ...Saturn, and other digital services.Snap Engineering teams build fun and technically...  ...privacy at the forefront.We’re looking for a Machine Learning Engineer to join Snap Inc!What... 
    Suggested
    Full time
    Live in
    Work at office
    Local area

    Snap

    San Francisco, CA
    3 days ago
  • $151.8k - $265.35k

     ...& Entertainment, marketing, and consumer retail, and is expanding rapidly into adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services that turn Firefly Foundry’s models into reliable, enterprise-grade products. You will... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    4 days ago
  • $163.42k - $285.98k

     ...in our recruiting process here.With more than 600 million users around the world and 300 billion ideas saved, Pinterest Machine Learning engineers build personalized experiences to help Pinners create a life they love. With just over 4,000 global employees, our teams... 
    Local area
    Relocation package

    Pinterest

    San Francisco, CA
    5 days ago
  • $148.7k - $199.4k

    Job Posting Title:Senior Machine Learning Engineer - ESPNReq ID:10150610Job Description:Disney Entertainment & ESPN TechnologyOn any given day at Disney Entertainment & ESPN Technology, we’re reimagining ways to create magical viewing experiences for the world’s most beloved... 
    Full time
    Worldwide

    Hulu

    San Francisco, CA
    4 days ago
  • $180k - $220k

    At Ouster, we build sensors and tools for engineers, roboticists, and researchers, so they can make the world safer and more efficient...  ...and need your help!We are looking for a highly technical Machine Learning Engineer to lead our efforts in Object Detection and... 
    Work experience placement
    Local area

    Ouster

    San Francisco, CA
    5 days ago
  • $138.91k - $285.98k

     ...group is dedicated to the development and research of applied machine learning. Our initiatives span a diverse range of AI/ML fields,...  ...core visual pod, a collaborative group of approximately six engineers and a product prototyping team, to create specialized evaluation... 
    Work at office
    Local area
    Remote work
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    3 days ago
  • $189.72k - $332.01k

     ...ofthe Advanced Technologies Group (ATG), Pinterest’s advanced machine learning team. ATG’s goal is to keep Pinterest at the forefront of...  ...that technology to the product in collaboration with product engineering teams. The team also publishes its work in applied research... 
    Work experience placement
    Work at office
    Local area
    Remote work
    Relocation package

    Pinterest

    San Francisco, CA
    2 days ago
  •  ...and phone orders—using DoorDash's logistics network. The Drive Machine Learning team builds the prediction and intelligence systems that...  ...consumer, and dasher outcomes.About the RoleAs a Machine Learning Engineer on the Drive team, you'll own machine learning systems end-to... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Relocation
    Flexible hours

    Doordash

    San Francisco, CA
    20 hours ago
  • $112k - $269k

    SummaryYelp engineering culture is driven by our values: we’re a cooperative team that values...  ...requires the use of cutting-edge Machine Learning (ML) and Artificial Intelligence (AI) to...  ...spanning various geographical locations. As a Staff-level ML Engineer on the Content... 
    Work experience placement
    Local area
    Remote work

    Yelp

    San Francisco, CA
    3 days ago
  • $140k - $200k

     ...GV, and Accel and enjoy multi-year runway.About the RoleWe’re looking for an ML Engineer to build the production systems that train, deploy, monitor, retrain, and serve our machine-learning models reliably. You sit between software engineering, data engineering, and modeling... 
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    2 days ago
  • $244k - $320k

     ...email, and push notifications, our AI-powered personalization engine delivers bespoke experiences that drive performance, revenue,...  ...Foundation's Corporate Equality Index!About the RoleOur Machine Learning Engineering team powers personalized experiences for hundreds... 
    Full time

    Attentive

    San Francisco, CA
    5 days ago
  • $117k - $152k

     ...streaming data, supporting analytics, product intelligence, machine learning pipelines, and business operations. As data volume and complexity...  ...production ML systems.We’re looking for a Machine Learning Engineer to join our Offline Infrastructure team. This is an ideal... 
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    3 days ago
  •  ...We leverage artificial intelligence and advanced ML, deep learning techniques to power decision-making in real time — from optimizing...  ...loop marketplace. About the RoleWe’re looking for a Machine Learning Engineer to help design, build, optimize and scale large-scale ML... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    5 days ago
  • $290k - $359.6k

     ...DescriptionDUTIES: Research, design, develop and test robotic and machine learning applications for an AV company. Duties may include:...  ...telecommute.REQUIREMENTS: Master's degree in Data Science, Aerospace Engineering, Automotive Engineering, Computer Engineering, Robotics or a... 
    Full time
    Temporary work
    Local area
    Remote work
    Work from home

    General Motors

    San Francisco, CA
    5 days ago
  • $162.8k - $203.5k

     ...immersive transportation to improve people's lives warrants modern ML utilizing peta-byte scale data. Our highly motivated Machine Learning Engineers work on these challenging problems and define solutions to directly impact various aspects of our core business.If you... 
    Hourly pay
    Work at office
    Local area
    3 days per week

    Lyft

    San Francisco, CA
    5 days ago
  • $189.72k - $332.01k

     ...in our recruiting process here.With more than 500 million users around the world and 300 billion ideas saved, Pinterest Machine Learning engineers build personalized experiences to help Pinners create a life they love. With just over 4,000 global employees, our teams... 
    Local area
    Relocation package

    Pinterest

    San Francisco, CA
    5 days ago
  •  ...to explore new frontiers, create unforgettable experiences, and build a legacy that inspires future generations. Senior Machine Learning Engineer Primary: Bay Area (San Francisco / Peninsula) | Secondary: NYC   The Opportunity We're doing an AI-first... 
    Full time
    Work at office
    Relocation package
    Flexible hours
    2 days per week

    Mrbeast

    San Francisco, CA
    1 day ago
  •  ...have to know a.) precisely where they are, and b.) everything about the fruit they are seeing. We are looking for a Machine Learning Engineer to build creative, practical, and robust solutions to ML/CV software and infrastructure problems, relating to training edge... 
    Full time
    Work at office
    Flexible hours
    Weekend work

    Orchard Robotics

    San Francisco, CA
    1 day ago
  • $150k - $200k

     ...practice—real customer use-cases with clear success criteria. Required Qualifications ● PhD or MS degree in Computer Science, Machine Learning, Robotics, or equivalent technical discipline ● Deep expertise in machine learning fundamentals, reinforcement learning, and... 
    Full time

    Deft Ai, Inc.

    San Francisco, CA
    1 day ago
  •  ...on problems that sit at the frontier of AI, manufacturing, and healthcare, this is the place. The Role As a Senior Machine Learning Engineer, you will build the intelligence layer that automates complex healthcare compliance and document procurement workflows.... 
    Full time
    Work at office

    Hike Medical

    San Francisco, CA
    1 day ago
  • $225k - $300k

     ...Machine Learning Engineer About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access,...  ...direction along the way. We are primarily hiring for senior and staff-level engineers who are comfortable owning critical systems... 
    Full time
    Work at office
    Immediate start

    Latent

    San Francisco, CA
    1 day ago
  • $200k - $400k

     ...a team of the world’s top interpretability researchers and engineers from organizations like OpenAI and DeepMind. We’ve raised $5...  ...Clinic, and Rakuten. About the role We’re looking for Machine Learning Engineers to help build our platform for training,... 
    Full time

    Goodfire

    San Francisco, CA
    1 day ago
  • $250k - $295k

     ...current delivery mechanism. The real product is a scalable risk engine, our Stand World Model . We stay when traditional insurers...  ...operate with far less friction. The Opportunity: As a Machine Learning Engineer on the Applied Science team, you will design, train... 
    Full time
    Temporary work
    H1b
    Work at office
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    Stand Insurance

    San Francisco, CA
    1 day ago
  • $160k - $250k

     ...San Francisco, Seattle, and Delhi offices. Please reach out if you are interested in joining the future of AI! Senior Machine Learning Engineer In order to execute our vision, we need to grow our team of best-in-class machine learning engineers. We are looking... 
    Full time

    Hive

    San Francisco, CA
    1 day ago
  • $200k - $260k

     ...class latency and reliability. We're looking for a Senior ML Engineer to drive the model serving layer for voice workloads. You'll...  ...signal processing) is a strong plus but not required — you can learn this quickly if you have strong ML engineering fundamentals.... 
    Full time

    Together Ai

    San Francisco, CA
    1 day ago
  •  ...Senior Machine Learning Engineer Location: SF or Waterloo, with ability to travel Start Date: Flexible, ideally Q3 2025 About Hum.ai Hum.ai is building planetary superintelligence. Backed by top funds, we’ve raised $10M+ and are now heads down building... 
    Remote job
    Full time
    Work experience placement
    Flexible hours

    hum.ai

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Machine Learning Engineer. Be the first to apply!