Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead Machine Learning Inference Engineer, Advertising

$246.5k

Roku, Building C

Teamwork makes the stream work.Roku is changing how the world watches TVRoku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.About the team The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across large-scale distributed systems with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning and Inference Platform that powers the entire landscape. About the role In this role, you will architect, design, and lead the development of a SOTA Inference platform that can handle Advertising-level low latencies, scale, throughput, and availability with optimizations that span across hardware, software, and models. We’re looking for a strong technical leader with deep experience in ML serving, high-performance computing, and industry standard frameworks - someone excited to mentor engineers, innovate at scale, and shape the future of machine learning at Roku.For California Only - The estimated annual salary for this position is between $246,500 - $486,100 annually. Compensation packages are based on factors unique to each candidate, including but not limited to skill set, certifications, and specific geographical location. This role is eligible for health insurance, equity awards, life insurance, disability benefits, parental leave, wellness benefits, and paid time off. What you’ll be doingLead the design and development of a SOTA Inference platform Oversee the development of monitoring, observability, and other tooling to ensure system and model performance, reliability, and scalability of online inference servicesIdentify and resolve system inefficiencies, performance bottlenecks, and reliability issues, ensuring optimized end-to-end performance Stay at the forefront of advancements in inference frameworks, ML hardware acceleration, and distributed systems, and incorporate innovations where and when they are impactfulWe’re excited if you haveM.S. or above in CS, ECE, or a related field 10+ years of experience in developing and deploying large-scale, distributed systems, with at least 5 years in a leadership or technical lead role Strong programming skills in high-performance languagesDeep understanding of inference frameworks and ML system deploymentProven experience optimizing performance for large-scale machine learning systems, including a deep knowledge of SOTA model optimizations, hardware-software co-design, GPU acceleration, and HPC techniquesExcellent communication and collaboration skillsExperience leading teams working on high-throughput, low-latency ML serving systemsExperience collaborating with and leading global, cross-functional teamsContributions to open-source ML or systems projects#LI-DH2What's Roku's approach to hybrid working?Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance.What are some of the benefits?Roku is committed to offering a diverse range of benefits as part of our compensation package to support our employees and their families. Our comprehensive benefits include global access to mental health and financial wellness support and resources. Local benefits include statutory and voluntary benefits which may include healthcare (medical, dental, and vision), life, accident, disability, commuter, and retirement options (401(k)/pension). Employees are supported in taking time off, in accordance with local leave policies and other personal needs to support their evolving work and life needs. It's important to note that not every benefit is available in all locations or for every role. For details specific to your location, please consult with your recruiter.AccommodationsRoku welcomes applicants of all backgrounds and provides reasonable accommodations and adjustments in accordance with applicable law. If you require reasonable accommodation at any point in the hiring process, please direct your inquiries to View email address on click.appcast.io should I know about Roku's culture?Roku is a great place for people who want to work in a fast-paced environment where everyone is focused on the company's success rather than their own. We try to surround ourselves with people who are great at their jobs, who are easy to work with, and who keep their egos in check. We appreciate a sense of humor. We believe a fewer number of very talented folks can do more for less cost than a larger number of less talented teams. We're independent thinkers with big ideas who act boldly, move fast and accomplish extraordinary things through collaboration and trust. In short, at Roku you'll be part of a company that's changing how the world watches TV. We have a unique culture that we are proud of. We think of ourselves primarily as problem-solvers, which itself is a two-part idea. We come up with the solution, but the solution isn't real until it is built and delivered to the customer. That penchant for action gives us a pragmatic approach to innovation, one that has served us well since 2002. To learn more about Roku, our global footprint, and how we've grown, visit .By providing your information, you acknowledge that you want Roku to contact you about job roles, that you have read Roku's Applicant Privacy Notice, and understand that Roku will use your information as described in that notice. If you do not wish to receive any communications from Roku regarding this role or similar roles in the future, you may unsubscribe at any time by emailing View email address on click.appcast.io.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Lead Machine Learning Inference Engineer, Advertising in San Jose, CA vacancy
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising...  ...building and operating production machine learning systems at scale.Experience...  ...:Experience with causal inference, experimentation, or... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising...  ...experiences by utilizing advanced machine learning models for identity...  ...end ML model deployment and inference infra for low-latency real-... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  • $224k - $356.5k

    NVIDIA is looking for a Machine Learning Engineer to join the GPU accelerated Apache Spark team.Apache...  ...ETL, SQL, and ML/DL model training and inference pipelines, spanning many domains and...  ...solutions.3+ experience as technical lead in ML model development.Proven hands-on... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $193.3k - $261.5k

    The Product: AWS Machine Learning accelerators are at the forefront of...  ...delivers best-in-class ML inference performance at the lowest cost...  ...including silicon engineering, hardware design and verification...  ...language experience- 5+ years of leading design or architecture (... 
    Suggested
    Internship
    Local area
    Work from home
    Relocation
    Flexible hours

    Amazon

    Cupertino, CA
    16 hours ago
  • $342.7k

     ...internet is managed. As a Distinguished Machine Learning Engineer, you will have a seat at the table...  ...senior executive leadership, you will lead the architecture for Cisco Cloud Control...  ...supervised fine-tuning.Deep knowledge of inference serving optimizations, such as KV/... 
    Suggested
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    Shift work

    CISCO Systems

    San Jose, CA
    3 days ago
  • $148.75k - $361k

     ...audiences, and provide advertisers unique capabilities...  ...low latency. We use Machine Learning, Reinforcement Learning...  ...Experimentation, and Inference Platform that powers...  ...experienced Senior Software Engineer, MLOps/DevOps, to...  ...What you’ll be doing Lead the design and... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    2 days ago
  • $229.5k - $360k

     ...large audiences, and provide advertisers unique capabilities to...  ...leveraging state-of-the-art machine learning. Our mission is to deliver...  ...Our work blends innovation, engineering excellence, and a deep commitment...  ...: feature store, real-time inference services, vector DBs, etc.,... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    1 day ago
  • $174.72k - $295.68k

    XPENG is a leading smart technology company at the forefront of...  ...cutting-edge R&D in AI, machine learning, and smart connectivity.Our...  ...tuning, PTQ, QAT, on-vehicle inference and related fields.Key ResponsibilitiesDevelop...  ...programming and software engineering skills.Ability to work... 
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $174.72k - $295.68k

    XPENG is a leading smart technology company at the forefront of...  ...through cutting-edge R&D in AI, machine learning, and smart connectivity.We...  ...full-time Machine Learning Engineer - AI Foundation, with deep knowledge...  ...accelerating model training/inference.Our mission is to solve the... 
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $151.8k - $265.35k

     ...adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services...  ...ownership of production ML or inference services at scale. Strong Python and...  ...customer experiences. Adobe’s industry-leading offerings including Adobe Acrobat Studio... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    1 day ago
  • $149.37k - $275k

     ...Title and Location : Sr Machine Learning Engineer in Santa Clara, CA. Job Responsibilities Implement deep-learning models and frameworks...  ...-scale image datasets (1 yr, 6 mos). Evaluate model inference performance and runtime behavior across heterogeneous... 
    Work at office
    Local area
    Remote work

    Blue River Technology

    Santa Clara, CA
    1 day ago
  • $160k - $180k

     ...160K – 180K • Offers Equity Job Title Machine Learning Engineer About Us UnitX builds the world's leading physical AI systems to automate repetitive visual...  ...technical field. Experience with model inference optimization Experience with non‑ML CV algorithms... 
    Full time

    UnitX

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...Machine Learning Engineer (Finance) NVIDIA is looking for a talented Machine Learning Engineer to drive the development, evaluation, deployment...  ...-level tuning for high-throughput, low-latency AI inference workflows. GitLab CI/CD & Security Automation: Advanced... 
    Flexible hours

    Nvidia Corporation in

    Santa Clara, CA
    2 days ago
  • $160k - $225k

     ...on a mission to democratize advanced advertising technology. We believe that cutting‑edge...  ...be used to expand our product and engineering teams, bringing our vision of intelligent...  ...we’re writing the manual. As an early Machine Learning Engineer at MAI, you will be the architect... 

    MAI Agents

    Mountain View, CA
    2 days ago
  • $174.3k - $200k

     ...Machine Learning Engineer UnitX builds the world's leading physical AI systems to automate repetitive visual tasks in factories. UnitX is a fast-moving startup...  ...) to ensure pixel-level precision and real-time inference. Build Automated Pipelines: Develop and... 

    UnitX

    Milpitas, CA
    1 day ago
  • $184.7k - $324.8k

     ...Services Apple Services Engineering (ASE) builds experiences that...  ...applied research, advanced machine learning, and large language models converge...  ...understanding, behavioral inference, discovery, and growth...  ...Apple’s services ecosystem. Lead research in areas such as... 
    Relocation

    Apple

    Cupertino, CA
    1 day ago
  • $160k - $200k

    Santa Clara, CAData Engineering - Machine Learning and Data Engineer /Full-time /HybridPlusAI is a Physical AI company pioneering AI-based virtual...  ...while ensuring optimal performance for both training and inference phases. You will build robust pipelines for managing... 
    Full time

    Plus.ai

    Santa Clara, CA
    2 days ago
  • $147k - $237.5k

     ....Job SummaryWe are seeking a highly experienced ML Engineer with a good understanding of machine learning principles and algorithms to join our team and drive...  ...model training, validation, and real-time inference.Optimize the existing ML models and pipelineEnsure... 
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    1 day ago
  • $190.2k - $345.65k

     ...into adjacent verticals.We are hiring a Senior Staff Machine Learning Engineer to architect and lead the data processing, indexing, and search infrastructure...  ...; hands-on familiarity with embedding models and the inference paths that produce them (PyTorch).A track record... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    3 days ago
  • $201.3k - $352.3k

    Description de l'entrepriseIt all started when engineer Fred Luddy wrote code that automated a...  .... Exposure to LLM fine-tuning or inference optimization in productionWhy join us Intelligence...  ...work and their assigned work location. Learn more here. To determine eligibility for... 
    Work experience placement
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    3 days ago
  • $206.4k - $379.1k

     ...Express, Stock, and Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical lead for our GenAI Services area. This is not a model-...  ...and served at enterprise scale. You will set the inference architecture and technical standards that a... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    3 days ago
  • $232k - $310k

     ...collaboration, and high standards. Our engineers, product leaders, and go-to-...  ...the boundaries of applied machine learning. We work with massive...  ...languages.Responsibilities:Lead the team in: research,...  ...strategies (QLORA, DPO) and inference optimization (vLLM, TensorRT... 
    Work experience placement
    Work at office
    Remote work
    Flexible hours
    3 days per week

    Eightfold

    Santa Clara, CA
    3 days ago
  • $152k - $190k

     ...cybersecurity.RoleWe are looking for a Staff Software Engineer to join our team. This is a hybrid, based in San...  ...production features, utilizing LLMs, various machine learning models, data processing, fine-tuning, and inference optimizationWork with the world class cloud... 
    Full time
    Work at office
    Local area
    Worldwide
    3 days per week

    Zscaler

    San Jose, CA
    2 days ago
  •  ...grown to 4,000+ people servicing a global network of prestigious advertising, public relations, media, healthcare and digital marketing...  ...deliver award-winning campaigns for their clients.OverviewThe Team Lead - Accounts Payable, will be responsible for supervising and... 
    Local area

    Publicis Media

    San Jose, CA
    2 days ago
  • $220k - $300k

     ...infrastructure, delivering a full-stack inference platform for customers worldwide. At...  ...About the role As a Senior Principal Machine Learning Engineer, you will be responsible for designing...  ...RDU and broader hardware ecosystem Lead hardware-software co-design efforts... 
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    SambaNova

    San Jose, CA
    2 days ago
  •  ...Job Description: We are looking for a Machine Learning Engineer to join our core research and...  ...scale training, high-throughput video inference, and reliable production pipelines over...  ...large-scale video. Publications at leading venues (CVPR, ICCV, ECCV, NeurIPS, SIGGRAPH... 

    Maxinsights

    Santa Clara, CA
    14 days ago
  • $2,000 per month

     ...throughput, and/or a mix of these metrics. Implement model-specific inference-time acceleration techniques such as speculative decoding,...  ...We are a fully in-person team in Cupertino, and greatly value engineering skills. We do not have boundaries between engineering and... 
    Work at office
    Relocation package

    ETCHED LLC

    Cupertino, CA
    2 days ago
  • $197.5k - $272k

     ...and improve continuously. That's why leading OEMs trust Sonatus to accelerate this...  ...Edge. We are looking for a great Staff Machine Learning Engineer to join our seasoned AI team and...  ...knowledge of modern C++ (C++14/17 for inference). ~ Deep proficiency with PyTorch or... 
    Work at office
    Worldwide
    Flexible hours
    Shift work
    3 days per week

    Sonatus

    San Jose, CA
    22 days ago
  • $152k - $241.5k

    We are now looking for a Senior Machine Learning Applications and Compiler Engineer!NVIDIA is seeking engineers to develop algorithms and optimizations for our LPX inference and compiler stack. You will work at the intersection of large-scale systems, compilers, and deep... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $117.7k - $221.4k

     ...sits at the intersection of machine learning, data infrastructure, and developer...  ...model reflects how Cola engineers think: build durable...  ...processing, featurization, and inference foundations that power scalable...  ...the responsibility to lead the change that will make our... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead Machine Learning Inference Engineer, Advertising. Be the first to apply!