Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead Machine Learning Inference Engineer, Advertising

$246.5k

Roku, Building C

Teamwork makes the stream work.Roku is changing how the world watches TVRoku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.About the team The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across large-scale distributed systems with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning and Inference Platform that powers the entire landscape. About the role In this role, you will architect, design, and lead the development of a SOTA Inference platform that can handle Advertising-level low latencies, scale, throughput, and availability with optimizations that span across hardware, software, and models. We’re looking for a strong technical leader with deep experience in ML serving, high-performance computing, and industry standard frameworks - someone excited to mentor engineers, innovate at scale, and shape the future of machine learning at Roku.For California Only - The estimated annual salary for this position is between $246,500 - $486,100 annually. Compensation packages are based on factors unique to each candidate, including but not limited to skill set, certifications, and specific geographical location. This role is eligible for health insurance, equity awards, life insurance, disability benefits, parental leave, wellness benefits, and paid time off. What you’ll be doingLead the design and development of a SOTA Inference platform Oversee the development of monitoring, observability, and other tooling to ensure system and model performance, reliability, and scalability of online inference servicesIdentify and resolve system inefficiencies, performance bottlenecks, and reliability issues, ensuring optimized end-to-end performance Stay at the forefront of advancements in inference frameworks, ML hardware acceleration, and distributed systems, and incorporate innovations where and when they are impactfulWe’re excited if you haveM.S. or above in CS, ECE, or a related field 10+ years of experience in developing and deploying large-scale, distributed systems, with at least 5 years in a leadership or technical lead role Strong programming skills in high-performance languagesDeep understanding of inference frameworks and ML system deploymentProven experience optimizing performance for large-scale machine learning systems, including a deep knowledge of SOTA model optimizations, hardware-software co-design, GPU acceleration, and HPC techniquesExcellent communication and collaboration skillsExperience leading teams working on high-throughput, low-latency ML serving systemsExperience collaborating with and leading global, cross-functional teamsContributions to open-source ML or systems projects#LI-DH2What's Roku's approach to hybrid working?Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance.What are some of the benefits?Roku is committed to offering a diverse range of benefits as part of our compensation package to support our employees and their families. Our comprehensive benefits include global access to mental health and financial wellness support and resources. Local benefits include statutory and voluntary benefits which may include healthcare (medical, dental, and vision), life, accident, disability, commuter, and retirement options (401(k)/pension). Employees are supported in taking time off, in accordance with local leave policies and other personal needs to support their evolving work and life needs. It's important to note that not every benefit is available in all locations or for every role. For details specific to your location, please consult with your recruiter.AccommodationsRoku welcomes applicants of all backgrounds and provides reasonable accommodations and adjustments in accordance with applicable law. If you require reasonable accommodation at any point in the hiring process, please direct your inquiries to View email address on click.appcast.io should I know about Roku's culture?Roku is a great place for people who want to work in a fast-paced environment where everyone is focused on the company's success rather than their own. We try to surround ourselves with people who are great at their jobs, who are easy to work with, and who keep their egos in check. We appreciate a sense of humor. We believe a fewer number of very talented folks can do more for less cost than a larger number of less talented teams. We're independent thinkers with big ideas who act boldly, move fast and accomplish extraordinary things through collaboration and trust. In short, at Roku you'll be part of a company that's changing how the world watches TV. We have a unique culture that we are proud of. We think of ourselves primarily as problem-solvers, which itself is a two-part idea. We come up with the solution, but the solution isn't real until it is built and delivered to the customer. That penchant for action gives us a pragmatic approach to innovation, one that has served us well since 2002. To learn more about Roku, our global footprint, and how we've grown, visit .By providing your information, you acknowledge that you want Roku to contact you about job roles, that you have read Roku's Applicant Privacy Notice, and understand that Roku will use your information as described in that notice. If you do not wish to receive any communications from Roku regarding this role or similar roles in the future, you may unsubscribe at any time by emailing View email address on click.appcast.io.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Lead Machine Learning Inference Engineer, Advertising in San Jose, CA vacancy
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising...  ...building and operating production machine learning systems at scale.Experience...  ...:Experience with causal inference, experimentation, or... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    1 day ago
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising systems...  ...end ML model deployment and inference infra for low-latency real-...  ...culture and environment. Learn more here.Inclusion is a Netflix... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  •  ...creating a compelling path for advertisers to reach audiences that are...  ...Team The Ads Platform Engineering teams build advertising systems...  ...end ML model deployment and inference infra for low-latency real-...  ...culture and environment. Learn more here . Inclusion... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    20 hours ago
  •  ...of areas including e-commerce, advertising, and fulfillment. We use machine learning and Internet-scale data to elevate...  ...modeling, and general causal inference. Search & Discovery ML : The...  ...Instacart works alongside world-class engineers, data scientists, and product... 
    Suggested
    Remote job
    Permanent employment
    Work experience placement
    Internship
    Work at office
    Work from home
    Flexible hours

    Instacart

    San Jose, CA
    1 day ago
  • $160k - $225k

     ...Machine Learning Engineer Location: Mountain View, CA Company Stage of Funding...  ...the $1 trillion digital advertising industry. Their platform...  ...support model training and inference. Build customer-facing...  ...Preferred Experience at leading technology companies or... 
    Suggested
    H1b
    Work at office
    Visa sponsorship

    Recruiting from Scratch

    Mountain View, CA
    2 days ago
  • $278.1k - $347.6k

     ...accelerated entirely within that runtime. As our Principal Engineer for On-Device AI Inference & Systems, you will be the foremost engineering authority...  ...worth the engineering cost. Engineering Leadership Lead and mentor a team of engineers; set engineering best... 
    Work at office
    Worldwide
    Relocation package

    Unity

    Mountain View, CA
    2 days ago
  • $184k - $287.5k

    Intelligent machines powered by Artificial Intelligence computers that can learn, reason, and interact with people are no longer...  ...seeking the best Machine Learning Engineers with a background in computer...  ...optimization for real-time inference on embedded or automotive platforms... 
    Full time
    Worldwide
    Night shift

    NVIDIA

    Santa Clara, CA
    12 hours ago
  • $151.8k - $265.35k

     ...adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services...  ...ownership of production ML or inference services at scale. Strong Python and...  ...customer experiences. Adobe’s industry-leading offerings including Adobe Acrobat Studio... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    1 day ago
  • $148.75k - $361k

     ...audiences, and provide advertisers unique capabilities...  ...low latency. We use Machine Learning, Reinforcement Learning...  ...Experimentation, and Inference Platform that powers...  ...experienced Senior Software Engineer, MLOps/DevOps, to...  ...What you’ll be doing Lead the design and... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    2 days ago
  • $229.5k - $360k

     ...large audiences, and provide advertisers unique capabilities to...  ...leveraging state-of-the-art machine learning. Our mission is to deliver...  ...Our work blends innovation, engineering excellence, and a deep commitment...  ...: feature store, real-time inference services, vector DBs, etc.,... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    1 day ago
  • $193.3k - $261.5k

    The Product: AWS Machine Learning accelerators are at the forefront of...  ...delivers best-in-class ML inference performance at the lowest cost...  ...including silicon engineering, hardware design and verification...  ...language experience- 5+ years of leading design or architecture (... 
    Internship
    Local area
    Work from home
    Relocation
    Flexible hours

    Amazon

    Cupertino, CA
    16 hours ago
  • $224k - $356.5k

    We are looking for outstanding Machine Learning Engineers to join our Physical AI teams. As the pioneers...  ...metric designs.SOTA Data Engineering: Lead the generation of massive training...  ...architecture to improve the performance during inference/training.Familiarity with simulation... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $174.72k - $295.68k

    XPENG is a leading smart technology company at the forefront of...  ...through cutting-edge R&D in AI, machine learning, and smart connectivity.We...  ...full-time Machine Learning Engineer - AI Foundation, with deep knowledge...  ...accelerating model training/inference. Our mission is to solve the... 
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $128k - $260.5k

     ...platform, its Zero Trust Engine, and the powerful...  ...education and mentorship, and lead with transparency and...  ...Careers at Netskope to learn more. Follow us on...  ...intelligence (AI) and machine learning (ML) to protect...  ...the bleeding edge of LLM inference optimization, utilizing... 

    Netskope

    Santa Clara, CA
    1 day ago
  • $160k - $225k

     ...on a mission to democratize advanced advertising technology. We believe that cutting-edge...  ...be used to expand our product and engineering teams, bringing our vision of intelligent...  ...re writing the manual. As an early Machine Learning Engineer at MAI, you won't just be writing... 
    Full time

    Mai

    Mountain View, CA
    20 hours ago
  • $272k - $431.25k

     ...NVIDIA is looking for a Machine Learning (ML) Engineer to join the GPU accelerated Apache Spark team. Apache...  ..., SQL, and ML/DL model training and inference pipelines, spanning many domains and...  ...solutions.5+ experience as technical lead in ML model development.Proven hands-... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $160k - $200k

    Santa Clara, CAData Engineering - Machine Learning and Data Engineer /Full-time /HybridPlusAI is a Physical AI company pioneering AI-based virtual...  ...while ensuring optimal performance for both training and inference phases. You will build robust pipelines for managing... 
    Full time

    Plus.ai

    Santa Clara, CA
    2 days ago
  • $214.67k - $322k

     ...individual's freedom.OKX is a leading crypto exchange, and the...  ...About The OpportunityBuilding machine learning systems for risk at a...  ...different from conventional ML engineering. The data spans on-chain...  ...offline training and online inference.Ensure models and decision... 

    OKX

    San Jose, CA
    1 day ago
  • $215.28k - $364.32k

    XPENG is a leading smart technology company at the forefront of...  ...through cutting-edge R&D in AI, machine learning, and smart connectivity....  ...for a strong Machine Learning Engineer / Computer Vision Engineer...  .../ TensorRT / quantization / inference acceleration.Work with deployment... 
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $232k - $310k

     ...collaboration, and high standards. Our engineers, product leaders, and go-to-...  ...the boundaries of applied machine learning. We work with massive...  ...languages.Responsibilities:Lead the team in: research,...  ...strategies (QLORA, DPO) and inference optimization (vLLM, TensorRT... 
    Work experience placement
    Work at office
    Remote work
    Flexible hours
    3 days per week

    Eightfold

    Santa Clara, CA
    3 days ago
  • $152k - $190k

     ...cybersecurity.RoleWe are looking for a Staff Software Engineer to join our team. This is a hybrid, based in San...  ...production features, utilizing LLMs, various machine learning models, data processing, fine-tuning, and inference optimizationWork with the world class cloud... 
    Full time
    Work at office
    Local area
    Worldwide
    3 days per week

    Zscaler

    San Jose, CA
    2 days ago
  • Lead the team in: research, design, development, and deployment of advanced AI agents...  ...experiences. Knowledge and passion in machine learning algorithms, Gen AI, LLMs, and natural...  ...fine-tuning strategies (QLORA, DPO) and inference optimization (vLLM, TensorRT-LLM). Research... 
    Full time
    Work experience placement

    Eightfold

    Santa Clara, CA
    20 hours ago
  •  ...We are now looking for a Deep Learning Software Engineer, TensorRT Performance! NVIDIA is seeking an...  ...improving the performance of NVIDIA’s inference ecosystem! NVIDIA is rapidly growing...  ...learning has provided the foundation for machines to learn, perceive, reason and solve... 

    NVIDIA

    Santa Clara, CA
    1 day ago
  • $174.3k - $200k

     ...Machine Learning Engineer UnitX builds the world's leading physical AI systems to automate repetitive visual tasks in factories. UnitX is a fast-moving startup...  ...) to ensure pixel-level precision and real-time inference. Build Automated Pipelines: Develop and... 

    UnitX

    Milpitas, CA
    1 day ago
  •  ...We are seeking a highly skilled Machine Learning Engineer to design and build a low-latency query understanding...  ..., with a focus on sub-second inference, CPU-based execution, and scalable...  ...entities, applications, evidence). Lead data collection, synthesis, analysis,... 
    Local area

    Sparktek

    San Jose, CA
    2 days ago
  • $110k - $145k

     ...Overview We are looking for a talented and experienced Deep Learning Engineer specializing in Large Language Models (LLMs) to join our dynamic...  ...and apply risk mitigation techniques, including safe inference strategies to ensure reliable AI behavior. Identify vulnerabilities... 

    A10 Networks

    San Jose, CA
    1 day ago
  • $164k - $313.3k

     ...OpportunityPhotoshop ART is seeking a Senior Machine Learning (ML) Systems & Efficiency Engineer to join our R&D team focused on...  ...production-ready improvements in inference performance, latency, and cost...  ...approaches when they directly lead to more efficient inference,... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    1 day ago
  • $361.3k

     ...audiences, and provide advertisers unique capabilities...  ...LLMs, reinforcement learning, multi-objective...  ...looking for a Senior Machine Learning Tech Lead to drive ranking and...  ...across the agentic engineering toolchain — coding harnesses...  ..., and causal inference techniquesBuild multi... 
    Part time
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    5 days ago
  • $152k - $241.5k

    We are now looking for a Senior Machine Learning Applications and Compiler Engineer!NVIDIA is seeking engineers to develop algorithms and optimizations for our LPX inference and compiler stack. You will work at the intersection of large-scale systems, compilers, and deep... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $120k

     ...Overview Netflix is one of the world's leading entertainment services with over 247 million...  .... The Ads Platform team builds the advertising systems that power the delivery of ads...  ...person will partner with Product, TPM, Engineering, Data Science, and other cross‑functional... 
    Flexible hours

    Netflix Global, LLC

    Los Gatos, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead Machine Learning Inference Engineer, Advertising. Be the first to apply!