Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead Machine Learning Inference Engineer, Advertising

$246.5k

Roku, Building C

Teamwork makes the stream work.Roku is changing how the world watches TVRoku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.About the team The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across large-scale distributed systems with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning and Inference Platform that powers the entire landscape. About the role In this role, you will architect, design, and lead the development of a SOTA Inference platform that can handle Advertising-level low latencies, scale, throughput, and availability with optimizations that span across hardware, software, and models. We’re looking for a strong technical leader with deep experience in ML serving, high-performance computing, and industry standard frameworks - someone excited to mentor engineers, innovate at scale, and shape the future of machine learning at Roku.For California Only - The estimated annual salary for this position is between $246,500 - $486,100 annually. Compensation packages are based on factors unique to each candidate, including but not limited to skill set, certifications, and specific geographical location. This role is eligible for health insurance, equity awards, life insurance, disability benefits, parental leave, wellness benefits, and paid time off. What you’ll be doingLead the design and development of a SOTA Inference platform Oversee the development of monitoring, observability, and other tooling to ensure system and model performance, reliability, and scalability of online inference servicesIdentify and resolve system inefficiencies, performance bottlenecks, and reliability issues, ensuring optimized end-to-end performance Stay at the forefront of advancements in inference frameworks, ML hardware acceleration, and distributed systems, and incorporate innovations where and when they are impactfulWe’re excited if you haveM.S. or above in CS, ECE, or a related field 10+ years of experience in developing and deploying large-scale, distributed systems, with at least 5 years in a leadership or technical lead role Strong programming skills in high-performance languagesDeep understanding of inference frameworks and ML system deploymentProven experience optimizing performance for large-scale machine learning systems, including a deep knowledge of SOTA model optimizations, hardware-software co-design, GPU acceleration, and HPC techniquesExcellent communication and collaboration skillsExperience leading teams working on high-throughput, low-latency ML serving systemsExperience collaborating with and leading global, cross-functional teamsContributions to open-source ML or systems projects#LI-DH2What's Roku's approach to hybrid working?Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance.What are some of the benefits?Roku is committed to offering a diverse range of benefits as part of our compensation package to support our employees and their families. Our comprehensive benefits include global access to mental health and financial wellness support and resources. Local benefits include statutory and voluntary benefits which may include healthcare (medical, dental, and vision), life, accident, disability, commuter, and retirement options (401(k)/pension). Employees are supported in taking time off, in accordance with local leave policies and other personal needs to support their evolving work and life needs. It's important to note that not every benefit is available in all locations or for every role. For details specific to your location, please consult with your recruiter.AccommodationsRoku welcomes applicants of all backgrounds and provides reasonable accommodations and adjustments in accordance with applicable law. If you require reasonable accommodation at any point in the hiring process, please direct your inquiries to View email address on us.fitly.work should I know about Roku's culture?Roku is a great place for people who want to work in a fast-paced environment where everyone is focused on the company's success rather than their own. We try to surround ourselves with people who are great at their jobs, who are easy to work with, and who keep their egos in check. We appreciate a sense of humor. We believe a fewer number of very talented folks can do more for less cost than a larger number of less talented teams. We're independent thinkers with big ideas who act boldly, move fast and accomplish extraordinary things through collaboration and trust. In short, at Roku you'll be part of a company that's changing how the world watches TV. We have a unique culture that we are proud of. We think of ourselves primarily as problem-solvers, which itself is a two-part idea. We come up with the solution, but the solution isn't real until it is built and delivered to the customer. That penchant for action gives us a pragmatic approach to innovation, one that has served us well since 2002. To learn more about Roku, our global footprint, and how we've grown, visit .By providing your information, you acknowledge that you want Roku to contact you about job roles, that you have read Roku's Applicant Privacy Notice, and understand that Roku will use your information as described in that notice. If you do not wish to receive any communications from Roku regarding this role or similar roles in the future, you may unsubscribe at any time by emailing View email address on us.fitly.work.

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the Lead Machine Learning Inference Engineer, Advertising in San Jose, CA vacancy
  • $160k - $225k

    Machine Learning Engineer Location: Mountain View, CA Company Stage of Funding...  ...for the $1 trillion digital advertising industry. Their platform...  ...support model training and inference. Build customer-facing AI...  .... Preferred Experience at leading technology companies or high... 
    Suggested
    H1b
    Work at office
    Visa sponsorship

    Recruiting from Scratch

    Mountain View, CA
    5 days ago
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising...  ...building and operating production machine learning systems at scale.Experience...  ...:Experience with causal inference, experimentation, or... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  •  ...scalable ML platform services including feature stores, real-time inference services, and vector databases serving millions of...  ...KPIs to optimize recommendation performance. Collaborate with engineering and cross-functional teams to translate business requirements... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work
    Monday to Friday
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    a month ago
  • $193.3k - $261.5k

    The Product: AWS Machine Learning accelerators are at the forefront of...  ...delivers best-in-class ML inference performance at the lowest cost...  ...including silicon engineering, hardware design and verification...  ...language experience- 5+ years of leading design or architecture (... 
    Suggested
    Internship
    Local area
    Work from home
    Relocation
    Flexible hours

    Amazon

    Cupertino, CA
    13 days ago
  • $224k - $356.5k

    NVIDIA is looking for a Machine Learning Engineer to join the GPU accelerated Apache Spark team.Apache...  ...ETL, SQL, and ML/DL model training and inference pipelines, spanning many domains and...  ...solutions.3+ experience as technical lead in ML model development.Proven hands-on... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $148.75k - $361k

     ...audiences, and provide advertisers unique capabilities...  ...low latency. We use Machine Learning, Reinforcement Learning...  ...Experimentation, and Inference Platform that powers...  ...experienced Senior Software Engineer, MLOps/DevOps, to...  ...What you’ll be doing Lead the design and... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    a month ago
  •  ...types. Accelerate cloud model inference for closed-loop simulation, reinforcement learning, and enterprise LLM/VLM...  ...in computer science, computer engineering, or electrical engineering, or...  ...work. Opportunity to work with leading talent on cutting-edge technology... 
    Full time

    XPENG

    Santa Clara, CA
    2 days ago
  • $174.72k - $295.68k

    XPENG is a leading smart technology company at the forefront of...  ...cutting-edge R&D in AI, machine learning, and smart connectivity.Our...  ...tuning, PTQ, QAT, on-vehicle inference and related fields.Key ResponsibilitiesDevelop...  ...programming and software engineering skills.Ability to work... 
    Full time

    XPENG Motors

    Santa Clara, CA
    22 days ago
  •  ...Responsibilities Take machine learning models from prototype to production by building reliable...  ..., scalable training, serving, and inference pipelines. Design and maintain data...  .... Work with Data Scientists, Data Engineers, and Software Engineers to turn business... 
    Full time
    Work at office
    Flexible hours

    Pure Storage

    Santa Clara, CA
    2 days ago
  • $170k - $240k

    Machine Learning Engineer Santa Clara, CA About the role We are seeking a high-impact, technically deep Machine Learning Engineer to develop,...  ...trained models into C++-based autonomy systems and optimize inference for production vehicle hardware. Model Optimization:... 
    Work at office

    Gatik AI

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    Intelligent machines powered by Artificial Intelligence computers that can learn, reason, and interact with people are no longer...  ...seeking the best Machine Learning Engineers with a background in computer...  ...optimization for real-time inference on embedded or automotive platforms... 
    Worldwide
    Night shift

    NVIDIA

    Santa Clara, CA
    2 days ago
  • $149.37k - $275k

    Title and Location : Sr Machine Learning Engineer in Santa Clara, CA. Job Responsibilities Implement deep-learning models and frameworks for...  ...large-scale image datasets (1 yr, 6 mos). Evaluate model inference performance and runtime behavior across heterogeneous hardware... 
    Work at office
    Local area
    Remote work

    Blue River Technology

    Santa Clara, CA
    5 days ago
  • $200k - $280k

    Machine Learning Engineer Responsibilities: Develop, optimize, and deploy lightweight machine learning...  ...for embedded AI applications. Improve inference efficiency and model compression techniques...  ...managing a team, serving in a Team Lead role, or demonstrating strong... 
    Local area

    TetraMem - Accelerate The World

    San Jose, CA
    5 days ago
  • $184.7k - $324.8k

    Senior Machine Learning Engineer Apple Services Engineering (ASE) builds experiences that touch hundreds...  ...language understanding, behavioral inference, discovery, and growth optimization—while...  ...across Apple's services ecosystem. Lead research in areas such as large-scale... 
    Relocation

    Apple

    Cupertino, CA
    3 days ago
  • $160k - $200k

     ...As a Senior ML Infrastructure Engineer at Plus, you will design...  ...performance for both training and inference phases. You will build robust...  ...with state-of-the-art deep learning frameworks like PyTorch or TensorFlow...  ...of what's possible in machine learning infrastructure and contribute... 

    PlusAI

    Santa Clara, CA
    9 days ago
  • $214.67k - $322k

     ...individual's freedom.OKX is a leading crypto exchange, and the...  ...About The OpportunityBuilding machine learning systems for risk at a...  ...different from conventional ML engineering. The data spans on-chain...  ...offline training and online inference.Ensure models and decision... 

    OKX

    San Jose, CA
    a month ago
  • $206.4k - $379.1k

     ...Express, Stock, and Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical lead for our GenAI Services area. This is not a model-...  ...and served at enterprise scale. You will set the inference architecture and technical standards that a... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    18 days ago
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising...  ...experiences by utilizing advanced machine learning models for identity...  ...end ML model deployment and inference infra for low-latency real-... 
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  • $220k - $300k

    Senior Principal Machine Learning Engineer San Jose, California, United States The era of pervasive AI...  ...architecture, training and fine-tuning, inference optimization, evaluation, and data...  ...'s RDU and broader hardware ecosystem Lead hardware-software co-design efforts in... 
    Full time
    Temporary work
    Local area
    Flexible hours

    SambaNova Systems

    San Jose, CA
    5 days ago
  • $2,000 per month

    Machine Learning Research Engineer Cupertino, CA Etched is building AI chips that are hard-coded for individual model architectures. Our first product...  ...and/or a mix of these metrics. Implement model-specific inference-time acceleration techniques such as speculative... 
    Work at office
    Relocation package

    ETCHED LLC

    Cupertino, CA
    5 days ago
  • $152k - $241.5k

    We are now looking for a Senior Machine Learning Applications and Compiler Engineer!NVIDIA is seeking engineers to develop algorithms and optimizations for our LPX inference and compiler stack. You will work at the intersection of large-scale systems, compilers, and deep... 
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  •  ...the future of autonomy, Plus is looking for talented individuals to join its fast-growing teams. We are seeking a Senior Machine Learning Engineer with expertise in deep learning and data analysis. In this role, you will apply data-driven techniques to develop high-... 

    PlusAI

    Santa Clara, CA
    23 days ago
  •  ...Responsibilities Develop and improve machine learning and large language model systems for...  ...generation, debugging, and related software engineering tasks. Build AI-powered developer...  .... Experience building or leading AI systems for software engineering, such... 
    Full time

    ByteDance

    San Jose, CA
    2 days ago
  • $184k - $287.5k

     ...scale. We seek a Senior ML Engineer to compose and deliver next-...  ...support the wider ecosystem.Lead the open-sourcing of solutions...  ...experience in applied machine learning or AI research.Deep expertise...  ...quantization, and real-time inference optimization for production... 
    Full time

    Nvidia

    Santa Clara, CA
    18 days ago
  • $150k - $230k

     ...the Role We are looking for a hands-on Machine Learning Engineer to drive the post-training of our...  ...read about it. Responsibilities Lead post-training of our LLMs across the full...  .../Accelerate, DeepSpeed or FSDP, and inference engines like vLLM. Solid understanding... 
    Full time
    Local area
    Work from home

    NewsBreak

    Mountain View, CA
    17 days ago
  •  ...Responsibilities Architect, train, and optimize multitask deep learning models across multimodal sensor streams. Build automated...  ...Profile and optimize models for efficient onboard accelerator inference. Requirements ~2–5+ years of experience training and... 
    Full time

    Waymo

    Mountain View, CA
    19 days ago
  • $165.2k - $223.6k

     ...development kit used to accelerate deep learning and GenAI workloads on Amazon’s custom machine learning accelerators,...  ...hardware-software boundary, our engineers craft high-performance kernels for...  ...PyTorch, enabling unparalleled ML inference and training performance.As part... 
    Internship
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    a month ago
  • $195k - $230k

     ..., visit  About the Role We are looking for a Senior Machine Learning Engineer to help evolve our large-scale recommendation systems and...  ...metrics. Own systems from offline training → online inference → A/B experimentation → metric analysis . Identify and... 
    Full time
    Local area
    Work from home

    NewsBreak

    Mountain View, CA
    17 days ago
  • $193.3k - $261.5k

     ...revolution? At AWS our vision is to make deep learning pervasive for everyday developers and to...  ...role is for a senior software engineer in the Compiler team for AWS Neuron. As...  ...professional.Basic qualifications- 5+ years of leading design or architecture (design patterns,... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    a month ago
  • $184k - $287.5k

     ...Intelligent machines powered by Artificial Intelligence computers that can learn, reason and interact with people are no longer...  ...extraordinary Senior Perception Engineer to develop and productize NVIDIA...  ...by technical publications in leading conferences/journals.Expertise... 
    Odd job
    Full time
    Work experience placement
    Remote work

    Nvidia

    Santa Clara, CA
    a month ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead Machine Learning Inference Engineer, Advertising. Be the first to apply!