Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead Machine Learning Inference Engineer, Advertising

$246.5k

Roku, Building C

Teamwork makes the stream work.Roku is changing how the world watches TVRoku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.About the team The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across large-scale distributed systems with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning and Inference Platform that powers the entire landscape. About the role In this role, you will architect, design, and lead the development of a SOTA Inference platform that can handle Advertising-level low latencies, scale, throughput, and availability with optimizations that span across hardware, software, and models. We’re looking for a strong technical leader with deep experience in ML serving, high-performance computing, and industry standard frameworks - someone excited to mentor engineers, innovate at scale, and shape the future of machine learning at Roku.For California Only - The estimated annual salary for this position is between $246,500 - $486,100 annually. Compensation packages are based on factors unique to each candidate, including but not limited to skill set, certifications, and specific geographical location. This role is eligible for health insurance, equity awards, life insurance, disability benefits, parental leave, wellness benefits, and paid time off. What you’ll be doingLead the design and development of a SOTA Inference platform Oversee the development of monitoring, observability, and other tooling to ensure system and model performance, reliability, and scalability of online inference servicesIdentify and resolve system inefficiencies, performance bottlenecks, and reliability issues, ensuring optimized end-to-end performance Stay at the forefront of advancements in inference frameworks, ML hardware acceleration, and distributed systems, and incorporate innovations where and when they are impactfulWe’re excited if you haveM.S. or above in CS, ECE, or a related field 10+ years of experience in developing and deploying large-scale, distributed systems, with at least 5 years in a leadership or technical lead role Strong programming skills in high-performance languagesDeep understanding of inference frameworks and ML system deploymentProven experience optimizing performance for large-scale machine learning systems, including a deep knowledge of SOTA model optimizations, hardware-software co-design, GPU acceleration, and HPC techniquesExcellent communication and collaboration skillsExperience leading teams working on high-throughput, low-latency ML serving systemsExperience collaborating with and leading global, cross-functional teamsContributions to open-source ML or systems projects#LI-DH2What's Roku's approach to hybrid working?Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance.What are some of the benefits?Roku is committed to offering a diverse range of benefits as part of our compensation package to support our employees and their families. Our comprehensive benefits include global access to mental health and financial wellness support and resources. Local benefits include statutory and voluntary benefits which may include healthcare (medical, dental, and vision), life, accident, disability, commuter, and retirement options (401(k)/pension). Employees are supported in taking time off, in accordance with local leave policies and other personal needs to support their evolving work and life needs. It's important to note that not every benefit is available in all locations or for every role. For details specific to your location, please consult with your recruiter.AccommodationsRoku welcomes applicants of all backgrounds and provides reasonable accommodations and adjustments in accordance with applicable law. If you require reasonable accommodation at any point in the hiring process, please direct your inquiries to View email address on click.appcast.io should I know about Roku's culture?Roku is a great place for people who want to work in a fast-paced environment where everyone is focused on the company's success rather than their own. We try to surround ourselves with people who are great at their jobs, who are easy to work with, and who keep their egos in check. We appreciate a sense of humor. We believe a fewer number of very talented folks can do more for less cost than a larger number of less talented teams. We're independent thinkers with big ideas who act boldly, move fast and accomplish extraordinary things through collaboration and trust. In short, at Roku you'll be part of a company that's changing how the world watches TV. We have a unique culture that we are proud of. We think of ourselves primarily as problem-solvers, which itself is a two-part idea. We come up with the solution, but the solution isn't real until it is built and delivered to the customer. That penchant for action gives us a pragmatic approach to innovation, one that has served us well since 2002. To learn more about Roku, our global footprint, and how we've grown, visit .By providing your information, you acknowledge that you want Roku to contact you about job roles, that you have read Roku's Applicant Privacy Notice, and understand that Roku will use your information as described in that notice. If you do not wish to receive any communications from Roku regarding this role or similar roles in the future, you may unsubscribe at any time by emailing View email address on click.appcast.io.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Lead Machine Learning Inference Engineer, Advertising in San Jose, CA vacancy
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising systems...  ...end ML model deployment and inference infra for low-latency real-...  ...culture and environment. Learn more here.Inclusion is a Netflix... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    15 hours ago
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising...  ...building and operating production machine learning systems at scale.Experience...  ...:Experience with causal inference, experimentation, or... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    15 hours ago
  • $128k - $256k

     ...Machine Learning Engineer Graduate (Ads Signal & Measurement) - 2027 StartLocation...  ...owns the full stack of advertising effectiveness — from...  ...machine learning and causal inference, applied at massive scale...  ...About TikTokTikTok is the leading destination for short-form... 
    Suggested
    Temporary work
    Internship
    Local area
    Worldwide

    Tik Tok

    San Jose, CA
    16 hours ago
  •  ...of areas including e-commerce, advertising, and fulfillment. We use machine learning and Internet-scale data to elevate...  ...modeling, and general causal inference. Search & Discovery ML : The...  ...Instacart works alongside world-class engineers, data scientists, and product... 
    Suggested
    Remote job
    Permanent employment
    Work experience placement
    Internship
    Work at office
    Work from home
    Flexible hours

    Instacart

    San Jose, CA
    4 days ago
  •  ...creating a compelling path for advertisers to reach audiences that are...  ....Our TeamThe Ads Platform Engineering teams build advertising...  ...experiences by utilizing advanced machine learning models for identity...  ...end ML model deployment and inference infra for low-latency real-... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    15 hours ago
  • $45 per hour

     ...Responsibilities The Lead Ads team empowers TikTok...  ...s personalized online advertising. We are looking for...  ...participate in social events, learning programs, and...  ...Ads algorithms by using Machine Learning. Work on NLP...  ...analysis, modeling, feature engineering Research and develop... 
    Hourly pay
    Full time
    Summer work
    Internship
    Local area

    Tik Tok

    San Jose, CA
    5 days ago
  •  ...TikTok is seeking a PhD-level Ads Core ML Engineer to advance a cutting-edge global advertising delivery system. You will work on ML/DL, RL, LLM, and scaling laws to optimize ad ranking, format, and revenue across US markets. The role focuses on building scalable ML pipelines... 

    Tik Tok

    San Jose, CA
    2 days ago
  • $215k - $285k

     ...intelligence to every moving machine on the planet. Applied Intuition...  ...Bangalore; Seoul; and Tokyo. Learn more at applied.co. We are an...  ...looking for a performance engineer who specializes in making large...  ..., and high-throughput batch inference sweeping petabytes of real-... 
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    NLP PEOPLE

    Sunnyvale, CA
    1 day ago
  • $184.7k - $324.8k

    Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference Santa Clara, California, United States Machine Learning and AI We are the...  ...organization. Minimum Qualifications 5+ years of experience leading complex, ambiguous technical projects from end to... 
    Worldwide
    Relocation

    Apple Inc.

    Santa Clara, CA
    1 day ago
  • $203.5k - $299.3k

     ...a causal question. About the Role We are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash...  ...you because you have… Deep practical experience with causal inference, econometrics, experimentation, or causal ML . Experience... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Visa Hunt

    Sunnyvale, CA
    3 days ago
  • $193.3k - $261.5k

    The Product: AWS Machine Learning accelerators are at the forefront of...  ...delivers best-in-class ML inference performance at the lowest cost...  ...including silicon engineering, hardware design and verification...  ...language experience- 5+ years of leading design or architecture (... 
    Internship
    Local area
    Work from home
    Relocation
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $174.72k - $295.68k

    XPENG is a leading smart technology company at the forefront of...  ...through cutting-edge R&D in AI, machine learning, and smart connectivity.We...  ...full-time Machine Learning Engineer - AI Foundation, with deep knowledge...  ...accelerating model training/inference. Our mission is to solve the... 
    Full time

    XPENG Motors

    Santa Clara, CA
    15 hours ago
  • $151.8k - $265.35k

     ...team is seeking SeniorMachine Learning Engineers for our GenAI Services area....  ...and develop efficient inference pipelines, optimize models for...  ...or PhD in Computer Science, Machine Learning, or a related field...  ...deployments.2+ years of experience leading large-scale, GPU-intensive... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    4 days ago
  • $128k - $260.5k

     ...platform, its Zero Trust Engine, and the powerful...  ...education and mentorship, and lead with transparency and...  ...Careers at Netskope to learn more. Follow us on...  ...intelligence (AI) and machine learning (ML) to protect...  ...the bleeding edge of LLM inference optimization, utilizing... 

    Netskope

    Santa Clara, CA
    4 days ago
  • $229.5k - $360k

     ...large audiences, and provide advertisers unique capabilities to...  ...leveraging state-of-the-art machine learning. Our mission is to deliver...  ...Our work blends innovation, engineering excellence, and a deep commitment...  ...: feature store, real-time inference services, vector DBs, etc.,... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    4 days ago
  • $148.75k - $361k

     ...audiences, and provide advertisers unique capabilities...  ...low latency. We use Machine Learning, Reinforcement Learning...  ...Experimentation, and Inference Platform that powers...  ...experienced Senior Software Engineer, MLOps/DevOps, to...  ...What you’ll be doing Lead the design and... 
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    15 hours ago
  • $184k - $287.5k

    Intelligent machines powered by Artificial Intelligence computers that can learn, reason, and interact with people are no longer...  ...seeking the best Machine Learning Engineers with a background in computer...  ...optimization for real-time inference on embedded or automotive platforms... 
    Full time
    Worldwide
    Night shift

    Nvidia

    Santa Clara, CA
    1 day ago
  • $174.72k - $295.68k

    XPENG is a leading smart technology company at the forefront of...  ...cutting-edge R&D in AI, machine learning, and smart connectivity.Our...  ...tuning, PTQ, QAT, on-vehicle inference and related fields.Key ResponsibilitiesDevelop...  ...programming and software engineering skills.Ability to work... 
    Full time

    XPENG Motors

    Santa Clara, CA
    11 hours ago
  • NVIDIA is seeking a Senior Product Manager for AI Platform Inference to build tools, SDKs, and libraries that enable developers to deploy inference on NVIDIA GPUs. You will craft product strategy, roadmaps, and go-to-market plans while partnering with developers to shape... 

    NVIDIA

    Santa Clara, CA
    4 days ago
  • $179.4k - $303.6k

    XPENG is a leading smart technology company at the forefront of...  ...through cutting-edge R&D in AI, machine learning, and smart connectivity....  ...for a strong Machine Learning Engineer / Computer Vision Engineer...  .../ TensorRT / quantization / inference acceleration.Work with deployment... 
    Full time

    XPENG Motors

    Santa Clara, CA
    15 hours ago
  • $272k - $431.25k

     ...NVIDIA is looking for a Machine Learning (ML) Engineer to join the GPU accelerated Apache Spark team. Apache...  ..., SQL, and ML/DL model training and inference pipelines, spanning many domains and...  ...solutions.5+ experience as technical lead in ML model development.Proven hands-... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $190.2k - $345.65k

     ...into adjacent verticals.We are hiring a Senior Staff Machine Learning Engineer to architect and lead the data processing, indexing, and search infrastructure...  ...; hands-on familiarity with embedding models and the inference paths that produce them (PyTorch).A track record... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    11 hours ago
  • $206.4k - $379.1k

     ...Express, Stock, and Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical lead for our GenAI Services area. This is not a model-...  ...and served at enterprise scale. You will set the inference architecture and technical standards that a... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    15 hours ago
  • $160k - $200k

    Santa Clara, CAData Engineering - Machine Learning and Data Engineer /Full-time /HybridPlusAI is a Physical AI company pioneering AI-based virtual...  ...while ensuring optimal performance for both training and inference phases. You will build robust pipelines for managing... 
    Full time

    Plus.ai

    Santa Clara, CA
    15 hours ago
  • $232k - $310k

     ...collaboration, and high standards. Our engineers, product leaders, and go-to-...  ...the boundaries of applied machine learning. We work with massive...  ...languages.Responsibilities:Lead the team in: research,...  ...strategies (QLORA, DPO) and inference optimization (vLLM, TensorRT... 
    Work experience placement
    Work at office
    Remote work
    Flexible hours
    3 days per week

    Eightfold

    Santa Clara, CA
    1 day ago
  • $201.3k - $352.3k

    Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a...  ...distributed systems. Formal grounding in machine learning fundamentals — modeling, training,...  ...retrieval. Exposure to LLM fine-tuning or inference optimization in production. Why join us... 
    Work experience placement
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    15 hours ago
  • $214.67k - $322k

     ...individual's freedom.OKX is a leading crypto exchange, and the...  ...About The OpportunityBuilding machine learning systems for risk at a...  ...different from conventional ML engineering. The data spans on-chain...  ...offline training and online inference.Ensure models and decision... 

    OKX

    San Jose, CA
    4 days ago
  • $124k - $195.5k

     ...Deep Learning Software Engineer, TensorRT PerformanceNVIDIA is seeking an experienced Deep Learning...  ...improving the performance of NVIDIA's inference ecosystem! NVIDIA is rapidly growing...  ...learning has provided the foundation for machines to learn, perceive, reason and solve... 

    NVIDIA

    Santa Clara, CA
    5 days ago
  • $110k - $145k

     ...Overview We are looking for a talented and experienced Deep Learning Engineer specializing in Large Language Models (LLMs) to join our dynamic...  ...and apply risk mitigation techniques, including safe inference strategies to ensure reliable AI behavior. Identify vulnerabilities... 

    A10 Networks

    San Jose, CA
    5 days ago
  • $184.7k - $324.8k

     ...AIML - ML Prototyping Engineer, Machine Learning ResearchThe Machine Learning Research Prototyping team sits at the intersection of cutting-edge...  ...with LLMs: prompting, fine-tuning, RLHF, inference optimizationExperience reproducing results from ML papersFamiliarity... 
    Relocation

    Apple

    Cupertino, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead Machine Learning Inference Engineer, Advertising. Be the first to apply!