Lead Machine Learning Inference Engineer, Advertising
$246.5kRoku, Building C
Teamwork makes the stream work.Roku is changing how the world watches TVRoku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.About the team The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across large-scale distributed systems with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning and Inference Platform that powers the entire landscape. About the role In this role, you will architect, design, and lead the development of a SOTA Inference platform that can handle Advertising-level low latencies, scale, throughput, and availability with optimizations that span across hardware, software, and models. We’re looking for a strong technical leader with deep experience in ML serving, high-performance computing, and industry standard frameworks - someone excited to mentor engineers, innovate at scale, and shape the future of machine learning at Roku.For California Only - The estimated annual salary for this position is between $246,500 - $486,100 annually. Compensation packages are based on factors unique to each candidate, including but not limited to skill set, certifications, and specific geographical location. This role is eligible for health insurance, equity awards, life insurance, disability benefits, parental leave, wellness benefits, and paid time off. What you’ll be doingLead the design and development of a SOTA Inference platform Oversee the development of monitoring, observability, and other tooling to ensure system and model performance, reliability, and scalability of online inference servicesIdentify and resolve system inefficiencies, performance bottlenecks, and reliability issues, ensuring optimized end-to-end performance Stay at the forefront of advancements in inference frameworks, ML hardware acceleration, and distributed systems, and incorporate innovations where and when they are impactfulWe’re excited if you haveM.S. or above in CS, ECE, or a related field 10+ years of experience in developing and deploying large-scale, distributed systems, with at least 5 years in a leadership or technical lead role Strong programming skills in high-performance languagesDeep understanding of inference frameworks and ML system deploymentProven experience optimizing performance for large-scale machine learning systems, including a deep knowledge of SOTA model optimizations, hardware-software co-design, GPU acceleration, and HPC techniquesExcellent communication and collaboration skillsExperience leading teams working on high-throughput, low-latency ML serving systemsExperience collaborating with and leading global, cross-functional teamsContributions to open-source ML or systems projects#LI-DH2What's Roku's approach to hybrid working?Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance.What are some of the benefits?Roku is committed to offering a diverse range of benefits as part of our compensation package to support our employees and their families. Our comprehensive benefits include global access to mental health and financial wellness support and resources. Local benefits include statutory and voluntary benefits which may include healthcare (medical, dental, and vision), life, accident, disability, commuter, and retirement options (401(k)/pension). Employees are supported in taking time off, in accordance with local leave policies and other personal needs to support their evolving work and life needs. It's important to note that not every benefit is available in all locations or for every role. For details specific to your location, please consult with your recruiter.AccommodationsRoku welcomes applicants of all backgrounds and provides reasonable accommodations and adjustments in accordance with applicable law. If you require reasonable accommodation at any point in the hiring process, please direct your inquiries to View email address on click.appcast.io should I know about Roku's culture?Roku is a great place for people who want to work in a fast-paced environment where everyone is focused on the company's success rather than their own. We try to surround ourselves with people who are great at their jobs, who are easy to work with, and who keep their egos in check. We appreciate a sense of humor. We believe a fewer number of very talented folks can do more for less cost than a larger number of less talented teams. We're independent thinkers with big ideas who act boldly, move fast and accomplish extraordinary things through collaboration and trust. In short, at Roku you'll be part of a company that's changing how the world watches TV. We have a unique culture that we are proud of. We think of ourselves primarily as problem-solvers, which itself is a two-part idea. We come up with the solution, but the solution isn't real until it is built and delivered to the customer. That penchant for action gives us a pragmatic approach to innovation, one that has served us well since 2002. To learn more about Roku, our global footprint, and how we've grown, visit .By providing your information, you acknowledge that you want Roku to contact you about job roles, that you have read Roku's Applicant Privacy Notice, and understand that Roku will use your information as described in that notice. If you do not wish to receive any communications from Roku regarding this role or similar roles in the future, you may unsubscribe at any time by emailing View email address on click.appcast.io.
- ...creating a compelling path for advertisers to reach audiences that are... ....Our TeamThe Ads Platform Engineering teams build advertising... ...building and operating production machine learning systems at scale.Experience... ...:Experience with causal inference, experimentation, or...SuggestedHourly payFull timeImmediate startFlexible hours
- ...creating a compelling path for advertisers to reach audiences that are... ....Our TeamThe Ads Platform Engineering teams build advertising systems... ...end ML model deployment and inference infra for low-latency real-... ...culture and environment. Learn more here.Inclusion is a Netflix...SuggestedHourly payFull timeImmediate startFlexible hours
- ...creating a compelling path for advertisers to reach audiences that are... ...Team The Ads Platform Engineering teams build advertising systems... ...end ML model deployment and inference infra for low-latency real-... ...culture and environment. Learn more here . Inclusion...SuggestedHourly payFull timeImmediate startFlexible hours
- ...of areas including e-commerce, advertising, and fulfillment. We use machine learning and Internet-scale data to elevate... ...modeling, and general causal inference. Search & Discovery ML : The... ...Instacart works alongside world-class engineers, data scientists, and product...SuggestedRemote jobPermanent employmentWork experience placementInternshipWork at officeWork from homeFlexible hours
$160k - $225k
...Machine Learning Engineer Location: Mountain View, CA Company Stage of Funding... ...the $1 trillion digital advertising industry. Their platform... ...support model training and inference. Build customer-facing... ...Preferred Experience at leading technology companies or...SuggestedH1bWork at officeVisa sponsorship$278.1k - $347.6k
...accelerated entirely within that runtime. As our Principal Engineer for On-Device AI Inference & Systems, you will be the foremost engineering authority... ...worth the engineering cost. Engineering Leadership Lead and mentor a team of engineers; set engineering best...Work at officeWorldwideRelocation package$184k - $287.5k
Intelligent machines powered by Artificial Intelligence computers that can learn, reason, and interact with people are no longer... ...seeking the best Machine Learning Engineers with a background in computer... ...optimization for real-time inference on embedded or automotive platforms...Full timeWorldwideNight shift$151.8k - $265.35k
...adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services... ...ownership of production ML or inference services at scale. Strong Python and... ...customer experiences. Adobe’s industry-leading offerings including Adobe Acrobat Studio...Full timeTemporary workLocal areaWorldwide$148.75k - $361k
...audiences, and provide advertisers unique capabilities... ...low latency. We use Machine Learning, Reinforcement Learning... ...Experimentation, and Inference Platform that powers... ...experienced Senior Software Engineer, MLOps/DevOps, to... ...What you’ll be doing Lead the design and...Work at officeLocal areaRemote workMonday to ThursdayFlexible hours$229.5k - $360k
...large audiences, and provide advertisers unique capabilities to... ...leveraging state-of-the-art machine learning. Our mission is to deliver... ...Our work blends innovation, engineering excellence, and a deep commitment... ...: feature store, real-time inference services, vector DBs, etc.,...Work at officeLocal areaRemote workMonday to ThursdayFlexible hours$193.3k - $261.5k
The Product: AWS Machine Learning accelerators are at the forefront of... ...delivers best-in-class ML inference performance at the lowest cost... ...including silicon engineering, hardware design and verification... ...language experience- 5+ years of leading design or architecture (...InternshipLocal areaWork from homeRelocationFlexible hours$224k - $356.5k
We are looking for outstanding Machine Learning Engineers to join our Physical AI teams. As the pioneers... ...metric designs.SOTA Data Engineering: Lead the generation of massive training... ...architecture to improve the performance during inference/training.Familiarity with simulation...Full time$174.72k - $295.68k
XPENG is a leading smart technology company at the forefront of... ...through cutting-edge R&D in AI, machine learning, and smart connectivity.We... ...full-time Machine Learning Engineer - AI Foundation, with deep knowledge... ...accelerating model training/inference. Our mission is to solve the...Full time$128k - $260.5k
...platform, its Zero Trust Engine, and the powerful... ...education and mentorship, and lead with transparency and... ...Careers at Netskope to learn more. Follow us on... ...intelligence (AI) and machine learning (ML) to protect... ...the bleeding edge of LLM inference optimization, utilizing...$160k - $225k
...on a mission to democratize advanced advertising technology. We believe that cutting-edge... ...be used to expand our product and engineering teams, bringing our vision of intelligent... ...re writing the manual. As an early Machine Learning Engineer at MAI, you won't just be writing...Full time$272k - $431.25k
...NVIDIA is looking for a Machine Learning (ML) Engineer to join the GPU accelerated Apache Spark team. Apache... ..., SQL, and ML/DL model training and inference pipelines, spanning many domains and... ...solutions.5+ experience as technical lead in ML model development.Proven hands-...Full time$160k - $200k
Santa Clara, CAData Engineering - Machine Learning and Data Engineer /Full-time /HybridPlusAI is a Physical AI company pioneering AI-based virtual... ...while ensuring optimal performance for both training and inference phases. You will build robust pipelines for managing...Full time$214.67k - $322k
...individual's freedom.OKX is a leading crypto exchange, and the... ...About The OpportunityBuilding machine learning systems for risk at a... ...different from conventional ML engineering. The data spans on-chain... ...offline training and online inference.Ensure models and decision...$215.28k - $364.32k
XPENG is a leading smart technology company at the forefront of... ...through cutting-edge R&D in AI, machine learning, and smart connectivity.... ...for a strong Machine Learning Engineer / Computer Vision Engineer... .../ TensorRT / quantization / inference acceleration.Work with deployment...Full time$232k - $310k
...collaboration, and high standards. Our engineers, product leaders, and go-to-... ...the boundaries of applied machine learning. We work with massive... ...languages.Responsibilities:Lead the team in: research,... ...strategies (QLORA, DPO) and inference optimization (vLLM, TensorRT...Work experience placementWork at officeRemote workFlexible hours3 days per week$152k - $190k
...cybersecurity.RoleWe are looking for a Staff Software Engineer to join our team. This is a hybrid, based in San... ...production features, utilizing LLMs, various machine learning models, data processing, fine-tuning, and inference optimizationWork with the world class cloud...Full timeWork at officeLocal areaWorldwide3 days per week- Lead the team in: research, design, development, and deployment of advanced AI agents... ...experiences. Knowledge and passion in machine learning algorithms, Gen AI, LLMs, and natural... ...fine-tuning strategies (QLORA, DPO) and inference optimization (vLLM, TensorRT-LLM). Research...Full timeWork experience placement
- ...We are now looking for a Deep Learning Software Engineer, TensorRT Performance! NVIDIA is seeking an... ...improving the performance of NVIDIA’s inference ecosystem! NVIDIA is rapidly growing... ...learning has provided the foundation for machines to learn, perceive, reason and solve...
$174.3k - $200k
...Machine Learning Engineer UnitX builds the world's leading physical AI systems to automate repetitive visual tasks in factories. UnitX is a fast-moving startup... ...) to ensure pixel-level precision and real-time inference. Build Automated Pipelines: Develop and...- ...We are seeking a highly skilled Machine Learning Engineer to design and build a low-latency query understanding... ..., with a focus on sub-second inference, CPU-based execution, and scalable... ...entities, applications, evidence). Lead data collection, synthesis, analysis,...Local area
$110k - $145k
...Overview We are looking for a talented and experienced Deep Learning Engineer specializing in Large Language Models (LLMs) to join our dynamic... ...and apply risk mitigation techniques, including safe inference strategies to ensure reliable AI behavior. Identify vulnerabilities...$164k - $313.3k
...OpportunityPhotoshop ART is seeking a Senior Machine Learning (ML) Systems & Efficiency Engineer to join our R&D team focused on... ...production-ready improvements in inference performance, latency, and cost... ...approaches when they directly lead to more efficient inference,...Full timeTemporary workLocal areaWorldwide$361.3k
...audiences, and provide advertisers unique capabilities... ...LLMs, reinforcement learning, multi-objective... ...looking for a Senior Machine Learning Tech Lead to drive ranking and... ...across the agentic engineering toolchain — coding harnesses... ..., and causal inference techniquesBuild multi...Part timeWork at officeLocal areaRemote workMonday to ThursdayFlexible hours$152k - $241.5k
We are now looking for a Senior Machine Learning Applications and Compiler Engineer!NVIDIA is seeking engineers to develop algorithms and optimizations for our LPX inference and compiler stack. You will work at the intersection of large-scale systems, compilers, and deep...Full time$120k
...Overview Netflix is one of the world's leading entertainment services with over 247 million... .... The Ads Platform team builds the advertising systems that power the delivery of ads... ...person will partner with Product, TPM, Engineering, Data Science, and other cross‑functional...Flexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead Machine Learning Inference Engineer, Advertising. Be the first to apply!
- lead operating engineer San Jose, CA
- lead engineer San Jose, CA
- senior ml engineer San Jose, CA
- machine learning engineer San Jose, CA
- ai ml engineer San Jose, CA
- machine learning software engineer San Jose, CA
- computer vision machine learning engineer San Jose, CA
- machine learning ai engineer San Jose, CA
- machine learning research scientist San Jose, CA
- data engineer machine learning San Jose, CA




