Lead Machine Learning Inference Engineer, Advertising
$246.5kRoku, Building C
Teamwork makes the stream work.Roku is changing how the world watches TVRoku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.About the team The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across large-scale distributed systems with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning and Inference Platform that powers the entire landscape. About the role In this role, you will architect, design, and lead the development of a SOTA Inference platform that can handle Advertising-level low latencies, scale, throughput, and availability with optimizations that span across hardware, software, and models. We’re looking for a strong technical leader with deep experience in ML serving, high-performance computing, and industry standard frameworks - someone excited to mentor engineers, innovate at scale, and shape the future of machine learning at Roku.For California Only - The estimated annual salary for this position is between $246,500 - $486,100 annually. Compensation packages are based on factors unique to each candidate, including but not limited to skill set, certifications, and specific geographical location. This role is eligible for health insurance, equity awards, life insurance, disability benefits, parental leave, wellness benefits, and paid time off. What you’ll be doingLead the design and development of a SOTA Inference platform Oversee the development of monitoring, observability, and other tooling to ensure system and model performance, reliability, and scalability of online inference servicesIdentify and resolve system inefficiencies, performance bottlenecks, and reliability issues, ensuring optimized end-to-end performance Stay at the forefront of advancements in inference frameworks, ML hardware acceleration, and distributed systems, and incorporate innovations where and when they are impactfulWe’re excited if you haveM.S. or above in CS, ECE, or a related field 10+ years of experience in developing and deploying large-scale, distributed systems, with at least 5 years in a leadership or technical lead role Strong programming skills in high-performance languagesDeep understanding of inference frameworks and ML system deploymentProven experience optimizing performance for large-scale machine learning systems, including a deep knowledge of SOTA model optimizations, hardware-software co-design, GPU acceleration, and HPC techniquesExcellent communication and collaboration skillsExperience leading teams working on high-throughput, low-latency ML serving systemsExperience collaborating with and leading global, cross-functional teamsContributions to open-source ML or systems projects#LI-DH2What's Roku's approach to hybrid working?Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance.What are some of the benefits?Roku is committed to offering a diverse range of benefits as part of our compensation package to support our employees and their families. Our comprehensive benefits include global access to mental health and financial wellness support and resources. Local benefits include statutory and voluntary benefits which may include healthcare (medical, dental, and vision), life, accident, disability, commuter, and retirement options (401(k)/pension). Employees are supported in taking time off, in accordance with local leave policies and other personal needs to support their evolving work and life needs. It's important to note that not every benefit is available in all locations or for every role. For details specific to your location, please consult with your recruiter.AccommodationsRoku welcomes applicants of all backgrounds and provides reasonable accommodations and adjustments in accordance with applicable law. If you require reasonable accommodation at any point in the hiring process, please direct your inquiries to View email address on us.fitly.work should I know about Roku's culture?Roku is a great place for people who want to work in a fast-paced environment where everyone is focused on the company's success rather than their own. We try to surround ourselves with people who are great at their jobs, who are easy to work with, and who keep their egos in check. We appreciate a sense of humor. We believe a fewer number of very talented folks can do more for less cost than a larger number of less talented teams. We're independent thinkers with big ideas who act boldly, move fast and accomplish extraordinary things through collaboration and trust. In short, at Roku you'll be part of a company that's changing how the world watches TV. We have a unique culture that we are proud of. We think of ourselves primarily as problem-solvers, which itself is a two-part idea. We come up with the solution, but the solution isn't real until it is built and delivered to the customer. That penchant for action gives us a pragmatic approach to innovation, one that has served us well since 2002. To learn more about Roku, our global footprint, and how we've grown, visit .By providing your information, you acknowledge that you want Roku to contact you about job roles, that you have read Roku's Applicant Privacy Notice, and understand that Roku will use your information as described in that notice. If you do not wish to receive any communications from Roku regarding this role or similar roles in the future, you may unsubscribe at any time by emailing View email address on us.fitly.work.
$160k - $225k
Machine Learning Engineer Location: Mountain View, CA Company Stage of Funding... ...for the $1 trillion digital advertising industry. Their platform... ...support model training and inference. Build customer-facing AI... .... Preferred Experience at leading technology companies or high...SuggestedH1bWork at officeVisa sponsorship- ...creating a compelling path for advertisers to reach audiences that are... ....Our TeamThe Ads Platform Engineering teams build advertising... ...building and operating production machine learning systems at scale.Experience... ...:Experience with causal inference, experimentation, or...SuggestedHourly payFull timeImmediate startFlexible hours
- ...scalable ML platform services including feature stores, real-time inference services, and vector databases serving millions of... ...KPIs to optimize recommendation performance. Collaborate with engineering and cross-functional teams to translate business requirements...SuggestedFull timeWork at officeLocal areaRemote workMonday to FridayMonday to ThursdayFlexible hours
$193.3k - $261.5k
The Product: AWS Machine Learning accelerators are at the forefront of... ...delivers best-in-class ML inference performance at the lowest cost... ...including silicon engineering, hardware design and verification... ...language experience- 5+ years of leading design or architecture (...SuggestedInternshipLocal areaWork from homeRelocationFlexible hours$224k - $356.5k
NVIDIA is looking for a Machine Learning Engineer to join the GPU accelerated Apache Spark team.Apache... ...ETL, SQL, and ML/DL model training and inference pipelines, spanning many domains and... ...solutions.3+ experience as technical lead in ML model development.Proven hands-on...SuggestedFull time$148.75k - $361k
...audiences, and provide advertisers unique capabilities... ...low latency. We use Machine Learning, Reinforcement Learning... ...Experimentation, and Inference Platform that powers... ...experienced Senior Software Engineer, MLOps/DevOps, to... ...What you’ll be doing Lead the design and...Work at officeLocal areaRemote workMonday to ThursdayFlexible hours- ...types. Accelerate cloud model inference for closed-loop simulation, reinforcement learning, and enterprise LLM/VLM... ...in computer science, computer engineering, or electrical engineering, or... ...work. Opportunity to work with leading talent on cutting-edge technology...Full time
$174.72k - $295.68k
XPENG is a leading smart technology company at the forefront of... ...cutting-edge R&D in AI, machine learning, and smart connectivity.Our... ...tuning, PTQ, QAT, on-vehicle inference and related fields.Key ResponsibilitiesDevelop... ...programming and software engineering skills.Ability to work...Full time- ...Responsibilities Take machine learning models from prototype to production by building reliable... ..., scalable training, serving, and inference pipelines. Design and maintain data... .... Work with Data Scientists, Data Engineers, and Software Engineers to turn business...Full timeWork at officeFlexible hours
$170k - $240k
Machine Learning Engineer Santa Clara, CA About the role We are seeking a high-impact, technically deep Machine Learning Engineer to develop,... ...trained models into C++-based autonomy systems and optimize inference for production vehicle hardware. Model Optimization:...Work at office$184k - $287.5k
Intelligent machines powered by Artificial Intelligence computers that can learn, reason, and interact with people are no longer... ...seeking the best Machine Learning Engineers with a background in computer... ...optimization for real-time inference on embedded or automotive platforms...WorldwideNight shift$149.37k - $275k
Title and Location : Sr Machine Learning Engineer in Santa Clara, CA. Job Responsibilities Implement deep-learning models and frameworks for... ...large-scale image datasets (1 yr, 6 mos). Evaluate model inference performance and runtime behavior across heterogeneous hardware...Work at officeLocal areaRemote work$200k - $280k
Machine Learning Engineer Responsibilities: Develop, optimize, and deploy lightweight machine learning... ...for embedded AI applications. Improve inference efficiency and model compression techniques... ...managing a team, serving in a Team Lead role, or demonstrating strong...Local area$184.7k - $324.8k
Senior Machine Learning Engineer Apple Services Engineering (ASE) builds experiences that touch hundreds... ...language understanding, behavioral inference, discovery, and growth optimization—while... ...across Apple's services ecosystem. Lead research in areas such as large-scale...Relocation$160k - $200k
...As a Senior ML Infrastructure Engineer at Plus, you will design... ...performance for both training and inference phases. You will build robust... ...with state-of-the-art deep learning frameworks like PyTorch or TensorFlow... ...of what's possible in machine learning infrastructure and contribute...$214.67k - $322k
...individual's freedom.OKX is a leading crypto exchange, and the... ...About The OpportunityBuilding machine learning systems for risk at a... ...different from conventional ML engineering. The data spans on-chain... ...offline training and online inference.Ensure models and decision...$206.4k - $379.1k
...Express, Stock, and Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical lead for our GenAI Services area. This is not a model-... ...and served at enterprise scale. You will set the inference architecture and technical standards that a...Full timeTemporary workLocal areaWorldwide- ...creating a compelling path for advertisers to reach audiences that are... ....Our TeamThe Ads Platform Engineering teams build advertising... ...experiences by utilizing advanced machine learning models for identity... ...end ML model deployment and inference infra for low-latency real-...Hourly payFull timeImmediate startFlexible hours
$220k - $300k
Senior Principal Machine Learning Engineer San Jose, California, United States The era of pervasive AI... ...architecture, training and fine-tuning, inference optimization, evaluation, and data... ...'s RDU and broader hardware ecosystem Lead hardware-software co-design efforts in...Full timeTemporary workLocal areaFlexible hours$2,000 per month
Machine Learning Research Engineer Cupertino, CA Etched is building AI chips that are hard-coded for individual model architectures. Our first product... ...and/or a mix of these metrics. Implement model-specific inference-time acceleration techniques such as speculative...Work at officeRelocation package$152k - $241.5k
We are now looking for a Senior Machine Learning Applications and Compiler Engineer!NVIDIA is seeking engineers to develop algorithms and optimizations for our LPX inference and compiler stack. You will work at the intersection of large-scale systems, compilers, and deep...Full time- ...the future of autonomy, Plus is looking for talented individuals to join its fast-growing teams. We are seeking a Senior Machine Learning Engineer with expertise in deep learning and data analysis. In this role, you will apply data-driven techniques to develop high-...
- ...Responsibilities Develop and improve machine learning and large language model systems for... ...generation, debugging, and related software engineering tasks. Build AI-powered developer... .... Experience building or leading AI systems for software engineering, such...Full time
$184k - $287.5k
...scale. We seek a Senior ML Engineer to compose and deliver next-... ...support the wider ecosystem.Lead the open-sourcing of solutions... ...experience in applied machine learning or AI research.Deep expertise... ...quantization, and real-time inference optimization for production...Full time$150k - $230k
...the Role We are looking for a hands-on Machine Learning Engineer to drive the post-training of our... ...read about it. Responsibilities Lead post-training of our LLMs across the full... .../Accelerate, DeepSpeed or FSDP, and inference engines like vLLM. Solid understanding...Full timeLocal areaWork from home- ...Responsibilities Architect, train, and optimize multitask deep learning models across multimodal sensor streams. Build automated... ...Profile and optimize models for efficient onboard accelerator inference. Requirements ~2–5+ years of experience training and...Full time
$165.2k - $223.6k
...development kit used to accelerate deep learning and GenAI workloads on Amazon’s custom machine learning accelerators,... ...hardware-software boundary, our engineers craft high-performance kernels for... ...PyTorch, enabling unparalleled ML inference and training performance.As part...InternshipLocal areaWork from homeFlexible hours$195k - $230k
..., visit About the Role We are looking for a Senior Machine Learning Engineer to help evolve our large-scale recommendation systems and... ...metrics. Own systems from offline training → online inference → A/B experimentation → metric analysis . Identify and...Full timeLocal areaWork from home$193.3k - $261.5k
...revolution? At AWS our vision is to make deep learning pervasive for everyday developers and to... ...role is for a senior software engineer in the Compiler team for AWS Neuron. As... ...professional.Basic qualifications- 5+ years of leading design or architecture (design patterns,...Local areaFlexible hours$184k - $287.5k
...Intelligent machines powered by Artificial Intelligence computers that can learn, reason and interact with people are no longer... ...extraordinary Senior Perception Engineer to develop and productize NVIDIA... ...by technical publications in leading conferences/journals.Expertise...Odd jobFull timeWork experience placementRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead Machine Learning Inference Engineer, Advertising. Be the first to apply!
- lead engineer San Jose, CA
- lead operating engineer San Jose, CA
- machine learning engineer San Jose, CA
- senior ml engineer San Jose, CA
- machine learning ai engineer San Jose, CA
- machine learning software engineer San Jose, CA
- ai ml engineer San Jose, CA
- computer vision machine learning engineer San Jose, CA
- machine learning part time San Jose, CA
- data engineer machine learning San Jose, CA


