Senior ML Scientist - Inference & Hardware Acceleration
Netskope
Netskope is seeking a Senior Staff Machine Learning Scientist in Santa Clara to own the inference and optimization layer for AI in agentic workflows. You will fine-tune models, push latency and throughput on real hardware, and build a runtime that executes bounded AI tasks with real customer data signals. You will work on quantization, KV-cache optimization, and hardware acceleration, partnering with systems and backend engineers to ship end-to-end capabilities in production environments. #J-18808-Ljbffr Netskope
$182.5k - $260.5k
...era. We secure and accelerate cloud, data, and AI... ...Positions are available at Senior Staff and above.... ...Machine Learning Scientist, you own the inference and optimization layer... ...throughput on real hardware, and build the runtime... ...+ years hands-on in ML/AI (model...Senior$148k - $235.75k
We are looking for a Senior Technical Product Marketing Manager.... ...business and pivotal in our inference marketing. You will be focused... ...years of experience in LLM, AI/ML development in an engineering... ...data center architectures, accelerated computing, distributed inference...SeniorFull time$195.2k - $262.2k
...building large in-house AI/ML infrastructure. Built... ...GPU orchestration to inference optimization, we own... ...deep expertise across hardware, software and AI R&D.... ...Nebius Token Factory needs scientists who can turn frontier inference... ...capabilities. A Senior Applied Scientist owns...SeniorTemporary workImmediate startRemote work$192.2k - $260k
We are looking for a Senior Applied Scientist to help drive the research and development... ..., and work closely with inference engineers to ensure your... ...architectures informed by hardware constraints and inference... ...to open-source speech/audio ML systems or widely used...SeniorLocal areaFlexible hours- d-Matrix in Santa Clara, CA is seeking a Sr. Staff ML Researcher to advance LLM algorithmic optimization on our DNN accelerators. You will design and implement efficient inference algorithms, collaborating with mathematicians, ML researchers and engineers on high-impact...Senior3 days per week
- Voltai in California (Palo Alto) seeks a senior formal verification researcher to develop new methods for formal proofs of... ...correctness. You will collaborate with RTL, verification, and ML teams to scale AI-hardware verification, prototype ideas on real RTL, and turn...Senior
$184k - $287.5k
...develop cutting-edge technologies in AI inference systems. You will design and optimize kernel technologies to accelerate workloads for NVIDIA's hardware architecture. The ideal candidate... ...and has over 6 years of experience in ML/DL systems. A competitive salary is offered...Senior- Nebius Token Factory seeks a PhD-level researcher to lead focused ML research projects from hypothesis to production handoff. You... ...to make prototypes production-ready and drive efficient LLM/VLM inference with measurable impact. The role emphasizes publishing results,...Senior
- Google Cloud in Sunnyvale, CA seeks a Senior TPU Co-Design Engineer to shape AI/ML hardware acceleration and drive TPU architecture across training and inference. You will bridge research, software, and hardware teams to maximize scalability, quality, and efficiency of...Senior
- ...industry-leading training and inference speeds; over 10 times faster than... ...running on Cerebras hardware.Test and verify deployment infrastructure... ...tools.Experience with ML inference infrastructure, model serving systems, or GPU-accelerated workloadsLocation: Toronto / SunnyvaleTeam...SeniorWork at office
$332k
...graphics, PC gaming, and accelerated computing for more... ....We are looking for a Senior leader to orchestrate... ...NVIDIA platform, including hardware and software. "Embed... ...a leadership Inference go-to-market strategy!... ...pre-sales with an AI/ML focus.Passion for transforming...SeniorFull timeWorldwide- ...build great products that accelerate next-generation computing... ...ROLE:AMD is seeking a Senior Product Manager to drive... ...focus on large-scale model inference on AMD Instinct™ and Radeon™ hardware. This is a key, central role... ...in the open-source AI/ML community: monitor GitHub...SeniorRemote work
$184k - $287.5k
...engineers to join us and build AI inference systems that serve large-... ...to push the frontier of accelerated computing for AI.What you’ll... ...models with the latest NVIDIA GPU hardware features; profile and optimize... ...frontier for the field of ML Systems; survey recent publications...SeniorFull time$193.3k - $261.5k
...machine learning training and inference clusters. Our organization... ...life — drivers that expose the hardware to the OS, runtime libraries... ...infrastructure to enable SoC validation, accelerate system software development,... ...exploration. As part of the ML accelerator systems modeling...SeniorLocal areaFlexible hours$192k - $304.75k
We are now looking for a Senior Research Scientist for Human‑AI Perception & Interaction! NVIDIA has... ...transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’... ...multi-node, multi-GPU training and inference workflows.Familiarity with and...SeniorFull time$183k - $247.6k
...designs silicon and software that accelerates innovation. Customers choose us... ...the world.We are seeking a Hardware Design Engineer with role in the... ...validation of AWS next generation ML Chips, Cards and server integration. As a senior member of our hardware team, you...SeniorLocal areaFlexible hours$192k - $304.75k
...problems with our unique approach to accelerated computing. We're looking for a passionate scientist at the intersection of quantum... ...for fault-tolerant quantum hardware.At NVIDIA, we want to help... ...performance prediction and parameter inference without full experimental...SeniorFull timeRemote work$148.75k - $361k
...At the core of this is our Machine Learning, Experimentation and Inference Platform that powers the entire landscape which we continuously... ...QualificationsExperience in the Advertising domainContributions to open-source ML projects #LI-DH2What's Roku's approach to hybrid working?Roku...SeniorWork at officeLocal areaRemote workMonday to ThursdayFlexible hours$192k - $304.75k
...our unique approach to accelerated computing. We're looking for a passionate scientist at the intersection of... ...fault-tolerant quantum hardware.At NVIDIA, we want to help... ...and parameter inference without full experimental... ...quantum systems and AI/ML research.Hands-on expertise...SeniorFull timeRemote work$192.2k - $260k
...talented, and inventive Applied Scientist with a strong machine... ...scale computing resources to accelerate advances in machine learning... ...Join our dynamic team of AI/ML practitioners and applied scientists... ....Key job responsibilitiesThe Senior Applied Scientist will lead the...SeniorLocal areaWorldwideFlexible hours$173k
...build for travelers everywhere.Senior Machine Learning ScientistThe Senior Machine Learning Scientist is responsible for building and... ...trip management. Owns end-to-end ML and GenAI projects—from problem... ...technical challenges, from inference problems on long-tail traveler...SeniorFull timeWorldwide- ...to build great products that accelerate next-generation computing experiences... ...Language Models (LLMs) and ML workloads on emerging... ...infrastructure, and model-to-hardware optimization, with a strong focus... ...movement optimization for ML inference workloads• Define and...
- ...to build great products that accelerate next-generation computing experiences... ...seeking a Principal GenAI Inference Optimization Engineer to join... ...working across the software-hardware stack.THE PERSONThe ideal... ...architectures.- Experience with ML frameworks (PyTorch, JAX, or...
$193.3k - $261.5k
...development kit used to accelerate deep learning and... ...and Trainium ML accelerators. This... ...enabling unparalleled ML inference and training... ...PyTorch till the hardware-software boundary,... ...functional team of applied scientists, system engineers,... ...mentorship. Our senior members enjoy one-...SeniorWork experience placementInternshipLocal areaFlexible hours- ...industry-leading training and inference speeds and allows users to run large-scale ML applications with less hardware management. Cerebras’... ...applications. About The Role Senior Director of Technical Product... ...a focus on infrastructure (accelerators, clusters, interconnect,...Senior
- ...-leading training and inference speeds; over 10 times... ...This is one of the most senior IC roles on the team,... ...partner closely with ML, Product and Infrastructure... ...systems, or GPU-accelerated workloads is a plus. Why... ...software make their own hardware. At Cerebras, we have...
$192k - $304.75k
...problems with our unique approach to accelerated computing. We're looking for a passionate AI research scientist with deep quantum computing... ...error-correcting codes and hardware platforms, while collaborating... ...doing:Design and architect AI/ML models—including deep neural...SeniorFull time$192.2k - $260k
...AWS) is assembling an elite team of world-class scientists and engineers to pioneer the next generation of... ...and present your pioneering work at premier ML and NLP conferences (NeurIPS, ICML, ICLR , ACL, EMNLP)- Accelerate innovation by working directly with customers to...SeniorWork at officeLocal areaWorldwideFlexible hours$119.8k - $234.7k
...representations, and heterogeneous event streams to infer user intent and advertiser value, even... .... The team owns end-to-end ML systems, including large-scale data and label... ...marketplace dynamics. Engineers and scientists on the team work at the intersection of deep...SeniorOngoing contractWork at officeLocal areaShift work$203.5k - $275.3k
...programs behind AWS's AI accelerators — Trainium and Inferentia... ...programs that span silicon, hardware, software platforms, and ML frameworks, and you drive... ..., and you are the leader senior executives rely on for a... ...for training and inference. Our programs sit at the...Local areaWorldwideFlexible hoursDay shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior ML Scientist - Inference & Hardware Acceleration. Be the first to apply!
- scientist ii Santa Clara, CA
- scientist 1 Santa Clara, CA
- image scientist Santa Clara, CA
- qc scientist Santa Clara, CA
- research scientist Santa Clara, CA
- analytical scientist Santa Clara, CA
- research scientist - biology Santa Clara, CA
- quality control scientist Santa Clara, CA
- applied scientist Santa Clara, CA
- manufacturing scientist Santa Clara, CA

