Senior ML Scientist - Inference & Hardware Acceleration
Netskope
Netskope is seeking a Senior Staff Machine Learning Scientist in Santa Clara to own the inference and optimization layer for AI in agentic workflows. You will fine-tune models, push latency and throughput on real hardware, and build a runtime that executes bounded AI tasks with real customer data signals. You will work on quantization, KV-cache optimization, and hardware acceleration, partnering with systems and backend engineers to ship end-to-end capabilities in production environments. #J-18808-Ljbffr Netskope
$124.5k - $272k
...era. We secure and accelerate cloud, data, and AI... ...Positions are available at Senior Staff and above.... ...Machine Learning Scientist, you own the inference and optimization layer... ...throughput on real hardware, and build the runtime... ...+ years hands-on in ML/AI (model...Senior$148k - $235.75k
We are looking for a Senior Technical Product Marketing Manager.... ...business and pivotal in our inference marketing. You will be focused... ...years of experience in LLM, AI/ML development in an engineering... ...data center architectures, accelerated computing, distributed inference...SeniorFull time$195.2k - $262.2k
...building large in-house AI/ML infrastructure. Built... ...GPU orchestration to inference optimization, we own... ...deep expertise across hardware, software and AI R&D.... ...Nebius Token Factory needs scientists who can turn frontier inference... ...capabilities. A Senior Applied Scientist owns...SeniorTemporary workImmediate startRemote work- d-Matrix in Santa Clara, CA is seeking a Sr. Staff ML Researcher to advance LLM algorithmic optimization on our DNN accelerators. You will design and implement efficient inference algorithms, collaborating with mathematicians, ML researchers and engineers on high-impact...Senior3 days per week
$192.2k - $260k
We are looking for a Senior Applied Scientist to help drive the research and development... ..., and work closely with inference engineers to ensure your... ...architectures informed by hardware constraints and inference... ...to open-source speech/audio ML systems or widely used...SeniorLocal areaFlexible hours- Voltai in California (Palo Alto) seeks a senior formal verification researcher to develop new methods for formal proofs of... ...correctness. You will collaborate with RTL, verification, and ML teams to scale AI-hardware verification, prototype ideas on real RTL, and turn...Senior
- Nebius Token Factory seeks a PhD-level researcher to lead focused ML research projects from hypothesis to production handoff. You... ...to make prototypes production-ready and drive efficient LLM/VLM inference with measurable impact. The role emphasizes publishing results,...Senior
- ...industry-leading training and inference speeds; over 10 times faster than... ...running on Cerebras hardware.Test and verify deployment infrastructure... ...tools.Experience with ML inference infrastructure, model serving systems, or GPU-accelerated workloadsLocation: Toronto / SunnyvaleTeam...SeniorWork at office
$332k
...graphics, PC gaming, and accelerated computing for more... ....We are looking for a Senior leader to orchestrate... ...NVIDIA platform, including hardware and software. "Embed... ...a leadership Inference go-to-market strategy!... ...pre-sales with an AI/ML focus.Passion for transforming...SeniorFull timeWorldwide- ...build great products that accelerate next-generation computing... ...ROLE:AMD is seeking a Senior Product Manager to drive... ...focus on large-scale model inference on AMD Instinct™ and Radeon™ hardware. This is a key, central role... ...in the open-source AI/ML community: monitor GitHub...SeniorRemote work
- ...PhD researcher to drive groundbreaking work at the intersection of AI HW/SW Co-Design, AI Hardware Accelerator Architecture, IC Design Methodology, and VLSI Design. The role spans ML fundamentals, quantization, digital VLSI circuits, and generative AI for hardware design...
$184k - $287.5k
...engineers to join us and build AI inference systems that serve large-... ...to push the frontier of accelerated computing for AI.What you’ll... ...models with the latest NVIDIA GPU hardware features; profile and optimize... ...frontier for the field of ML Systems; survey recent publications...SeniorFull time$193.3k - $261.5k
...machine learning training and inference clusters. Our organization... ...life — drivers that expose the hardware to the OS, runtime libraries... ...infrastructure to enable SoC validation, accelerate system software development,... ...exploration. As part of the ML accelerator systems modeling...SeniorLocal areaFlexible hours$183k - $247.6k
...designs silicon and software that accelerates innovation. Customers choose us... ...the world.We are seeking a Hardware Design Engineer with role in the... ...validation of AWS next generation ML Chips, Cards and server integration. As a senior member of our hardware team, you...SeniorLocal areaFlexible hours$192k - $304.75k
We are now looking for a Senior Research Scientist for Human‑AI Perception & Interaction! NVIDIA has... ...transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’... ...multi-node, multi-GPU training and inference workflows.Familiarity with and...SeniorFull time$192k - $304.75k
...problems with our unique approach to accelerated computing. We're looking for a passionate scientist at the intersection of quantum... ...for fault-tolerant quantum hardware.At NVIDIA, we want to help... ...performance prediction and parameter inference without full experimental...SeniorFull timeRemote work$148.75k - $361k
...At the core of this is our Machine Learning, Experimentation and Inference Platform that powers the entire landscape which we continuously... ...QualificationsExperience in the Advertising domainContributions to open-source ML projects #LI-DH2What's Roku's approach to hybrid working?Roku...SeniorWork at officeLocal areaRemote workMonday to ThursdayFlexible hours$192k - $304.75k
...our unique approach to accelerated computing. We're looking for a passionate scientist at the intersection of... ...fault-tolerant quantum hardware.At NVIDIA, we want to help... ...and parameter inference without full experimental... ...quantum systems and AI/ML research.Hands-on expertise...SeniorFull timeRemote work$192.2k - $260k
...talented, and inventive Applied Scientist with a strong machine... ...scale computing resources to accelerate advances in machine learning... ...Join our dynamic team of AI/ML practitioners and applied scientists... ....Key job responsibilitiesThe Senior Applied Scientist will lead the...SeniorLocal areaWorldwideFlexible hours$193.3k - $261.5k
...development kit used to accelerate deep learning and... ...and Trainium ML accelerators. This... ...enabling unparalleled ML inference and training... ...PyTorch till the hardware-software boundary,... ...functional team of applied scientists, system engineers,... ...mentorship. Our senior members enjoy one-...SeniorWork experience placementInternshipLocal areaFlexible hours- ...to build great products that accelerate next-generation computing experiences... ...seeking a Principal GenAI Inference Optimization Engineer to join... ...working across the software-hardware stack.THE PERSONThe ideal... ...architectures.- Experience with ML frameworks (PyTorch, JAX, or...
- ...to build great products that accelerate next-generation computing experiences... ...Language Models (LLMs) and ML workloads on emerging... ...infrastructure, and model-to-hardware optimization, with a strong focus... ...movement optimization for ML inference workloads• Define and...
- ...-leading training and inference speeds; over 10 times... ...This is one of the most senior IC roles on the team,... ...partner closely with ML, Product and Infrastructure... ...systems, or GPU-accelerated workloads is a plus. Why... ...software make their own hardware. At Cerebras, we have...
$192k - $304.75k
...problems with our unique approach to accelerated computing. We're looking for a passionate AI research scientist with deep quantum computing... ...error-correcting codes and hardware platforms, while collaborating... ...doing:Design and architect AI/ML models—including deep neural...SeniorFull time$192k - $304.75k
...NVIDIA is searching for an outstanding Senior Researcher working on efficient deep learning... ...architecture design, adaptive/dynamic inference, resource-efficient training and fine-... ...characteristic protected by law. NVIDIA pioneered accelerated computing. Today, our AI infrastructure...SeniorFull time$183.83k - $275.98k
...combining cutting-edge AI with automotive-grade hardware. Nuro licenses its core technology, the... ...on bringing advancements in the field of ML and large-scale learning to the AV domain... ...to improve model optimization and inference speeds. You will use your applied research...Senior$192.2k - $260k
...AWS) is assembling an elite team of world-class scientists and engineers to pioneer the next generation of... ...and present your pioneering work at premier ML and NLP conferences (NeurIPS, ICML, ICLR , ACL, EMNLP)- Accelerate innovation by working directly with customers to...SeniorWork at officeLocal areaWorldwideFlexible hours$147k - $210k
...horizontal agents to accelerate researcher velocity and... ...Publish breakthroughs at ML conferences, establish... ...developing test-time/inference-time compute scaling... ...of work. As a Research Scientist, you'll setup large-... ...language processing, hardware and software performance...$119.8k - $234.7k
...representations, and heterogeneous event streams to infer user intent and advertiser value, even... .... The team owns end-to-end ML systems, including large-scale data and label... ...marketplace dynamics. Engineers and scientists on the team work at the intersection of deep...SeniorOngoing contractWork at officeLocal areaShift work- ...Clara, CA, seeks a Principal System Software Engineer for AI Inference Execution. You will join the software team to productize... ..., developing deployment software and collaborating with ML, compiler, and hardware experts. Required: strong background in system software...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior ML Scientist - Inference & Hardware Acceleration. Be the first to apply!
- materials scientist Santa Clara, CA
- health scientist Santa Clara, CA
- quality control scientist Santa Clara, CA
- deep learning scientist Santa Clara, CA
- application scientist Santa Clara, CA
- decision scientist Santa Clara, CA
- scientist 1 Santa Clara, CA
- lab scientist Santa Clara, CA
- scientist Santa Clara, CA
- regulatory scientist Santa Clara, CA

