Senior Applied Scientist — Efficient LLM Inference & Optimization
Nebius
Nebius is seeking a Senior Applied Scientist to turn frontier inference bottlenecks into production-ready solutions. You will design rigorous experiments, write high-quality code in Python and PyTorch, and collaborate with ML engineers to ship research into production. You will own well-scoped research projects, publish credible work, mentor engineers, and define evaluation methods for latency, throughput, and cost per token across LLM/VLM inference workloads. #J-18808-Ljbffr Nebius
$195.2k - $262.2k
...GPU orchestration to inference optimization, we own the hard... ...storage, networking and applied AI. Listed on... ...Token Factory needs scientists who can turn frontier... ...inference capabilities. A Senior Applied Scientist... ...research programs in efficient LLM and VLM inference...SeniorTemporary workImmediate startRemote work$192k - $304.75k
...are now looking for an Applied Deep Learning Research Scientist, Efficiency!Join our ADLR - Efficiency... ...and algorithms to optimize neural networks for training... ...effect on neural network inference and training accuracy. This... ..., optimizers and LLM training.Experience with...SeniorFull time- d-Matrix in Santa Clara, CA is seeking a Sr. Staff ML Researcher to advance LLM algorithmic optimization on our DNN accelerators. You will design and implement efficient inference algorithms, collaborating with mathematicians, ML researchers and engineers on high-impact...Senior3 days per week
$139k - $229k
...for an ambitious data scientist to have an impact. We... ...exhibit technical acumen on inference and algorithms, and... ...centered on trust and optimized for culture,... ...a thought partner to senior leaders to prioritize/... ...Informatics, Engineering, Applied Mathematics, Economics...SeniorFull timeFor contractorsWork experience placementWork at officeFlexible hours- ...projects from hypothesis to production handoff. You will partner with MLEs to make prototypes production-ready and drive efficient LLM/VLM inference with measurable impact. The role emphasizes publishing results, sharing technical reports, and mentoring teams on rigorous...Senior
$192.2k - $260k
...practical experience to join the Modeling and Optimization (MOP) Routing Science team. Your main... ...data structures, particularly as it applies to vehicle routing and related problems... ..., and analysis that will improve the efficiency and cost effectiveness of global fulfillment...SeniorLocal areaFlexible hours$119.8k - $234.7k
...end-to-end systems for AI inference and agent workflows on Windows... ...We are looking for a Senior Applied Scientist to develop and ship machine... ...model quality and the efficiency of AI workloads on a diverse... ...solve challenges in model optimization, search and retrieval, inference...SeniorOngoing contractLocal area$192.2k - $260k
...Amazon Advertising, we apply Machine Learning at massive scale to optimize programmatic... ...looking for a talented Senior Applied Scientist to join our team of scientists... ...) serving at inference latencies under 10ms... ...stores) to deploy models efficiently at scaleMentor scientists...SeniorLocal areaWorldwideFlexible hoursShift work$184k - $230k
Senior Applied Scientist, Inference (MCM) Join to apply for the Senior Applied Scientist, Inference (MCM) role at Moloco . About Moloco Moloco builds... ...to our ML models. Validate and quantify the efficiency and performance gain from hypotheses with advanced statistical...SeniorFull time$192.2k - $260k
...an elite team of world-class scientists and engineers to pioneer the... ...tools. Join the Amazon Kiro LLM-Training team and help create... ...any scale with unprecedented efficiency.Broadly, AWS Utility Computing... ...Master's degree and 6+ years of applied research experience-...SeniorWork at officeLocal areaWorldwideFlexible hours- ...days per week. The role Senior Staff ML Researcher - LLM Algorithmic Optimization What You Will Do d-... ...design, and implement efficient algorithms that will be... ...optimize large language model inference on DNN accelerators we... ...who create and apply advanced algorithmic and...Senior3 days per week
$182.5k - $260.5k
...Positions are available at Senior Staff and above.... ...Machine Learning Scientist, you own the inference and optimization layer that makes AI... ...workflows fast, efficient, and production-grade... .../SGLang, TensorRT-LLM, ONNX Runtime,... ...applicable laws, the range applies to candidates in...Senior- ...: Sr. Staff, ML Researcher - LLM Algorithmic OptimizationWhat... ...invent, design, and implement efficient algorithms that will be used to optimize large language model inference on DNN accelerators we develop... ...ML engineers who create and apply advanced algorithmic and numerical...Senior3 days per week
$192k - $304.75k
...looking for a passionate scientist at the intersection of... .... As a Sr. Quantum Applied Research Scientist, you... ...modeling, and co-optimized calibration-decoding pipelines... ...and parameter inference without full experimental... ...tuning—including parameter-efficient methods (LoRA, QLoRA,...SeniorFull timeRemote work$192.2k - $260k
We are looking for a Senior Applied Scientist to help drive the research and... ...roadmap, and work closely with inference engineers to ensure your... ...- Advance the scaling and efficiency of conversational models,... ...SFT through RL alignment, optimized for real-time multimodal...SeniorLocal areaFlexible hours$192.2k - $260k
...discover new products they love, be the most efficient way for advertisers to meet their... ...into specific plans for research and applied scientists, as well as engineering and product... ...perform proof-of-concept, experiment, optimize, and deploy your models into production...SeniorLocal areaFlexible hours- ...Sr. Staff, ML Researcher - LLM Algorithmic Optimization What You Will Do d-Matrix... ...invent, design, and implement efficient algorithms that will be... ...large language model inference on DNN accelerators we develop... ...engineers who create and apply advanced algorithmic and numerical...Senior3 days per week
$143.63k - $299.38k
...information. You will apply your insights on the data... ...just follow the latest LLM trends; you understand... ...architectures and how to optimize them for massive scale.... ...models using parameter‑efficient techniques (LoRA,... ...distributed training and inference. Publications, patents...SeniorWork at officeFlexible hours- ...what’s possible with LLM inference on heterogeneous hardware... ...patterns to deep optimization of inference kernels,... ...computational fabric. We are an applied research and... ...AI models run efficiently and cost-effectively... ...real hardware.• Small, senior team with high autonomy...Senior
$228.7k - $309.4k
As a Principal Applied Scientist for Full-Funnel Campaign optimization, you will invent the models that jointly allocate budget across sponsored ad products to maximize advertiser outcomes and long-term customer value.This is a rare charter to build foundational optimization...Local areaFlexible hoursDay shift$228.7k - $309.4k
...from ad creation and optimization to performance analysis... ...You will be the Gen AI applied science leader that... ...expertise in the area of ML, LLM and GenAI models. You... ...our team of applied scientists and engineers.Key job... ...of autonomy and efficiency. You'll be responsible...Local areaFlexible hours$171.6k - $222.2k
...lifecycle from ad creation and optimization to performance analysis... ...and motivated Applied Scientist with machine learning engineering... ..., from training to inference, including emerging LLM-based systems, that deliver... ..., automation, and efficiency of large-scale training...Local areaWorldwideFlexible hours$192.2k - $260k
...Science team is seeking an experienced Applied Scientist who will join a team of experts in the... ...of Amazon's data to help automate and optimize key processes - Design, development... ...feature creations - Establish scalable, efficient, automated processes for large scale data...SeniorLocal areaFlexible hours$192k - $304.75k
NVIDIA is searching for an outstanding Senior Researcher working on efficient deep learning to join our learning... ...methods for post-training model optimization (pruning, quantization, NAS),... ...architecture design, adaptive/dynamic inference, resource-efficient training and...SeniorFull time$165k - $238k
...About the role:As a Senior Research Scientist you will be a key architect... ...and high-fidelity inference over extremely large... ...the gap between applied research and... ...building blocks, such as LLM-as-a-judge evaluation... ...state management, and optimization for multi-step...SeniorFull timeWork at office3 days per week$183.83k - $275.98k
...with state-of-the-art architectures quickly and efficiently, collaborating with other teams to determine data... ...infrastructure support needs, and working to improve model optimization and inference speeds. You will use your applied research skills to think through the creation and...SeniorImmediate startFlexible hours$119.8k - $234.7k
...power ad ranking, pricing, and optimization across large-scale consumer... ...heterogeneous event streams to infer user intent and advertiser... ...marketplace dynamics. Engineers and scientists on the team work at the... ...from you. #MicrosoftAI Applied Sciences IC4 - The typical...SeniorOngoing contractWork at officeLocal areaShift work$192.2k - $260k
...Description Amazon is seeking an exceptional Sr. Applied Scientist to lead the development of perception systems that harness... ...DENSE, Astyx, RADDet, Boreas) Experience with real-time inference, model optimization (TensorRT, ONNX), and edge deployment Experience...SeniorLocal areaFlexible hoursNight shift$192.2k - $260k
...advertising lifecycle from ad creation and optimization to performance analysis and customer... ...and motivated Machine Learning Applied Scientist who loves to innovate at the intersection... ...discover new products they love, be the most efficient way for advertisers to meet their...SeniorLocal areaWorldwideFlexible hours- d-Matrix inc. is looking for a Senior Staff ML Researcher to join our Algo team in Santa Clara, CA. This hybrid position involves... ...week. The successful candidate will develop algorithms for optimizing LLM inference on our DNN accelerators. Ideal applicants should have a MSc...Senior3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Applied Scientist — Efficient LLM Inference & Optimization. Be the first to apply!
- scientist 1 Palo Alto, CA
- image scientist Palo Alto, CA
- qc scientist Palo Alto, CA
- research scientist Palo Alto, CA
- analytical scientist Palo Alto, CA
- research scientist - biology Palo Alto, CA
- quality control scientist Palo Alto, CA
- applied scientist Palo Alto, CA
- manufacturing scientist Palo Alto, CA
- machine learning research scientist Palo Alto, CA

