Senior DL Performance Engineer - LLM/Transformer Optimizations
NVIDIA AI
NVIDIA is seeking software engineers to join a Deep Learning models performance engineering team tasked with building and optimizing libraries and tools for efficient AI applications. The team accelerates DL frameworks like PyTorch and JAX, enabling researchers and engineers to design, develop, and deploy state-of-the-art AI. We work across NVIDIA and the broader open-source community to deliver performance gains on the world’s leading AI platform. #J-18808-Ljbffr NVIDIA AI
- NVIDIA is seeking a Senior DL Algorithms Engineer to optimize LLM/Omni models and enhance performance across its software stack. The ideal candidate will have a PhD and 3+ years of experience in deep learning, specifically in inference. This role involves profiling, analyzing...SeniorPerformance
$152k - $241.5k
...Deep Learning models performance engineering team at NVIDIA is hiring... ...levels to build and optimize the libraries and... ...doing:Build and support Transformer Engine, the open-... ...PyTorch, JAX, or any other DL framework.Experience... ....Knowledge of modern LLM architectures,...TransformerSeniorPerformanceFull time- ...generative AI to power the transformation of technology. We are... ..., ML Researcher - LLM Algorithmic OptimizationWhat... ...that will be used to optimize large language model... ...researchers, and ML engineers who create and apply... ...individual and company performance. This is in addition to...TransformerSeniorPerformance3 days per week
$184k - $287.5k
We are now looking for a Senior DL Algorithms Engineer! NVIDIA is seeking senior engineers who are mindful of performance analysis and optimization to help us squeeze every last clock cycle out... ...and deliver production code to TRT-LLM, NVIDIA’s open-source inference serving...SeniorPerformanceFull time$184k - $287.5k
...We are now looking for a Senior High-Performance LLM Training Engineer!NVIDIA is seeking experienced engineers specializing... ...in performance analysis and optimization to improve the efficiency of LLM... ...platform stack, from drivers to DL frameworks.Build and support NVIDIA...SeniorPerformanceFull timeWork experience placement$195.2k - $262.2k
...infrastructure. Built by engineers, for engineers.... ...to inference optimization, we own the hard... ...systems, GPU performance, model training... ...engineering. A Senior MLE owns... .... Optimize LLM and VLM endpoints... ...high-throughput transformer inference systems...TransformerSeniorPerformanceFull timeTemporary workImmediate startRemote work- ...magnitude increase in speed is transforming the user experience of AI applications... ....About The RoleWe are hiring a Senior Performance Engineer to join our Product team. You... ...stacks (vLLM, SGLang, TensorRT-LLM), GPU kernel-level optimization toolchains (CUDA, Triton), and an...TransformerSeniorPerformanceContract workShift work
- ...Staff or Principal level engineer who is passionate... ...Opportunities: As a Senior member on the team, you... ...and inference optimizations across a variety of applications... ...including innovative transformer architectures,... ...team to E2E co-optimize performance on current and future...TransformerSeniorPerformance
$184k - $287.5k
...and low-level hardware optimization has never been more... ...unlock maximum hardware performance for emerging AI... ...intersection of high-level DL frameworks and low-... ...Science, Computer Engineering, Electrical Engineering... ...Learning with a focus on transformers.Strong understanding...TransformerSeniorPerformanceFull time$193.3k - $261.5k
We are looking for a Senior Inference Engineer to own inference for... ...in• Implement and optimize the inference path for... ...and tune high-performance kernels for critical... ...fall outside standard LLM serving patterns — sustained... ...architectures (transformers, attention mechanisms...TransformerSeniorPerformanceInternshipLocal areaFlexible hours$184k - $287.5k
We are now looking for a Senior High-Performance LLM Training Engineer! NVIDIA is seeking experienced engineers specializing... ...in performance analysis and optimization to improve the efficiency of LLM... ...platform stack, from drivers to DL frameworks. Build and support...SeniorPerformanceWork experience placement$272k - $431.25k
...NVIDIA is seeking a Principal Engineer to drive the performance of large-scale AI training and post... .... You will analyze and optimize frontier-scale LLM workloads running on thousands of... ...learning workloads, especially transformer-based models, LLM pre-training,...TransformerPerformanceFull time- ...NVIDIA Corporation is seeking a Senior Software Engineer for the TensorRT Edge-LLM team in the US. You will develop a high-performance inference framework in modern C++ that... ...collaborate across CUDA and robotics teams, optimize transformer components, and contribute to kernel...TransformerSeniorPerformance
$174.72k - $295.68k
...strong foundation for LLM deployment and... ...research team to establish performance estimates and prove model... ...understanding of Transformer architectures and LLM... ...programming and software engineering skills.Ability to... ...Contributions to model optimization, inference, compiler,...TransformerSeniorPerformanceFull time$152k - $241.5k
...Join NVIDIA’s TensorRT Edge-LLM team and help shape the... ...compiler and runtime optimizations tailored for transformer-based models running on constrained... ...robotics to deliver high-performance, production-ready... ...Science, Electrical/Computer Engineering, or a closely related...TransformerSeniorPerformanceFull time$182.5k - $260.5k
..., its Zero Trust Engine, and the powerful... ...control without performance trade-offs.At Netskope... ...are available at Senior Staff and above.... ...inference and optimization layer that makes... ...SGLang, TensorRT-LLM, ONNX Runtime, llama... ....Solid grasp of transformer internals and the...TransformerSeniorPerformance$184k - $287.5k
We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance ceiling? Our LLM Inference Performance Analysis and Optimization team builds the answer from the ground up. We develop...SeniorPerformanceFull time$184k - $287.5k
NVIDIA has been transforming computer graphics, PC gaming, and accelerated... ...an experienced Compiler Optimization Engineer for an exciting role in our... ...in action as HPC and DL developers use features and... ...optimizations to achieve the best performance of their applications. If...SeniorPerformanceFull time- ...Sr. Staff ML Researcher to advance LLM algorithmic optimization on our DNN accelerators. You will design... ...mathematicians, ML researchers and engineers on high-impact research at the math-... ..., plus Python and OOP skills. Transformer knowledge is a plus; hybrid on-site...TransformerSenior3 days per week
$184k - $356.5k
...leading technology company in California is seeking a Senior DL Algorithms Engineer to drive inference performance for Deep Learning workloads. The role involves... ...inference and collaborating with co-design teams to optimize performance across hardware and software...SeniorPerformance- ...headquarters 3 days per week. The role Senior Staff ML Researcher - LLM Algorithmic Optimization What You Will Do d-Matrix is... ..., ML researchers, and ML engineers who create and apply advanced... ...OOP code design Experience with transformer architecture is advantageous...TransformerSenior3 days per week
$184k - $287.5k
...interact with people—are transforming every industry. GPU-accelerated... ...seeking an exceptional Senior Perception Engineer to help design and... ...frameworks to quantify perception performance; analyze large-scale real... ...platforms, including optimization for latency, memory, and...TransformerSeniorPerformanceFull time$224k - $356.5k
...interact with people—are transforming every industry. GPU-... ...seeking an exceptional Senior Radar Perception Engineer to help design and productize... ..., selection, and layout optimization to support L2-L4... ...quantify radar perception performance; analyze large-scale real...TransformerSeniorPerformanceFull time$194.6k - $285.7k
...for optical communications products. We optimize design that will integrate into the... ...Passive component design: inductors,* transformers, transmission-lines, etc.* Floorplanning... ...policies.Employees on sales plans earn performance-based incentive pay on top of their base...TransformerSeniorPerformanceFull timeTemporary workWork at officeLocal areaFlexible hours3 days per week$170.6k - $261.3k
...transportation on a global scale.As a Senior Machine Learning Engineer on the State Estimation... ...to improve model performance against those metrics.... ...pipelines, including model optimization techniques (e.g., pruning... ...as convolutional and transformer‑based architectures for:...TransformerSeniorPerformanceFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours$152k - $241.5k
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for... ...on the world.We are hiring software engineers for the CUDA Tile team. NVIDIA GPUs are... ...dialects and lowering passes, and optimize the performance of tile-based kernels to ensure they...SeniorPerformanceFull timeRemote work$184k - $287.5k
...Solutions Architect with a performance engineering background who can... ....Engaging in deep optimization of high-performance operators... ...-related distributed transformer workloads by... ...hands-on validated ML/DL performance engineering... ...one of these areas: LLM and HPC. Having...TransformerSeniorPerformanceFull timeRemote work$195.2k - $361.2k
...MissionAt Intel, our journey is to transform AI into something safer,... ...people actually own. You optimize inference engines (llama.cpp, vLLM) for... ...hardware tiers and publish honest performance comparisonsUpstream fixes... ...-level codeExperience with LLM inference. (attention, KV...SeniorPerformanceFull timeInternshipLocal areaImmediate startShift work$2,000 per month
...product (Sohu) only supports transformers, but has an order of... ...seeking a skilled PCB Layout Engineer to join our dynamic team and contribute to our high-performance data center products. As a PCB... ...crucial role in designing and optimizing printed circuit boards (PCBs...TransformerSeniorPerformanceContract workFor contractorsFor subcontractorWork at officeRelocation package$136k - $218.5k
We’re looking for a Senior Power Architecture & Optimization Engineer to push the limits of energy efficiency using advanced analytics and AI, including LLMs... ...principles alone.Partner closely with Architects, Performance, Software, ASIC Design, and Physical Design teams to...SeniorPerformanceFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior DL Performance Engineer - LLM/Transformer Optimizations. Be the first to apply!
- senior lead project manager Santa Clara, CA
- senior robotics software engineer Santa Clara, CA
- senior devops engineer remote Santa Clara, CA
- senior sas administrator Santa Clara, CA
- senior IT manager Santa Clara, CA
- sr project manager Santa Clara, CA
- senior windows systems engineer Santa Clara, CA
- senior researcher Santa Clara, CA
- senior manager data science Santa Clara, CA
- senior principal engineer Santa Clara, CA

