Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior DL Performance Engineer - LLM/Transformer Optimizations

NVIDIA AI

NVIDIA is seeking software engineers to join a Deep Learning models performance engineering team tasked with building and optimizing libraries and tools for efficient AI applications. The team accelerates DL frameworks like PyTorch and JAX, enabling researchers and engineers to design, develop, and deploy state-of-the-art AI. We work across NVIDIA and the broader open-source community to deliver performance gains on the world’s leading AI platform. #J-18808-Ljbffr NVIDIA AI

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior DL Performance Engineer - LLM/Transformer Optimizations in Santa Clara, CA vacancy
  • NVIDIA is seeking a Senior DL Algorithms Engineer to optimize LLM/Omni models and enhance performance across its software stack. The ideal candidate will have a PhD and 3+ years of experience in deep learning, specifically in inference. This role involves profiling, analyzing... 
    Senior
    Performance

    NVIDIA

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...Deep Learning models performance engineering team at NVIDIA is hiring...  ...levels to build and optimize the libraries and...  ...doing:Build and support Transformer Engine, the open-...  ...PyTorch, JAX, or any other DL framework.Experience...  ....Knowledge of modern LLM architectures,... 
    Transformer
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...generative AI to power the transformation of technology. We are...  ..., ML Researcher - LLM Algorithmic OptimizationWhat...  ...that will be used to optimize large language model...  ...researchers, and ML engineers who create and apply...  ...individual and company performance. This is in addition to... 
    Transformer
    Senior
    Performance
    3 days per week

    d-Matrix

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    We are now looking for a Senior DL Algorithms Engineer! NVIDIA is seeking senior engineers who are mindful of performance analysis and optimization to help us squeeze every last clock cycle out...  ...and deliver production code to TRT-LLM, NVIDIA’s open-source inference serving... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    14 hours ago
  • $184k - $287.5k

     ...We are now looking for a Senior High-Performance LLM Training Engineer!NVIDIA is seeking experienced engineers specializing...  ...in performance analysis and optimization to improve the efficiency of LLM...  ...platform stack, from drivers to DL frameworks.Build and support NVIDIA... 
    Senior
    Performance
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    4 days ago
  • $195.2k - $262.2k

     ...infrastructure. Built by engineers, for engineers....  ...to inference optimization, we own the hard...  ...systems, GPU performance, model training...  ...engineering. A Senior MLE owns...  .... Optimize LLM and VLM endpoints...  ...high-throughput transformer inference systems... 
    Transformer
    Senior
    Performance
    Full time
    Temporary work
    Immediate start
    Remote work

    Nebius

    Palo Alto, CA
    1 day ago
  •  ...magnitude increase in speed is transforming the user experience of AI applications...  ....About The RoleWe are hiring a Senior Performance Engineer to join our Product team. You...  ...stacks (vLLM, SGLang, TensorRT-LLM), GPU kernel-level optimization toolchains (CUDA, Triton), and an... 
    Transformer
    Senior
    Performance
    Contract work
    Shift work

    Cerebras Systems

    Sunnyvale, CA
    2 days ago
  •  ...Staff or Principal level engineer who is passionate...  ...Opportunities: As a Senior member on the team, you...  ...and inference optimizations across a variety of applications...  ...including innovative transformer architectures,...  ...team to E2E co-optimize performance on current and future... 
    Transformer
    Senior
    Performance

    AMD

    San Jose, CA
    2 days ago
  • $184k - $287.5k

     ...and low-level hardware optimization has never been more...  ...unlock maximum hardware performance for emerging AI...  ...intersection of high-level DL frameworks and low-...  ...Science, Computer Engineering, Electrical Engineering...  ...Learning with a focus on transformers.Strong understanding... 
    Transformer
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $193.3k - $261.5k

    We are looking for a Senior Inference Engineer to own inference for...  ...in• Implement and optimize the inference path for...  ...and tune high-performance kernels for critical...  ...fall outside standard LLM serving patterns — sustained...  ...architectures (transformers, attention mechanisms... 
    Transformer
    Senior
    Performance
    Internship
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    2 days ago
  • $184k - $287.5k

    We are now looking for a Senior High-Performance LLM Training Engineer! NVIDIA is seeking experienced engineers specializing...  ...in performance analysis and optimization to improve the efficiency of LLM...  ...platform stack, from drivers to DL frameworks. Build and support... 
    Senior
    Performance
    Work experience placement

    NVIDIA

    Santa Clara, CA
    1 day ago
  • $272k - $431.25k

     ...NVIDIA is seeking a Principal Engineer to drive the performance of large-scale AI training and post...  .... You will analyze and optimize frontier-scale LLM workloads running on thousands of...  ...learning workloads, especially transformer-based models, LLM pre-training,... 
    Transformer
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ...NVIDIA Corporation is seeking a Senior Software Engineer for the TensorRT Edge-LLM team in the US. You will develop a high-performance inference framework in modern C++ that...  ...collaborate across CUDA and robotics teams, optimize transformer components, and contribute to kernel... 
    Transformer
    Senior
    Performance

    NVIDIA Corporation

    Santa Clara, CA
    18 hours ago
  • $174.72k - $295.68k

     ...strong foundation for LLM deployment and...  ...research team to establish performance estimates and prove model...  ...understanding of Transformer architectures and LLM...  ...programming and software engineering skills.Ability to...  ...Contributions to model optimization, inference, compiler,... 
    Transformer
    Senior
    Performance
    Full time

    XPENG Motors

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...Join NVIDIA’s TensorRT Edge-LLM team and help shape the...  ...compiler and runtime optimizations tailored for transformer-based models running on constrained...  ...robotics to deliver high-performance, production-ready...  ...Science, Electrical/Computer Engineering, or a closely related... 
    Transformer
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $182.5k - $260.5k

     ..., its Zero Trust Engine, and the powerful...  ...control without performance trade-offs.At Netskope...  ...are available at Senior Staff and above....  ...inference and optimization layer that makes...  ...SGLang, TensorRT-LLM, ONNX Runtime, llama...  ....Solid grasp of transformer internals and the... 
    Transformer
    Senior
    Performance

    Netskope

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance ceiling? Our LLM Inference Performance Analysis and Optimization team builds the answer from the ground up. We develop... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

    NVIDIA has been transforming computer graphics, PC gaming, and accelerated...  ...an experienced Compiler Optimization Engineer for an exciting role in our...  ...in action as HPC and DL developers use features and...  ...optimizations to achieve the best performance of their applications. If... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...Sr. Staff ML Researcher to advance LLM algorithmic optimization on our DNN accelerators. You will design...  ...mathematicians, ML researchers and engineers on high-impact research at the math-...  ..., plus Python and OOP skills. Transformer knowledge is a plus; hybrid on-site... 
    Transformer
    Senior
    3 days per week

    Entrada Ventures

    Santa Clara, CA
    2 days ago
  • $184k - $356.5k

     ...leading technology company in California is seeking a Senior DL Algorithms Engineer to drive inference performance for Deep Learning workloads. The role involves...  ...inference and collaborating with co-design teams to optimize performance across hardware and software... 
    Senior
    Performance

    NVIDIA Corporation

    Santa Clara, CA
    1 day ago
  •  ...headquarters 3 days per week. The role Senior Staff ML Researcher - LLM Algorithmic Optimization What You Will Do d-Matrix is...  ..., ML researchers, and ML engineers who create and apply advanced...  ...OOP code design Experience with transformer architecture is advantageous... 
    Transformer
    Senior
    3 days per week

    d-Matrix inc.

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...interact with people—are transforming every industry. GPU-accelerated...  ...seeking an exceptional Senior Perception Engineer to help design and...  ...frameworks to quantify perception performance; analyze large-scale real...  ...platforms, including optimization for latency, memory, and... 
    Transformer
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $224k - $356.5k

     ...interact with people—are transforming every industry. GPU-...  ...seeking an exceptional Senior Radar Perception Engineer to help design and productize...  ..., selection, and layout optimization to support L2-L4...  ...quantify radar perception performance; analyze large-scale real... 
    Transformer
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $194.6k - $285.7k

     ...for optical communications products. We optimize design that will integrate into the...  ...Passive component design: inductors,* transformers, transmission-lines, etc.* Floorplanning...  ...policies.Employees on sales plans earn performance-based incentive pay on top of their base... 
    Transformer
    Senior
    Performance
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    CISCO Systems

    San Jose, CA
    4 days ago
  • $170.6k - $261.3k

     ...transportation on a global scale.As a Senior Machine Learning Engineer on the State Estimation...  ...to improve model performance against those metrics....  ...pipelines, including model optimization techniques (e.g., pruning...  ...as convolutional and transformer‑based architectures for:... 
    Transformer
    Senior
    Performance
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  • $152k - $241.5k

    NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for...  ...on the world.We are hiring software engineers for the CUDA Tile team. NVIDIA GPUs are...  ...dialects and lowering passes, and optimize the performance of tile-based kernels to ensure they... 
    Senior
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    14 hours ago
  • $184k - $287.5k

     ...Solutions Architect with a performance engineering background who can...  ....Engaging in deep optimization of high-performance operators...  ...-related distributed transformer workloads by...  ...hands-on validated ML/DL performance engineering...  ...one of these areas: LLM and HPC. Having... 
    Transformer
    Senior
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $195.2k - $361.2k

     ...MissionAt Intel, our journey is to transform AI into something safer,...  ...people actually own. You optimize inference engines (llama.cpp, vLLM) for...  ...hardware tiers and publish honest performance comparisonsUpstream fixes...  ...-level codeExperience with LLM inference. (attention, KV... 
    Senior
    Performance
    Full time
    Internship
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    2 days ago
  • $2,000 per month

     ...product (Sohu) only supports transformers, but has an order of...  ...seeking a skilled PCB Layout Engineer to join our dynamic team and contribute to our high-performance data center products. As a PCB...  ...crucial role in designing and optimizing printed circuit boards (PCBs... 
    Transformer
    Senior
    Performance
    Contract work
    For contractors
    For subcontractor
    Work at office
    Relocation package

    Etched

    Cupertino, CA
    14 hours ago
  • $136k - $218.5k

    We’re looking for a Senior Power Architecture & Optimization Engineer to push the limits of energy efficiency using advanced analytics and AI, including LLMs...  ...principles alone.Partner closely with Architects, Performance, Software, ASIC Design, and Physical Design teams to... 
    Senior
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior DL Performance Engineer - LLM/Transformer Optimizations. Be the first to apply!