Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Engineer, CUDA UMD - GPU Kernel Scheduling

$272k - $431.25k

NVIDIA

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. We're looking to grow our company, and form teams with the smartest people in the world. Join us at the forefront of technological advancement.Are you a motivated system software engineer with a deep understanding of device drivers who has phenomenal C/C++ skills? If so, this role might be for you. We are looking for a seasoned software professional to work on the CUDA Driver, a core component of our platform for accelerating general purpose computation on the GPU. You will be an integral part of a team that delivers features and improvements to better realize the potential of NVIDIA hardware for a growing range of computational workloads, ranging from deep learning, scientific computation, data science and self-driving cars to video games and virtual reality.What you'll be doing:As a member of our team, you will use your design abilities, coding expertise, and creativity to deliver the best compute platform in the world. You will craft elegant solutions to exciting problems and shape the future direction of CUDA as you collaborate with your peers across NVIDIA.Evangelize, architect, and implement new featuresCoordinate and drive development efforts across multiple teamsHelp define forward-looking improvements to the CUDA APIs and programming modelExtend important CUDA programming models and functionality such as CUDA GraphsExplore ways to use Graphs to improve the scheduling of AI/ML workloads on our GPUS to be more efficient and faster. Write effective, maintainable, and well-tested codeDevelop code for multiple operating systemsWhat we need to see:BS or MS degree in Computer Science, Electrical Engineering​ or related field (or equivalent experience)Strong C and C++ programming skillsMinimum of 15+ years of related development experience (multiple positions for varying experience levels open)Experience driving projects across multiple teamsExperience working with large codebasesBackground with operating system interfaces for threads, process control, and virtual memoryExperience writing and debugging multithreaded programsGood written communication as well as presentation skillsWays to stand out from the crowd:Prior experience with parallel computing - preferably writing CUDA Programs or Libraries that use CUDAUnderstanding of system level architecture, such as interconnects, memory hierarchy, interrupts, and memory-mapped IOKnowledge of memory coherence and consistency modelsBackground with kernel mode developmentExperience with Linux Systems Software development as well as experience maintaining and extending programming models or higher-level language support for similar environmentsYour base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until July 14, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa ClaraType: Full time

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Principal Engineer, CUDA UMD - GPU Kernel Scheduling in Santa Clara, CA vacancy
  • $184k - $287.5k

    We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push...  ..., compiler decisions, and runtime scheduling.Direct experience with LLM inference frameworks...  ...of GPU kernel optimization — CUDA, CUTLASS, Triton, or equivalent — and... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    13 hours ago
  • $184k - $287.5k

    We are now looking for a Senior Formal Verification Engineer for GPU Kernels! Modern AI performance relies on highly optimized GPU kernels — performance...  ...from the crowd:Knowledge of CPU and/or GPU architecture. CUDA or OpenCL experience is a plus.Background in the... 
    Suggested
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...THE ROLE:We are looking for a Senior GPU Inference Performance Engineer to own end-to-end performance...  ...rocProfiler, Omniperf) and NVIDIA (CUDA, Nsight Systems/Compute, DCGM) GPUs...  ...HBM bandwidth, compute utilization, kernel scheduling, memory allocation, and PCIe/Infinity... 
    Suggested

    AMD

    Santa Clara, CA
    4 days ago
  •  ...ROLEWe are seeking a Principal GenAI Inference Optimization Engineer to join our Models...  ...workloads on AMD GPU platforms. You will...  ...multiple layers—from kernels and runtimes to...  ...memory bandwidth, scheduling).- Contribute to cross...  ...systems language (C++/CUDA/HIP).- Experience... 
    Suggested

    AMD

    San Jose, CA
    3 days ago
  •  ...a performance-obsessed engineer to drive AI inference performance...  ..., diagnosing a kernel-level bottleneck, and...  ...Engineering. You understand GPU kernel performance...  ...dispatch to framework-level scheduling and multi-node...  ...development or integration (HIP, CUDA, Triton, CK, or similar... 
    Suggested

    AMD

    San Jose, CA
    4 days ago
  • $165k - $242k

     ...skilled and motivated Systems Kernel Engineer to join the HAVOCK Team,...  ...networking, storage, virtualization, GPU/DPU enablement). Stack‑Wide...  ..., kubelet) HPC/AI workloads (CUDA, GPUDirect, RoCE/InfiniBand)...  ...(memory management, scheduling, networking, storage, drivers... 
    Permanent employment
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    3 days ago
  • $272k - $431.25k

     ...transforming every industry. GPU-accelerated deep learning...  ...are seeking an exceptional Principal Perception Engineer to lead the design and productization...  ...pipelines.Experience with CUDA development and optimizing...  ...through custom CUDA kernels or other GPU-accelerated... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $221.7k - $364.8k

     ...market segments) consumed by millions of people around the world. Come build with us!Role and ResponsibilitiesAs a Principal GPU Design Verification Engineer - Subsystems, you will lead the verification of one or more GPU subsystems for Samsung’s next-generation mobile... 
    Hourly pay
    Full time
    Relocation

    Samsung Semiconductor

    San Jose, CA
    1 day ago
  • $152k - $287.5k

     ...NVIDIA is seeking a Senior Engineer to join their team in Santa Clara to work on the CUDA driver, essential for GPU computing. In this role, you'll collaborate with architects and specialists to drive innovative solutions across various fields including deep learning... 

    NVIDIA

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...seeking a self‑motivated senior engineer for the Aerial Omniverse...  ...implementation of a real‑time, GPU‑accelerated propagation engine...  ...experience.Hands‑on proficiency with CUDA and at least one GPU ray‑...  ...vs‑bandwidth trade‑offs at the kernel level.Working knowledge of electromagnetic... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...era of computing. An era in which our GPU acts as the brains of computers, robots...  ...impact on the world.We are hiring software engineers for the CUDA Tile team. NVIDIA GPUs are at the...  ...optimize the performance of tile-based kernels to ensure they execute efficiently across... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...NVIDIA is hiring a Senior Compiler Engineer to join our team driving the next generation of GPU systems programming. We are...  ...tooling of Rust to native GPU and CUDA development. On this team, you will...  ...-safe, high-performance GPU kernels in idiomatic Rust.What you’ll be... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $207k - $340k

     ...approval. We’re hiring a Principal Staff Software Engineer to lead LinkedIn’s GPU-Based Retrieval Platform,...  ...distributed serving, GPU scheduling, memory efficiency, batching, and kernel-level optimization. You will...  ...Optimize GPU performance across CUDA, Triton, memory... 
    For contractors
    Work at office
    Remote work
    Work from home
    Flexible hours

    Linkedin

    Mountain View, CA
    4 days ago
  •  ...ROLE: We are seeking a Principal Software Engineer to serve as the...  ...qualification on AMD Instinct™ GPU platforms. You will...  ..., scientific HPC kernels, MLPerf-class benchmarks...  ...stacks (ROCm, CUDA, oneAPI, SYCL)AI/ML frameworks...  ..., and topology-aware scheduling.Experience supporting... 
    Contract work
    Shift work

    AMD

    San Jose, CA
    13 hours ago
  • $278.1k - $417.1k

     ...that runtime. As our Principal Engineer for On-Device AI Inference...  ..., optimization, and kernel-level tuning, to a...  ...tuning across NPU, mobile GPU, and desktop/laptop...  ...SPIR-V compute, D3D12, CUDA); profile with browser...  ...game engine: real-time scheduling, threading, memory pooling... 
    Work at office
    Worldwide
    Relocation package

    Unity

    Mountain View, CA
    1 day ago
  • $272k - $431.25k

     ...to increase, we are seeking outstanding engineers to join our team and help shape the future...  ...understanding of computer architecture, and GPU/parallel datacenter computing...  ...modeling.GPU programming experience with CUDA or OpenCL.NVIDIA has been transforming computer... 
    Full time

    Nvidia

    Santa Clara, CA
    13 hours ago
  • $224k - $356.5k

     ...combining optimized inference engines, model profiles/recipes...  ...by anticipating schedule, staffing, and dependency...  ..., inference engine/kernel, performance optimization...  ...containers, Kubernetes, GPU, or inference...  ...GPU technologies such as CUDA, cuDNN, CUTLASS, cuBLAS... 

    Socket.dev

    Santa Clara, CA
    17 hours ago
  • $124k - $195.5k

     ...projects at the intersection of CUDA and Deep Learning Systems. As...  ...in model optimization, custom kernel development, and cluster-scale...  ...systems from a single GPU to supercomputer clusters, we...  ...in Computer Science, Computer Engineering, Electrical Engineering, or related... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $272k - $431.25k

     ...efficient, and secure for millions globally. We seek a Senior Engineer to lead technical efforts in deploying advanced AI agent...  ...of LLM inference pipelines (Ollama, Llama.cpp, vLLM), GPU-accelerated computing (CUDA, TensorRT), and experience running local models on... 
    Full time
    Local area
    Shift work

    Nvidia

    Santa Clara, CA
    13 hours ago
  • $272k - $431.25k

     ...NVIDIA is seeking a Principal Engineer to drive the performance of large...  ...of distributed training, GPU architecture, systems...  ...framework/runtime internals, CUDA libraries, communication...  ..., memory, communication, scheduling, parallelism strategy, kernel efficiency, framework... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $140k - $224.25k

     ...creative, and hands-on software engineer with a test to failure...  ...optimize the testing workflows in GPU domain.Write maintainable, reliable...  ...create a realistic delivery schedule and work on challenging...  ...DLSS, Frame Generation, Reflex, CUDA, G-Sync, etc.The ability to collaborate... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ...on AMD Instinct™ GPUs. You’ll work across kernels, distributed training, and framework...  ...ideal candidate is passionate about software engineering and the craft of training performance. You...  ..., and optimizer steps.Optimize multi‑GPU/multi‑node training and communication patterns... 

    AMD

    San Jose, CA
    1 day ago
  • $100k

     ...for contributors of all seniorities.We are seeking an Senior Engineer to develop and optimize the software stack for our IP...  ...Parallel processing, SIMD architectures (SSE, AVX, Arm Neon), GPU programming (CUDA, OpenCL) a plus.Background in firmware, device drivers, embedded... 
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    4 days ago
  •  ...we advance your career. THE ROLE:We are looking for a Principal Machine Learning Engineer to join our Models and Applications team. If you are excited...  ..., especially large models, is a plus.Experience with GPU kernel optimization is a plus.Excellent Python or C++... 

    AMD

    San Jose, CA
    1 day ago
  • $272k - $431.25k

     ...NVIDIA is looking for a Machine Learning (ML) Engineer to join the GPU accelerated Apache Spark team. Apache Spark is the most popular data processing...  ...related to Apache Spark.Familiarity with NVIDIA GPUs and CUDA.Experience coding in Scala, Java, and/or C++.NVIDIA is... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $272k - $431.25k

     ...We're hiring a Principal Developer Relations Manager to...  ...Relations Managers and engineers, running internal deep-...  ...-on experience inCUDA, CUDA C++, GNN, SNN, CUTLASS,...  ...NVIDIA software including GPU programming fluency...  ...and reason about custom kernels, Tensor Cores, and low-... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $206k - $333k

     ...this role We're looking for a Principal Engineer to be the technical lead of...  ...tracks with NVIDIA (CUDA, cuDNN, TensorRT/TensorRT-LLM...  ...6/FP8/FP4), batch sizes, and GPU types. Maintain a corpus of representative...  ...s controllers/operators to schedule benchmarks at scale;... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    18 days ago
  • $307k - $427k

     ...hardware deployment and software scheduling systems with Google’s carbon...  ...directly into CPU, TPU, GPU, and storage platform roadmaps...  ...Partner with chip design, platform engineering, software infrastructure,...  .... We are looking for a Principal Engineer, Data Center Sustainability... 
    Worldwide

    Google

    Sunnyvale, CA
    2 days ago
  • $224k - $356.5k

    NVIDIA's invention of the GPU sparked the growth of the PC gaming market, redefined...  ...world.We are now looking for a Compiler Engineering Manager! NVIDIA’s HPC Compiler team is looking...  ...team objectives to meet schedules and goals.Act as a mentor and advisor to... 
    Full time
    Immediate start
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $102.3k - $209.5k

    What You’ll DoOwn and deliver GPU infrastructure programs spanning datacenter enablement...  ...and execution across multiple teams (engineering, networking, datacenter ops, supply...  ...moving. Define program scope, milestones, schedules, and success metrics, with a focus on regular... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Santa Clara, CA
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Engineer, CUDA UMD - GPU Kernel Scheduling. Be the first to apply!