Principal Engineer, CUDA UMD - GPU Kernel Scheduling
$272k - $431.25kNVIDIA
NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. We're looking to grow our company, and form teams with the smartest people in the world. Join us at the forefront of technological advancement.Are you a motivated system software engineer with a deep understanding of device drivers who has phenomenal C/C++ skills? If so, this role might be for you. We are looking for a seasoned software professional to work on the CUDA Driver, a core component of our platform for accelerating general purpose computation on the GPU. You will be an integral part of a team that delivers features and improvements to better realize the potential of NVIDIA hardware for a growing range of computational workloads, ranging from deep learning, scientific computation, data science and self-driving cars to video games and virtual reality.What you'll be doing:As a member of our team, you will use your design abilities, coding expertise, and creativity to deliver the best compute platform in the world. You will craft elegant solutions to exciting problems and shape the future direction of CUDA as you collaborate with your peers across NVIDIA.Evangelize, architect, and implement new featuresCoordinate and drive development efforts across multiple teamsHelp define forward-looking improvements to the CUDA APIs and programming modelExtend important CUDA programming models and functionality such as CUDA GraphsExplore ways to use Graphs to improve the scheduling of AI/ML workloads on our GPUS to be more efficient and faster. Write effective, maintainable, and well-tested codeDevelop code for multiple operating systemsWhat we need to see:BS or MS degree in Computer Science, Electrical Engineering or related field (or equivalent experience)Strong C and C++ programming skillsMinimum of 15+ years of related development experience (multiple positions for varying experience levels open)Experience driving projects across multiple teamsExperience working with large codebasesBackground with operating system interfaces for threads, process control, and virtual memoryExperience writing and debugging multithreaded programsGood written communication as well as presentation skillsWays to stand out from the crowd:Prior experience with parallel computing - preferably writing CUDA Programs or Libraries that use CUDAUnderstanding of system level architecture, such as interconnects, memory hierarchy, interrupts, and memory-mapped IOKnowledge of memory coherence and consistency modelsBackground with kernel mode developmentExperience with Linux Systems Software development as well as experience maintaining and extending programming models or higher-level language support for similar environmentsYour base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until July 14, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa ClaraType: Full time
$184k - $287.5k
We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push... ..., compiler decisions, and runtime scheduling.Direct experience with LLM inference frameworks... ...of GPU kernel optimization — CUDA, CUTLASS, Triton, or equivalent — and...SuggestedFull time$184k - $287.5k
We are now looking for a Senior Formal Verification Engineer for GPU Kernels! Modern AI performance relies on highly optimized GPU kernels — performance... ...from the crowd:Knowledge of CPU and/or GPU architecture. CUDA or OpenCL experience is a plus.Background in the...SuggestedFull timeWork experience placement- ...Clara, CA is seeking a Senior Software Engineer to advance the CUDA driver and Unified Memory. You will... ...and implement features across chips and kernels, collaborating with cross-functional teams... ...and contribute to high-performance GPU software. #J-18808-Ljbffr NVIDIA...Suggested
$272k - $431.25k
...Analysis Manager to lead an engineering team responsible for supervising... ...role involves analyzing kernel authoring flows from DSLs to... ...testlists, and coordinate with CUDA release schedules. Additionally, you will... ...continuous performance tracking for GPU software products from pre-...SuggestedFull timeShift work- ...ROLEWe are seeking a Principal GenAI Inference Optimization Engineer to join our Models... ...workloads on AMD GPU platforms. You will... ...multiple layers—from kernels and runtimes to... ...memory bandwidth, scheduling).- Contribute to cross... ...systems language (C++/CUDA/HIP).- Experience...Suggested
- ...THE ROLE:We are looking for a Senior GPU Inference Performance Engineer to own end-to-end performance... ...rocProfiler, Omniperf) and NVIDIA (CUDA, Nsight Systems/Compute, DCGM) GPUs... ...HBM bandwidth, compute utilization, kernel scheduling, memory allocation, and PCIe/Infinity...
- ...California is looking for a Member of Technical Staff for Kernel/Compiler/Communication. This critical role requires strong expertise in CUDA and GPU optimization, along with 5+ years of experience in performance engineering. The ideal candidate will design high-performance...
- CoreWeave is hiring a Senior Engineer for its Benchmarking & Performance team to write, profile, and optimize GPU kernels on the LLM inference path. You will improve latency and throughput and collaborate with product, orchestration, and hardware teams to achieve strict...
$272k - $431.25k
...transforming every industry. GPU-accelerated deep learning... ...are seeking an exceptional Principal Perception Engineer to lead the design and productization... ...pipelines.Experience with CUDA development and optimizing... ...through custom CUDA kernels or other GPU-accelerated...Full time$221.7k - $364.8k
...market segments) consumed by millions of people around the world. Come build with us!Role and ResponsibilitiesAs a Principal GPU Design Verification Engineer - Subsystems, you will lead the verification of one or more GPU subsystems for Samsung’s next-generation mobile...Hourly payFull timeRelocation$165k - $242k
...skilled and motivated Systems Kernel Engineer to join the HAVOCK Team,... ...networking, storage, virtualization, GPU/DPU enablement). Stack‑Wide... ...kubelet) HPC/AI workloads (CUDA, GPUDirect, RoCE/InfiniBand)... ...(memory management, scheduling, networking, storage, drivers...Permanent employmentTemporary workCasual workWork at officeRemote workFlexible hours- NVIDIA is seeking a senior software engineer to design and implement Rust libraries and CUDA-backed GPU APIs. You will build safe Rust abstractions over C/C++ interfaces and maintain the necessary C/C++ components to support Rust-facing functionality. You will optimize...
- CoreWeave is seeking a Senior Software Engineer for the Systems Engineering team to own kernel tracing and patching across Kubernetes and container runtimes.... ...teams. You will work on cgroups, namespaces, the CFS scheduler, memory management, and the orchestration stack to...
$272k - $431.25k
...computing. An era in which our GPU acts as the brains of... ...doing: Own and drive the Linux kernel strategy for NVIDIA CPU platforms... ...including memory management, scheduling, virtualization, PCIe, ACPI,... ...firmware teams, and software engineers to influence hardware architecture...$241.8k - $409.2k
GPGPU Software Architect/ Principal Engineer XPENG is a leading... ...towards General Purpose GPU (GPGPU) architecture.... ...and embracing the CUDA ecosystem. Our goal is... ...build fully reusable kernel libraries Partner with... ...SM architecture, Warp Scheduler, Shared-Memory conflicts...Full time- ..., Mirantis empowers platform engineering teams to deliver composable,... ...Mirantis delivers the automation, GPU orchestration, and policy-... ..., memory bandwidth, I/O, scheduling). Comfortable in the frameworks... ...ecosystem (e.g., NCCL, CUDA-level concepts, containers, schedulers...Full time
$206.4k - $379.1k
...Premiere. We are hiring a Principal Machine Learning Engineer to serve as the technical... ...across model families and GPU fleets. The system design... ...using tools such as PyTorch, CUDA, Triton, and TensorRT.... ...runtimes, quantization, GPU scheduling — to improve engineering velocity...Full timeTemporary workLocal areaWorldwide$152k - $241.5k
...NVIDIA is hiring a Senior Compiler Engineer to join our team driving the next generation of GPU systems programming. We are... ...tooling of Rust to native GPU and CUDA development. On this team, you will... ...-safe, high-performance GPU kernels in idiomatic Rust.What you’ll be...Full timeRemote work$207k - $340k
...approval. We’re hiring a Principal Staff Software Engineer to lead LinkedIn’s GPU-Based Retrieval Platform,... ...distributed serving, GPU scheduling, memory efficiency, batching, and kernel-level optimization. You will... ...Optimize GPU performance across CUDA, Triton, memory...For contractorsWork at officeRemote workWork from homeFlexible hours$184k - $287.5k
...seeking a self‑motivated senior engineer for the Aerial Omniverse... ...implementation of a real‑time, GPU‑accelerated propagation engine... ...experience.Hands‑on proficiency with CUDA and at least one GPU ray‑... ...vs‑bandwidth trade‑offs at the kernel level.Working knowledge of electromagnetic...Full time$152k - $241.5k
...helped craft as a member of the GPU Foundations Developer Tools... ...Collaborate with developer tools engineers, software library developers,... ...and both user-mode and kernel-mode drivers.Proven knowledge... ...monitor hardware.Experience in CUDA or GPU programming models.Prior...Full time$206k - $333k
...this role We're looking for a Principal Engineer to be the technical lead of... ...tracks with NVIDIA (CUDA, cuDNN, TensorRT/TensorRT-LLM... ...6/FP8/FP4), batch sizes, and GPU types. Maintain a corpus of representative... ...s controllers/operators to schedule benchmarks at scale;...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours- NVIDIA seeks a Principal Linux Kernel Engineer in Santa Clara, CA to lead upstream-first kernel development for NVIDIA CPU platforms. You will guide... ...kernel strategy, coordinate across memory management, scheduling, virtualization, PCIe, ACPI, and related subsystems, and...
$248k - $396.75k
...accelerated computing, inventing the GPU and pioneering the... ...our time. NVIDIA's IT Storage Engineering team architects, designs, deploys... ...performance, including kernel-level concepts around I/O subsystems... ...Familiarity with HPC job schedulers such as IBM LSF or Slurm — understanding...Full timeShift work$200k - $260k
...for talented individuals to join its fast-growing teams.As a Principal Engineer on our Mapping and Localization team, you will design and... ...experience with LOAM, lego-LOAM, ORB-SLAM, or VINS.Experience with CUDA/GPU programming or optimization for ARM-based embedded platforms...Odd jobFull timeShift work$206k - $303k
...CoreWeave runs some of the largest GPU clusters in the world. The AI infrastructure... ...under constant pressure. As a Principal Engineer in AI Infrastructure, you will lead the... .... Act as a technical authority on scheduling, quota enforcement, fairness, pre-emption...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$272k - $431.25k
...to increase, we are seeking outstanding engineers to join our team and help shape the future... ...understanding of computer architecture, and GPU/parallel datacenter computing... ...modeling.GPU programming experience with CUDA or OpenCL.NVIDIA has been transforming computer...Full time- ...Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines... ...You will collaborate across teams, work on CUDA C/C++, Triton, and cutting-edge MLIR-based...
$272k - $431.25k
...efficient, and secure for millions globally. We seek a Senior Engineer to lead technical efforts in deploying advanced AI agent... ...of LLM inference pipelines (Ollama, Llama.cpp, vLLM), GPU-accelerated computing (CUDA, TensorRT), and experience running local models on...Full timeLocal areaShift work- NVIDIA Gruppe in Santa Clara is seeking a motivated system software engineer with strong C/C++ skills to work on the CUDA Driver. This integral role involves designing and implementing features for NVIDIA hardware, collaborating with teams to enhance the platform for various...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Engineer, CUDA UMD - GPU Kernel Scheduling. Be the first to apply!
- principal network engineer Santa Clara, CA
- principal application developer Santa Clara, CA
- principal engineer Santa Clara, CA
- principal infrastructure engineer Santa Clara, CA
- director quality engineering Santa Clara, CA
- senior civil engineer project manager Santa Clara, CA
- senior chief engineer Santa Clara, CA
- engineering director Santa Clara, CA
- senior director engineering Santa Clara, CA
- director systems engineering Santa Clara, CA


