Compiler Engineer - AI Inference
$152k - $241.5kNVIDIA
NVIDIA's invention of the GPU 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company”.NVIDIA is seeking top-tier AI Compiler Engineers to drive innovation within our world-class compiler organization. In this role, you will push the boundaries of what is possible in AI performance and help build the technology that powers the next generation of computing. Join us and make a tangible impact on a global scale.What you’ll be doing:Drive technical innovation: Participating in hands-on development focusing on kernel generation and computational graph optimizations for next-generation NVIDIA GPUs.Advance the state-of-the-art: Solve complex compilation problems for AI workloads (both inference and training) and successfully transition these breakthroughs into enterprise and consumer products.Collaborate on hardware/software co-design: Partner with leading experts across our software, hardware, and research divisions to architect and co-design future silicon.Scale AI to the datacenter: Participating in the advancement and optimization of datacenter-scale AI workload deployments.What we need to see:BS or MS in Computer Science, Computer Engineering, or a related field (or equivalent experience). A PhD is strongly preferred.Compiler Experience: 3+ years of relevant industry experience specializing in compiler optimizations, synthesis, and placement.MLIR Knowledge: Demonstrated, hands-on experience working with MLIR.Programming Excellence: Exceptional C/C++ and Python programming and software design skills, including rigorous debugging, performance analysis, and test design.Team Dynamics: Strong communication and interpersonal skills, with the ability to collaborate effectively in a dynamic, fast-paced, and product-oriented environment.Ways to stand out from the crowd:Hardware Implementation: Hands-on experience implementing complex AI workloads on CPU, GPU, and/or custom AI accelerator architectures.LLM Knowledge: Deep understanding of Large Language Model (LLM) inference and its profound implications on computer architecture.Architecture & Design: Demonstrated understanding in the designing and architecting of comprehensive compiler frameworks from the ground up.With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing. If you're a creative and autonomous engineer with a real passion for technology, we want to hear from you.Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until April 28, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa ClaraType: Full time
$184k - $287.5k
We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency... ...inference stacks, optimize GPU kernels and compilers, drive industry benchmarks, and scale workloads across...SuggestedFull time$152k - $241.5k
...recently, GPU deep learning ignited modern AI — the next era of computing — with the... ...are looking for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers... ...DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers, personal...SuggestedFull timeRemote work$152k - $241.5k
...eager to work on cutting-edge AI technology for safety-... ...TensorRT team as a Senior Software Engineer, and be at the forefront of technology... ...enabling high-performance AI inference solutions for automotive... ...into TensorRT's compiler and runtime for specialized and...SuggestedFull time$140k - $224.25k
...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our... ...lasting impact on the world.NVIDIA is looking for a world-class engineer to join its Compiler Build & DevOps team. The position would be part of a team...SuggestedFull timeWork experience placement$184k - $287.5k
...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts... ...can make a lasting impact on the world.We are seeking an AI Compiler Engineer with deep expertise in compiler technologies to join our team....SuggestedFull timeRemote work- ...generation computing experiences—from AI and data centers, to PCs,... ...and SOTA LLM and Multimodal inference at scale across multi-GPU and... ...and optimize cutting-edge compiler technologies and drive... ...ecosystem. THE PERSON: Skilled engineer with strong technical and analytical...
$152k - $241.5k
...are now looking for a Senior Software Engineer for Deep Learning Inference! Would you like to make a big impact... ..., and optimizations.Background in compiler developmentExperience in working with... ...an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA...Full time- ...Do you want to be part of the AI revolution? Do you want to think out of the box... ...are looking for general or deep learning compiler engineers to join our compiler team at Baidu’s... ...similar Experience with neural networks inference on dedicated SOC or GPU is preferred...Full timeWork at officeRemote workAfternoon shift
$193.3k - $261.5k
...comprehensive toolkit includes an ML compiler, runtime, and application... ...JAX enabling unparalleled ML inference and training performance.The... ...-software boundary, our engineers build systematic infrastructure... ...boundaries of what's possible in AI acceleration.As part of the...Work experience placementInternshipLocal areaFlexible hours$193.3k - $261.5k
...really fast on the Trainium hardware.As a Software Development Engineer on the Inference Model Enablement team, you will onboard and optimize state... ...issues across the stack in collaboration with the compiler and runtime teams* Mentoring junior engineers on system design...InternshipLocal areaFlexible hours- ...generation computing experiences—from AI and data centers, to PCs,... ...the next generation of compiler and software infrastructure to... ...Principal Software Development Engineer to lead technical development... ...movement optimization for ML inference workloads• Define and...
$152k - $241.5k
NVIDIA's GPUs are at the core of modern AI infrastructure, from training large-scale models to running inference in production. That position depends on software as much as hardware, and compiler engineering is a big part of what makes it work.We are looking for outstanding...Full timeRemote work- ...unleashing the potential of generative AI to power the transformation of technology... ....The role: Principal System Software Engineer, AI Inference ExecutionWhat you will do:The role requires... ...closely with other software (ML and compilers) and hardware experts in the company....3 days per week
$152k - $241.5k
...tapping into the unlimited potential of AI to define the next era of computing. An... ...are looking for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for... ...DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers, personal...Full timeRemote work- ...California or Austin, TexasNXP is searching for a hands-on AI Compiler Engineer who thrives at the convergence of cutting-edge AI, compiler... ...Excellent C/C++ and Python skillsSolid understanding of AI inference workloads (CNNs, transformers, perception or generative models...Full timeWork at officeLocal area
$184k - $287.5k
...hardware performance for emerging AI workloads. You will be a... ...bottlenecks in both training and inference pipelines.Collaborate closely... ...SW architects, kernel and compiler authors and CUDA driver... ...in Computer Science, Computer Engineering, Electrical Engineering, or related...Full time$152k - $241.5k
...Senior Machine Learning Applications and Compiler Engineer!NVIDIA is seeking engineers to develop... ...and optimizations for our LPX inference and compiler stack. You will work at the... ...or similar.Experience with large-scale AI distributed inference or training systems...Full time- ...generation computing experiences—from AI and data centers, to PCs,... ...ROLEWe are hiring a senior engineering leader to define, build, and... ...encode, pre/post-processing, inference, tracking, streaming,... .../pip (where relevant), cross‑compile toolchains, reproducible builds...
$164k - $313.3k
...Learning (ML) Systems & Efficiency Engineer to join our R&D team focused... ...-ready improvements in inference performance, latency, and cost... ...as Artificial Intelligence (AI), ML systems, and computer vision... ...applying inference compilation and optimization tools (e.g.,...Full timeTemporary workLocal areaWorldwide$184k - $287.5k
We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to... ...assembly layer. Our team works closely with compiler, kernel, hardware, and framework... ...agentic kernel optimization: applying AI-driven analysis to diagnose performance...Full time$184k - $287.5k
...recently, GPU deep learning ignited modern AI — the next era of computing — with the... ...We're looking for a Senior Performance Compiler Engineer to join our team and work on the open-... ..., accelerating both training and inference. You will be immersed in a diverse, supportive...Full timeRemote work- ...generation computing experiences—from AI and data centers, to PCs, gaming and... ...ROLE:We are looking for a Senior GPU Inference Performance Engineer to own end-to-end performance analysis... ...libraries, operator fusion, graph compilation), or configuration differences.Multi-...
$100k
...is leading the industry on cutting-edge AI technology, revolutionizing performance expectations... ...to unify innovations in software models, compilers, platforms, networking, and... ...posting.Who You AreA passionate software engineer eager to work on compiler technologies and...Permanent employment$152k - $241.5k
NVIDIA is the platform upon which every new AI‑powered application is built. We are seeking a Senior Software Engineer - AI Inference to advance open‑source LLM serving by contributing directly to upstream inference engines like vLLM and SGLang-ensuring they run best‑in...Full timeRemote work$152k - $241.5k
We are now looking for a Senior Software Engineer for Quantized Inference! NVIDIA is seeking software engineers to accelerate the discovery and deployment... ...fundamentals: concise, well-tested code; fluent with AI-assisted toolingExperience with ML accelerators with a basic...Full time$165.2k - $223.6k
...research and development company that designs and engineers high-profile consumer products like the Kindle... ...Astro. We are building the next generation of edge AI capabilities through our advanced compression platform, compiler and custom neural accelerator silicon. Come join...InternshipLocal areaFlexible hoursNight shiftDay shift- ...generation computing experiences—from AI and data centers, to PCs, gaming and... ...ROLEWe are seeking a Principal GenAI Inference Optimization Engineer to join our Models and Applications team... ....- Collaborate with hardware, compiler, and framework teams to improve overall...
$168k - $264.5k
...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our... ...looking for best-in-class Senior Physical Design Methodology Engineer, PPA Fusion Compiler to join our outstanding Networking Silicon engineering team...Full time$2,000 per month
...About Etched Etched is building AI chips that are hard-coded for individual model architectures... ...agents. Job Summary Etched’s Inference SW team enables optimal mapping of models... ...seeking a highly skilled and motivated engineer to join our team as we work towards...Full timeWork at officeRelocation package$100k
...is leading the industry on cutting-edge AI technology, revolutionizing performance expectations... ...to unify innovations in software models, compilers, platforms, networking, and... ...experienced and highly skilled Software Engineer with expertise in compilers and semiconductor...Permanent employmentFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Compiler Engineer - AI Inference. Be the first to apply!

