Compiler Engineer - AI Inference
$152k - $241.5kNVIDIA Gruppe
NVIDIA's invention of the GPU 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI — the next era of computing — with the GPU acting as the brain of computers, robots, and self‑driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company”. NVIDIA is seeking top‑tier AI Compiler Engineers to drive innovation within our world‑class compiler organization. In this role, you will push the boundaries of what is possible in AI performance and help build the technology that powers the next generation of computing. Join us and make a tangible impact on a global scale. What you’ll be doing: Drive technical innovation: Participating in hands‑on development focusing on kernel generation and computational graph optimizations for next‑generation NVIDIA GPUs. Advance the state‑of‑the‑art: Solve complex compilation problems for AI workloads (both inference and training) and successfully transition these breakthroughs into enterprise and consumer products. Collaborate on hardware/software co‑design: Partner with leading experts across our software, hardware, and research divisions to architect and co‑design future silicon. Scale AI to the datacenter: Participating in the advancement and optimization of datacenter‑scale AI workload deployments. What we need to see: BS or MS in Computer Science, Computer Engineering, or a related field (or equivalent experience). A PhD is strongly preferred. Compiler Experience: 3+ years of relevant industry experience specializing in compiler optimizations, synthesis, and placement. MLIR Knowledge: Demonstrated, hands‑on experience working with MLIR. Programming Excellence: Exceptional C/C++ and Python programming and software design skills, including rigorous debugging, performance analysis, and test design. Team Dynamics: Strong communication and interpersonal skills, with the ability to collaborate effectively in a dynamic, fast‑paced, and product‑oriented environment. Ways to stand out from the crowd: Hardware Implementation: Hands‑on experience implementing complex AI workloads on CPU, GPU, and/or custom AI accelerator architectures. LLM Knowledge: Deep understanding of Large Language Model (LLM) inference and its profound implications on computer architecture. Architecture & Design: Demonstrated understanding in the designing and architecting of comprehensive compiler frameworks from the ground up. With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward‑thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing. If you’re a creative and autonomous engineer with a real passion for technology, we want to hear from you. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is $152,000 USD - $241,500 USD. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until April 28, 2026. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law. #J-18808-Ljbffr NVIDIA Gruppe
$184k - $287.5k
We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency... ...inference stacks, optimize GPU kernels and compilers, drive industry benchmarks, and scale workloads across...SuggestedFull time$152k - $241.5k
...recently, GPU deep learning ignited modern AI — the next era of computing — with the... ...are looking for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers... ...DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers, personal...SuggestedFull timeRemote work$152k - $241.5k
...eager to work on cutting-edge AI technology for safety-... ...TensorRT team as a Senior Software Engineer, and be at the forefront of technology... ...enabling high-performance AI inference solutions for automotive... ...into TensorRT's compiler and runtime for specialized and...SuggestedFull time- ...headquartered in Santa Clara, CA, seeks a Principal System Software Engineer for AI Inference Execution. You will join the software team to productize... ...developing deployment software and collaborating with ML, compiler, and hardware experts. Required: strong background in...Suggested
- NVIDIA Gruppe in Santa Clara, California is seeking AI Compiler Engineers to drive technological innovation within their compiler organization. The role involves working on kernel generation and optimization for next-generation NVIDIA GPUs and solving complex compilation...Suggested
$184k - $287.5k
...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts... ...can make a lasting impact on the world.We are seeking an AI Compiler Engineer with deep expertise in compiler technologies to join our team....Full timeRemote work$165.2k - $223.6k
...comprehensive toolkit includes an ML compiler, runtime, and application... ...JAX enabling unparalleled ML inference and training performance.The... ...-software boundary, our engineers build systematic infrastructure... ...boundaries of what's possible in AI acceleration.As part of the...Work experience placementInternshipLocal areaFlexible hours$193.3k - $261.5k
...fast on the Trainium hardware.As a Sr. Software Development Engineer on the Inference Model Enablement team, you will onboard and optimize state... ...issues across the stack in collaboration with the compiler and runtime teams* Mentoring junior engineers on system design...InternshipLocal areaFlexible hours- ...generation computing experiences—from AI and data centers, to PCs,... ...the next generation of compiler and software infrastructure to... ...Principal Software Development Engineer to lead technical development... ...movement optimization for ML inference workloads• Define and...
- ...high-performance computing, cloud, and AI. Whether you’re designing next-gen processors... ...and technically exceptional PMTS AI/ML Compiler Engineer to join AMD's AI Software organization.... ...state-of-the-art AI training and inference workloads on AMD GPU platforms. You will...
$184k - $287.5k
...forefront of the generative AI revolution! The Algorithmic Model... ...diffusion models for maximal inference efficiency using techniques... ...externally by research and engineering teams alike developing best-in... ...TorchDynamo, torch.export, torch.compile, etc...) to analyze and...Full time- NVIDIA is seeking a Senior Agentic AI Software Engineer (Finance) to build agentic systems and advance AI inference workloads in real-world finance contexts. You will design and implement scalable software that drives experimental agents, optimize performance, and contribute...
$152k - $241.5k
...recently, GPU deep learning ignited modern AI — the next era of computing — with the... ...company”.NVIDIA is hiring a Senior AI Compiler Engineer. GPUs are driving rapid progress in... ...based AI compiler that powers NVIDIA’s inference engine end to end, with a focus on performance...Full timeRemote work$152k - $241.5k
...tapping into the unlimited potential of AI to define the next era of computing. An... ...are looking for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for... ...DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers, personal...Full timeRemote work- ...unleashing the potential of generative AI to power the transformation of technology... ....The role: Principal System Software Engineer, AI Inference ExecutionWhat you will do:The role requires... ...closely with other software (ML and compilers) and hardware experts in the company....3 days per week
- NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models...
- CoreWeave is hiring a Senior Engineer for its Benchmarking & Performance team to write, profile, and optimize GPU kernels on the LLM inference path. You will improve latency and throughput and collaborate with product, orchestration, and hardware teams to achieve strict...
- NVIDIA AI in Santa Clara is seeking a highly capable software engineer to advance an advanced inference framework using modern C++. The role focuses on extending TensorRT with autoregressive... ...across CUDA, kernel libraries, compilers, and robotics teams to deliver high-...
- NVIDIA is seeking a Senior Agentic AI Software Engineer to advance agentic AI systems and workloads from scalable research to production-grade solutions. You will build agentic components, analyze inference dynamics, and collaborate with teams owning evaluation pipelines...
$152k - $241.5k
## Senior Software Engineer, Deep Learning Inference - TensorRTApplylocations: US, CA, Santa Claratime type:... ...profiling, and optimizations.* Background in compiler development* Experience in working... ...for an existing vacancy.NVIDIA uses AI tools in its recruiting processes....- ...California or Austin, TexasNXP is searching for a hands-on AI Compiler Engineer who thrives at the convergence of cutting-edge AI, compiler... ...Excellent C/C++ and Python skillsSolid understanding of AI inference workloads (CNNs, transformers, perception or generative models...Full timeWork at officeLocal area
$124k - $195.5k
...hardware performance for emerging AI workloads. You will be a... ...bottlenecks in both training and inference pipelines.Collaborate closely... ...SW architects, kernel and compiler authors and CUDA driver... ...in Computer Science, Computer Engineering, Electrical Engineering, or related...Full time$150k - $275k
A cutting-edge tech company in San Jose is seeking a Supercomputing Engineer to ensure the reliability of its inference servers. This role involves designing and executing test suites, analyzing performance, and collaborating with engineering teams. Ideal candidates will...$152k - $241.5k
...Senior Machine Learning Applications and Compiler Engineer!NVIDIA is seeking engineers to develop... ...and optimizations for our LPX inference and compiler stack. You will work at the... ...or similar.Experience with large-scale AI distributed inference or training systems...Full time- ...Systems Software Engineer — Marvis Minis & Edge AIThe future of networking is autonomous and AI-driven. HPE Mist Networking is building... ...embedded Linux development — cross-compilation, on-device debugging,... ...with AI/ML concepts — model inference, data preprocessing, or signal...
$152k - $241.5k
NVIDIA's GPUs are at the core of modern AI infrastructure, from training large-scale models to running inference in production. That position depends on software as much as hardware, and compiler engineering is a big part of what makes it work. We are looking for outstanding...- ...generation computing experiences—from AI and data centers, to PCs,... ...ROLEWe are hiring a senior engineering leader to define, build, and... ...encode, pre/post-processing, inference, tracking, streaming,... .../pip (where relevant), cross‑compile toolchains, reproducible builds...
$152k - $241.5k
...deep learning ignited modern AI — the next era of computing —... ...looking for versatile software engineers for our XLA team. NVIDIA is... ...doing:In this role, develop compiler optimization algorithms for deep... ...workloads. You will optimize inference and training performance for...Full timeRemote work- ...scientific discovery to powering AI and the technologies people... ...knowledge platform that lets engineers ask natural-language... ...systems, evaluation, and efficient inference, and evaluate promising techniques... ...microarchitecture, firmware, compilers, or chip verification....Worldwide
$184k - $287.5k
We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to... ...assembly layer. Our team works closely with compiler, kernel, hardware, and framework... ...agentic kernel optimization: applying AI-driven analysis to diagnose performance...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Compiler Engineer - AI Inference. Be the first to apply!

