Engineering Manager, Inference Benchmarking — AI Perf
$224k - $356.5kNVIDIA
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an outstanding legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.NVIDIA’s open-source benchmarking platform, AIPerf, is the growing standard for assessing LLM serving performance across various inference frameworks. Hyperscalers, cloud providers, and enterprises use AIPerf to inform decisions on production inference. This includes choosing GPUs, optimizing costs, reducing latency, improving efficiency, and scaling. As Technical Lead Manager, you will lead the engineering team within NVIDIA’s Dynamo organization. Your responsibility is to build and advance the platform so AIPerf becomes the leading benchmarking tool for datacenter, local, and edge use cases. This span LLM, multimodal, diffusion, and computer vision inference. This position combines hands-on leadership with expertise in systems engineering, inference infrastructure, and open-source communities. It has a direct effect on how AI performance is measured and pushed forward.What you'll be doing:Driving the technical roadmap for AIPerf's core infrastructure: load generation, ZMQ-based microservices, GPU telemetry (DCGM/PyNVML, Prometheus metrics, statistical confidence intervals, and Kubernetes-native deployment.Taking ownership for the accuracy and statistical soundness of benchmark results that engineering groups throughout the industry depend on to inform production infrastructure decisions.Advising upstream engine integrations involving vLLM, TRT-LLM, and SGLang in partnership with NVIDIA's Dynamo and NIM teams to maintain AIPerf's relevance across emerging hardware, workload categories, and inference configurations.Hiring, mentoring, and growing a team of senior engineers operating in a high-velocity open-source environment with active external contributors worldwide.What we need to see:Bachelor's degree in Computer Science, Electrical Engineering, or related field, or equivalent experience.8+ overall years of software engineering experience building performance-critical infrastructure, ML tooling, or distributed systems.3+ years of engineering leadership experience as a tech lead, TLM, or engineering manager.Deep understanding of LLM inference mechanics — TTFT, ITL, KV caching, Prefill/Decode, speculative decoding — and the ability to reason about measurement correctness and reproducibility.Proven track record of collaborating across multi-functional groups and delivering production-quality output in high-velocity, high-external-visibility environments.Ways to stand out from the crowd:Extensive experience with vLLM, TRT-LLM or SGLang internals along with contributions to their upstream projects.Experience building Kubernetes-native infrastructure including operators, Helm charts, and GPU observability tooling (DCGM, dcgm-exporter, PyNVML).Background in competitive benchmarking frameworks such as MLPerf or equivalent industry-standard evaluation systems.History leading or making meaningful contributions to active open-source projects with external communities.Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until June 1, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa Clara; US, TX, Austin; US, AL, Remote; US, CO, Remote; US, WA, Remote; US, CA, RemoteType: Full time
$207k - $301k
People Management and Talent Development: Lead, mentor,... ...team of systems and ML engineers. Drive a culture of... ...analysis, profiling, and benchmarking of LLM models on GPU... ...broader organizational AI priorities.Minimum qualifications... ...Cloud (DSC) AI Inference Platform team operates...Suggested$201.3k - $352.3k
...DescriptionIt all started when engineer Fred Luddy wrote code... ..., ServiceNow is the AI control tower for... ..., and product managers with a dual mission. We... ...Manager, Agentic & GenAI Benchmarking and Evaluations to establish... ...efficiency, and inference costs.Lead a High-Performing...SuggestedWork experience placementWork at officeImmediate startRemote workFlexible hoursShift work- ...generation computing experiences—from AI and data centers, to PCs, gaming and... ...ROLEWe are seeking a Principal GenAI Inference Optimization Engineer to join our Models and Applications team... ....- Develop and use profiling, benchmarking, and performance analysis tools for inference...Suggested
$206k - $333k
...The Essential Cloud for AI™. Built for pioneers by... ...for a Principal Engineer to be the technical lead of CoreWeave's Benchmarking & Performance team. You... ...If MLPerf (Training & Inference), Working closely with... ..., and audit trails. Perf Ownership - Lead end-to...SuggestedPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$224k - $356.5k
NVIDIA is seeking an exceptional Manager, Deep Learning Inference Software, to lead a world-class engineering team advancing the state of AI model deployment. You will shape the software powering today’s most sophisticated AI systems — from large language models to multimodal...SuggestedFull timeRemote workWorldwide$168k - $270.25k
...parallel computing, and modern AI. Today, NVIDIA high-... ...GPU programs for training and inference. As AI models, GPU architectures... ...reasoning, and large-scale systems engineering. To address these complex... ...we are seeking an Engineering Manager to spearhead our strategy for...Full time- Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This... ...to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based... ....About The RoleWe're hiring a Principal Engineer for our Inference Cloud Platform. This team...
$224k - $356.5k
...is the platform upon which every new AI-powered application is built. We are... ...seeking a deeply technical software manager to lead production AI inference for NVIDIA Inference Microservices (... ..., combining optimized inference engines, model profiles/recipes, validated runtime...$272k - $431.25k
...into the unlimited potential of AI to define the next era of computing... ....The AI Networking Codesign and Benchmarking R&D group requires a senior software engineer. In this exciting role, you will... ...distributed Deep Learning LLM training and inference. Your primary focus will be...Full timeRemote work$272k - $431.25k
...piece of infrastructure that stores, manages, and serves exabytes of data... ...the storage backbone for NVIDIA's AI infrastructure, enabling researchers and engineers to reliably store massive datasets... ...and accelerating training and inference pipelines.We are seeking a seasoned...Full time- ...TexasNXP is searching for a hands-on AI Compiler Engineer who thrives at the convergence of cutting... ...efficiency.Level up validation, benchmarking, and regression pipelines by harnessing... ...skillsSolid understanding of AI inference workloads (CNNs, transformers, perception...Full timeWork at officeLocal area
$320k
...is a world leader in physical AI, powering self-driving cars,... .... This team is the execution engine behind NVIDIA’s Vision AI strategy... ...robust, low-latency inference at scale. You have led teams... ...customers and partners.Performance Benchmarking: Orchestrate efforts to...Full time$272k - $431.25k
As a Senior Engineering Manager for Agentic Systems & Platform Architecture, you will... ...quality—backed by evaluations, benchmarking, and feedback loopsAssess and... ...and GPU-optimized training and inference workflows.Lead integration of the AI Data Platform into NVIDIA’s on-...Full time$220k - $265k
...Sonatus, we’re driving the transformation to AI-enabled software-defined vehicles.... ...is looking for an experienced Senior Engineering Manager to build and lead our AI Validation function... ...scalable evaluation frameworks and benchmarking platforms for AI systems, including multi...Work at officeWorldwideFlexible hoursShift work$224k - $356.5k
NVIDIA is seeking an Engineering Manager to lead the development of an agentic platform for observing... ...visibility into model behavior, inference performance, reliability, and cost across... ...production.You will work across the NVIDIA AI software stack with teams focused on...Full time$224k - $356.5k
...we aren't just powering the AI revolution—we're accelerating... ...it. We are accelerating LLM inference across the stack and across... ...a highly skilled and driven Engineering Manager to take the lead in accelerating... ...models. Work with inference benchmark teams to help tune...Full time$272k - $431.25k
...and secure for millions globally. We seek a Senior Engineer to lead technical efforts in deploying advanced AI agent frameworks and local runtimes on Windows and... ...on consumer PCs. By combining powerful local inference (Nemotron models) with strong privacy routers and...Full timeLocal areaShift work$219k - $351k
..., and communities.Job Title: Principal engineer, AI Serving Framework Architect (Software)The... ...methodologies for maximizing AI inference performance in multi-rack scale memory-... ...by our human recruiting team and hiring managers to ensure every candidate is evaluated...Work at officeFlexible hours- ...DescriptionWe are seeking an Application Engineering Manager to lead a customer-facing systems and... ...supporting high-performance compute (HPC), AI, hyperscaler, and advanced digital... ...engagements from requirements definition, benchmarks, and bring-up through production...
$272k - $431.25k
...NVIDIA IT’s Enterprise AI & Automation team to... ...business results across engineering, IT, supply chain,... ...from Kubernetes to GPU inference stacks and translate new... ...quality through telemetry, benchmarking, automated evaluation,... ...models, multi-agent management (e.g., LangChain,...Full time$195.2k - $391.2k
...OpportunityWe are looking for a Senior Engineering Manager to lead the design, development, and scaling... ...will serve as the foundation for AI/ML workloads, GPU infrastructure, and enterprise... ...space of GPU scheduling, training, and inference systemsExecution & Operational...Work at officeLocal areaRemote workRelocation package3 days per week$296.3k
...times a week, at minimum.The Role:We are seeking a Principal AI Engineer to lead the design and advancement of our AI platform. You will... ...the infrastructure that powers large-scale training and cloud inference. This includes accelerating training throughput, scaling multi...Full timeLocal areaRemote workWork from homeFlexible hours$272k - $431.25k
...NVIDIA is looking for a Machine Learning (ML) Engineer to join the GPU accelerated Apache Spark... ...ETL, SQL, and ML/DL model training and inference pipelines, spanning many domains and use... ...with GPUs. You will apply the latest ML/AI methods to empower enterprises to migrate...Full time$156.82k - $234.9k
...Across enterprise, cloud and AI, and carrier architectures, our... ...a Sr. Principal Product Engineer, you will play a critical leadership... ...-generation AI training and inference infrastructure.High... ...collaboration, and stakeholder management skills.Preferred QualificationsExperience...Permanent employmentFull timeInternshipWork from home$146.7k - $339.3k
...network infrastructure and data center operations supporting over 10,000 employees and AI/ML training infrastructure. The manager oversees a team of 12 globally distributed network engineers and data center technicians. The team is operating in a follow-the-sun model to...Full timeWork at officeRemote work$207k - $301k
Lead a team of software engineers to design, build, deploy, and in some cases, operate the critical software systems that directly manage the global data center networking infrastructure.Cultivate... ...disaster recovery protocols.The AI and Infrastructure team is redefining...Worldwide$250.6k - $362.6k
...innovative hardware platforms central to the AI era, powering Cisco’s core Switching,... ..., floorplan suggestions.Evaluate and benchmark LLM-generated outputs against ground truth... ...QualificationsBachelor’s degree in Engineering and 15+ years of ASIC related experience...Full timeTemporary workLocal areaFlexible hours$255.85k - $361.2k
...As a Software Enabling and Optimization Engineer, you will play a pivotal role in driving... ...differentiation across enterprise, cloud, client,, AI, gaming, graphics and analytics. Your... ...collateral, training materials, and benchmarking tools to highlight Intel's competitive...Full timeLocal areaImmediate startShift work$224k - $356.5k
...across the entire processing pipeline. We are seeking a seasoned Engineering Manager to lead our rapidly growing NvStreams team. If you are... ...1, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering...Full time$219k - $351k
..., and communities.Job Title: Principal Engineer, AI System Architect (Hardware)The Architecture... ..., DLRMs, and large-scale training and inference systems.Demonstrated ability to... ...by our human recruiting team and hiring managers to ensure every candidate is evaluated...Work at officeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Engineering Manager, Inference Benchmarking — AI Perf. Be the first to apply!

