Engineering Manager, Inference Benchmarking — AI Perf
$224k - $356.5kNVIDIA
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an outstanding legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.NVIDIA’s open-source benchmarking platform, AIPerf, is the growing standard for assessing LLM serving performance across various inference frameworks. Hyperscalers, cloud providers, and enterprises use AIPerf to inform decisions on production inference. This includes choosing GPUs, optimizing costs, reducing latency, improving efficiency, and scaling. As Technical Lead Manager, you will lead the engineering team within NVIDIA’s Dynamo organization. Your responsibility is to build and advance the platform so AIPerf becomes the leading benchmarking tool for datacenter, local, and edge use cases. This span LLM, multimodal, diffusion, and computer vision inference. This position combines hands-on leadership with expertise in systems engineering, inference infrastructure, and open-source communities. It has a direct effect on how AI performance is measured and pushed forward.What you'll be doing:Driving the technical roadmap for AIPerf's core infrastructure: load generation, ZMQ-based microservices, GPU telemetry (DCGM/PyNVML, Prometheus metrics, statistical confidence intervals, and Kubernetes-native deployment.Taking ownership for the accuracy and statistical soundness of benchmark results that engineering groups throughout the industry depend on to inform production infrastructure decisions.Advising upstream engine integrations involving vLLM, TRT-LLM, and SGLang in partnership with NVIDIA's Dynamo and NIM teams to maintain AIPerf's relevance across emerging hardware, workload categories, and inference configurations.Hiring, mentoring, and growing a team of senior engineers operating in a high-velocity open-source environment with active external contributors worldwide.What we need to see:Bachelor's degree in Computer Science, Electrical Engineering, or related field, or equivalent experience.8+ overall years of software engineering experience building performance-critical infrastructure, ML tooling, or distributed systems.3+ years of engineering leadership experience as a tech lead, TLM, or engineering manager.Deep understanding of LLM inference mechanics — TTFT, ITL, KV caching, Prefill/Decode, speculative decoding — and the ability to reason about measurement correctness and reproducibility.Proven track record of collaborating across multi-functional groups and delivering production-quality output in high-velocity, high-external-visibility environments.Ways to stand out from the crowd:Extensive experience with vLLM, TRT-LLM or SGLang internals along with contributions to their upstream projects.Experience building Kubernetes-native infrastructure including operators, Helm charts, and GPU observability tooling (DCGM, dcgm-exporter, PyNVML).Background in competitive benchmarking frameworks such as MLPerf or equivalent industry-standard evaluation systems.History leading or making meaningful contributions to active open-source projects with external communities.Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until June 1, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa Clara; US, TX, Austin; US, AL, Remote; US, CO, Remote; US, WA, Remote; US, CA, RemoteType: Full time
$206k - $333k
...The Essential Cloud for AI™. Built for pioneers by... ...for a Principal Engineer to be the technical lead of CoreWeave's Benchmarking & Performance team. You... ...If MLPerf (Training & Inference), Working closely with... ..., and audit trails. Perf Ownership - Lead end-to...SuggestedPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$201.3k - $352.3k
...DescriptionIt all started when engineer Fred Luddy wrote code... ..., ServiceNow is the AI control tower for... ..., and product managers with a dual mission. We... ...Manager, Agentic & GenAI Benchmarking and Evaluations to establish... ...efficiency, and inference costs.Lead a High-Performing...SuggestedWork experience placementWork at officeImmediate startRemote workFlexible hoursShift work- ...generation computing experiences—from AI and data centers, to PCs, gaming and... ...ROLEWe are seeking a Principal GenAI Inference Optimization Engineer to join our Models and Applications team... ....- Develop and use profiling, benchmarking, and performance analysis tools for inference...Suggested
$224k - $356.5k
NVIDIA is seeking an exceptional Manager, Deep Learning Inference Software, to lead a world-class engineering team advancing the state of AI model deployment. You will shape the software powering today’s most sophisticated AI systems — from large language models to multimodal...SuggestedFull timeRemote workWorldwide$240k - $290k
...Sonatus, we're driving the transformation to AI-enabled software-defined vehicles.... ...is looking for an experienced Senior Engineering Manager to build and lead our AI Validation function... ...scalable evaluation frameworks and benchmarking platforms for AI systems, including multi...SuggestedWork at officeWorldwideFlexible hoursShift work$224k - $356.5k
NVIDIA is seeking an Engineering Manager to lead the development of an agentic platform for observing... ...visibility into model behavior, inference performance, reliability, and cost across... ...production.You will work across the NVIDIA AI software stack with teams focused on...Full time$207k - $300k
...years of experience in a people management or team leadership role.... ...qualifications:Master’s degree or PhD in Engineering, Computer Science, or a... ...supporting training and inference across Alphabet. You will support... ...business delivery needs.The AI and Infrastructure team is...Immediate startWorldwide$168k - $270.25k
...parallel computing, and modern AI. Today, NVIDIA high-... ...GPU programs for training and inference. As AI models, GPU architectures... ...reasoning, and large-scale systems engineering. To address these complex... ...we are seeking an Engineering Manager to spearhead our strategy for...Full time- ...intelligence . As the only vertically integrated AI infrastructure company built from the... ...Crusoe. About the Role: As an Engineering Manager on the Managed AI team at Crusoe, you... ...with CPU & GPU performance, inference frameworks, or LLM systems is a strong...Temporary workWork at office
- Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This... ...to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based... ....About The RoleWe're hiring a Principal Engineer for our Inference Cloud Platform. This team...
$114.6k - $234.6k
...innovations to life-saving care. And with AI embedded across our products and... ...topology-aware measurement, analytics, and inference systems for data center and WAN networks... ...operational analysis while maintaining engineering, security, and quality standards.Write robust...Temporary workFlexible hours$140k - $215k
...’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems,... ...you.About the Role:CrowdStrike is seeking an experienced SDET Engineering Manager to lead our quality engineering efforts within the AI Detection...Full timeContract workWork experience placementWork at officeLocal areaWorldwide2 days per week3 days per week$272k - $431.25k
...NVIDIA IT’s Enterprise AI & Automation team to... ...business results across engineering, IT, supply chain,... ...from Kubernetes to GPU inference stacks and translate new... ...quality through telemetry, benchmarking, automated evaluation,... ...models, multi-agent management (e.g., LangChain,...Full time$206.4k - $379.1k
...Foundry is Adobe's enterprise managed-service offering for custom multimedia generative AI — deep-tuned image, video, and... ...hiring a Principal Machine Learning Engineer to serve as the technical lead... ...scale. You will set the inference architecture and technical standards...Full timeTemporary workLocal areaWorldwide$296.3k
...times a week, at minimum.The Role:We are seeking a Principal AI Engineer to lead the design and advancement of our AI platform. You will... ...the infrastructure that powers large-scale training and cloud inference. This includes accelerating training throughput, scaling multi...Full timeLocal areaRemote workWork from homeFlexible hours- ...NVIDIA is looking for a creative SW engineering manager to lead the Spectrum-X team, delivering an end-to-end Ethernet networking solution for AI-scale deployments. You will manage a US-based engineering group, engage directly with customers, and collaborate with engineers...
$278.1k - $417.1k
...building the next generation of AI-driven game experiences,... .... As our Principal Engineer for On-Device AI Inference & Systems, you will be the... ...path, and frame-budget management alongside the renderer. Architect... ...and automated on-device benchmarking in CI. Research...Work at officeWorldwideRelocation package$272k - $431.25k
...into the unlimited potential of AI to define the next era of computing... ....The AI Networking Codesign and Benchmarking R&D group requires a senior software engineer. In this exciting role, you will... ...distributed Deep Learning LLM training and inference. Your primary focus will be...Full timeRemote work$220k - $300k
Senior Principal Machine Learning Engineer San Jose, California, United States The era of pervasive AI has arrived. In this era,... ...architecture, training and fine-tuning, inference optimization, evaluation, and... ...direction without direct management authority Track record of...Full timeTemporary workLocal areaFlexible hours$182k - $242k
...Description Job Description CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a... ...more at What You'll Do CoreWeave is looking for an Engineering Manager to lead a team building and operating Kubernetes infrastructure...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$171k - $247k
...DFx.Provide pre-production build support, manage line bring-up, deliver product debug... ...new product technology, collaborating with Engineering and NPI Operations to influence design, highlight... ...will sit at the epicenter of Google's AI infrastructure ambitions, bridging the...Contract workWorldwide$180k - $225k
...Senior Manager Software Engineering San Jose, California, USA Zscaler accelerates digital transformation... ...the future of work is Human + AI and are building an AI-native enterprise... ...Zscaler's salary ranges are benchmarked and are determined by role and level....Full timeWork at officeLocal area$292k
...Director, Software Engineering (Finance) In this role, you will own the enterprise platforms... ..., deployment, operation, and scaling of AI agents and GPU-accelerated workloads for... ...for the AI Factory, offering scalable AI inference and secure environments for agents to...- ...investing deeply in Generative AI — pioneering advanced... ...Principal Machine Learning Systems Engineer (P60) to lead technical directions... ...large-scale model training, inference pipelines, or search/... ...NDCG, groundedness, latency benchmarks).Experience with monitoring,...Work at officeLocal area
$245k - $325k
...Software Engineering Manager for SambaStack Platform SambaNova is hiring a Software Engineering Manager for the SambaStack platform. We help enterprises and service providers host their own AI inference platforms, powered by our state-of-the-art RDU (Reconfigurable...Full timeTemporary workLocal areaFlexible hours$248k - $396.75k
Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline... ..., databases, capacity management, continuous delivery, and observability... ...direction of NVIDIA’s AI Platform Runtime and lead reliability... ..., GPU infrastructure, inference systems, training environments...Full time- ...We Are Synopsys is the leader in engineering solutions from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon... ...to expand Prime Power RTL footprint, lead benchmark engagements that influence competitive...
$220k - $245k
...Title: Sr Data Engineering Manager This position is based in our Campbell, California offices. This position is on-site, full-time.... ...foundation of Imperative Care’s modern data architecture and AI capability. This includes technology such as Modern Data Platform...Full timeContract workWork experience placement$144.6k - $263.8k
...Senior Engineering Program Manager, On-Device MLAt Apple, we work every day to build products that enrich people's lives. The On-Device Machine... ...or engineering experience in Machine Learning training or inference technologies.Experience using objectives and key results to...Relocation$200k
...Senior Engineering Program ManagerVelaura is designing the compute silicon... .... As Senior Engineering Program Manager, you will drive execution of our Physical AI System-on-Chip (SoC) program across... ...and transparency and regularly benchmarks compensation to ensure we remain...Contract workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Engineering Manager, Inference Benchmarking — AI Perf. Be the first to apply!


