Senior AI Inference Platform PM Performance & Scale
NVIDIA
NVIDIA is seeking a highly technical Product Manager to own AI inference optimization on NVIDIA hardware, from single-GPU workstations to large data centers. You will translate deep optimization techniques into capabilities that customers can adopt across the inference stack. The role spans framework integration with TensorRT-LLM, vLLM, SGLang, and NVIDIA Dynamo, plus benchmarking, release readiness, and cross-functional collaboration to deliver measurable performance gains and compelling #J-18808-Ljbffr NVIDIA
- NVIDIA seeks a Product Manager for Inference to enable developers to deploy high-performance AI on NVIDIA GPUs. You will shape product strategy, roadmaps, and go-to-market plans, collaborating with internal teams and external developers to build model-optimization software...PlatformSeniorPerformance
- NVIDIA Corporation in Santa Clara, CA is seeking a Senior Product Manager for AI Platform Inference to lead the development of tools, SDKs, and libraries... ...collaborating with developers to optimize model deployment performance, with strong emphasis on GenAI concepts and GPU-...PlatformSeniorPerformance
- ...Santa Clara is seeking a highly technical Product Manager to own AI inference products that optimize latency, throughput, and cost per... ...deep optimization techniques into shipped capabilities, set performance strategy for agentic workloads, and partner with TensorRT-LLM...SeniorPerformance
- ...builds the world’s largest AI chip, 56 times larger... ...-leading training and inference speeds; over 10 times... ...deploy 750 megawatts of scale, transforming key workloads... ...infrastructure, high-performance kernel enablement, and... ...engineers, cloud platform teams, AI researchers,...PlatformSeniorPerformance
$184k - $287.5k
...stack for distributed inference which will be used to... ...the world by applying AI inference aware technology... ...them to AI-Grid platforms and SDKs by providing... ...customers design high-performance and secure workload aware... ...deploying networks at scale as a Software Engineer...PlatformSeniorPerformanceFull timeWork experience placement- ...builds the world's largest AI chip, 56 times larger... ...-leading training and inference speeds; over 10 times... ...deploy 750 megawatts of scale, transforming key workloads... ...for an Inference Platform SDET to join the Inference... ...to validate platform performance, scalability, and reliability...PlatformSeniorPerformanceWork at office
$152k - $241.5k
...are now looking for a Senior Software Engineer for Deep Learning Inference! Would you like to... ...software that can be scaled to multiple platforms for functionality and... ...NVIDIA’s SDK for high-performance deep learning inference... ...vacancy. NVIDIA uses AI tools in its...PlatformSeniorPerformanceFull time$208k - $327.75k
...unlimited potential of AI to define the next era... ...the best possible performance from AI models and applications... ...hardware. Every inference deployment — from a single... ...what we retire.Build platforms, not one-offs. Deliver... ...experience at scale counts too: capacity planning...PlatformSeniorPerformanceFull time$184k - $287.5k
...software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency. You’ll architect and implement high-performance inference stacks, optimize GPU... .../CRI-O/CRIU.Experience with cloud platforms (AWS/GCP/Azure), infrastructure as...PlatformSeniorPerformanceFull time$184k - $287.5k
...unlimited potential of AI to define the next... ...re searching for a Senior Systems Software... ...containers, and systems performance and scalability.... ..., distributed inference serving, and major cloud platforms. You’ll own hard technical... ...problems at large scale and help shape how...PlatformSeniorPerformanceFull timeRemote work$182.5k - $260.5k
...for the cloud and AI era. We secure and... ...the Netskope One platform, its Zero Trust Engine... ...control without performance trade-offs.At... ...are available at Senior Staff and above. Candidates... ..., you own the inference and optimization layer... ...constraints.Real scale to build against....PlatformSeniorPerformance$152k - $241.5k
...Intelligence, High-Performance Computing and... ...highly motivated Senior Software Engineers... ...focus on NVLink Rack-Scale Systems Stability... ...first-of-their-kind platforms into stable, reliable... ..., and large-scale AI infrastructure,... ...distributed training/inference systems.Experience...PlatformSeniorPerformanceFull timeRemote work$193.3k - $261.5k
...enabling unparalleled ML inference and training performance.The Inference... ...what's possible in AI acceleration.As part... ...peak performance at scale for customers and developers... ...and mentorship. Our senior members enjoy one-on... ...TensorRT or similar platforms in production...PlatformSeniorPerformanceWork experience placementInternshipLocal areaFlexible hours- ...builds the world's largest AI chip, 56 times larger... ...GPUs. Our novel wafer‑scale architecture provides... ...industry‑leading training and inference speeds and empowers... ..., solution briefs, performance analyses, and architecture... ...a breakthrough AI platform beyond the constraints...PlatformSeniorPerformance
$165k - $242k
...Apply for the Senior Software Engineer II, Inference role at CoreWeave. CoreWeave... ...Essential Cloud for AI™. Built for pioneers... ...delivers a platform of technology, tools... ...innovators to build and scale AI with confidence.... ...superior infrastructure performance with deep technical...PlatformSeniorPerformancePermanent employmentTemporary workCasual workWork at officeRemote workFlexible hoursShift work$168k - $258.75k
Inference is the fastest growing and most competitive area in Generative AI today. It is where AI models impact our daily... ...ever bit of accuracy and performance matters for quality,... ...deployment techniques. As a Senior Product Manager for AI Platform Inference you will be...PlatformSeniorPerformanceFull time- ...Accelloris is an AI-native services firm focused on... ...operationalizing AI at scale, delivering measurable business... ...Architect — AI Systems, Inference & Platform Internals is sought to... ...workflows. This senior role emphasizes GPU-level performance, distributed inference,...PlatformSeniorPerformance
$152k - $241.5k
...driving advancements in AI and machine learning to... ...-leading deep learning inference software for NVIDIA AI accelerators. As a Senior Software Engineer in... ...Knowledge of close-to-metal performance analysis, optimization... ...-effective computing platform driving our success in...PlatformSeniorPerformanceFull time$224k - $356.5k
We are now looking for a Senior System Software Engineer to work... ...to power a revolution in AI, enabling breakthroughs in problems... ...team building Generative AI inference platform to make design and... ...build robust, scalable, high performance software components to support...PlatformSeniorPerformanceFull time$148k - $235.75k
We are looking for a Senior Technical Product Marketing... ...and pivotal in our inference marketing. You will be... ...position in AI inference.Want to join... ...intelligence and high performance computing. Come grow your... ...drive NVIDIA’s inference platform technical go-to-market...PlatformSeniorPerformanceFull time$184k - $287.5k
...into the unlimited potential of AI to define the next era of... ...a dedicated engineer for the Senior Systems Software Engineer role, focusing on GPU Performance at Scale. At NVIDIA, this role is uniquely... ...optimize large-scale performance platforms.What you'll be doing:Lead the...PlatformSeniorPerformanceFull timeRemote work- ...NVIDIA Corporation is seeking a Senior Software Engineer for the TensorRT Edge-LLM... ...team in the US. You will develop a high-performance inference framework in modern C++ that extends... ...management, within embedded and edge platforms. You will collaborate across CUDA and...PlatformSeniorPerformance
- NVIDIA in Santa Clara, CA, seeks a senior product leader to own the inference performance roadmap, shaping how models are represented, memory/state is managed... ...and tokens are generated. You will build scalable platforms across model families and deployment topologies for...PlatformSeniorPerformance
- ...the world's largest AI chip, 56 times... ...leading training and inference speeds; over 10... ...750 megawatts of scale, transforming key... ...models ship, how they perform, and how the world... ...every model on our platform delivers... ...above the level of Senior PM.5+ years of total...PlatformSeniorPerformanceWork experience placementWork at officeRemote workShift work
- NVIDIA is seeking a Senior Product Manager for AI Inference Performance (Finance) to own optimization strategies across the inference stack and drive platform capabilities for diverse deployments. You will translate deep optimization into broadly adoptable products, collaborate...PlatformSeniorPerformance
$184k - $287.5k
NVIDIA has become the platform upon which every new AI-powered application is built. We are seeking a Sr. HPC Performance engineer to join our team of scientists and engineers passionate... ...performant features for large scale, CUDA-backed ML training frameworks, using...PlatformSeniorPerformanceFull time$193.93k - $352.29k
...profound opportunity for AI to drive positive... ...building a universal autonomy platform: self-driving for all... ...to deploy autonomy at scale, from robotaxis and logistics... ...an in-house ML inference platform to serve large... ...with a high standard for performance, scalability, and code...PlatformSeniorPerformanceImmediate startFlexible hours$152k - $241.5k
...into the unlimited potential of AI to define the next era of... ...Join a team that analyzes large-scale datacenter workloads on GPU-accelerated... ...to find application and platform improvement opportunities.Work... ...and HPC / large-scale or performance-sensitive environmentsExperience...PlatformSeniorPerformanceFull timeRemote work$152k - $241.5k
...recently, GPU deep learning ignited modern AI — the next era of computing — with the... ...DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers, personal... ...compiler must deliver leading inference performance, fast build time, reduced memory...PlatformSeniorPerformanceFull timeRemote work- NVIDIA is seeking a Senior Product Manager for AI Platform Inference to lead the development of tools, SDKs, and libraries that enable developers to deploy inference workloads efficiently on NVIDIA GPUs. You will craft product strategy, roadmaps, and go-to-market plans...PlatformSenior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior AI Inference Platform PM Performance & Scale. Be the first to apply!
- senior lead project manager Santa Clara, CA
- senior robotics software engineer Santa Clara, CA
- senior devops engineer remote Santa Clara, CA
- senior sas administrator Santa Clara, CA
- senior IT manager Santa Clara, CA
- sr project manager Santa Clara, CA
- senior windows systems engineer Santa Clara, CA
- senior researcher Santa Clara, CA
- senior manager data science Santa Clara, CA
- senior principal engineer Santa Clara, CA

