Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Technical Lead Manager, AI Inference Benchmarking

Jobleads-US

NVIDIA in Santa Clara, CA is seeking a Technical Lead Manager to lead the AIPerf benchmark platform engineering. You will guide the core infrastructure, oversee load generation, GPU telemetry, and Kubernetes deployments, and ensure data-driven, reproducible results for industry partners.

You will mentor senior engineers, collaborate with Dynamo and NIM teams, and drive open-source contributions while aligning with the company's advanced AI hardware and software efforts.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 15 hours ago
Similar jobs that could be interesting for youBased on the Technical Lead Manager, AI Inference Benchmarking in Santa Clara, CA vacancy
  • $184k - $287.5k

     ...upon which every new AI-powered...  ...Software Engineer - AI Inference Performance to advance...  ...you'll be doing:Lead end-to-end analysis...  ...batching, KV-cache management, quantization,...  ...Establish repeatable benchmarks, canonical run records...  ...presentations, technical talks, or blog... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $182k - $242k

     ...The Essential Cloud for AI™. Built for pioneers...  ...confidence. Trusted by leading AI labs, startups, and...  ...performance with deep technical expertise to...  ...Engineer for CoreWeave's Benchmarking & Performance team. You...  ...end MLPerf Training and Inference runs, including workload... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    a month ago
  • $240k - $333k

     ...Technical Lead Manager, Machine Learning, Memory Subsystem Design Share Technical Lead Manager, Machine Learning, Memory Subsystem Design...  ...including DRAM manufacturers and third party IP providers. The AI and Infrastructure team is redefining what's possible. We empower... 
    Suggested
    Worldwide

    Google LLC

    Sunnyvale, CA
    16 hours ago
  • $280k - $304k

     ...providers modernize and manage infrastructure for the AI era without adding...  ...experience testing AI inference workloads: GPU-based...  ..., throughput/latency benchmarking, and inference accuracy...  ...experience leading test architecture and technical direction across multiple... 
    Suggested
    Full time
    Contract work
    Work at office
    Immediate start
    Remote work
    Work visa
    Flexible hours

    Spectro Cloud

    San Jose, CA
    10 days ago
  • $197.3k - $245.6k

    AI Engineer 4 (AI Foundations: Benchmarking, Evaluation, and Explainability) At Capital...  ...deliver our industry leading capabilities with breakthrough...  ...research scientists, technical program managers, and product managers...  ...large language model inference, agents and multi-... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    3 days ago
  • $194k

     ...unparalleled reputation for being leading and reliable force in...  ...Team The Search AI Product & Mobile team...  ...We're looking for a Manager - Mobile Enginerring,...  ...team members), drive the technical direction of the team...  ...of ML models (on-device inference, server-side ranking... 
    Temporary work
    Flexible hours

    Coupang, Inc.

    Mountain View, CA
    1 day ago
  •  ...NVIDIA is seeking a Software Engineer to bring up, triage, benchmark, and optimize distributed training and inference workloads across GPU platforms at scale. You will work on multi-GPU and multi-node LLM workloads, build benchmarking tooling, and collaborate with... 

    Jobleads-US

    Santa Clara, CA
    1 day ago
  • $207k - $300k

    Analyze and optimize AI inference workloads across the application, model, and distributed...  ...Engineering, Applied Mathematics, or a related technical field, or equivalent practical...  .... By combining deep systems profiling, benchmarking, and first-principles problem solving,... 

    Google

    Mountain View, CA
    19 hours ago
  • $193.3k - $261.5k

     ...Senior Software Development Engineer, AI/ML, AWS Neuron, Model Inference Job ID: 10505474 | Amazon.com...  ...scientists, system engineers, and product managers to deliver state-of-the-art...  ...drive business decisions with your technical input. You will work in a startup-like... 
    Work experience placement
    Internship
    Flexible hours

    Amazon Inc.

    Cupertino, CA
    16 hours ago
  •  ...discovery to powering AI and the technologies...  ...breakthroughs, or bringing leading edge products to...  ...:We are seeking a Technical Program Manager to drive execution across...  ...learning, inference, evaluation infrastructure...  ...initiatives involving benchmarking, evaluation, experiment... 
    Shift work

    AMD

    Santa Clara, CA
    19 hours ago
  • $165.2k - $223.6k

     ...at the right moment. As AI models outgrow any single...  ...frontier of disaggregated inference: splitting LLM serving into...  ...to grow your scope and technical depth — on a team at the leading edge of AI/ML infrastructure...  ...reviews, source control management, build processes, testing... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    19 hours ago
  • $235.03k - $352.29k

     ...and profound opportunity for AI to drive positive change in the...  ...Fidelity, T. Rowe Price, and other leading investors.About the TeamL4...  ...deployment.Lead and grow a technical team dedicated to this problem...  ...technical projects and mentoring or managing engineers; able to set... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    2 days ago
  • $207k - $300k

     ...including the adoption of new AI features, user engagement metrics...  ..., and subscription products.Lead the design and delivery of the...  ...the end-to-end user lifecycle.Manage, mentor, and grow a team of...  ...engineering excellence to solve complex technical problems.Build user-facing... 

    Google

    Mountain View, CA
    1 day ago
  • $230k - $350k

     ...Links LLC is seeking an experienced Member of Technical Staff, Inference Systems, to build and optimize a high-performance AI inference platform from the ground up....  ...GPU and multi-node environments. Profile, benchmark, and optimize latency, throughput, reliability... 
    Work at office

    Premier Global Links

    Palo Alto, CA
    14 days ago
  •  ...deployment, and high-performance inference services. The ideal candidate will build and manage scalable enterprise GenAI...  ...and cost efficiency. Develop AI platform services and APIs using...  ...Perform performance tuning, benchmarking, capacity planning, and production... 
    Contract work

    2T Consulting

    Santa Clara, CA
    a month ago
  • $180k

     ...SpaceXAI's mission is to create AI systems that can accurately...  ...the high-performance inference platform that serves Grok to...  ...reliability. As a Member of Technical Staff - Inference, you will design...  ...and excellent tail latency. Benchmark, fine-tune, and accelerate... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    3 days ago
  •  ...seeks an experienced software architect to own the architecture and design of scalable AI inference and training infrastructure. You will partner with development teams and tech leads to deliver cost-effective, high-performance systems for production use. You will lead... 

    Jobleads-US

    Sunnyvale, CA
    15 hours ago
  • $167.7k - $245.2k

     ...runtimes. Beyond serving, We benchmark, evaluate, monitor, and...  ...evaluation. You'll lead incident response and...  ...production experience in AI infrastructure, model...  ...engineers, and raises technical standards. The platform...  ...work serving models on inference runtimes including vLLM... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    2 days ago
  • $150k - $183k

    FortiAIGate is Fortinet's AI security and governance gateway....  ...Deploy and tune them on Triton Inference Server and vLLM — batching,...  ...governance detectors. Build benchmark and suites measuring detection...  ...communicate effectively across technical and non-technical teams.... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    1 day ago
  •  ...discovery to powering AI and the technologies people...  ..., or bringing leading edge products to market...  ...and LLVM/MLIR.Provide technical leadership, code reviews...  ...distributed training, inference, tensor parallelism, expert...  ..., debugging, benchmarking, and kernel performance... 

    AMD

    San Jose, CA
    2 days ago
  • $235.03k - $352.29k

     ...and profound opportunity for AI to drive positive change in the...  ..., T. Rowe Price, and other leading investors. About the Role...  ...Autonomy Leader to drive the technical roadmap for the systems that...  ...technical execution. People Management: Experience in hiring,... 
    Full time
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    3 days ago
  •  ...Amazon is hiring a Senior Software Development Engineer for AI/ML workloads on AWS Neuron to accelerate inference on Inferentia and Trainium accelerators. You will design, implement, and optimize ML models and kernels across the stack, from PyTorch to hardware-bound performance... 

    Jobleads-US

    Cupertino, CA
    15 hours ago
  • $152k - $241.5k

     ...performance computing platforms are powering the AI revolution across many applications and...  ...computations on NVIDIA GPUs. What you'll be doing: Benchmark the performance of state-of-the-art deep learning models' inference and training passes to identify key GPU kernel... 

    NVIDIA Corporation

    Santa Clara, CA
    1 hour ago
  • $160k - $190k

     ...is a leader in next-generation AI infrastructure, delivering a full-stack inference platform for customers worldwide...  ...to the market by turning complex technical ideas into usable narratives, content...  ..., customer conversations, benchmarking discussions, and technical field... 
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    SambaNova

    San Jose, CA
    3 days ago
  •  ...builds the world's largest AI chip, 56 times larger than...  ...to deliver industry-leading training and inference speeds; over 10 times faster...  ...performance modeling, empirical benchmarking, and hands-on prototyping...  ...recommendations through technical reports, presentations, and... 

    Cerebras Systems

    Sunnyvale, CA
    4 days ago
  •  ...discovery to powering AI and the technologies people...  ..., or bringing leading edge products to market...  ...both model training and inference.Our team supports a broad...  ...stack, solving technically challenging performance...  ...performance tooling, benchmarks, automation, and observability... 

    AMD

    San Jose, CA
    19 hours ago
  • $272k - $431.25k

     ...NVIDIA IT's Enterprise AI & Automation team to...  ...for a deeply involved technical leader writing code daily...  ...Kubernetes to GPU inference stacks and translate new...  ...quality through telemetry, benchmarking, automated evaluation,...  ...models, multi-agent management (e.g., LangChain,... 

    NVIDIA Corporation

    Santa Clara, CA
    1 hour ago
  • $139k - $257.55k

     ...evolve it, develop new features and benchmark possible solutions, in terms of...  ...in building, deploying and managing infrastructures in public clouds...  ...that powers Machine Learning and AI across Adobe cloud Product lines by catering inference and training needs of thousands... 
    Temporary work
    Local area

    Adobe

    San Jose, CA
    1 hour ago
  • $168k - $258.75k

     ...unlimited potential of AI to define the next era...  ..., for training and inference Find opportunities...  ...with previous product management, AI related engineering...  ...computer science, or another technical field. MBA a plus. ~...  ...in building or leading cloud computing infrastructure... 

    NVIDIA Corporation

    Santa Clara, CA
    1 hour ago
  • $168k - $258.75k

     ...into the unlimited potential of AI to define the next era of...  ...are looking for a passionate Technical Program Manager to be part of this journey. This role will be leading all steps of the development...  ...training, engineering for optimized inference and deployment. What you... 
    Shift work

    NVIDIA Corporation

    Santa Clara, CA
    1 hour ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Technical Lead Manager, AI Inference Benchmarking. Be the first to apply!