Senior Inference Performance & Pricing Analyst
Cerebras Systems
A leading AI technology firm in Sunnyvale, California, is hiring a Senior Performance Analyst. You will focus on performance benchmarking and competitive pricing for their inference product. The role requires deep knowledge of ML systems and open-source frameworks, with at least 5 years of experience. Responsibilities include designing benchmarking suites and translating insights into pricing recommendations. The company fosters an inclusive environment and values diverse perspectives. Flexibility and strong communication skills are essential. #J-18808-Ljbffr Cerebras Systems
- ...Sunnyvale, California, is seeking a Senior Performance Analyst to join their Product team. The ideal... ...systems and deep knowledge of open-source inference frameworks. Responsibilities include... ...performance and maintaining pricing models for competitive analysis. This...Senior
$208k - $327.75k
...technical Product Manager to own the products that help customers extract the best possible performance from AI models and applications running on NVIDIA hardware. Every inference deployment — from a single-GPU workstation to a multi-thousand-GPU data center — lives or...SeniorFull time- NVIDIA is seeking a Senior Product Manager for AI Inference Performance (Finance) to own optimization strategies across the inference stack and drive platform capabilities for diverse deployments. You will translate deep optimization into broadly adoptable products, collaborate...Senior
$182.5k - $260.5k
...gain full visibility and control without performance trade-offs.At Netskope, our technology... ...and Instagram.Positions are available at Senior Staff and above. Candidates are assessed... ...Machine Learning Scientist, you own the inference and optimization layer that makes AI in...Senior- ...Senior Enterprise Sales Executive AI Inference Infrastructure Tensordyne (formerly Recogni) is building the next generation of AI inference infrastructure... ...architecture delivers breakthrough inference performance with radically lower energy consumption and total cost...Senior
- GlobalFoundries is seeking a Senior Analyst, Transfer Pricing to automate intercompany calculations, build TP models, and manage data flows across U.S., EMEA, and APAC time zones. You will extract P&L data from FCCS, ingest budgets from the LRP and Anaplan, develop Alteryx...SeniorLocal area
$168k - $258.75k
Inference is the fastest growing and most competitive area in Generative AI today. It is where... ..., and where ever bit of accuracy and performance matters for quality, safety, and cost.... ...usecases, and deployment techniques. As a Senior Product Manager for AI Platform...SeniorFull time- CoreWeave is hiring a Senior Engineer for its Benchmarking & Performance team to write, profile, and optimize GPU kernels on the LLM inference path. You will improve latency and throughput and collaborate with product, orchestration, and hardware teams to achieve strict...Senior
- CoreWeave is seeking a Senior Engineer for its Benchmarking & Performance team to own kernel-level optimization for LLM inference and end-to-end model serving, focusing on CUDA kernels and throughput/latency improvements. You will lead kernel design reviews, mentor engineers...Senior
- Cerebras Systems, Inc. is looking for a Senior Performance Engineer to enhance the performance benchmarking and competitive pricing models for their AI chip. The ideal candidate... ...extensive experience with open-source inference frameworks and an understanding of ML systems...Senior
$184k - $287.5k
We are now looking for a Senior DL Algorithms Engineer! NVIDIA is seeking senior engineers who are mindful of performance analysis and optimization to help us squeeze every last... ...Implement language and multimodal model inference as part of NVIDIA Inference Microservices...SeniorFull time- Apple Inc. is seeking a Sr. Machine Learning Engineer for the Foundation Models Inference team in Santa Clara, CA. You will collaborate with research and external partners to bring cutting-edge model architectures from prototype to planetary-scale deployment, owning hard...Senior
- Accellor is seeking a Technical Architect — AI Systems, Inference & Platform Internals to design, scale, and optimize internal AI systems powering ChatGPT and OpenAI API workloads. The role focuses on inference runtime, model serving, GPU infrastructure, and distributed...Senior
- ...Technical Program Manager in Sunnyvale, California. In this role, you will oversee capacity planning and fleet strategy for the AI Inference Service organization, collaborating closely with teams across Engineering, Product, and Operations. Your responsibilities include...Senior
- Tensordyne seeks an experienced Sr./Director of Technical Product Management to own AI inference compute in datacenters—from silicon to software. You will report to the VP of Product Management, guiding product efforts across hardware, software, and infrastructure to shape...Senior
- Intel Corporation is seeking a software engineer to make models fast on hardware people own, optimizing inference engines for edge environments. You will work with llama.cpp and vLLM, tuning KV cache, batching, and quantization while reducing CPU overhead and startup costs...SeniorLocal area
$155.4k - $272k
...together. And we're just getting started.Join us to put AI to work for people.Job DescriptionWhat you get to do in this role:As a Senior Product Pricing Strategy Manager, you will lead and advise on strategic Pricing and Packaging recommendations to drive growth for...SeniorWork at officeImmediate startRemote workFlexible hours- NVIDIA is seeking a senior leader to shape the global strategy for scaled-out AI inference. You will architect high-throughput, low-latency distributed pipelines and model serving strategies for massive scale and reliability on NVIDIA hardware. You will drive the technical...Senior
- Advanced Micro Devices is seeking a Senior GPU Inference Performance Engineer to own end-to-end profiling of GPU-accelerated AI inference workloads. You will analyze workloads across AMD Instinct and NVIDIA GPUs, profile AI serving frameworks, and explain performance gaps...Senior
- NVIDIA is seeking a Senior Product Manager for AI Platform Inference to lead the development of tools, SDKs, and libraries that enable developers to deploy inference workloads efficiently on NVIDIA GPUs. You will craft product strategy, roadmaps, and go-to-market plans...Senior
- NVIDIA seeks a Senior Product Manager for AI Platform Inference (Finance) in Santa Clara to lead tooling, SDKs, and libraries enabling GPU-based inference deployments. You will craft strategy, roadmaps, and market plans, collaborating with developers to optimize models...Senior
- NVIDIA is seeking a Senior Agentic AI Software Engineer (Finance) to build agentic systems and advance AI inference workloads in real-world finance contexts. You will design and... ...drives experimental agents, optimize performance, and contribute to cutting-edge AI research...Senior
- NVIDIA is seeking a Senior Agentic AI Software Engineer to advance agentic AI systems and workloads from scalable research to production... ...-grade solutions. You will build agentic components, analyze inference dynamics, and collaborate with teams owning evaluation...Senior
- AMD in Santa Clara is seeking a Senior Product Manager to drive strategy and execution for ROCm, AMD’s open-source GPU software stack. This role involves managing the product roadmap for inference capabilities and engaging with the open-source community to enhance AMD'...SeniorRemote job
- Advanced Micro Devices is looking for a Senior Product Manager in Santa Clara, CA, to drive the strategy for ROCm, AMD's GPU software... ...strong background in product management, GPU computing, and AI inference technologies. This role provides an opportunity to work in a collaborative...Senior
- ...features that enhance system resiliency and high availability across distributed environments. The role includes developing scalable AI inference services and deploying cloud-based workflows. Ideal candidates have a master's degree and significant experience with deployment...Senior
- NVIDIA Corporation in Santa Clara, CA seeks a Senior Software Engineer specializing in Quantized Inference to speed up LLM deployment. You will implement quantized... ...paths, and collaborating with cross-team inference groups to push performance. #J-18808-Ljbffr NVIDIASenior
- NVIDIA in Santa Clara, CA, seeks a senior product leader to own the inference performance roadmap, shaping how models are represented, memory/state is managed, and tokens are generated. You will build scalable platforms across model families and deployment topologies for...Senior
- ...Clara, CA, seeks a Principal System Software Engineer for AI Inference Execution. You will join the software team to productize the AI... ...architecture, proficiency in C/C++/Python on Linux, and experience with distributed, high-performance software. #J-18808-Ljbffr Jobleads-USSenior
- Netskope is seeking a Senior Staff Machine Learning Scientist in Santa Clara to own the inference and optimization layer for AI in agentic workflows. You will fine-tune models, push latency and throughput on real hardware, and build a runtime that executes bounded AI tasks...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Inference Performance & Pricing Analyst. Be the first to apply!
- senior lead project manager Sunnyvale, CA
- senior robotics software engineer Sunnyvale, CA
- senior devops engineer remote Sunnyvale, CA
- senior sas administrator Sunnyvale, CA
- senior IT manager Sunnyvale, CA
- sr project manager Sunnyvale, CA
- senior windows systems engineer Sunnyvale, CA
- senior manager data science Sunnyvale, CA
- senior ui ux designer Sunnyvale, CA
- senior principal engineer Sunnyvale, CA
