Senior GenAI Inference Architect
Adobe
Adobe’s Firefly Foundry is seeking a Principal ML Engineer to lead enterprise-scale GenAI inference architecture and service delivery for Adobe’s flagship products. You will establish the inference architecture, optimize multi-model pipelines, and co-develop high‑performance APIs and backends supporting internal and external integrations. You will mentor engineers across teams, drive evaluation of cutting‑edge inference and MLOps technologies, and set standards for reliability, scalability, and #J-18808-Ljbffr Adobe
$184k - $287.5k
We are now looking for a Senior Deep Learning Architect, LLM Inference!NVIDIA is at the forefront of the generative AI revolution. The Inference Benchmarking (IB) team specifically focuses on inference server performance optimization for Large Language Models (LLMs). If...SeniorFull time- Adobe is seeking a Principal Machine Learning Engineer to serve as the technical lead for GenAI Services. You will own the architecture, depth, and hands-on execution of inference pipelines across heterogeneous generative models and APIs that power Adobe products. You will...Suggested
- NVIDIA is seeking a senior leader to shape the global strategy for scaled-out AI inference. You will architect high-throughput, low-latency distributed pipelines and model serving strategies for massive scale and reliability on NVIDIA hardware. You will drive the technical...Senior
- ...Firefly's Generative AI Services team is seeking Senior Machine Learning Engineers to help build scalable GenAI systems powering features across Adobe products like... ..., Express, Stock, and Premiere. You will design inference pipelines, optimize models for latency, and...Senior
$152k - $241.5k
We are now looking for a Senior Performance Architect for Nemotron! At NVIDIA, we are redefining the future of... ...end-to-end performance impact of emerging GenAI workflows - such as Speculative Decoding, Agentic Pipelines, Inference-time compute scaling, RL etc. - to understand...SeniorFull time- ...perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. THE ROLEWe are seeking a Principal GenAI Inference Optimization Engineer to join our Models and Applications team. This role focuses on improving performance, efficiency, and...
- Accellor is seeking a Technical Architect — AI Systems, Inference & Platform Internals to design, scale, and optimize internal AI systems powering ChatGPT and OpenAI API workloads. The role focuses on inference runtime, model serving, GPU infrastructure, and distributed...Senior
$231.1k - $358.2k
SiMa.ai is seeking a Sr. Principal SoC Architect in San Jose, CA. This full-time position focuses on driving SoC architecture and software development for next-generation AI applications. The ideal candidate will have over 20 years of experience in ML product development...SeniorFull time- SambaNova in San Jose is looking for a Software Architect to lead the SambaStack platform's technical direction. You will ensure the architecture meets enterprise standards while collaborating with cross-functional teams to drive innovation in AI. We seek someone with over...Senior
- ...we advance your career. THE ROLE: We are seeking a Robotics AI Architect to define and scale next-generation Physical AI systems, with... ...influencing external and internal stakeholdersDeep understanding of:AI inference runtimes and deployment tradeoffsSystem architecture level CPU...Senior
$208k - $327.75k
...scalable software-defined architectures.We are looking for a Senior AI Architect to help define the next generation of AI model paradigms for... ...of distributed training systems, scaling laws, and inference optimization techniques.Experience with model optimization methods...SeniorFull timeWorldwide$184k - $287.5k
...accelerates workloads at scale, and we’re seeking a visionary Product Architect with strong expertise in systems architecture to join our... ...with AI workloads, including agentic & RAG-based workflows, inference at scale, large scale training & fine-tuning, and model...SeniorFull time$184k - $287.5k
We are now looking for a Senior Deep Learning Performance Architect! NVIDIA is seeking outstanding Performance Analysis Architects to help analyze and... ...Learning ASIC architecture evaluation for training and/or inference.Strong programming skills in Python and C++. Ways to...SeniorFull time$184k - $287.5k
...You'll Be Doing:The software architecture group at NVIDIA has openings for a Deep Learning Communication Architect. We scale the DNN models and training/inference frameworks to systems with hundreds of thousands of nodes.Optimizing communication performance: Identify and...SeniorFull timeWork experience placement$184k - $287.5k
We are now looking for a Senior GPU & Deep Learning Architect!The NVIDIA GPU Architecture group is looking for world class architects and software... ...especially for deep learning workloads, both training and inference, and maintain our leadership by developing new parallel...SeniorFull time$224k - $356.5k
...make a lasting impact on the world.As an AI Storage Platform Architect at NVIDIA, this position will be the linchpin between cutting-... ...Architect end-to-end reference architectures for disaggregated inference (aligned with NVIDIA Dynamo), large-scale foundation model training...SeniorFull time$224k - $356.5k
...can make a lasting impact on the world.We’re looking for a Senior Architect to help shape the next generation of Agentic AI platforms for... ...technical tradeoffs across model quality, retrieval precision, inference cost, throughput, latency, personalization, data freshness,...SeniorFull time- ...mining infrastructure company, seeks a Senior AI Storage Infrastructure Engineer in San... ...fabric for our AI-native NeoCloud. You will architect high-performance storage ensuring GPUs... ...saturated with data during training and inference across distributed systems. You will...Senior
- ...Samsung Semiconductor seeks a Sr Staff Engineer to lead AI-inference workloads and storage-system software programs. You will translate workload insights into architecture, drive data-placement strategies, and mentor engineers across the organization. The role emphasizes...Senior
$168k - $258.75k
Inference is the fastest growing and most competitive area in Generative AI today. It is where... ..., and deployment techniques. As a Senior Product Manager for AI Platform Inference... ...TorchAO, etc.)Demonstrable knowledge of GenAI or machine learning concepts, particularly...SeniorFull time- NVIDIA is seeking a Senior Software Engineer for Agent Architecture and Evaluation to build agentic AI systems and analyze inference workloads in realistic settings. You will work with teams advancing evaluation across software and hardware to improve accuracy and efficiency...Senior
- NVIDIA seeks a Product Manager for Inference to enable developers to deploy high-performance AI on NVIDIA GPUs. You will shape product strategy... ...management, strong communication, and hands-on knowledge of inference deployment and GenAI concepts. #J-18808-Ljbffr NVIDIA GruppeSenior
- NVIDIA Corporation in Santa Clara, CA is seeking a Senior Product Manager for AI Platform Inference to lead the development of tools, SDKs, and libraries enabling... ...deployment performance, with strong emphasis on GenAI concepts and GPU-aware software delivery. #J-18808-...Senior
- Micron Technology, Inc in San Jose invites applications for an LPDDR Product Architect to shape memory solutions for AI inference at the edge. You will collaborate with customers, researchers, and internal teams to translate system insights into real-world LPDDR module...
- NVIDIA AI in Santa Clara is seeking a highly capable software engineer to advance an advanced inference framework using modern C++. The role focuses on extending TensorRT with autoregressive model serving capabilities and requires collaboration across CUDA, kernel libraries...
- Micron Technology seeks an MCBU Product Architecture Engineer for DRAM to shape memory solutions for AI inference at the edge. You will collaborate with customers, researchers and internal teams to translate system-level insights into practical memory features. Applicants...
- Etched in San Jose is seeking a talented Computer Architect to join our architecture team and design next-generation AI accelerators for inference workloads. You will work on compute architectures, performance modeling, and cross-functional collaboration to bring chip...
$182.5k - $260.5k
...Follow us on LinkedIn and Instagram.Positions are available at Senior Staff and above. Candidates are assessed individually and leveled... ...roleAs a Senior Staff Machine Learning Scientist, you own the inference and optimization layer that makes AI in agentic workflows fast,...Senior$231.1k - $358.2k
...Robots market. We are looking for a Chief SoC Architect to help define next generation SoC... ...Responsibilities: SiMa.ai is looking for a senior architect to lead its SoC architecture. Drive... ...architectural innovations required for GenAI (transformers, multimodal models),...SeniorFull timeWork at office- ...Micron Technology is seeking a Senior / Principal Layout Engineer to drive pathfinding and architecture exploration for next-generation memory modules and form factors. This role sits in the Module Architecture Group, using layout-driven insights to influence design decisions...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior GenAI Inference Architect. Be the first to apply!
- senior associate architect San Jose, CA
- senior dynamics crm developer San Jose, CA
- senior application security San Jose, CA
- senior account director San Jose, CA
- senior supervisor San Jose, CA
- senior plumbing designer San Jose, CA
- senior advisor San Jose, CA
- senior cloud data engineer San Jose, CA
- senior customer service manager San Jose, CA
- senior San Jose, CA
