Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior GenAI Inference Architect

Adobe

Adobe’s Firefly Foundry is seeking a Principal ML Engineer to lead enterprise-scale GenAI inference architecture and service delivery for Adobe’s flagship products. You will establish the inference architecture, optimize multi-model pipelines, and co-develop high‑performance APIs and backends supporting internal and external integrations. You will mentor engineers across teams, drive evaluation of cutting‑edge inference and MLOps technologies, and set standards for reliability, scalability, and #J-18808-Ljbffr Adobe

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Senior GenAI Inference Architect in San Jose, CA vacancy
  • $184k - $287.5k

    We are now looking for a Senior Deep Learning Architect, LLM Inference!NVIDIA is at the forefront of the generative AI revolution. The Inference Benchmarking (IB) team specifically focuses on inference server performance optimization for Large Language Models (LLMs). If... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • Adobe is seeking a Principal Machine Learning Engineer to serve as the technical lead for GenAI Services. You will own the architecture, depth, and hands-on execution of inference pipelines across heterogeneous generative models and APIs that power Adobe products. You will... 
    Suggested

    Adobe

    San Jose, CA
    4 hours ago
  • NVIDIA is seeking a senior leader to shape the global strategy for scaled-out AI inference. You will architect high-throughput, low-latency distributed pipelines and model serving strategies for massive scale and reliability on NVIDIA hardware. You will drive the technical... 
    Senior

    NVIDIA

    Santa Clara, CA
    1 day ago
  •  ...Firefly's Generative AI Services team is seeking Senior Machine Learning Engineers to help build scalable GenAI systems powering features across Adobe products like...  ..., Express, Stock, and Premiere. You will design inference pipelines, optimize models for latency, and... 
    Senior

    Adobe Inc.

    San Jose, CA
    3 days ago
  • $152k - $241.5k

    We are now looking for a Senior Performance Architect for Nemotron! At NVIDIA, we are redefining the future of...  ...end-to-end performance impact of emerging GenAI workflows - such as Speculative Decoding, Agentic Pipelines, Inference-time compute scaling, RL etc. - to understand... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. THE ROLEWe are seeking a Principal GenAI Inference Optimization Engineer to join our Models and Applications team. This role focuses on improving performance, efficiency, and... 

    AMD

    San Jose, CA
    14 hours ago
  • Accellor is seeking a Technical Architect — AI Systems, Inference & Platform Internals to design, scale, and optimize internal AI systems powering ChatGPT and OpenAI API workloads. The role focuses on inference runtime, model serving, GPU infrastructure, and distributed... 
    Senior

    Accellor

    Mountain View, CA
    2 days ago
  • $231.1k - $358.2k

    SiMa.ai is seeking a Sr. Principal SoC Architect in San Jose, CA. This full-time position focuses on driving SoC architecture and software development for next-generation AI applications. The ideal candidate will have over 20 years of experience in ML product development... 
    Senior
    Full time

    SiMa.ai

    San Jose, CA
    4 days ago
  • SambaNova in San Jose is looking for a Software Architect to lead the SambaStack platform's technical direction. You will ensure the architecture meets enterprise standards while collaborating with cross-functional teams to drive innovation in AI. We seek someone with over... 
    Senior

    SambaNova

    San Jose, CA
    1 day ago
  •  ...we advance your career. THE ROLE: We are seeking a Robotics AI Architect to define and scale next-generation Physical AI systems, with...  ...influencing external and internal stakeholdersDeep understanding of:AI inference runtimes and deployment tradeoffsSystem architecture level CPU... 
    Senior

    AMD

    San Jose, CA
    2 days ago
  • $208k - $327.75k

     ...scalable software-defined architectures.We are looking for a Senior AI Architect to help define the next generation of AI model paradigms for...  ...of distributed training systems, scaling laws, and inference optimization techniques.Experience with model optimization methods... 
    Senior
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...accelerates workloads at scale, and we’re seeking a visionary Product Architect with strong expertise in systems architecture to join our...  ...with AI workloads, including agentic & RAG-based workflows, inference at scale, large scale training & fine-tuning, and model... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    We are now looking for a Senior Deep Learning Performance Architect! NVIDIA is seeking outstanding Performance Analysis Architects to help analyze and...  ...Learning ASIC architecture evaluation for training and/or inference.Strong programming skills in Python and C++. Ways to... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...You'll Be Doing:The software architecture group at NVIDIA has openings for a Deep Learning Communication Architect. We scale the DNN models and training/inference frameworks to systems with hundreds of thousands of nodes.Optimizing communication performance: Identify and... 
    Senior
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    We are now looking for a Senior GPU & Deep Learning Architect!The NVIDIA GPU Architecture group is looking for world class architects and software...  ...especially for deep learning workloads, both training and inference, and maintain our leadership by developing new parallel... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $224k - $356.5k

     ...make a lasting impact on the world.As an AI Storage Platform Architect at NVIDIA, this position will be the linchpin between cutting-...  ...Architect end-to-end reference architectures for disaggregated inference (aligned with NVIDIA Dynamo), large-scale foundation model training... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $224k - $356.5k

     ...can make a lasting impact on the world.We’re looking for a Senior Architect to help shape the next generation of Agentic AI platforms for...  ...technical tradeoffs across model quality, retrieval precision, inference cost, throughput, latency, personalization, data freshness,... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...mining infrastructure company, seeks a Senior AI Storage Infrastructure Engineer in San...  ...fabric for our AI-native NeoCloud. You will architect high-performance storage ensuring GPUs...  ...saturated with data during training and inference across distributed systems. You will... 
    Senior

    Bitdeer

    San Jose, CA
    2 days ago
  •  ...Samsung Semiconductor seeks a Sr Staff Engineer to lead AI-inference workloads and storage-system software programs. You will translate workload insights into architecture, drive data-placement strategies, and mentor engineers across the organization. The role emphasizes... 
    Senior

    Jobleads-US

    San Jose, CA
    4 hours ago
  • $168k - $258.75k

    Inference is the fastest growing and most competitive area in Generative AI today. It is where...  ..., and deployment techniques. As a Senior Product Manager for AI Platform Inference...  ...TorchAO, etc.)Demonstrable knowledge of GenAI or machine learning concepts, particularly... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • NVIDIA is seeking a Senior Software Engineer for Agent Architecture and Evaluation to build agentic AI systems and analyze inference workloads in realistic settings. You will work with teams advancing evaluation across software and hardware to improve accuracy and efficiency... 
    Senior

    NVIDIA Gruppe

    Santa Clara, CA
    2 days ago
  • NVIDIA seeks a Product Manager for Inference to enable developers to deploy high-performance AI on NVIDIA GPUs. You will shape product strategy...  ...management, strong communication, and hands-on knowledge of inference deployment and GenAI concepts. #J-18808-Ljbffr NVIDIA Gruppe
    Senior

    NVIDIA Gruppe

    Santa Clara, CA
    14 hours ago
  • NVIDIA Corporation in Santa Clara, CA is seeking a Senior Product Manager for AI Platform Inference to lead the development of tools, SDKs, and libraries enabling...  ...deployment performance, with strong emphasis on GenAI concepts and GPU-aware software delivery. #J-18808-... 
    Senior

    NVIDIA Corporation

    Santa Clara, CA
    3 days ago
  • Micron Technology, Inc in San Jose invites applications for an LPDDR Product Architect to shape memory solutions for AI inference at the edge. You will collaborate with customers, researchers, and internal teams to translate system insights into real-world LPDDR module... 

    Micron Technology

    San Jose, CA
    2 days ago
  • NVIDIA AI in Santa Clara is seeking a highly capable software engineer to advance an advanced inference framework using modern C++. The role focuses on extending TensorRT with autoregressive model serving capabilities and requires collaboration across CUDA, kernel libraries... 

    NVIDIA AI

    Santa Clara, CA
    14 hours ago
  • Micron Technology seeks an MCBU Product Architecture Engineer for DRAM to shape memory solutions for AI inference at the edge. You will collaborate with customers, researchers and internal teams to translate system-level insights into practical memory features. Applicants... 

    Micron

    San Jose, CA
    4 days ago
  • Etched in San Jose is seeking a talented Computer Architect to join our architecture team and design next-generation AI accelerators for inference workloads. You will work on compute architectures, performance modeling, and cross-functional collaboration to bring chip... 

    The Consensus

    San Jose, CA
    2 days ago
  • $182.5k - $260.5k

     ...Follow us on LinkedIn and Instagram.Positions are available at Senior Staff and above. Candidates are assessed individually and leveled...  ...roleAs a Senior Staff Machine Learning Scientist, you own the inference and optimization layer that makes AI in agentic workflows fast,... 
    Senior

    Netskope

    Santa Clara, CA
    2 days ago
  • $231.1k - $358.2k

     ...Robots market. We are looking for a Chief SoC Architect to help define next generation SoC...  ...Responsibilities: SiMa.ai is looking for a senior architect to lead its SoC architecture. Drive...  ...architectural innovations required for GenAI (transformers, multimodal models),... 
    Senior
    Full time
    Work at office

    SiMa Technologies

    San Jose, CA
    14 hours ago
  •  ...Micron Technology is seeking a Senior / Principal Layout Engineer to drive pathfinding and architecture exploration for next-generation memory modules and form factors. This role sits in the Module Architecture Group, using layout-driven insights to influence design decisions... 
    Senior

    Jobleads-US

    San Jose, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior GenAI Inference Architect. Be the first to apply!