Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Chief Architect, AI Inference Supercomputing

Etched

A pioneering AI hardware company is seeking a Head of Supercomputing to define and lead the architecture and software for its cluster-scale AI compute systems. This role involves deep systems expertise and managing a talented engineering team. Responsibilities include setting the technical vision, overseeing development, and ensuring system reliability and efficiency. The ideal candidate will have extensive experience in system software and infrastructure, along with proven leadership abilities in a fast-paced environment. #J-18808-Ljbffr Etched

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Chief Architect, AI Inference Supercomputing in San Jose, CA vacancy
  • Etched, building at-scale AI inference supercomputers powered by its own chips, is seeking a Head of Supercomputing to define and lead the architecture, software stack, and operational model for cluster-scale AI compute systems. This leader will own the end-to-end system... 
    Suggested

    The Consensus

    San Jose, CA
    1 day ago
  • NVIDIA is seeking a senior leader to shape the global strategy for scaled-out AI inference. You will architect high-throughput, low-latency distributed pipelines and model serving strategies for massive scale and reliability on NVIDIA hardware. You will drive the technical... 
    Suggested

    NVIDIA

    Santa Clara, CA
    1 day ago
  • NVIDIA AI in Santa Clara is seeking a highly capable software engineer to advance an advanced inference framework using modern C++. The role focuses on extending TensorRT with autoregressive model serving capabilities and requires collaboration across CUDA, kernel libraries... 
    Suggested

    NVIDIA AI

    Santa Clara, CA
    12 hours ago
  • Micron Technology, Inc in San Jose invites applications for an LPDDR Product Architect to shape memory solutions for AI inference at the edge. You will collaborate with customers, researchers, and internal teams to translate system insights into real-world LPDDR module... 
    Suggested

    Micron Technology

    San Jose, CA
    2 days ago
  • Etched in San Jose is seeking a talented Computer Architect to join our architecture team and design next-generation AI accelerators for inference workloads. You will work on compute architectures, performance modeling, and cross-functional collaboration to bring chip... 
    Suggested

    The Consensus

    San Jose, CA
    2 days ago
  •  ...protocols and StorageGRID. You will guide architectural decisions, align with product strategy, and balance enterprise workloads with AI/ML demands at scale. You will build and mentor senior leaders across multiple domains, partner with executive teams, and drive high-... 
    Remote work

    NetApp

    San Jose, CA
    3 days ago
  • Accellor is seeking a Technical Architect — AI Systems, Inference & Platform Internals to design, scale, and optimize internal AI systems powering ChatGPT and OpenAI API workloads. The role focuses on inference runtime, model serving, GPU infrastructure, and distributed... 

    Accellor

    Mountain View, CA
    2 days ago
  • SambaNova in San Jose is looking for a Software Architect to lead the SambaStack platform's technical direction. You will ensure the architecture...  ...with cross-functional teams to drive innovation in AI. We seek someone with over 12 years of experience in software engineering... 

    SambaNova

    San Jose, CA
    1 day ago
  •  ...Technology seeks an MCBU Product Architecture Engineer for DRAM in San Jose, California. This role involves shaping memory solutions for AI at the edge, analyzing architectures, and leading innovations. Candidates should have a Bachelor's in Electrical Engineering and over... 

    Micron Technology

    San Jose, CA
    1 day ago
  • $208k - $327.75k

     ...the forefront of accelerated computing, AI, and autonomous machines. From generative...  ...architectures.We are looking for a Senior AI Architect to help define the next generation of AI...  ...training systems, scaling laws, and inference optimization techniques.Experience with model... 
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...Parallel Compute Infrastructure, and Agentic AI - the biggest technology breakthroughs of...  ..., and we’re seeking a visionary Product Architect with strong expertise in systems...  ...including agentic & RAG-based workflows, inference at scale, large scale training & fine-tuning... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $272k - $431.25k

     ...Software Architecture group is solving some of AI’s hardest infrastructure problems. The...  ...computing interconnects.This Principal Architect role leads the research agenda and...  ...parallelism, or distributed training and inference patterns.Proficiency in programming languages... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...next-generation computing experiences—from AI and data centers, to PCs, gaming and...  .... THE ROLE: We are seeking a Robotics AI Architect to define and scale next-generation Physical...  ...internal stakeholdersDeep understanding of:AI inference runtimes and deployment tradeoffsSystem... 

    AMD

    San Jose, CA
    2 days ago
  • $224k - $356.5k

     ...tapping into the unlimited potential of AI to define the next era of computing. An era...  ...on the world.We’re looking for a Senior Architect to help shape the next generation of Agentic...  ...model quality, retrieval precision, inference cost, throughput, latency, personalization... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    We are now looking for a Senior Deep Learning Architect, LLM Inference!NVIDIA is at the forefront of the generative AI revolution. The Inference Benchmarking (IB) team specifically focuses on inference server performance optimization for Large Language Models (LLMs). If... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $206.4k - $379.1k

     ...effortlessly produce impressive content. The AI Foundations team constructs a flexible,...  ....We're looking for a Principal Architect to build and implement the AI framework for...  ...to life — spanning model orchestration, inference systems, data pipelines, caching and storage... 
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    Adobe Systems

    San Jose, CA
    1 day ago
  • A leading AI research company in Mountain View is seeking a visionary Design Leader to define the user experience of the Gemini app. You will architect the AI-UI bridge, optimize mobile UI performance, and mentor teams in establishing best practices for emerging technologies... 

    DeepMind

    Mountain View, CA
    3 days ago
  •  ...Job Summary The AI/ML ASIC Architect will lead the design and optimization of advanced ASIC and SoC architectures to support large-scale...  ...-bandwidth memory hierarchies. Optimize LLM training/inference including Dense, Mixture of Experts (MoE), and multi-modal... 

    Compunnel

    Milpitas, CA
    2 days ago
  • $219k - $351k

     ...partners, and communities.Job Title: Principal engineer, AI Serving Framework Architect (Software)The Architecture Research Lab (ARL) focuses on...  ...directionResearch on dynamic scheduling methodologies for maximizing AI inference performance in multi-rack scale memory-centric systems,... 
    Work at office
    Flexible hours

    Samsung Semiconductor

    San Jose, CA
    1 day ago
  •  ...moving forward. Job Description In this AI/ML ASIC Architecture position, you will...  ...ML Accelerator product. As an AI/ML ASIC Architect you will help drive new architecture...  ...bandwidth Architect memory-efficient inference/training systems utilizing techniques like... 
    Temporary work
    Remote work
    Flexible hours
    Shift work
    Night shift

    SanDisk

    Milpitas, CA
    4 days ago
  • Updated Role | Now hiring: Full-Stack AI Compute Architect About OXMIQ OXMIQ provides complete hardware and software GPU IP that lets our customers...  ...to production — hardening the platform and scaling it for inference at real customer scale. You'll join the OXMIQ architecture... 
    Immediate start
    Shift work

    Oxmiq Labs

    Campbell, CA
    2 days ago
  • Bitdeer, a world-leading AI and Bitcoin mining infrastructure company, seeks a Senior...  ...fabric for our AI-native NeoCloud. You will architect high-performance storage ensuring GPUs...  ...saturated with data during training and inference across distributed systems. You will build... 

    Bitdeer

    San Jose, CA
    2 days ago
  • Adobe’s Firefly Foundry is seeking a Principal ML Engineer to lead enterprise-scale GenAI inference architecture and service delivery for Adobe’s flagship products. You will establish the inference architecture, optimize multi-model pipelines, and co-develop high‑performance... 

    Adobe

    San Jose, CA
    4 days ago
  • Oxmiq Labs in California seeks a Full-Stack AI Compute Architect to define the software stack—from model ingestion and graph capture through...  ...and collaborate with silicon teams to scale workloads for inference at real customer scale. This role requires fluency across ML... 

    Oxmiq Labs

    Campbell, CA
    2 days ago
  • Micron Technology in San Jose is seeking an experienced LPDDR Product Architect to shape memory solutions for AI inference at the edge. You will analyze system architectures and define features to boost DRAM performance in AI workloads. The role requires 12+ years in engineering... 

    Micron Memory Malaysia Sdn Bhd

    San Jose, CA
    2 days ago
  • Mandatory Skills Agentic AI/ADK/Python We are looking for a skilled MLOps Architect to join our team and help us build, deploy, and maintain robust and scalable...  ...using Vertex AI Endpoints for real-time inference. Qualifications Strong experience with Google Cloud... 

    TechDigital Group

    San Jose, CA
    3 days ago
  •  ...About Opplane Inc. Opplane Inc. is a pioneering Fintech AI company dedicated to building the future of data-driven banking. Our...  ...powered compliance. Job Summary Opplane Inc. is seeking a visionary Chief Architect to lead the strategic technical direction of our Fintech AI... 
    Permanent employment
    Flexible hours

    Opplane

    Santa Clara, CA
    2 days ago
  •  ...Velaura seeks a Principal AI SoC Runtime Software Architect to own the software architecture for turning Velaura’s heterogeneous AI SoC into a coherent...  ...‑to‑end runtime across sensor ingest, preprocessing, AI inference, postprocessing, and delivery to robotics and edge AI... 

    Jobleads-US

    Santa Clara, CA
    4 days ago
  •  ...d-Matrix is seeking a Principal Software Engineer to advance performance analysis and modeling across its AI inference accelerators. You will analyze ML workloads, build analytical models, and extend simulators, collaborating with hardware design, compiler, inference... 
    Remote work
    3 days per week

    Jobleads-US

    Santa Clara, CA
    1 day ago
  •  ...Advanced Micro Devices is seeking a Robotics AI Architect to define and scale next‑generation Physical AI systems for complex robotic...  ...grade performance. The role requires deep expertise in AI inference runtimes, robot perception, planning and real‑time control, with... 

    Jobleads-US

    San Jose, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Chief Architect, AI Inference Supercomputing. Be the first to apply!