Chief Architect, Cluster-Scale AI Inference
The Consensus
Etched, building at-scale AI inference supercomputers powered by its own chips, is seeking a Head of Supercomputing to define and lead the architecture, software stack, and operational model for cluster-scale AI compute systems. This leader will own the end-to-end system software and control-plane strategy—from orchestration, telemetry, provisioning, networking, to fleet reliability—working with ASIC, hardware, kernel, runtime, and infrastructure teams to deliver the highest-performance #J-18808-Ljbffr The Consensus
- A pioneering AI hardware company is seeking a Head of Supercomputing to define and lead the architecture and software for its cluster-scale AI compute systems. This role involves deep systems expertise and managing a talented engineering team. Responsibilities include...Suggested
- NVIDIA is seeking a senior leader to shape the global strategy for scaled-out AI inference. You will architect high-throughput, low-latency distributed pipelines and model serving strategies for massive scale and reliability on NVIDIA hardware. You will drive the technical...Suggested
- ...and StorageGRID. You will guide architectural decisions, align with product strategy, and balance enterprise workloads with AI/ML demands at scale. You will build and mentor senior leaders across multiple domains, partner with executive teams, and drive high-impact...SuggestedRemote work
$184k - $287.5k
...Infrastructure, and Agentic AI - the biggest technology... ...workloads at scale, and we’re seeking a visionary Product Architect with strong expertise in... ...world’s most powerful AI clusters. As a Product Architect... ...& RAG-based workflows, inference at scale, large scale training...SuggestedFull time- NVIDIA AI in Santa Clara is seeking a highly capable software engineer to advance an advanced inference framework using modern C++. The role focuses on extending TensorRT with autoregressive model serving capabilities and requires collaboration across CUDA, kernel libraries...Suggested
- Micron Technology, Inc in San Jose invites applications for an LPDDR Product Architect to shape memory solutions for AI inference at the edge. You will collaborate with customers, researchers, and internal teams to translate system insights into real-world LPDDR module...
- Accellor is seeking a Technical Architect — AI Systems, Inference & Platform Internals to design, scale, and optimize internal AI systems powering ChatGPT and OpenAI API workloads. The role focuses on inference runtime, model serving, GPU infrastructure, and distributed...
- Lyten is seeking a Sr Legal Operations Specialist to shape how the legal department works as Lyten rapidly scales. You will define Legal's AI roadmap, uplevel operating performance, and run systems that enable a lean team to scale with the business. Responsibilities include...
- Etched in San Jose is seeking a talented Computer Architect to join our architecture team and design next-generation AI accelerators for inference workloads. You will work on compute architectures, performance modeling, and cross-functional collaboration to bring chip...
$272k - $431.25k
...Architecture group is solving some of AI’s hardest infrastructure... ...interconnects.This Principal Architect role leads the research agenda... ...’s AI systems communicate at scale—across GPUs, DPUs, NICs, and heterogeneous... ..., or distributed training and inference patterns.Proficiency in...Full timeRemote work$184k - $287.5k
...recently, GPU deep learning ignited modern AI — the next era of computing. NVIDIA is... ...looking for an outstanding hands-on architect/engineer for a Senior HPC architect... ...support deployment and bringup of large-scale GPU compute clusters. Be a key player to enable the most...Full timeRemote work- ...generation computing experiences—from AI and data centers, to PCs, gaming and... ...THE ROLE: We are seeking a Robotics AI Architect to define and scale next-generation Physical AI systems,... ...stakeholdersDeep understanding of:AI inference runtimes and deployment...
$224k - $356.5k
...into the unlimited potential of AI to define the next era of... ...world.We’re looking for a Senior Architect to help shape the next generation... ..., governance, and scale.What You’ll Be Doing:Lead the... ...quality, retrieval precision, inference cost, throughput, latency, personalization...Full time$208k - $327.75k
...forefront of accelerated computing, AI, and autonomous machines. From... ...are looking for a Senior AI Architect to help define the next... ...modern AI architectures and large-scale model systemsExperience... ...training systems, scaling laws, and inference optimization techniques.Experience...Full timeWorldwide$206.4k - $379.1k
...produce impressive content. The AI Foundations team constructs a... ...that drives creativity at scale in design, imaging, motion, and... ...We're looking for a Principal Architect to build and implement the AI... ...spanning model orchestration, inference systems, data pipelines, caching...Full timeTemporary workLocal areaWorldwideFlexible hours- ...Job Summary The AI/ML ASIC Architect will lead the design and optimization of advanced ASIC... ...and SoC architectures to support large-scale machine learning workloads. This role... ...hierarchies. Optimize LLM training/inference including Dense, Mixture of Experts (MoE...
- ...Job Description In this AI/ML ASIC Architecture position,... ...Accelerator product. As an AI/ML ASIC Architect you will help drive new... ...Architect memory-efficient inference/training systems utilizing techniques... ...experience optimizing large-scale ML systems, GPU architectures...Temporary workRemote workFlexible hoursShift workNight shift
$219k - $351k
...communities.Job Title: Principal engineer, AI Serving Framework Architect (Software)The Architecture Research... ...memory capacity/bandwidth and system-scale communication. By leveraging Samsung’... ...methodologies for maximizing AI inference performance in multi-rack scale memory...Work at officeFlexible hours- Updated Role | Now hiring: Full-Stack AI Compute Architect About OXMIQ OXMIQ provides complete hardware and software GPU IP that lets... ...development to production — hardening the platform and scaling it for inference at real customer scale. You'll join the OXMIQ architecture...Immediate startShift work
- Adobe’s Firefly Foundry is seeking a Principal ML Engineer to lead enterprise-scale GenAI inference architecture and service delivery for Adobe’s flagship products. You will establish the inference architecture, optimize multi-model pipelines, and co-develop high‑performance...
$261.8k - $379.1k
A leading technology company seeks a Principal Architect to build the AI framework for Adobe Express in San Jose, California. The role involves architecting the AI stack, developing data infrastructure, and mentoring teams. Qualified candidates have over 10 years of experience...$180k
...is seeking a Member of Technical Staff, Infrastructure Compute in San Jose, California to lead the development of large-scale GPU computing clusters. The ideal candidate should have significant experience in systems engineering and machine learning. Your responsibilities...- ...Technology seeks an MCBU Product Architecture Engineer for DRAM in San Jose, California. This role involves shaping memory solutions for AI at the edge, analyzing architectures, and leading innovations. Candidates should have a Bachelor's in Electrical Engineering and over...
- SambaNova in San Jose is looking for a Software Architect to lead the SambaStack platform's technical direction. You will ensure the architecture... ...with cross-functional teams to drive innovation in AI. We seek someone with over 12 years of experience in software engineering...
- Oxmiq Labs in California seeks a Full-Stack AI Compute Architect to define the software stack—from model ingestion and graph capture... ...strategies, and collaborate with silicon teams to scale workloads for inference at real customer scale. This role requires fluency across...
- ...Inc. Opplane Inc. is a pioneering Fintech AI company dedicated to building the future... ...Opplane Inc. is seeking a visionary Chief Architect to lead the strategic technical direction... ...intelligent systems that process petabyte-scale data, leverage advanced Machine Learning...Permanent employmentFlexible hours
- ...Advanced Micro Devices is seeking a Robotics AI Architect to define and scale next‑generation Physical AI systems for complex robotic platforms.... ...performance. The role requires deep expertise in AI inference runtimes, robot perception, planning and real‑time control...
- ...Samsung Semiconductor seeks a Sr Staff Engineer to lead AI-inference workloads and storage-system software programs. You will translate... ...benchmarking, and collaboration with product and hardware teams to scale AI data paths in NAND/SSD storage, with emphasis on...
- ...Chief Blockchain Architect (CBA) About the Company Forward-thinking blockchain & gaming ecosystem... ...support massive parallel execution for large-scale gaming and digital entertainment... ...party computation, and a background in AI-driven systems or developer tool ecosystems...Contract work
- ...You will own the architecture, depth, and hands-on execution of inference pipelines across heterogeneous generative models and APIs that... ...define the roadmap and standards, driving performance, reliability, and scalable deployment at enterprise scale. #J-18808-Ljbffr Adobe
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Chief Architect, Cluster-Scale AI Inference. Be the first to apply!
- search executive San Jose, CA
- executive search company San Jose, CA
- chief digital officer San Jose, CA
- technical executive San Jose, CA
- chief of police San Jose, CA
- chief revenue officer San Jose, CA
- board member San Jose, CA
- hospital ceo San Jose, CA
- managing director San Jose, CA
- assisted living executive director San Jose, CA

