Chief Architect, AI Inference Supercomputing
Etched
A pioneering AI hardware company is seeking a Head of Supercomputing to define and lead the architecture and software for its cluster-scale AI compute systems. This role involves deep systems expertise and managing a talented engineering team. Responsibilities include setting the technical vision, overseeing development, and ensuring system reliability and efficiency. The ideal candidate will have extensive experience in system software and infrastructure, along with proven leadership abilities in a fast-paced environment. #J-18808-Ljbffr Etched
- Etched, building at-scale AI inference supercomputers powered by its own chips, is seeking a Head of Supercomputing to define and lead the architecture, software stack, and operational model for cluster-scale AI compute systems. This leader will own the end-to-end system...Suggested
- NVIDIA is seeking a senior leader to shape the global strategy for scaled-out AI inference. You will architect high-throughput, low-latency distributed pipelines and model serving strategies for massive scale and reliability on NVIDIA hardware. You will drive the technical...Suggested
- NVIDIA AI in Santa Clara is seeking a highly capable software engineer to advance an advanced inference framework using modern C++. The role focuses on extending TensorRT with autoregressive model serving capabilities and requires collaboration across CUDA, kernel libraries...Suggested
- Micron Technology, Inc in San Jose invites applications for an LPDDR Product Architect to shape memory solutions for AI inference at the edge. You will collaborate with customers, researchers, and internal teams to translate system insights into real-world LPDDR module...Suggested
- Etched in San Jose is seeking a talented Computer Architect to join our architecture team and design next-generation AI accelerators for inference workloads. You will work on compute architectures, performance modeling, and cross-functional collaboration to bring chip...Suggested
- ...protocols and StorageGRID. You will guide architectural decisions, align with product strategy, and balance enterprise workloads with AI/ML demands at scale. You will build and mentor senior leaders across multiple domains, partner with executive teams, and drive high-...Remote work
- Accellor is seeking a Technical Architect — AI Systems, Inference & Platform Internals to design, scale, and optimize internal AI systems powering ChatGPT and OpenAI API workloads. The role focuses on inference runtime, model serving, GPU infrastructure, and distributed...
- SambaNova in San Jose is looking for a Software Architect to lead the SambaStack platform's technical direction. You will ensure the architecture... ...with cross-functional teams to drive innovation in AI. We seek someone with over 12 years of experience in software engineering...
- ...Technology seeks an MCBU Product Architecture Engineer for DRAM in San Jose, California. This role involves shaping memory solutions for AI at the edge, analyzing architectures, and leading innovations. Candidates should have a Bachelor's in Electrical Engineering and over...
$208k - $327.75k
...the forefront of accelerated computing, AI, and autonomous machines. From generative... ...architectures.We are looking for a Senior AI Architect to help define the next generation of AI... ...training systems, scaling laws, and inference optimization techniques.Experience with model...Full timeWorldwide$184k - $287.5k
...Parallel Compute Infrastructure, and Agentic AI - the biggest technology breakthroughs of... ..., and we’re seeking a visionary Product Architect with strong expertise in systems... ...including agentic & RAG-based workflows, inference at scale, large scale training & fine-tuning...Full time$272k - $431.25k
...Software Architecture group is solving some of AI’s hardest infrastructure problems. The... ...computing interconnects.This Principal Architect role leads the research agenda and... ...parallelism, or distributed training and inference patterns.Proficiency in programming languages...Full timeRemote work- ...next-generation computing experiences—from AI and data centers, to PCs, gaming and... .... THE ROLE: We are seeking a Robotics AI Architect to define and scale next-generation Physical... ...internal stakeholdersDeep understanding of:AI inference runtimes and deployment tradeoffsSystem...
$224k - $356.5k
...tapping into the unlimited potential of AI to define the next era of computing. An era... ...on the world.We’re looking for a Senior Architect to help shape the next generation of Agentic... ...model quality, retrieval precision, inference cost, throughput, latency, personalization...Full time$184k - $287.5k
We are now looking for a Senior Deep Learning Architect, LLM Inference!NVIDIA is at the forefront of the generative AI revolution. The Inference Benchmarking (IB) team specifically focuses on inference server performance optimization for Large Language Models (LLMs). If...Full time$206.4k - $379.1k
...effortlessly produce impressive content. The AI Foundations team constructs a flexible,... ....We're looking for a Principal Architect to build and implement the AI framework for... ...to life — spanning model orchestration, inference systems, data pipelines, caching and storage...Full timeTemporary workLocal areaWorldwideFlexible hours- A leading AI research company in Mountain View is seeking a visionary Design Leader to define the user experience of the Gemini app. You will architect the AI-UI bridge, optimize mobile UI performance, and mentor teams in establishing best practices for emerging technologies...
- ...Job Summary The AI/ML ASIC Architect will lead the design and optimization of advanced ASIC and SoC architectures to support large-scale... ...-bandwidth memory hierarchies. Optimize LLM training/inference including Dense, Mixture of Experts (MoE), and multi-modal...
$219k - $351k
...partners, and communities.Job Title: Principal engineer, AI Serving Framework Architect (Software)The Architecture Research Lab (ARL) focuses on... ...directionResearch on dynamic scheduling methodologies for maximizing AI inference performance in multi-rack scale memory-centric systems,...Work at officeFlexible hours- ...moving forward. Job Description In this AI/ML ASIC Architecture position, you will... ...ML Accelerator product. As an AI/ML ASIC Architect you will help drive new architecture... ...bandwidth Architect memory-efficient inference/training systems utilizing techniques like...Temporary workRemote workFlexible hoursShift workNight shift
- Updated Role | Now hiring: Full-Stack AI Compute Architect About OXMIQ OXMIQ provides complete hardware and software GPU IP that lets our customers... ...to production — hardening the platform and scaling it for inference at real customer scale. You'll join the OXMIQ architecture...Immediate startShift work
- Bitdeer, a world-leading AI and Bitcoin mining infrastructure company, seeks a Senior... ...fabric for our AI-native NeoCloud. You will architect high-performance storage ensuring GPUs... ...saturated with data during training and inference across distributed systems. You will build...
- Adobe’s Firefly Foundry is seeking a Principal ML Engineer to lead enterprise-scale GenAI inference architecture and service delivery for Adobe’s flagship products. You will establish the inference architecture, optimize multi-model pipelines, and co-develop high‑performance...
- Oxmiq Labs in California seeks a Full-Stack AI Compute Architect to define the software stack—from model ingestion and graph capture through... ...and collaborate with silicon teams to scale workloads for inference at real customer scale. This role requires fluency across ML...
- Micron Technology in San Jose is seeking an experienced LPDDR Product Architect to shape memory solutions for AI inference at the edge. You will analyze system architectures and define features to boost DRAM performance in AI workloads. The role requires 12+ years in engineering...
- Mandatory Skills Agentic AI/ADK/Python We are looking for a skilled MLOps Architect to join our team and help us build, deploy, and maintain robust and scalable... ...using Vertex AI Endpoints for real-time inference. Qualifications Strong experience with Google Cloud...
- ...About Opplane Inc. Opplane Inc. is a pioneering Fintech AI company dedicated to building the future of data-driven banking. Our... ...powered compliance. Job Summary Opplane Inc. is seeking a visionary Chief Architect to lead the strategic technical direction of our Fintech AI...Permanent employmentFlexible hours
- ...Velaura seeks a Principal AI SoC Runtime Software Architect to own the software architecture for turning Velaura’s heterogeneous AI SoC into a coherent... ...‑to‑end runtime across sensor ingest, preprocessing, AI inference, postprocessing, and delivery to robotics and edge AI...
- ...d-Matrix is seeking a Principal Software Engineer to advance performance analysis and modeling across its AI inference accelerators. You will analyze ML workloads, build analytical models, and extend simulators, collaborating with hardware design, compiler, inference...Remote work3 days per week
- ...Advanced Micro Devices is seeking a Robotics AI Architect to define and scale next‑generation Physical AI systems for complex robotic... ...grade performance. The role requires deep expertise in AI inference runtimes, robot perception, planning and real‑time control, with...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Chief Architect, AI Inference Supercomputing. Be the first to apply!
- search executive San Jose, CA
- executive search company San Jose, CA
- chief digital officer San Jose, CA
- technical executive San Jose, CA
- chief of police San Jose, CA
- chief revenue officer San Jose, CA
- board member San Jose, CA
- hospital ceo San Jose, CA
- managing director San Jose, CA
- assisted living executive director San Jose, CA

