Staff HPC Systems Architect
Lambda Labs
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.If you'd like to build the world's best AI cloud, join us.*Note: This position requires presence in our San Jose, San Francisco, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.What You’ll DoArchitect and define scalable compute platforms optimized for AI/ML, simulation, and high-throughput workloads.Develop compute system standards and design patterns to ensure consistency, performance, and maintainability across infrastructure.Evaluate emerging CPU, GPU, and accelerator technologies, owning architectural tradeoff decisions that impact compute density, power, cooling, and total cost.Collaborate with product and engineering teams to map workload requirements to compute platform capabilities across bare metal and cloud deployments.Experience converting ambiguous business or customer needs into measurable platform requirements, technical specifications, acceptance criteria, and architecture decisions.Define compute platform roadmaps and architectural reference designs that guide hardware selection, firmware baselines, rack-level, and cluster design.Act as a technical lead during new platform introductions, guiding validation and performance characterization efforts.Mentor systems engineers and cross-functional stakeholders on compute performance tuning, sizing, and architectural decisions.YouProven experience (7+ years) architecting large-scale 10k-100k+ GPU HPC or cloud compute platforms.Deep knowledge of CPU/GPU architectures, memory hierarchies, and accelerator topologies.Experience designing systems around high-bandwidth, low-latency fabrics (NVLink, InfiniBand, and RoCE).Strong understanding of system performance tuning, resource scheduling, thermal and power optimization, and compute lifecycle management.Comfortable working across hardware and software boundaries, especially at the intersection of compute architecture, OS behavior, and orchestration layers.Skilled at balancing architectural tradeoffs for density, power efficiency, cooling, and performance.Strong analytical and communication skills, with a track record of influencing technical strategy across teams.Strong ownership and can do attitude, self-starter who feels comfortable working in ambiguity.Nice to HaveHands-on experience with AI/ML workloads and their compute performance characteristics.Familiarity with orchestration tools used in HPC. (Slurm, Kubernetes, etc)Experience with virtualization technologies, specifically GPU virtualization.Exposure to hardware validation, vendor collaboration, and long-term OEM roadmap alignment.Background in compute telemetry, real-time performance profiling, or large-scale A/B infrastructure testing.Salary Range InformationThe annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.About LambdaFounded in 2012, with 500+ employees, and growing fastOur investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent CoveWe have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOGOur values are publicly available: We offer generous cash & equity compensationHealth, dental, and vision coverage for you and your dependentsWellness and commuter stipends for select roles401k Plan with 2% company match (USA employees)Flexible paid time off plan that we all actually useEqual Opportunity EmployerLambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.Compensation Range: $314K - $465KLocationSan Jose Office (First St); Bellevue Office; San Francisco Office (Fremont St)Employment TypeFull timeLocation TypeHybridDepartmentData Center BusinessCompensationSan Francisco / San JoseSan Francisco / San Jose $349K – $465KBellevueBellevue $314K – $419K
- ...). Learn more at you want to help us build the DNA of tech.? Vishay San Jose is currently seeking applicants for a Senior Staff System Architect, Power to architect core control IP for the development of digitally controlled power management solutions. This will involve...SuggestedFull time
$184k - $287.5k
...Senior System Architect: Heterogeneous EDA SystemsNVIDIA is seeking a Senior System Architect to solve a complex challenge in accelerated computing... ...building automated RCA (Root Cause Analysis) pipelines for HPC or cloud-scale environments.CPU Architecture Deep-Dive: Expert...Suggested$155k - $180k
...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide... ...leaders to join us.Job Summary:The Staff Data Center Solutions Architect will be part of a growing, dynamic,... ....Develop Bids: Develop and review systems solutions, technical bid responses...SuggestedWorldwide$195k - $225k
...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide... ...leaders to join us.Job Summary:As a Staff Data Center Solutions Architect on the Supermicro DCBBS Datacenter... ...focus integrating AI Cloud Systems monitoring and application software...SuggestedWorldwide- ...-site and hybrid environments.Partner with compute and storage architects to ensure seamless end-to-end data flow and fault tolerance.Guide... ...high-performance data center networks, preferably for HPC, AI/ML, or large-scale cloud infrastructure.Deep expertise with...SuggestedWork at officeLocal areaWork from homeFlexible hours
$224k - $356.5k
...impact on the world.NVIDIA is looking for a Datacenter Product Architect to help define & design products for AI, high performance computing... ...markets.What you'll be doing:Drive architecture for datacenter systems considering everything from the chip to the full datacenterWork...Full time- ...data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and... ...ROLE We are seeking a hands‑on System Architect to lead the end‑to‑end architecture for our... ...and protocols. Mentor engineering staff in system‑level thinking, architecture methods...
- ...significantly reducing the Total Cost of Ownership (TCO) of hardware systems — a critical barrier to scalable adoption. Kandou’s... ...hardware is built for the future. We are actively seeking a System Architect based in US (Bay Area or Austin preferred), EU considered....Shift work
$192k - $278k
...strategic hardware decisions through deep-dive analysis of full-system workloads and first-party software stacks.Shape software... ...teams to engineer category-defining features and user experiences.Architect end-to-end solutions that optimize for power, performance, and...Worldwide- KLA in Milpitas, CA is seeking a Sr. System Design Engineer to lead next-generation hardware development, from technology evaluation to architecture decisions, and to define system-level requirements. You will run experiments and simulations to predict performance, manage...
$184.7k - $324.8k
Systems Architect Apple's Ecosystem Products & Technologies team is looking for a highly skilled and motivated product-focused Systems Architect with a passion for creating great user experiences and high quality products. Join a collaborative development team that plays...Relocation$210k - $250k
..., Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We are the... .... Job Summary: The Principal Solution Architect leads a virtual team of product engineers, design engineers, system engineers and FAEs to develop best in class solutions...Work experience placementWorldwide$136k - $212.75k
...see how you can make a lasting impact on the world. We are looking for a broadly knowledgeable and highly capable hardware hacker / system prototyper for our Santa Clara, CA location. What you'll be doing: Create new and innovative prototypes that test future...Full time$224k - $356.5k
...experienced network infrastructure Solutions Architect. Do you want to be part of a team that... ...networking for AI, Machine Learning, and HPC. As part of the NVIDIA Solutions... ...TCP protocol tacks* Good understanding of system hardware architecture impact on network performance...Full time- Sandisk is seeking a System Architect to drive system-level solutions for advanced memory technologies and AI inference infrastructure in Milpitas, CA. You will optimize compute-memory bottlenecks, latency, and power while mapping workloads to hardware across chiplet,...
$208k - $327.75k
NVIDIA Enterprise Platforms Group is seeking a Senior System Architect to define, design, and validate enterprise AI factory reference architectures... ..., fine-tuning, inference, agentic AI, physical AI, and HPC workloadsEvaluate tradeoffs across performance, scalability, resiliency...Full time$184k - $287.5k
...computation horsepower that NVIDIA GPUs excel.We are seeking a GPU System Architect who will architect and design multi-GPU scale-up and scale-out systems for next-generation datacenter platforms for AI and HPC. The architect in this role will explore and define system...Full timeWorldwide- ...Jose Summary Celestica is seeking a Principal Hardware Systems Architect for Compute Platforms to provide senior technical leadership for... ..., with significant experience in server, compute, AI, HPC, or data-center platforms. Demonstrated experience defining...Local area
$224k - $356.5k
We are now looking for a Senior Deep Learning Systems Architect!NVIDIA is seeking architects like you to help design hardware accelerator and... ...Work experience with GPU computing (CUDA, OpenCL, OpenACC) and HPC (MPI, OpenMP) is a huge plus.Intelligent machines powered by AI...Full timeWork experience placementNight shift- ...reliable and scalable cooling architectures for high-density AI and HPC environments while ensuring alignment with North American... ...density thermal management solutions. Develop technical proposals, system architectures, equipment selections, sizing calculations, and TCO...Worldwide
$184k - $287.5k
...matter to the world. This is our life’s work, to amplify human creativity and intelligence.NVIDIA is looking for a Datacenter System Architect to help define & design products for AI, high performance computing, and other datacenter markets.What you will be doing:Drive...Full time$190k - $300k
...technology for aviation that will save lives. Automated aviation systems will enable a future where air transportation is safer, more... ...driving cars working to make this future a reality.As a Systems Architect at Reliable Robotics, you will be a part of the Systems team and...Permanent employmentCasual workRemote work$220k - $250k
...inspired, and seasoned builders that have created large-scale data systems and globally distributed platforms that sit at the heart of some... ...self-optimizing data lake platform! Job Description As a Staff Solutions Engineer at Onehouse, you will operate as a technical...Work at officeRemote work- Sandisk seeks an experienced system-level test engineer to lead development and qualification of test solutions for advanced NAND technologies. You will drive cross-functional efforts across Product and Test Engineering, align with customer expectations, and ensure high...
- ...understanding and segmentation. In this role, you will build the systems that let us search, decompose, and describe massive volumes of... ..., temporally-grounded descriptions of actions and scenes. Architect agentic systems and orchestration pipelines that chain embedding...Full time
$176.1k - $308.2k
...OpenAI, you’ll be the connective tissue between technical usage and financial discipline. About the role We are looking for a Staff FinOps AI Governance Lead to drive financial accountability and optimization across ServiceNow’s AI spend. This senior individual...Work at officeImmediate startRemote workFlexible hours- ...for advertisers to reach audiences that are deeply engaged. Our Team The Ads Platform Engineering teams build advertising systems and integrations that powers the delivery of ads using our world class content delivery ecosystem. We use a number of Netflix investments...Hourly payFull timeImmediate startFlexible hours
- Research, design, development, and deployment of advanced AI agents and agentic systems. Architect and implement complex multi-agent systems, including planning, decision-making, and execution capabilities. Develop and integrate large language models (LLMs) and other state...Full timeWork experience placement
- ...confidence-aware quality gating, and evaluation protocols that separate real accuracy gains from benchmark noise. Ship end-to-end systems at scale — large-scale training, high-throughput video inference, and reliable production pipelines over high-bandwidth multi-...Full time
$180k - $215k
...electrical, electronic and fiber optic connectors and interconnect systems, antennas, sensors and sensor-based products and coaxial and... ..., business equipment, and automotive. Position: Solutions Architect Location: Santa Clara, CA Amphenol High Speed and Commercial...Temporary workWork at officeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff HPC Systems Architect. Be the first to apply!



