AI Systems Engineer: HPC & GPU Clusters
AMD
AMD in San Jose, CA is seeking an AI Systems Engineer to join our IT compute platforms team. You will design, deploy, and manage HPC infrastructure, GPU clusters, and AI workload schedulers to enable scalable AI services on AMD hardware. You should have passion for large-scale distributed computing, experience across a globally distributed organization, and a drive to deliver end-to-end outcomes with high reliability and performance. #J-18808-Ljbffr AMD
Vacancy posted 5 hours ago
Similar jobs that could be interesting for youBased on the AI Systems Engineer: HPC & GPU Clusters in San Jose, CA vacancy
- AMD is seeking an AI Systems Engineer to design, deploy, and manage HPC/AI infrastructure, GPU clusters, and AI workload schedulers. You will collaborate across teams to deliver scalable, high-performance AI services on AMD hardware, with a focus on end-to-end reliability...Suggested
- Semiconductor Engineering in San Jose, CA seeks a senior AI Systems Engineer to lead the lifecycle of its AI infrastructure. This hands-on individual contributor... ...architect, build and operate high-performance GPU clusters, and drive deployment and optimization of advanced...Suggested
$176k - $276k
NVIDIA is looking for an experienced HPC-AI Engineer to join the Networking Clusters Solutions Infrastructure team. we are focused on... ...in artificial intelligence and GPU computing. Provide insights on at-scale system design and tuning mechanisms for large-scale...SuggestedFull time- ...high-performance computing, cloud, and AI. Whether you’re designing next-gen... ....THE ROLE:We are seeking an AI Systems Engineer to join our AMD IT compute platforms... ...administration of High-Performance Computing (HPC) infrastructure, GPU clusters, and AI workload schedulers. THE...Suggested
- ...computing, cloud, and AI. Whether you’re designing... ...of large-scale AI/ML clustered infrastructure. You... ...of multi-disciplined engineers that operates across industry... ...: Linux operating systems, networking,... ...patterns, Kubernetes for HPC/AI (GPU operators, device plugins...SuggestedFlexible hours
- Semiconductor Engineering seeks an AI Systems Engineer in San Jose to lead the design, deployment, and optimization of our AI infrastructure. This... ...architecture through operations and support, including GPU clusters and agentic services. You will build and optimize high-...
- AMD, Inc. is seeking a PMTS Systems Design Engineer to research, design, develop, and test operating... ..., and to integrate software for GPU clusters supporting AI inferencing and training. Remote... ...performance, RDMA networking, and HPC system design. #J-18808-Ljbffr Socket...Remote job
$160k - $198k
...members.What You’ll DoAs a Senior AI Systems Engineer, you will architect, deploy,... ...high-performance computing (HPC), or ML infrastructure.Multi-... ...AI-centric bare-metal and GPU clouds (Nebius AI Cloud).Cloud... ..., paired with cloud-agnostic cluster abstractors like SkyPilot to...Local area- ...Ubuntu, etc.), Windows Server operating systems, Windows Client operating systems, and VMWare... ...of current and next-generation HPE HPC products. Ensure development issues are... ...appropriate automated test execution to test engineers at various global locations. Provide training...Local areaRemote work
$190k - $237k
...members.What You’ll DoAs a Staff AI Systems Engineer, you will architect, deploy,... ...high-performance computing (HPC), or ML infrastructure.Multi-... ...AI-centric bare-metal and GPU clouds (Nebius AI Cloud).Cloud... ..., paired with cloud-agnostic cluster abstractors like SkyPilot to...Local area- AMD seeks an AI Systems Engineer to advance ML workloads on AMD AI accelerators, bridging hardware and software from kernel design to production inference across NPU and GPU platforms. You will collaborate with compiler, runtime, silicon, and architecture teams, delivering...
- ...of accelerated and distributed Python APIs for numerical computing. Python dominates AI, data science and HPC, with NumPy, SciPy, TensorFlow and PyTorch. Join our team to develop GPU-accelerated Python libraries, optimize performance, and enable distributed workflows from...
$144k - $180k
.... What You’ll Do As a Senior AI Systems Engineer, you will architect, deploy,... ...high-performance computing (HPC), or ML infrastructure. Multi... ...specialized AI-centric bare-metal and GPU clouds (Nebius AI Cloud).... ..., paired with cloud-agnostic cluster abstractors like SkyPilot to...Local area$152k - $241.5k
...computing, known for inventing the GPU and driving breakthroughs in... ...everything from generative AI to autonomous systems, and we continue to shape... ...that enable researchers and engineers to develop the next... ...are looking for a strong AI & HPC Observability Engineer to build...Full time- A leading AI technology firm in California is seeking an experienced Senior Software Engineer to develop and optimize AI infrastructure software using state-of-the-art GPU systems. Candidates should have a Bachelor's degree in a technical field and a minimum of 5 years...
$255k - $340k
...Cloud, is a leader in AI cloud infrastructure... .... One person, one GPU.If you'd like to build... ...Tuesday.Hardware Engineering at Lambda is responsible... ....What You’ll DoOwn system integration validation for new HPC AI/ML, general... ...Center Engineering, and Cluster Network Design to ensure...Work at officeLocal areaWork from homeFlexible hours- ...computing experiences—from AI and data centers,... ...and embedded systems. Grounded in a culture... ...Marketing Engineer (TME) within the Software... ...AMD’s Data Center GPU Business Unit, you... ...operate AMD-powered GPU clusters & networks to... ...performance and value across HPC & AI solutions....
- NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models...
- ...scientific discovery to powering AI and the technologies... ...platform that lets engineers ask natural-language questions across GPU design knowledge, including... ...retrieval, RAG, agentic systems, evaluation, and... ...reconciliation, observability, and cluster orchestration. Advance...Worldwide
- We are seeking a highly skilled and experienced AI Systems Engineer to join our team. This is a hands-on, senior individual contributor role... ...systems, from architecting and building high-performance GPU clusters to deploying and optimizing our most advanced AI models and...
- ...Austin, TX is seeking a Technical Marketing Engineer (TME) within the Software Product Management organization for AMD’s Data Center GPU Business Unit. You will shape the customer... ...teams design and deploy AMD-powered GPU clusters and networks, while producing high-quality...
- Bitdeer Technologies Group is seeking an L1 NOC/US Data Center operator to support NeoCloud's GPU DCs during 8AM-8PM PST shifts. You will monitor GPU clusters, networks, and storage, respond to alerts, and execute runbooks for common incidents across shore-to-APAC handoffs...Shift workNight shift
- Analytical Mechanics Associates, Inc. is seeking an Aerothermodynamics Engineer to lead CFD software development for aerothermodynamics on HPC platforms, including GPU clusters, at NASA Ames Research Center in Mountain View, CA. The role emphasizes scientific software development...
$152k - $241.5k
We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency efforts... ...experience with NSight Systems and NSight ComputeExperience with... ...systems like Lustre and GPFS for AI/HPC workloadsFamiliarity with deep learning...Full timeRemote work$136.3k - $231.7k
...without us. KLA invents systems and solutions for the manufacturing... ...teams of physicists, engineers, data scientists and... ...of their algorithms. AI, including several... ...class team of physicists, HPC system designers, machine... ...Computing - HPC (including GPU), Machine Learning, Deep...Minimum wageFull timeWork experience placementFlexible hours- AMD in San Jose, CA is hiring a hands-on ML engineer to lead GoldenEye’s retrieval, ranking, and answer-quality architecture. You will scale a prototype into a production platform, setting technical direction and mentoring engineers across hardware, software, and security...
$100k
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance... ...-performance computing, and emerging autonomous AI systems. As the RISC-V AI / HPC & Agentic Software Engineering Lead, you will operate at the hardware-software boundary...Permanent employment$176k - $333.5k
...Our invention of the GPU in 1999 fueled the... ...ignited modern AI and enabled the next... ...a Senior Software Engineer to join our mission... ...continue improving our HPC infrastructure. Our... ...distributed systems, and has the ability... ...demands of our HPC clusters Evaluate new and innovative...- ...next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration... ...challenges. As the Director of Cloud, HPC & Sovereign AI Customer Engineering within the Compute & Enterprise AI...Remote work
- ...AI SYSTEMS ENGINEERAt AMD, we believe technology can change lives for the better. It can heal... ...forward.THE ROLEAMD is seeking an AI Systems Engineer to help develop and optimize machine... ...inference performance across AMD NPU and GPU platforms.You will collaborate closely with...Worldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Systems Engineer: HPC & GPU Clusters. Be the first to apply!
Related searches
- machine learning ai engineer San Jose, CA
- ai developer San Jose, CA
- senior ai engineer San Jose, CA
- ai engineer San Jose, CA
- ai ml engineer San Jose, CA
- ai prompt engineer San Jose, CA
- ai engineer remote San Jose, CA
- system engineer remote San Jose, CA
- senior windows systems engineer San Jose, CA
- senior linux systems engineer San Jose, CA

