Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Systems Engineer: HPC & GPU Clusters

AMD

AMD in San Jose, CA is seeking an AI Systems Engineer to join our IT compute platforms team. You will design, deploy, and manage HPC infrastructure, GPU clusters, and AI workload schedulers to enable scalable AI services on AMD hardware. You should have passion for large-scale distributed computing, experience across a globally distributed organization, and a drive to deliver end-to-end outcomes with high reliability and performance. #J-18808-Ljbffr AMD

Vacancy posted 5 hours ago
Similar jobs that could be interesting for youBased on the AI Systems Engineer: HPC & GPU Clusters in San Jose, CA vacancy
  • AMD is seeking an AI Systems Engineer to design, deploy, and manage HPC/AI infrastructure, GPU clusters, and AI workload schedulers. You will collaborate across teams to deliver scalable, high-performance AI services on AMD hardware, with a focus on end-to-end reliability... 
    Suggested

    Advanced Micro Devices, Inc.

    San Jose, CA
    3 days ago
  • Semiconductor Engineering in San Jose, CA seeks a senior AI Systems Engineer to lead the lifecycle of its AI infrastructure. This hands-on individual contributor...  ...architect, build and operate high-performance GPU clusters, and drive deployment and optimization of advanced... 
    Suggested

    Semiconductor Engineering

    San Jose, CA
    1 day ago
  • $176k - $276k

    NVIDIA is looking for an experienced HPC-AI Engineer to join the Networking Clusters Solutions Infrastructure team. we are focused on...  ...in artificial intelligence and GPU computing. Provide insights on at-scale system design and tuning mechanisms for large-scale... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...high-performance computing, cloud, and AI. Whether you’re designing next-gen...  ....THE ROLE:We are seeking an AI Systems Engineer to join our AMD IT compute platforms...  ...administration of High-Performance Computing (HPC) infrastructure, GPU clusters, and AI workload schedulers. THE... 
    Suggested

    AMD

    San Jose, CA
    1 day ago
  •  ...computing, cloud, and AI. Whether you’re designing...  ...of large-scale AI/ML clustered infrastructure. You...  ...of multi-disciplined engineers that operates across industry...  ...: Linux operating systems, networking,...  ...patterns, Kubernetes for HPC/AI (GPU operators, device plugins... 
    Suggested
    Flexible hours

    AMD

    Santa Clara, CA
    1 day ago
  • Semiconductor Engineering seeks an AI Systems Engineer in San Jose to lead the design, deployment, and optimization of our AI infrastructure. This...  ...architecture through operations and support, including GPU clusters and agentic services. You will build and optimize high-... 

    Semiconductor Engineering

    San Jose, CA
    2 days ago
  • AMD, Inc. is seeking a PMTS Systems Design Engineer to research, design, develop, and test operating...  ..., and to integrate software for GPU clusters supporting AI inferencing and training. Remote...  ...performance, RDMA networking, and HPC system design. #J-18808-Ljbffr Socket... 
    Remote job

    Socket.dev

    Santa Clara, CA
    2 days ago
  • $160k - $198k

     ...members.What You’ll DoAs a Senior AI Systems Engineer, you will architect, deploy,...  ...high-performance computing (HPC), or ML infrastructure.Multi-...  ...AI-centric bare-metal and GPU clouds (Nebius AI Cloud).Cloud...  ..., paired with cloud-agnostic cluster abstractors like SkyPilot to... 
    Local area

    Archer Aviation

    San Jose, CA
    3 days ago
  •  ...Ubuntu, etc.), Windows Server operating systems, Windows Client operating systems, and VMWare...  ...of current and next-generation HPE HPC products. Ensure development issues are...  ...appropriate automated test execution to test engineers at various global locations. Provide training... 
    Local area
    Remote work

    Net2Source

    San Jose, CA
    1 day ago
  • $190k - $237k

     ...members.What You’ll DoAs a Staff AI Systems Engineer, you will architect, deploy,...  ...high-performance computing (HPC), or ML infrastructure.Multi-...  ...AI-centric bare-metal and GPU clouds (Nebius AI Cloud).Cloud...  ..., paired with cloud-agnostic cluster abstractors like SkyPilot to... 
    Local area

    Archer Aviation

    San Jose, CA
    1 day ago
  • AMD seeks an AI Systems Engineer to advance ML workloads on AMD AI accelerators, bridging hardware and software from kernel design to production inference across NPU and GPU platforms. You will collaborate with compiler, runtime, silicon, and architecture teams, delivering... 

    Socket.dev

    San Jose, CA
    3 days ago
  •  ...of accelerated and distributed Python APIs for numerical computing. Python dominates AI, data science and HPC, with NumPy, SciPy, TensorFlow and PyTorch. Join our team to develop GPU-accelerated Python libraries, optimize performance, and enable distributed workflows from... 

    NVIDIA

    Santa Clara, CA
    1 day ago
  • $144k - $180k

     .... What You’ll Do As a Senior AI Systems Engineer, you will architect, deploy,...  ...high-performance computing (HPC), or ML infrastructure. Multi...  ...specialized AI-centric bare-metal and GPU clouds (Nebius AI Cloud)....  ..., paired with cloud-agnostic cluster abstractors like SkyPilot to... 
    Local area

    AlleyCorp

    San Jose, CA
    2 days ago
  • $152k - $241.5k

     ...computing, known for inventing the GPU and driving breakthroughs in...  ...everything from generative AI to autonomous systems, and we continue to shape...  ...that enable researchers and engineers to develop the next...  ...are looking for a strong AI & HPC Observability Engineer to build... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • A leading AI technology firm in California is seeking an experienced Senior Software Engineer to develop and optimize AI infrastructure software using state-of-the-art GPU systems. Candidates should have a Bachelor's degree in a technical field and a minimum of 5 years... 

    Intelliswift

    Sunnyvale, CA
    1 day ago
  • $255k - $340k

     ...Cloud, is a leader in AI cloud infrastructure...  .... One person, one GPU.If you'd like to build...  ...Tuesday.Hardware Engineering at Lambda is responsible...  ....What You’ll DoOwn system integration validation for new HPC AI/ML, general...  ...Center Engineering, and Cluster Network Design to ensure... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  •  ...computing experiences—from AI and data centers,...  ...and embedded systems. Grounded in a culture...  ...Marketing Engineer (TME) within the Software...  ...AMD’s Data Center GPU Business Unit, you...  ...operate AMD-powered GPU clusters & networks to...  ...performance and value across HPC & AI solutions.... 

    AMD

    Santa Clara, CA
    4 days ago
  • NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models... 

    NVIDIA

    Santa Clara, CA
    1 day ago
  •  ...scientific discovery to powering AI and the technologies...  ...platform that lets engineers ask natural-language questions across GPU design knowledge, including...  ...retrieval, RAG, agentic systems, evaluation, and...  ...reconciliation, observability, and cluster orchestration. Advance... 
    Worldwide

    AMD

    San Jose, CA
    1 day ago
  • We are seeking a highly skilled and experienced AI Systems Engineer to join our team. This is a hands-on, senior individual contributor role...  ...systems, from architecting and building high-performance GPU clusters to deploying and optimizing our most advanced AI models and... 

    Semiconductor Engineering

    San Jose, CA
    2 days ago
  •  ...Austin, TX is seeking a Technical Marketing Engineer (TME) within the Software Product Management organization for AMD’s Data Center GPU Business Unit. You will shape the customer...  ...teams design and deploy AMD-powered GPU clusters and networks, while producing high-quality... 

    AMD

    Santa Clara, CA
    1 day ago
  • Bitdeer Technologies Group is seeking an L1 NOC/US Data Center operator to support NeoCloud's GPU DCs during 8AM-8PM PST shifts. You will monitor GPU clusters, networks, and storage, respond to alerts, and execute runbooks for common incidents across shore-to-APAC handoffs... 
    Shift work
    Night shift

    Bitdeer (NASDAQ: BTDR)

    San Jose, CA
    2 days ago
  • Analytical Mechanics Associates, Inc. is seeking an Aerothermodynamics Engineer to lead CFD software development for aerothermodynamics on HPC platforms, including GPU clusters, at NASA Ames Research Center in Mountain View, CA. The role emphasizes scientific software development... 

    Analytical Mechanics Associates, Inc.

    Mountain View, CA
    4 days ago
  • $152k - $241.5k

    We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency efforts...  ...experience with NSight Systems and NSight ComputeExperience with...  ...systems like Lustre and GPFS for AI/HPC workloadsFamiliarity with deep learning... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $136.3k - $231.7k

     ...without us. KLA invents systems and solutions for the manufacturing...  ...teams of physicists, engineers, data scientists and...  ...of their algorithms. AI, including several...  ...class team of physicists, HPC system designers, machine...  ...Computing - HPC (including GPU), Machine Learning, Deep... 
    Minimum wage
    Full time
    Work experience placement
    Flexible hours

    KLA-Tencor

    Milpitas, CA
    1 day ago
  • AMD in San Jose, CA is hiring a hands-on ML engineer to lead GoldenEye’s retrieval, ranking, and answer-quality architecture. You will scale a prototype into a production platform, setting technical direction and mentoring engineers across hardware, software, and security... 

    Socket.dev

    San Jose, CA
    4 days ago
  • $100k

    Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance...  ...-performance computing, and emerging autonomous AI systems. As the RISC-V AI / HPC & Agentic Software Engineering Lead, you will operate at the hardware-software boundary... 
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    11 hours ago
  • $176k - $333.5k

     ...Our invention of the GPU in 1999 fueled the...  ...ignited modern AI and enabled the next...  ...a Senior Software Engineer to join our mission...  ...continue improving our HPC infrastructure. Our...  ...distributed systems, and has the ability...  ...demands of our HPC clusters Evaluate new and innovative... 

    NVIDIA

    Santa Clara, CA
    2 days ago
  •  ...next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration...  ...challenges. As the Director of Cloud, HPC & Sovereign AI Customer Engineering within the Compute & Enterprise AI... 
    Remote work

    AMD

    Santa Clara, CA
    3 days ago
  •  ...AI SYSTEMS ENGINEERAt AMD, we believe technology can change lives for the better. It can heal...  ...forward.THE ROLEAMD is seeking an AI Systems Engineer to help develop and optimize machine...  ...inference performance across AMD NPU and GPU platforms.You will collaborate closely with... 
    Worldwide

    Advanced Micro Devices , Inc.

    San Jose, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Systems Engineer: HPC & GPU Clusters. Be the first to apply!