Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Sr. System Engineer/GPU Platforms

$137k - $156k
Full-time

Super Micro Computer

Job Req ID: 30083About Supermicro:Supermicro is a Top Tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon Valley Top 50 technology firms. Our unprecedented global expansion has provided us with the opportunity to offer a large number of new positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:Supermicro is seeking an experienced Senior Systems Engineer / GPU Platforms to support the bring-up, qualification, enablement, and customer deployment of advanced GPU computing platforms.This role focuses on multi-GPU server systems used for AI, HPC, enterprise computing, and accelerated workloads. The successful candidate will work across the product lifecycle, from initial system bring-up and qualification through product release, customer POC/EVAL support, debugging, and post-launch technical enablement.The ideal candidate combines strong server hardware knowledge with hands-on Linux and GPU software experience and can independently troubleshoot complex issues across hardware, firmware, operating systems, networking, and GPU software environments.Essential Duties and Responsibilities:Key ResponsibilitiesSupport system bring-up, configuration, integration, validation, and troubleshooting of advanced GPU server platforms.Execute and support GPU platform qualification activities, including NVIDIA NVQUAL or equivalent validation processes.Install, configure, and troubleshoot Linux, GPU drivers, CUDA environments, firmware, libraries, and related software components.Diagnose complex system issues using logs, telemetry, diagnostics, and vendor tools, and drive issues to resolution or appropriate engineering escalation.Support multi-GPU server platforms throughout qualification, product launch, and post-release engineering activities.Participate in customer-facing POC/EVAL engagements, including system preparation, technical calls, debugging, and issue resolution.Collaborate with internal Architecture, Systems, Software, Validation, Product Management, and other engineering teams, as well as external technology partners.Develop technical documentation, troubleshooting guides, and best practices.Deliver technical presentations, training sessions, and internal knowledge-sharing activities.Serve as a technical resource and mentor for other engineers when appropriate.Key CompetenciesStrong technical depth and systems-level troubleshooting ability.Ownership and accountability for complex technical issues.Ability to quickly learn new server, GPU, and software technologies.Effective cross-functional collaboration.Clear technical communication and documentation.Willingness to share knowledge and support team development5–15 years of relevant experienceCandidates should demonstrate the ability to independently support complex GPU platforms and technical customer environments. More senior candidates should additionally bring broad system-level expertise, technical leadership, mentoring experience, and ownership of complex platform or customer-facing initiatives.Qualifications:Required QualificationsBachelor’s degree in Computer Engineering, Electrical Engineering, Computer Science, Information Technology, or a related discipline, or equivalent practical experience.5–15 years of relevant industry experience in systems engineering, server engineering, platform engineering, validation, technical enablement, HPC, AI infrastructure, or a related field.Strong knowledge of enterprise server hardware and system architecture.Hands-on experience with Linux server environments.Experience installing, configuring, validating, and troubleshooting server hardware and software.Strong system-level troubleshooting and root-cause-analysis skills.Working knowledge of PCIe architectures and high-performance I/O.Experience with GPU computing, accelerators, or comparable high-performance computing technologies.Ability to independently manage complex technical assignments and drive issues toward resolution.Strong written and verbal communication skills.Ability to work effectively with cross-functional and geographically distributed engineering teams.Comfortable participating in customer-facing technical discussions.Preferred QualificationsHands-on experience with NVIDIA data center or professional GPU platforms.Experience with CUDA and NVIDIA GPU software environments.Experience with NVIDIA NVQUAL or similar platform qualification processes.Experience with 4-GPU or 8-GPU server platforms.Familiarity with NVIDIA Blackwell, B200, Rubin, or comparable accelerator architectures.Knowledge of PCIe topology, NUMA, DMA, IOMMU, and GPU-to-NIC communication.Experience with GPUDirect RDMA, InfiniBand, RoCE, or high-speed Ethernet.Familiarity with NCCL, NVML, DCGM, Fabric Manager, or similar GPU diagnostic and management tools.Experience with Docker, containers, Kubernetes, or related orchestration technologies.Experience supporting AI, machine learning, HPC, or accelerated computing environments.Experience with customer POCs, technical evaluations, or engineering escalations.Experience delivering technical training or knowledge-sharing sessions.Bash, Python, or other scripting experience is a plus.Salary Range​$137,000 - $156,000 The salary offered will depend on several factors, including your location, level, education, training, specific skills, years of experience, and comparison to other employees already in this role. In addition to a comprehensive benefits package, candidates may be eligible for other forms of compensation, such as participation in bonus and equity award programs.​EEO StatementSupermicro is an Equal Opportunity Employer and embraces diversity in our employee population. It is the policy of Supermicro to provide equal opportunity to all qualified applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, protected veteran status or special disabled veteran, marital status, pregnancy, genetic information, or any other legally protected status.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Sr. System Engineer/GPU Platforms in San Jose, CA vacancy
  • $137k - $156k

     ...technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:Supermicro is seeking an experienced Senior Systems Engineer / GPU Platforms to support the bring-up, qualification, enablement, and... 
    Senior
    Worldwide

    Super Micro Computer

    San Jose, CA
    4 days ago
  • $152k - $253k

    Join NVIDIA as a Senior System Mechanical Engineer and help build the systems powering the future of AI...  ...development of mechanical systems for GPU and LPU-based AI infrastructure—from early...  ...for AI and high-performance computing platforms.Translate product requirements into... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...boundaries of innovation and engineering? At NVIDIA, we lead the world...  ...graphics, and high‑performance systems.As a Senior Hardware Systems...  ...(Language Processing Unit) platforms that support the most demanding...  ...AI platforms such as LPU, GPU, TPU, or custom accelerators.... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $224k - $356.5k

     ...computing. An era in which our GPU acts as the brains of...  ...best experience. We own the platform — performance, CI/CD pipelines...  ...in Computer Science, Computer Engineering, Electrical Engineering, or equivalent...  ...depth in GPU computing, ML systems, or high-performance... 
    Senior
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    5 days ago
  • $168k - $270.25k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline...  ...large scale production systems with high efficiency and availability...  ..., AI Skills to accelerate platform operationsDrive automation...  ...Computing and Visualization. The GPU, our invention, serves as the... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $160k - $253k

     ...NVIDIA accelerated computing is the engine of artificial intelligence. Our data center platforms integrate high performance...  ...pivotal in showcasing NVIDIA's GPU architecture, server-level platforms...  ...for NVIDIA’s GPU and rack-scale systems. This role bridges architecture... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $173.9k - $235.2k

     ...Web Services (AWS) Hardware Engineering team creates compute and storage...  ....We are looking for a Sr System Development Engineer, Mfg Bring...  ...validation across compute, memory, GPU, networking, power, and...  ...manufacturing partners to ensure every platform is production-ready ahead of... 
    Senior
    Contract work
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    5 days ago
  •  ...more, visit Position Overview We are seeking a Senior GPU Systems & Fabric Engineer to serve as the critical bridge between our physical GPU/...  ...high-performance host networking stacks, including RDMA, SR-IOV, RoCEv2, and InfiniBand, ensuring line-rate throughput... 
    Senior
    Remote job
    Full time
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    17 days ago
  • $160.36k - $291.15k

     ...why we’re building a universal autonomy platform: self-driving for all roads and all rides...  ...leading investors.About the WorkLead system-level requirements definition, architecture...  ...testing. Hands on experience in systems engineering.Ability to read code in C++ and write... 
    Senior
    Odd job
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    4 days ago
  •  ...future team: Search PlatformThe Search Platform team is responsible for powering Rovo Search...  ...and build high quality agentic search systems, addressing tail latency, throughput,...  ...simplify fragmented systems, mentor senior engineers, and align technical and product leaders... 
    Senior
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    1 day ago
  • $168k - $258.75k

    Join NVIDIA's datacenter product engineering team in our Operations organization and be at the...  ...of technological advancement! As a Senior System Debug Engineer, you will drive failure...  ...industry to ensure the flawless transfer of GPU Server products from development... 
    Senior
    Full time
    Work experience placement
    Overseas

    Nvidia

    Santa Clara, CA
    4 days ago
  • $159k - $230k

    Gather system requirements, define architecture, execute hardware design, and product validation...  ...:Bachelor’s degree in Electrical Engineering, Computer Engineering, Physics, a...  ...that are deployed in the data center.Our Platforms Infrastructure Engineering team designs... 
    Senior
    Worldwide

    Google

    Sunnyvale, CA
    4 days ago
  • $136k - $212.75k

     ...next era of computing. An era in which our GPU acts as the brains of computers, robots,...  ...are now looking for a Senior Validation Engineer in the DGX Server Product Engineering Team...  ...products.What you will be doing:System architecture, design, performance modelling... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $168k - $264.5k

     ...sits at the crossroads of architecture, silicon, systems, and manufacturing where first-principles thinking and engineering judgement at the highest level translate...  ...product outcomes at scale. Our team gets every GPU, SoC, and CPU silicon program from first power-... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $137k - $156k

     ...technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:High-performance product team in Supermicro is seeking talented Sr. System Engineer who can lead the technical collateral development of... 
    Senior
    Worldwide

    Super Micro Computer

    San Jose, CA
    5 days ago
  • $224k - $356.5k

     ...Architect in the Agent Harness & Runtime Engineering team to build foundational systems for the next generation of agentic...  ..., HPC, cloud, Kubernetes, and GPU compute environments.Optimize...  ...Track record of building reusable platforms supporting heterogeneous AI workloads... 
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $163k - $224k

     ...global leader in materials science and engineering solutions that are at the foundation of...  ...troubleshooting of advanced RF power delivery systems used in semiconductor manufacturing...  ...specifications for new products and technology platforms. Support product releases by developing... 
    Senior
    Full time

    Applied Materials

    Santa Clara, CA
    2 days ago
  • $168k - $264.5k

     ...itself over two decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market...  ...generation, developing efficient and reliable systems is an imperative. We are looking for a System Reliability Engineer to join NVIDIA's existing Reliability... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...Amazon seeks a Sr Cloud Hardware Dev Engineer to define server architectures for AI training and inference at scale. You will lead ODM/JDM partners...  ...thermal, mechanical, power, and signal integrity across GPU platforms. You will collaborate with firmware, software, and... 
    Senior

    Amazon

    Cupertino, CA
    12 hours ago
  • $161k - $221k

    Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced...  ..., or wherever you may go. Learn more about our benefits. As a Systems Engineer, you’ll design, integrate, and optimize complex systems... 
    Full time
    Remote work

    Applied Materials

    Santa Clara, CA
    5 days ago
  • $124k - $271.2k

     ...You Can ExpectAs a Lead Staff Site Reliability Engineer, you will be one of the technical leads for our DevOps Platforms organization. This group is responsible for DevOps...  ..., and building the automation and reliability systems that underpin Zoom's global services. If you... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    4 days ago
  • $168k - $264.5k

     ...continuously reinvented itself over three decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined...  ...We are seeking an outstanding candidate for Silicon Reliability Engineer to drive and utilize cutting edge technologies to deliver high... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $207k - $311k

     ...communication skills and a proven track record in customer-facing technical roles? If so, you'll thrive on our team as a Sr. Systems Sales Engineer at Nutanix, where you'll have the opportunity to transform businesses through innovative cloud solutions while collaborating... 
    Senior
    Work at office
    Local area
    Remote work
    Relocation package

    Nutanix

    San Jose, CA
    2 days ago
  • $196k - $310.5k

     ...era of computing. An era in which our GPU acts as the brains of computers, robots...  ...world.We are now looking for a Senior System Level Test Engineer. NVIDIA is seeking Senior System Level...  ...what is possible today and define the platform for the future of computing!What you'... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $116k - $184k

     ...outstanding Senior HTOL Reliability Engineer to join our Santa Clara lab....  ...to burn-in boards, HTOL systems, and thermal interface...  ...board design for high power GPU or SoC devices.Pattern translations...  ...Familiarity with reliability analytics platforms (e.g., JMP) and statistical... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...superintelligence. One person, one GPU.If you'd like to build...  ...is currently Tuesday.Engineering at Lambda is...  ...website, cloud APIs and systems as well as internal tooling...  ...cloud networking platform and SDN infrastructureOperate...  ...technologies, SR-IOV, and DPDKUnderstanding... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    2 days ago
  • $184k - $287.5k

     ...era of computing. An era in which our GPU acts as the brains of computers,...  ...looking for an experienced Senior Software Engineer for the Embedded Platform team. This is an outstanding...  ..., IGX and DGX Spark Product Software system development within NVIDIA. Using your... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $253k

     ...Computing and Visualization. The GPU, our invention, serves as the...  ...now looking for a Senior Test Engineer to manage Datacenter product...  ...test solutions for Datacenter system products at NVIDIA.What you'll...  ...and support rack level tester platforms at contract manufacturers, internal... 
    Senior
    Full time
    Contract work

    Nvidia

    Santa Clara, CA
    5 days ago
  •  ...superintelligence. One person, one GPU.If you'd like to build the...  ...day is currently Tuesday.Engineering at Lambda is responsible for...  ...Lambda website, cloud APIs and systems as well as internal tooling for...  ...automate the validation of platform quality.Design, build, and maintain... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  • $267k - $356k

     ...superintelligence. One person, one GPU.If you'd like to build...  ...Tuesday.Lambda's Storage Engineering team is the backbone...  ...of Lambda's data platform services—from low-level storage systems to the APIs and tooling our...  ...drivers.Experience with SR-IOV and virtualization (KVM... 
    Senior
    Work experience placement
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Sr. System Engineer/GPU Platforms. Be the first to apply!