Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Software Engineer - HPC

$176k - $333.5k

NVIDIA

NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 fueled the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI and enabled the next era of computing. NVIDIA is a “learning machine” that constantly evolves by adapting to new opportunities that are hard to address, that matters to the world, and that only we can address. This is our life’s work, to amplify human imagination and intelligence, and expand what is possible. We’re seeking strategic, bold, hard-working, and creative individuals who are passionate about helping us tackle challenges no one else can solve. Make the choice to join us today.

We are looking for a Senior Software Engineer to join our mission to continue improving our HPC infrastructure. Our team builds and operates sophisticated infrastructure to enable business critical services and AI applications. You will be working with a team of passionate and skilled engineers that are continuously working to provide better tools to build and manage this infrastructure. Ideal candidate is strong in software development, designing and creating reliable distributed systems, and has the ability to implement well thought out long term maintenance strategy.

What you’ll be doing:

  • Design highly available and scalable systems to meet the demands of our HPC clusters
  • Evaluate new and innovative technologies as the landscape evolves
  • Continuously improve infrastructure provisioning and management using automation
  • Support a globally distributed, multi-cloud hybrid environment - AWS, GCP and On-prem
  • Build strong cross functional relationships and align with partners across various business units
  • Ensure the highest level of up-time and Quality of Service (QoS) to our users through operational excellence
  • Participate in team's on-call rotation and be a contact for service incidents

What we need to see:

  • 10+ years of experience in design, implementation, and delivery of large engineering projects
  • Comfortable with at least two of the following programming languages: Golang, Java, C/C++, Scala, Python, Elixir.
  • Understands scalability challenges and performance of server-side code. Able to craft and develop horizontally-scalable, resilient and performing-under-load systems.
  • Versatile technologist with experience in full software development lifecycle – from inception and design to deployment, operation, and iterative development.
  • Proficient in cloud computing and are hands‑on in at least one cloud platform: GCP, AWS, or Azure.
  • Proficient in modern CI/CD techniques, GitOps and Infrastructure as Code (IaC)
  • Strong work ethic and a passion for problem solving
  • B.S. degree in Computer Science or related technical field (or equivalent experience)
  • Detail oriented with great communication and collaboration skills

Ways to stand out from the crowd:

  • Prior experience building solutions for HPC clusters based on Slurm or Kubernetes
  • Strong understanding of Linux operation system and TCP/IP fundamentals

The base salary range is 176,000 USD - 333,500 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

#J-18808-Ljbffr
Vacancy posted 15 hours ago
Similar jobs that could be interesting for youBased on the Senior Software Engineer - HPC in Santa Clara, CA vacancy
  • $152k - $241.5k

     ...Come join the team and see how you can make a lasting impact on the world. We are looking for a Senior Software Engineer to join our mission to continue improving our HPC infrastructure. Our team builds and operates sophisticated infrastructure to enable business critical... 
    Senior

    NVIDIA Gruppe

    Santa Clara, CA
    3 days ago
  • $152k - $287.5k

     ...NVIDIA Gruppe is seeking a highly motivated Senior Software Engineer to join our communication libraries and network software team in Santa Clara, California. This role focuses on designing and maintaining software for complex computing systems used in High Performance... 
    Senior

    NVIDIA Gruppe

    Santa Clara, CA
    3 days ago
  • HPE Labs - Senior Software Engineer - Integrated HPC & Quantum SolutionsThis role has been designed as 'Hybrid' with a requirement that you will work on average 2 days per week from an HPE office.Who We Are:Hewlett Packard Enterprise is the global edge-to-cloud company... 
    Senior
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    2 days per week

    Hewlett Packard Enterprise

    Milpitas, CA
    10 hours ago
  • $152k - $241.5k

     ...infrastructure, platforms, and tools that enable researchers and engineers to develop the next generation of AI/ML systems. By joining us,...  ...heart of this transformation. We are looking for a strong AI & HPC Observability Engineer to build and scale next-generation Observability... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $176k - $333.5k

     ...NVIDIA is looking for an experienced HPC-AI Engineer to join the Networking Clusters Solutions Infrastructure team. we are focused on building...  ...be a key player to the most exciting computing hardware and software to contribute to the latest breakthroughs in artificial... 
    Senior
    Full time

    NVIDIA

    Santa Clara, CA
    1 day ago
  • $255k - $340k

     ...’s designated work from home day is currently Tuesday.Hardware Engineering at Lambda is responsible for building and scaling the physical...  ...the hands-on technical lead for integrating OEM and white-label HPC AI/ML, general purpose compute, storage, and network hardware into... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    10 hours ago
  • $193.3k - $261.5k

    We are seeking an experienced engineer to work on distributed AI/ML systems...  ...with high-speed networking or HPC interconnects is valued highly...  ...AWS and develops hardware and software components that are critical...  ..., you can both expect senior mentorship and will be expected... 
    Senior
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    10 hours ago
  •  ...ASML Germany GmbH is seeking an experienced engineer to develop and maintain HPC infrastructure supporting scalable workloads across products. You will design compute clusters, storage, and networking, delivering stable, production-ready platforms with strong observability... 
    Senior

    ASML Germany GmbH

    San Jose, CA
    15 hours ago
  • $184k - $287.5k

     ...NVIDIA. We build communication libraries like NCCL, NVSHMEM, and UCX that are crucial for scaling Deep Learning and HPC. We're seeking a Senior Software Architect to help co-design next-gen data center platforms and scalable communications software.DL and HPC applications... 
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...NVIDIA is seeking a Senior Software Engineer in Westford, Massachusetts to improve their HPC infrastructure. The role includes designing scalable systems and supporting multi-cloud environments. The ideal candidate will have 10+ years of experience, strong software... 
    Senior

    NVIDIA

    Santa Clara, CA
    15 hours ago
  • $152k - $241.5k

     ...NVIDIA Gruppe in Santa Clara is seeking a Senior Software Engineer to enhance their HPC infrastructure. The role involves applying distributed systems patterns, automation, and building scalable services in a hybrid multi-cloud environment. Candidates should have strong... 
    Senior

    NVIDIA Gruppe

    Santa Clara, CA
    15 hours ago
  • $152k - $241.5k

     ...artificial intelligence.We are looking for a highly motivated senior software engineer for an exciting role in our communication libraries and...  ...Learning frameworks (e.g. NCCL for TensorFlow/Pytorch) and HPC programming interfaces (e.g. UCX for MPI/OpenSHMEM) on GPU clusters... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $147.75k - $221.63k

    Role SummaryDevelop and maintain HPC infrastructure (compute, storage, networking, container platform) to support scalable, high-performance workloads across products.Key ResponsibilitiesDesign and evolve HPC infrastructure (compute clusters, storage, networking stack)Deliver... 
    Senior
    Full time

    ASML Holding

    San Jose, CA
    2 days ago
  • $184k - $287.5k

     ...next generation of biological discovery. We are looking for a Senior Software Engineer to join the cuEquivariance team — an NVIDIA library that...  ...computer science with a focus on geometric deep learning or HPC.Contributions to open-source geometric ML or GPU computing projects... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    10 hours ago
  • Crusoe Cloud seeks a Senior Cloud Support Engineer to empower customers using sustainable GPU compute power. You’ll be the main technical contact, diagnose...  .... The role emphasizes customer success, Kubernetes and HPC knowledge, and collaboration on onboarding materials and... 
    Senior

    Crusoe

    Sunnyvale, CA
    22 hours ago
  • $152k - $241.5k

     ...artificial intelligence.We are looking for highly motivated Senior Software Engineers to join our Fabric Networking team with a targeted focus on...  ...NVIDIA GPU systems, NVLink, NVSwitch, CUDA, and large-scale AI/HPC clusters such as NVIDIA GB200 NVL72.Strong understanding of... 
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    10 hours ago
  • $165.2k - $223.6k

     ...The Nitro Team is looking for engineers with systems knowledge and experience...  ....The Nitro High Memory and HPC team owns the purpose built...  ...-sharing and mentorship. Our senior members enjoy one-on-one mentoring...  ...non-internship professional software development experience- 2+... 
    Internship
    Local area
    Flexible hours

    AmazonWebServices

    Santa Clara, CA
    1 day ago
  • $147.75k - $221.63k

     ...Role Summary Develop and maintain HPC infrastructure (compute, storage, networking, container platform) to support scalable, high-performance workloads across products. Key Responsibilities Design and evolve HPC infrastructure (compute clusters, storage, networking stack... 
    Senior

    ASML US, LLC

    San Jose, CA
    4 days ago
  • $152k - $241.5k

     ...libraries like NCCL, NVSHMEM, UCX for Deep Learning and HPC. We are looking for a motivated Performance engineer to influence the roadmap of our communication...  ...interactions and operating systems principles (aka systems software fundamentals)Implement micro-benchmarks in C/C++,... 
    Senior

    NVIDIA

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    NVIDIA is seeking a skilled Senior Software Engineer to join our Subnet Manager team. This team develops network management software that configures...  ..., contributing directly to the world’s leading AI and HPC systems.What you’ll be doing:Design, develop and optimize server... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $160k - $200k

     ...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We...  ...talented, passionate, and committed engineers, technologists, and business leaders to...  ...Supermicro is seeking an experienced Senior Software Engineer to join the team. In this role,... 
    Senior
    Worldwide
    Flexible hours

    Supermicro

    San Jose, CA
    2 days ago
  • $184k - $287.5k

    We are now looking for a Senior Software Engineer for AI Resiliency!At NVIDIA, we are pushing the boundaries of what’s possible in AI. We are currently...  ...performance tuning large-scale AI workloads in cloud and HPC environments, ensuring seamless operation of AI training and... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...are looking for a motivated Deep Learning engineer to bring advanced CUDA features and...  ...and runtimes for scaling Deep Learning and HPC applications. Your customers will have diverse...  ...systems principles (aka systems software fundamentals)Adaptability and passion to... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...partner with OS, container, GPU, and systems engineers. When useful, you will apply machine...  ...classification/prediction) inside existing software workflows.What we need to see:5+ years...  ...the crowd:TensorFlow or PyTorchLinux and HPC / large-scale or performance-sensitive environmentsExperience... 
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ...home day is currently Tuesday.About the RoleWe are seeking a Senior Software Engineer to join our Managed Kubernetes (Mk8s) team. You will play a...  ..., Network Operator, NCCL tuning, or similarFamiliarity with HPC and traditional job schedulers (Slurm) and Kubernetes-native... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    10 hours ago
  • SpaceX is seeking a Sr. Software Engineer, High Performance Computing for the Starlink program in Palo Alto, CA. You will own the full software lifecycle—from development to testing and support—aiming to optimize beam formation for a low-latency, global satellite internet... 
    Senior

    InvestedintheMission

    Palo Alto, CA
    1 day ago
  • $184k - $287.5k

     ...accelerated computing platform is foundational to modern HPC and AI. At the center of this platform are CUDA Core...  ...build fast, reliable, and scalable GPU-accelerated software. We are hiring a Senior Software Engineer to develop the Rust experience for CUDA Core... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...lasting impact on the world.We are looking for an outstanding Senior Software Engineer to work on our security team focused on securing at scale...  ...management and Host-Based Access Control.Experience hardening HPC schedulers and storage, Slurm alongside parallel filesystems... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We’re searching for a highly motivated, technical...  ...simulation & emulation technology.Mentor architects and engineering teams to grow them into future leaders.Make key technical decisions... 
    Senior
    Full time
    Remote work
    Shift work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $266k - $395k

     ...Senior Software Engineer - Storage Control Plane Design and implement a vendor-agnostic storage control plane for AI workloads Location: San...  ..., Pure, NetApp, or IBM Storage Scale. Ceph at 100 PB+ in HPC or AI environments. CXL memory pooling, computational... 
    Senior
    Local area
    Flexible hours

    Front Door Defense

    San Jose, CA
    15 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Software Engineer - HPC. Be the first to apply!