Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior HPC Storage Engineer

$184k - $287.5k

NVIDIA

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.As a member of the HW Infrastructure Storage Strategy team, you will provide leadership in the research, design and implementation of ground breaking fast storage solutions to enable runs of demanding high performance computing, and computationally intensive workloads. We seek an expert to identify architectural changes encompassing file, block, and object storage, to cater to the scaling and performance requirements of an expanding cloud infrastructure. As an expert, you will help us with the next-gen storage solutions strategic challenges we encounter with storage design for large scale, high performance workloads, evolving our private/public cloud strategy, capacity modelling, and growth planning across our global computing environment.What you'll be doing:Research and analyze existing internal distributed storage services.Research, design, and implement scalable, next-gen distributed storage services for HPC workloads, optimizing both performance and cost-effectiveness to meet NVIDIA’s growing infrastructure needsDevelop tooling to automate management of large-scale infrastructure environments, to automate operational monitoring and alerting, and to enable self-service consumption of resources.Detail the general procedures and practices, perform technology evaluations, related to distributed file systems.Collaborate across teams to better understand developers' workflows and capture their infrastructure requirements.Influence and guide methodologies for building, testing, and deploying applications to ensure efficient performance and resource utilization.Supporting our researchers to run their flows on our clusters including performance analysis and optimizations of deep learning workflowsRoot cause analysis and suggest corrective action for problems large and small scalesWhat we need to see:Bachelor’s degree in Computer Science, Electrical Engineering or related field or equivalent experience.8+ years of experience designing and/or operating large scale storage infrastructure.Experience analyzing and tuning storage performance for a variety of workloads.Proficient in Centos/RHEL and/or Ubuntu Linux distros including Python programming and bash scriptingIn depth understanding of container technologies like Docker, EnrootWays to stand out from the crowd:Distributed Storage Expertise: Extensive experience with parallel and distributed filesystems (Ceph, Weka.io, Vast, Lustre, GPFS) and Linux storage kernel development.GPU & AI Infrastructure: Proficient with NVIDIA GPUs, CUDA programming, and NCCL, including performance benchmarking via MLPerf.Hardware & Storage Engineering: Deep familiarity with storage hardware (HDDs, SSDs, NVMe), enclosures, and specialized appliances like Network Appliance.Advanced Networking: Strong background in Software Defined Networking (SDN) and high-performance networking for AI/HPC clusters.Deep Learning Frameworks: Practical experience applying industry-standard frameworks, specifically PyTorch and TensorFlow.NVIDIA offers highly competitive salaries and a comprehensive benefits package. We have some of the most resourceful and talented people in the world working for us and, due to unprecedented growth, our extraordinary engineering teams are growing fast. If you're a creative and autonomous engineer with real passion for technology, we want to hear from you. Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until June 13, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa Clara; US, TX, AustinType: Full time

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior HPC Storage Engineer in Santa Clara, CA vacancy
  • $155k - $185k

     ...a Top Tier provider of advanced server, storage, and networking solutions for Data Center...  ...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We...  ...seek talented, passionate, and committed engineers, technologists, and business leaders to... 
    Senior
    Contract work
    Immediate start
    Worldwide

    Supermicro

    San Jose, CA
    5 days ago
  • Stanford University is seeking an HPC Systems Administrator to own the physical infrastructure for Sherlock and related platforms. You...  ...,500+ compute nodes, high-density GPU racks, and petabyte-scale storage, ensuring high availability and reliable performance for diverse... 
    Senior

    Stanfordlivetickets

    Palo Alto, CA
    3 days ago
  • $152k - $241.5k

     ...on the world.We are seeking a highly skilled and experienced HPC Cluster Engineer to design, deploy, and operate GPU Compute Clusters for EDA...  ...systems, including the deployment of compute, networking, and storage.Foster strong customer and multi-functional partnerships to... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $184k - $287.5k

    NVIDIA has become the platform upon which every new AI-powered application is built. We are seeking a Sr. HPC Performance engineer to join our team of scientists and engineers passionate about building the next generation of scientific machine learning (ML) frameworks.... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • $184k - $287.5k

    NVIDIA Math Libraries team is looking for a senior engineer to join our development efforts in the area of kernel generation for AI and HPC, specifically targeting matrix operations, JITing and fusions. Around the world, leading commercial and academic organizations are... 
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    5 days ago
  • $152k - $241.5k

     ...environment remains resilient, measurable, and aligned with long-term engineering demands.What you'll be doing:Manage, scale, and optimize job...  ...and tuning job scheduling systems (LSF, Slurm, etc.) in HPC or silicon design environmentsProficiency in Linux systems administration... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • NVIDIA has become the platform upon which every new AI-powered application is built. We are seeking a Sr. HPC Performance engineer to join our team of scientists and engineers passionate about building the next generation of scientific machine learning (ML) frameworks.... 
    Senior

    NVIDIA

    Santa Clara, CA
    4 days ago
  • $255k - $340k

     ...designated work from home day is currently Tuesday.Hardware Engineering at Lambda is responsible for building and scaling the...  ...You’ll DoOwn system integration validation for new HPC AI/ML, general purpose compute, storage, and network hardware platforms throughout hardware... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  • $168k - $270.25k

     ...people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance Computing (HPC)... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    7 days ago
  • $272k - $431.25k

     ...lasting impact on the world.At NVIDIA, our storage platforms support some of the most...  ...ready to scale. You will work closely with engineering and AI teams, help shape how data moves through...  ...building or scaling storage for AI/ML or HPC workloads, including hybrid or multi... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  • NVIDIA is seeking a Senior Software Engineer in Westford, Massachusetts to improve their HPC infrastructure. The role includes designing scalable systems and supporting multi-cloud environments. The ideal candidate will have 10+ years of experience, strong software development... 
    Senior

    NVIDIA

    Santa Clara, CA
    5 days ago
  • $267k - $356k

     ...designated work from home day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-class storage offerings...  ...years of experience operating Linux systems in production or HPC environments, with hands-on storage experience at scale on scale... 
    Senior
    Work experience placement
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  • $255k - $340k

     ...designated work from home day is currently Tuesday.Hardware Engineering at Lambda is responsible for building and scaling the...  ...technical lead for integrating OEM and white-label HPC AI/ML, general purpose compute, storage, and network hardware into Lambda’s HPC platform... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  • $176k - $276k

    Production engineering is a field that involves crafting, building, and maintaining large-scale...  ...and systems engineering practices, storage, data management, and services. Professionals...  ...systems and ensure low-latency data access for HPC and AI/ML workloads.Storage Production... 
    Senior
    Full time
    Flexible hours

    Nvidia

    Santa Clara, CA
    5 days ago
  • $182k - $319k

     ...City: San Jose General OverviewFunctional Area: Engineering (ENG)Career Stream: Engineering (ENG)Role: Senior Principal (SPR)Job Title: Senior Principal, Design...  ...leading-edge Hardware Platform Solutions in Networking, Storage, and Server solutions from general purpose to... 
    Senior
    Local area

    Celestica

    San Jose, CA
    4 days ago
  • DDN is seeking a Senior Platform Product Manager to lead the strategy, roadmap, and execution...  .... You will shape how our hardware and storage media evolve to support AI, HPC, and enterprise workloads at scale, working across engineering, marketing, sales, and customer success.... 
    Senior

    DDN

    Santa Clara, CA
    5 days ago
  • $135k - $160k

     ...a Top Tier provider of advanced server, storage, and networking solutions for Data Center...  ...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We...  ...seek talented, passionate, and committed engineers, technologists, and business leaders to... 
    Senior
    Worldwide

    Supermicro

    San Jose, CA
    5 days ago
  • $108k - $155k

     ...Solutions (VES) has a rich history of leadership in providing storage and storage server platforms to hyperscale and enterprise data...  ...countries on six continents. Job Purpose:In this role, the System Test Engineer will participate in product validation of cutting edge data... 
    Senior
    Work experience placement
    Worldwide

    Sanmina-SCI

    San Jose, CA
    2 days ago
  • CoreWeave is seeking a Senior Software Engineer II for its Storage team to build and operate high-performance, multi-tenant storage systems. You will work on the NFS and block storage paths, CSI integration, and the control plane, using Go and Rust within a Kubernetes environment... 
    Senior

    CoreWeave

    Sunnyvale, CA
    2 days ago
  • $95k - $150k

    Position: Senior Signal Integrity Engineer Location: Santa Clara, CA Amphenol High Speed Products Group is the market leader for high speed, high bandwidth...  ...for the Telecom/Datacom market (Mobile Networks, Storage, Servers, Routers, Switches, etc.). Our products help to... 
    Senior
    Temporary work

    Amphenol ICC

    Santa Clara, CA
    3 days ago
  • $140k - $165k

     ...a Top Tier provider of advanced server, storage, and networking solutions for Data Center...  ...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We...  ...seek talented, passionate, and committed engineers, technologists, and business leaders to... 
    Senior
    Worldwide

    Supermicro

    San Jose, CA
    1 day ago
  • $137k - $156k

     ...a Top Tier provider of advanced server, storage, and networking solutions for Data Center...  ...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We...  ...seek talented, passionate, and committed engineers, technologists, and business leaders to... 
    Senior
    Temporary work
    Work experience placement
    Work at office
    Worldwide

    Supermicro

    San Jose, CA
    2 days ago
  • $148.7k - $201.2k

     ...services such as Amazon’s Simple Storage Service (S3) and Amazon...  ...The Nitro Team is looking for engineers with systems knowledge and experience...  ....The Nitro High Memory and HPC team owns the purpose built platform...  ...-sharing and mentorship. Our senior members enjoy one-on-one... 
    Internship
    Local area
    Flexible hours

    Amazon

    Santa Clara, CA
    2 days ago
  • $182k - $242k

     ...with the internal and customer engineering teams, offering valuable insights...  ...development. About the role: As a Senior Specialist Field Engineer at...  ...offerings, focusing on storage technologies within high-performance compute (HPC) environments Collaborate closely... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    9 days ago
  • Oracle Cloud Infrastructure’s Object Storage Service team is seeking a senior engineer to own software design and development for major components in a large-scale, distributed storage platform. You will be a hands-on coder who values simplicity, scalability, and collaborative... 
    Senior

    Ll Oefentherapie

    Santa Clara, CA
    2 days ago
  •  ...cloud, and our mission is to provide industry leading compute, storage, networking, database, security, and foundational cloud-based services...  .... As part of the Object Storage Service team, we seek hands-on engineers to tackle distributed systems, storage, and high availability... 
    Senior

    Oracle

    Santa Clara, CA
    4 days ago
  • $100k - $150k

     ...a Top Tier provider of advanced server, storage, and networking solutions for Data Center...  ...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We...  ...seek talented, passionate, and committed engineers, technologists, and business leaders to... 
    Senior
    Worldwide

    Supermicro

    San Jose, CA
    1 day ago
  • $100k - $150k

     ...a Top Tier provider of advanced server, storage, and networking solutions for Data Center...  ...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We...  ...seek talented, passionate, and committed engineers, technologists, and business leaders to... 
    Senior
    Work experience placement
    Work at office
    Worldwide
    Shift work

    Supermicro

    San Jose, CA
    4 days ago
  • $112.5k - $163k

     ...Senior Mechanical Design Engineer QuantumScape is on a mission to transform energy storage with solid-state lithium-metal battery technology. The company's next-generation batteries are designed to enable greater energy density, faster charging and enhanced safety to... 
    Senior

    QuantumScape Corporation

    San Jose, CA
    4 days ago
  • $100k - $138k

     ...a Top Tier provider of advanced server, storage, and networking solutions for Data Center...  ...Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We...  ...seek talented, passionate, and committed engineers, technologists, and business leaders to... 
    Senior
    Work at office
    Worldwide

    Supermicro

    San Jose, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior HPC Storage Engineer. Be the first to apply!