Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Storage Engineer - AI Infra & GPU Clusters (Remote)

Jobleads-US

Hamilton Barnes is seeking a Senior Storage Engineer to own the high-performance storage layer for large-scale GPU clusters and AI workloads. You will design, deploy, and operate storage platforms, collaborating with infrastructure, compute, and networking teams to scale performance.

The role focuses on optimizing distributed storage, troubleshooting bottlenecks, and driving resiliency and data protection through automation and observability in petabyte-scale environments.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior Storage Engineer - AI Infra & GPU Clusters (Remote) in San Francisco, CA vacancy
  • $250k

     ...Join a rapidly scaling AI cloud infrastructure...  ...a next-generation GPU platform designed...  ...company is looking for a Senior / Staff Site Reliability Engineer to support and scale...  ...for GPU compute clusters Collaborate with...  ...options Bonus  Remote working option and allowance... 
    Remote work
    Senior
    Full time
    San Francisco, CA
    more than 2 months ago
  •  ...SpaceX in Hawthorne, CA seeks a Senior Software Engineer for AI infrastructure within Starshield. You will...  ...design, operate, and scale software and GPU infrastructure supporting critical...  ...missions, including on-prem Kubernetes clusters and AI services. You will lead... 
    Senior

    Jobleads-US

    Hawthorne, CA
    1 day ago
  • Inferact Inc. is building a world-class GPU compute platform powering vLLM. We seek a hands-on cluster administration engineer to own the high-performance infrastructure that...  ...provider deployments. This San Francisco-based, remote-friendly role offers strong compensation and... 
    Remote work
    Senior

    Inferact Inc.

    Brooklyn, NY
    1 day ago
  •  ...Hamilton Barnes is seeking a Senior Network Engineer to design, deploy, and...  ...throughput networks for large GPU clusters. You will work with NVIDIA...  ...environments optimized for AI and HPC traffic. The role...  ...Terraform). This position offers remote options and a competitive... 
    Remote job
    Senior

    Jobleads-US

    San Francisco, CA
    2 days ago
  •  ...NVIDIA Corporation is seeking a Senior Customer Success Engineer for the DGX Cloud organization. You will design and implement distributed cloud infrastructure...  ...code, and build tooling to automate workflows and manage GPU capacity. You will collaborate with internal research,... 
    Remote job
    Senior

    Jobleads-US

    Santa Clara, CA
    1 day ago
  • $168k - $200k

     ...records to powering the AI revolution in...  ...We're looking for a Senior Site Reliability Engineer to join our Data & ML...  ...pipelines, ML workflows, and infra components using...  ...including workspace setup, cluster/job management, and...  ...Feature Stores, and GPU workload... 
    Remote work
    Senior

    Datavant

    United States
    1 day ago
  • $15k

     ...company that applies state-of-the-art AI and machine learning techniques to...  ...catered lunches, and more.   As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our...  ...) ~ Experience with distributed storage technologies (Lustre, Ceph, S3)... 
    Remote work
    Senior
    Work at office
    Local area

    The Voleon Group

    United States
    4 days ago
  •  ...NVIDIA Corporation seeks a Senior AI/ML Performance and Efficiency Engineer for GPU Clusters to advance AI efficiency across research workloads. You will partner with...  ...demands 5+ years in managing large-scale compute infra, strong ML tooling knowledge, and hands-on... 
    Senior

    Jobleads-US

    Santa Clara, CA
    1 day ago
  • $175k - $250k

     ...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must...  ...with generative AI. They are the team behind...  ...Manage and automate GPU compute clusters using tools such as...  ..., distributed storage, and networking Ensure... 
    Remote work
    Senior
    Full time
    Relocation
    Relocation package

    The Recruiting Guy

    Washington DC
    1 day ago
  •  ...AI companies need inference that’s fast, reliable...  ...compatible API, we pool GPU capacity from providers...  ..., reliability is an engineering problem that spans the...  ...provisioning, networking, storage, and service deployment...  ...that cross application, cluster, network, and hardware... 
    Remote work
    Senior

    Parasail

    United States
    4 days ago
  •  ...application changes, infra, network changes...  ...and Mongo DB Seniority level ~...  ...Site Reliability Engineer” roles. Bellevue...  ...Reliability Engineer, AI/ML Platforms Senior...  ...- AI Research Clusters Redmond, WA $18...  ...Reliability Engineer (SRE, Remote US) Seattle, WA... 
    Remote work
    Senior
    Contract work

    Signature IT World Inc

    Washington DC
    1 day ago
  • $165k - $200k

     ...vertically integrated AI infrastructure...  ...energy, detail-oriented Senior Network Production Engineer to support the physical...  ...performance compute (HPC) and GPU-based AI...  ...comfortable both on-site and remote, follows and helps...  ...testing (SAT) on new clusters, chase down failures,... 
    Remote work
    Senior
    Full time
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  •  ...for enterprises and AI innovators around...  ...Compute, Cloud GPU, Bare Metal, and Cloud Storage solutions. In December...  ...$500 stipend for remote office setup in...  ...a highly skilled Senior Site Reliability Engineer, Databases to join...  ...spanning MySQL InnoDB Clusters, PostgreSQL, and... 
    Remote work
    Senior
    Full time
    Work at office
    Immediate start
    Flexible hours

    Vultr

    Remote
    8 days ago
  •  ...Senior Site Reliability Engineer Canonical is a leading provider...  ...cloud, data science, AI, engineering...  ...with pure Python infra-as-code, from bare...  ...and application clusters for customers across...  ...: Globally remote role The role...  ...software defined storage, and we enable devsecops... 
    Remote work
    Senior
    Work at office
    Local area
    Work from home
    Worldwide

    Canonical

    United States
    1 day ago
  •  ...Kubernetes-native AI infrastructure company...  ...empowers platform engineering teams to deliver...  ...the automation, GPU orchestration, and...  ...high-performance storage for GPU-accelerated...  ...storage, wiring it into clusters via CSI, and...  ...are looking for a senior DevOps engineer who... 
    Remote job
    Senior
    Local area

    grabjobs

    Oak Grove, MO
    2 days ago
  •  ...Senior Site Reliability Engineer We are looking for a senior Kubernetes...  ...plane for enterprise GPU infrastructure. You will...  ..., and pre-production clusters fast and reproducible...  ...on how k0rdent AI is deployed and operated...  ...workloads, networking, storage, RBAC, resource... 
    Remote work
    Senior
    Local area

    Mirantis

    United States
    3 days ago
  • $125k - $250k

     ...Senior Account Executive- GPU/AI Infrastructure Senior Account Executive - GPU and AI Infrastructure Location: Remote within the USA Compensation: $125k-$250k base + bonus + benefits...  ...Operating some of the largest GPU clusters globally, we deliver high-performance... 
    Remote work
    Senior
    Temporary work
    Flexible hours

    ESR Healthcare

    New York, NY
    5 days ago
  •  ...NVIDIA Corporation seeks a Production Storage Engineer to design, deploy, and optimize large-scale storage clusters for GPU-accelerated AI/ML workloads in Santa Clara, CA. You will build and maintain monitoring, logging, and alerting, ensuring data integrity, low latency... 
    Senior

    Jobleads-US

    Santa Clara, CA
    1 day ago
  • $140k - $215k

     ...the world's most advanced AI-native platform. We work...  ...seeking a highly skilled Senior Linux Systems Engineer to design, build, monitor...  ...scale cloud environments for clustered object storage solutions while...  ...business outcomes. #LI-LY1 #LI-Remote This role will require the... 
    Remote job
    Senior
    Work experience placement
    Work at office
    Local area

    CrowdStrike Holdings, Inc.

    Brooklyn, NY
    3 days ago
  • $153k - $191.3k

     ...processing, and software engineering, our office is a truly...  ...with employees working remotely world wide and joining...  ...optimization, management, and cluster tuning in a constrained...  ...with CUDA-based GPU programs Security expertise...  ...to assist you. AI in Our Interviewing... 
    Remote work
    Senior
    Full time
    Temporary work
    For contractors
    Work at office
    Local area
    Home office
    3 days per week

    Planet Labs PBC

    San Francisco, CA
    7 hours ago
  •  ...Job Description Job Description Senior Product Manager – GPU Products & AI Infrastructure Why This...  ...role in defining the GPU products, clusters, and services designed to support...  ...technology, you will partner with engineering and industry technology leaders to... 
    Remote work
    Senior

    Globalchannelmanagement

    Cambridge, MA
    a month ago
  •  ...Senior GPU Systems / AI Infrastructure Engineer (NYC) Location: New York City (Hybrid / On-site preferred) Comp...  ...equity (Series A-C / high-growth AI infra) About the Role We’re hiring...  ...training (multi-node, multi-GPU clusters) Improve memory bandwidth utilisation... 
    Full time
    New York, NY
    more than 2 months ago
  •  ...Job Description Senior Product Manager – Next...  ...Infrastructure & GPU Platforms We are...  ...work closely with engineering, architecture, sales...  ...products supporting AI, HPC, enterprise, and...  ...AI/HPC workloads, cluster scaling, and...  ...Linux, networking, storage, containers, multi-... 
    Remote work
    Senior

    Globalchannelmanagement

    Cambridge, MA
    a month ago
  •  ...SpaceX is seeking a Senior Software Engineer for AI Infrastructure (Starshield) to design, operate, and scale GPU/CPU infrastructures supporting critical national security missions...  ...across teams to deliver reliable AI clusters. The role emphasizes Kubernetes, Terraform... 
    Senior

    Jobleads-US

    Washington DC
    1 day ago
  •  ...SpaceX is hiring a Sr. Software Engineer for AI Infrastructure (Starshield) in Palo Alto to design, deploy, and scale software and GPU infrastructure supporting critical national security...  ...on-prem resources, build scalable AI clusters, and collaborate with AI engineers to... 
    Senior

    Jobleads-US

    Palo Alto, CA
    1 day ago
  •  ...SpaceX is seeking a Sr. Software Engineer focused on AI infrastructure for Starshield. This role designs, operates, and scales GPU-enabled on-premises infrastructure and Kubernetes-based AI clusters to support national security missions. You will automate deployments... 
    Senior

    Jobleads-US

    Kentucky
    1 day ago
  •  ...SpaceX is seeking a Sr. Software Engineer, AI Infrastructure (Starshield) in Redmond, WA to design, operate, and scale GPU-backed AI infrastructure supporting critical national...  ...building automated on-premise Kubernetes clusters, managing GPU services, and collaborating... 
    Senior

    Jobleads-US

    Redmond, WA
    1 day ago
  • $224k - $356.5k

     ...the boundless possibilities of AI to build the next era of computing. An era where our GPU functions as the intelligence behind...  ...for highly motivated, creative engineers to join the Platform Software...  ...Clara; US, OR, Hillsboro; US, CA, Remote; US, WA, RedmondType: Full time
    Remote work
    Senior
    Full time

    Nvidia

    Redmond, WA
    2 days ago
  •  ...Liftoff is a leading AI-powered performance marketing platform for the mobile app economy. As a Senior Staff Software Engineer on the Serving team, you will design, build, and operate high...  ...in real time, while optimizing GPU-powered inference pipelines. You will... 
    Remote job
    Senior

    Jobleads-US

    Kentucky
    2 days ago
  • $152k - $241.5k

     ..., CA, Santa Clara US, Remote Full time JR1997214...  ...and Visualization. The GPU, our invention, serves...  ...motivated Performance engineer to influence the roadmap...  ...multi-GPU and multi-node clusters. Study the interaction...  ...vacancy. NVIDIA uses AI tools in its recruiting... 
    Remote work
    Senior
    Full time

    NVIDIA

    Santa Clara, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Storage Engineer - AI Infra & GPU Clusters (Remote). Be the first to apply!