Senior Storage Engineer - AI Infra & GPU Clusters (Remote)
Jobleads-US
- Remote job
Hamilton Barnes is seeking a Senior Storage Engineer to own the high-performance storage layer for large-scale GPU clusters and AI workloads. You will design, deploy, and operate storage platforms, collaborating with infrastructure, compute, and networking teams to scale performance.
The role focuses on optimizing distributed storage, troubleshooting bottlenecks, and driving resiliency and data protection through automation and observability in petabyte-scale environments.
#J-18808-Ljbffr Jobleads-US$250k
...Join a rapidly scaling AI cloud infrastructure... ...a next-generation GPU platform designed... ...company is looking for a Senior / Staff Site Reliability Engineer to support and scale... ...for GPU compute clusters Collaborate with... ...options Bonus Remote working option and allowance...Remote workSeniorFull time- ...SpaceX in Hawthorne, CA seeks a Senior Software Engineer for AI infrastructure within Starshield. You will... ...design, operate, and scale software and GPU infrastructure supporting critical... ...missions, including on-prem Kubernetes clusters and AI services. You will lead...Senior
- Inferact Inc. is building a world-class GPU compute platform powering vLLM. We seek a hands-on cluster administration engineer to own the high-performance infrastructure that... ...provider deployments. This San Francisco-based, remote-friendly role offers strong compensation and...Remote workSenior
- ...Hamilton Barnes is seeking a Senior Network Engineer to design, deploy, and... ...throughput networks for large GPU clusters. You will work with NVIDIA... ...environments optimized for AI and HPC traffic. The role... ...Terraform). This position offers remote options and a competitive...Remote jobSenior
- ...NVIDIA Corporation is seeking a Senior Customer Success Engineer for the DGX Cloud organization. You will design and implement distributed cloud infrastructure... ...code, and build tooling to automate workflows and manage GPU capacity. You will collaborate with internal research,...Remote jobSenior
$168k - $200k
...records to powering the AI revolution in... ...We're looking for a Senior Site Reliability Engineer to join our Data & ML... ...pipelines, ML workflows, and infra components using... ...including workspace setup, cluster/job management, and... ...Feature Stores, and GPU workload...Remote workSenior$15k
...company that applies state-of-the-art AI and machine learning techniques to... ...catered lunches, and more. As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our... ...) ~ Experience with distributed storage technologies (Lustre, Ceph, S3)...Remote workSeniorWork at officeLocal area- ...NVIDIA Corporation seeks a Senior AI/ML Performance and Efficiency Engineer for GPU Clusters to advance AI efficiency across research workloads. You will partner with... ...demands 5+ years in managing large-scale compute infra, strong ML tooling knowledge, and hands-on...Senior
$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must... ...with generative AI. They are the team behind... ...Manage and automate GPU compute clusters using tools such as... ..., distributed storage, and networking Ensure...Remote workSeniorFull timeRelocationRelocation package- ...AI companies need inference that’s fast, reliable... ...compatible API, we pool GPU capacity from providers... ..., reliability is an engineering problem that spans the... ...provisioning, networking, storage, and service deployment... ...that cross application, cluster, network, and hardware...Remote workSenior
- ...application changes, infra, network changes... ...and Mongo DB Seniority level ~... ...Site Reliability Engineer” roles. Bellevue... ...Reliability Engineer, AI/ML Platforms Senior... ...- AI Research Clusters Redmond, WA $18... ...Reliability Engineer (SRE, Remote US) Seattle, WA...Remote workSeniorContract work
$165k - $200k
...vertically integrated AI infrastructure... ...energy, detail-oriented Senior Network Production Engineer to support the physical... ...performance compute (HPC) and GPU-based AI... ...comfortable both on-site and remote, follows and helps... ...testing (SAT) on new clusters, chase down failures,...Remote workSeniorFull timeTemporary work- ...for enterprises and AI innovators around... ...Compute, Cloud GPU, Bare Metal, and Cloud Storage solutions. In December... ...$500 stipend for remote office setup in... ...a highly skilled Senior Site Reliability Engineer, Databases to join... ...spanning MySQL InnoDB Clusters, PostgreSQL, and...Remote workSeniorFull timeWork at officeImmediate startFlexible hours
- ...Senior Site Reliability Engineer Canonical is a leading provider... ...cloud, data science, AI, engineering... ...with pure Python infra-as-code, from bare... ...and application clusters for customers across... ...: Globally remote role The role... ...software defined storage, and we enable devsecops...Remote workSeniorWork at officeLocal areaWork from homeWorldwide
- ...Kubernetes-native AI infrastructure company... ...empowers platform engineering teams to deliver... ...the automation, GPU orchestration, and... ...high-performance storage for GPU-accelerated... ...storage, wiring it into clusters via CSI, and... ...are looking for a senior DevOps engineer who...Remote jobSeniorLocal area
- ...Senior Site Reliability Engineer We are looking for a senior Kubernetes... ...plane for enterprise GPU infrastructure. You will... ..., and pre-production clusters fast and reproducible... ...on how k0rdent AI is deployed and operated... ...workloads, networking, storage, RBAC, resource...Remote workSeniorLocal area
$125k - $250k
...Senior Account Executive- GPU/AI Infrastructure Senior Account Executive - GPU and AI Infrastructure Location: Remote within the USA Compensation: $125k-$250k base + bonus + benefits... ...Operating some of the largest GPU clusters globally, we deliver high-performance...Remote workSeniorTemporary workFlexible hours- ...NVIDIA Corporation seeks a Production Storage Engineer to design, deploy, and optimize large-scale storage clusters for GPU-accelerated AI/ML workloads in Santa Clara, CA. You will build and maintain monitoring, logging, and alerting, ensuring data integrity, low latency...Senior
$140k - $215k
...the world's most advanced AI-native platform. We work... ...seeking a highly skilled Senior Linux Systems Engineer to design, build, monitor... ...scale cloud environments for clustered object storage solutions while... ...business outcomes. #LI-LY1 #LI-Remote This role will require the...Remote jobSeniorWork experience placementWork at officeLocal area$153k - $191.3k
...processing, and software engineering, our office is a truly... ...with employees working remotely world wide and joining... ...optimization, management, and cluster tuning in a constrained... ...with CUDA-based GPU programs Security expertise... ...to assist you. AI in Our Interviewing...Remote workSeniorFull timeTemporary workFor contractorsWork at officeLocal areaHome office3 days per week- ...Job Description Job Description Senior Product Manager – GPU Products & AI Infrastructure Why This... ...role in defining the GPU products, clusters, and services designed to support... ...technology, you will partner with engineering and industry technology leaders to...Remote workSenior
- ...Senior GPU Systems / AI Infrastructure Engineer (NYC) Location: New York City (Hybrid / On-site preferred) Comp... ...equity (Series A-C / high-growth AI infra) About the Role We’re hiring... ...training (multi-node, multi-GPU clusters) Improve memory bandwidth utilisation...Full time
- ...Job Description Senior Product Manager – Next... ...Infrastructure & GPU Platforms We are... ...work closely with engineering, architecture, sales... ...products supporting AI, HPC, enterprise, and... ...AI/HPC workloads, cluster scaling, and... ...Linux, networking, storage, containers, multi-...Remote workSenior
- ...SpaceX is seeking a Senior Software Engineer for AI Infrastructure (Starshield) to design, operate, and scale GPU/CPU infrastructures supporting critical national security missions... ...across teams to deliver reliable AI clusters. The role emphasizes Kubernetes, Terraform...Senior
- ...SpaceX is hiring a Sr. Software Engineer for AI Infrastructure (Starshield) in Palo Alto to design, deploy, and scale software and GPU infrastructure supporting critical national security... ...on-prem resources, build scalable AI clusters, and collaborate with AI engineers to...Senior
- ...SpaceX is seeking a Sr. Software Engineer focused on AI infrastructure for Starshield. This role designs, operates, and scales GPU-enabled on-premises infrastructure and Kubernetes-based AI clusters to support national security missions. You will automate deployments...Senior
- ...SpaceX is seeking a Sr. Software Engineer, AI Infrastructure (Starshield) in Redmond, WA to design, operate, and scale GPU-backed AI infrastructure supporting critical national... ...building automated on-premise Kubernetes clusters, managing GPU services, and collaborating...Senior
$224k - $356.5k
...the boundless possibilities of AI to build the next era of computing. An era where our GPU functions as the intelligence behind... ...for highly motivated, creative engineers to join the Platform Software... ...Clara; US, OR, Hillsboro; US, CA, Remote; US, WA, RedmondType: Full timeRemote workSeniorFull time- ...Liftoff is a leading AI-powered performance marketing platform for the mobile app economy. As a Senior Staff Software Engineer on the Serving team, you will design, build, and operate high... ...in real time, while optimizing GPU-powered inference pipelines. You will...Remote jobSenior
$152k - $241.5k
..., CA, Santa Clara US, Remote Full time JR1997214... ...and Visualization. The GPU, our invention, serves... ...motivated Performance engineer to influence the roadmap... ...multi-GPU and multi-node clusters. Study the interaction... ...vacancy. NVIDIA uses AI tools in its recruiting...Remote workSeniorFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Storage Engineer - AI Infra & GPU Clusters (Remote). Be the first to apply!
- senior living director San Francisco, CA
- senior php developer remote San Francisco, CA
- senior manager customer operations San Francisco, CA
- senior support engineer San Francisco, CA
- senior product manager mobile San Francisco, CA
- senior software engineer ruby on rails San Francisco, CA
- sr finance manager San Francisco, CA
- sr marketing manager San Francisco, CA
- senior customer service San Francisco, CA
- senior business manager San Francisco, CA



