Senior GPU Systems Engineer: Scale AI Clusters & HPC
Career Techniques
Career Techniques in New York seeks an experienced infrastructure engineer to design, deploy, and scale large-scale GPU clusters for AI research. You will work across compute, storage, OS, and automation to support hundreds of petabytes and thousands of nodes. You will profile GPU workloads, remove bottlenecks, and collaborate with researchers to translate findings into speedups. Expect to own end-to-end infrastructure projects from design through long-term support and vendor engagement. #J-18808-Ljbffr Career Techniques
- ...compute, storage, and automation behind large-scale trading workloads. You will join engineers responsible for hundreds of petabytes of storage and thousands of GPU-accelerated nodes, shaping architectures for AI clusters and performance profiling. Collaboration with researchers...Suggested
- ...a skilled professional in New York to design and operate large-scale GPU infrastructure for model inference and reinforcement learning. The role demands several years of experience in deploying GPU systems, optimizing model performance, and working with frameworks like...Senior
- Tower Research Capital seeks an accomplished engineer to design, deploy, and scale distributed GPU clusters, building out hardware selection, production operations, and monitoring across thousands of nodes. You will diagnose bottlenecks across compute, storage, and network...Suggested
$153k - $204k
...Essential Cloud for AI. Built for... ...innovators to build and scale AI with... ...You'll Do:The Systems Engineering team owns the host... ...of the largest GPU fleets in the... ...framework into HPC verification,... ...the role:As a Senior Software Engineer... ...experience.HPC or large-cluster experience —...SeniorPermanent employmentFull timeTemporary workCasual workLive inWork at officeFlexible hours- ...first enterprise AI company. We... ...are building AI systems. We believe that... ...of researchers, engineers, designers, and... ...looking for a senior engineer to help... ...our frontier-scale language models... ...distributed systems, and HPC infrastructure.... ...on multi-node clusters (e.g., GB200/30...SeniorFull timeWork at officeLocal areaRemote workHome office
- ...Senior GPU Systems / AI Infrastructure Engineer (NYC) Location: New York City (Hybrid / On-site... ...infrastructure powering large-scale model training and... ...training (multi-node, multi-GPU clusters) Improve memory... ...in systems engineering, HPC, GPU computing, or AI infrastructure...Full time
- ...SR. HIGH PERFORMANCE COMPUTING (HPC) SYSTEMS ENGINEER SpaceX is looking for an HPC Systems... ...: Administer and manage HPC clusters, storage systems, and high-speed... ...CFD, FEA) Familiarity with large scale AI training Familiarity with GPU usage in a compute cluster and Cuda...SeniorPermanent employmentFlexible hoursWeekend work
$108k - $172.5k
...unlimited potential of AI to define the... ...era in which our GPU acts as the... ...highly motivated Senior HPC Support Engineer focussing on InfiniBand... ...and supporting systems using Linux Operating... ...on large-scale networking and AI... ...and GPU Technology.Clustering or HPC Data-Center...SeniorFull timeWork experience placement- Scale AI is hiring for a senior software engineer to design, build, and scale full‑stack systems powering our GenAI data engine. You will work across front-end, back-end, and infrastructure, using React, TypeScript, Node.js, Python and data stores like MongoDB, Elasticsearch...Senior
$150k - $300k
Hudson River Trading (HRT) is looking for GPU Systems Engineers to help scale and evolve our exceptionally sophisticated HPC/AI research environment. Joining our Research and... ...-scale storage and massive CPU and GPU clusters in globally distributed data centers. As such...Work at officeLocal areaImmediate start$200k - $300k
...systematic trading and engineering talent. We empower... ...the economies of scale that come from a... ..., operating systems, and automation behind... ...and large CPU and GPU clusters spanning thousands... ...architecture of a new AI cluster, the next... ...Linux systems in HPC, AI, or...Work experience placementCasual workWork at office$125k - $250k
...Senior Account Executive- GPU/AI Infrastructure Senior Account Executive - GPU and AI Infrastructure... ...era. Operating some of the largest GPU clusters globally, we deliver high-performance... ...Enterprise GPU clusters for large-scale AI initiatives On-demand GPU compute...SeniorTemporary workRemote workFlexible hours$182k - $242k
...Essential Cloud for AI. Built for... ...innovators to build and scale AI with confidence... ...What You'll Do:The Systems Engineering team owns the... ...one of the largest GPU fleets in the world... ...About the role:As a Senior Software Engineer... ...hardware, platform, HPC, and Fleet teams...SeniorPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours- Principal Senior Systems Engineer - HPE Networking (Mid-Telco Accounts)This role has been designated as... ...IP, transport, data center, and AI-ready networking environments.The ideal... ...expertise designing and supporting large-scale service provider networks including:BGPIS...SeniorFull timeWork experience placementWork at officeLocal areaImmediate startRemote workWork from home
- ...entity. Responsibilities As a senior Machine Learning Systems Engineer on the Search Platform team, you... ...retrieval workflows. Partner with Rovo and AI platform teams to evolve search... ...quality, freshness, and relevance at scale.Operational Excellence & Cost DisciplineDrive...SeniorWork at officeLocal area
- ...innovation. Join us!About the RoleWe’re looking for a Senior Design Systems Engineer II to shape and scale the visual and interaction foundations of Blink’s... ...JavaScript, and modern front-end tooling.Expert use of modern AI-assisted tools (e.g., Cursor, GitHub Copilot, ChatGPT...Senior
$112.9k - $155.24k
.... SCOPE OF POSITION The Senior Systems Engineer is responsible for working directly... ...team to shape the future of AI infrastructure in hyperscale data... ...AI/ML fabric architectures (scale-up / scale-out), including GPU/XPU cluster designs and optical I/O requirements...SeniorFull timeWork experience placement- ...scalable worker platform in NYC. We’re seeking engineers with 5+ years in software development and deep experience with distributed systems and performance tuning. You’ll work... ...from first principles, enjoy operating large-scale bare-metal fleets, and are excited #J-188...Senior
- You are one of the first systems engineers at a founder‑led AI firm building the execution layer for enterprise... ...on tomorrow. You are coming in at a senior level, which means you are making... ...written the runbook. When something scales, it is because you planned for it. Day...Senior
$85k - $140k
Nebius is seeking a Data Center Support Engineer to manage large-scale bare metal GPU infrastructure supporting AI workloads. The role involves troubleshooting physical servers, Linux systems, networking components, and maintaining operational excellence across the data...Senior$160k - $220k
...Every Identity, from AI to HumanIdentity is... ...are too, let's talk.Senior Database Reliability Engineer (DBRE) Experience Level... ...in PostgreSQL at scale and solid experience... ...scale, mission-critical systems. You will work... ...available PostgreSQL clusters (physical replication...SeniorPermanent employmentWork at officeLocal areaWorldwideFlexible hours$152k - $241.5k
We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency... ...and operating large scale compute... ...experience with NSight Systems and NSight ComputeExperience... ...Lustre and GPFS for AI/HPC workloadsFamiliarity with...SeniorFull timeRemote work$184k - $287.5k
...computing platforms for AI and HPC. Because of our... ..., and engineers can push the boundaries... ...motivated Senior Solutions Architect... ...with a focus on GPU, NVLink, and... ...generation GPU-based clusters enabling the... ...designing large-scale distributed systems, AI clusters, or...SeniorFull timeRemote work$150k - $250k
...systematic trading and engineering talent. We empower portfolio... ...the economies of scale that come from a large,... ...and optimize Kubernetes clusters in cloud environments Automate... ...to-cloud migrations of HPC and distributed... ...Infrastructure as Code, Linux systems, cloud networking, and...SeniorCasual workWork at office$184k - $287.5k
...seeking outstanding AI Solutions... ...work across product, engineering, sales, developer... ...advisor for accelerated systems architecture, GPU and networking systems, cluster design,... ..., including large-scale clustersSupport infrastructure... ...customers on AI/HPC infrastructureYour...SeniorFull timeRemote work$150k - $250k
...This opportunity is ideal for an engineer looking to build and scale cloud platforms supporting high-... ...manage and optimize Kubernetes clusters and Linux-based platforms. You... ...high-performance computing (HPC) or distributed systems. You should have knowledge of...Full time$112.9k - $155.24k
Senior Systems Engineer Co-Packaged Optics Locations: Santa Clara and Bay Area (West Coast... ...team to shape the future of AI infrastructure in hyperscale data... ...with AI/ML fabric architectures (scale-up / scale-out), including GPU/XPU cluster designs and optical I/O requirements...SeniorFull timeWork experience placement- Viridien is seeking a highly experienced HPC Data Center Senior Linux IT Specialist to join the IT team, contributing to global HPC and Digital... .... The role focuses on Linux administration, storage systems, and IT service management within a high-availability environment...Senior
$224k - $356.5k
...Performance Computing and AI workloads across... ...with NVIDIA Engineering, Product, and... ...next-generation GPU architectures.... ...DL, recommender systems, GNN, monte-... ...ML/DL models at scale on on-prem or public... ...cloud computing clusters in... ...developers building AI, HPC, or data analytics...SeniorFull time$224k - $356.5k
...Performance Computing and AI workloads across... ...with NVIDIA Engineering, Product, and Sales... ...next-generation GPU architectures.Work... ...ML/DL, recommender systems, GNN, monte-carlo... ...deploying ML/DL models at scale on public cloud... ...and/or on-prem HPC clusters in productionWays...SeniorFull timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior GPU Systems Engineer: Scale AI Clusters & HPC. Be the first to apply!
- senior windows systems engineer New York, NY
- software system engineer New York, NY
- mission system engineer New York, NY
- healthcare systems engineer New York, NY
- operating system engineer New York, NY
- system engineer remote New York, NY
- application system engineer New York, NY
- advanced systems engineer New York, NY
- active directory systems engineer New York, NY
- systems engineer intern New York, NY


