Senior GPU Systems Engineer: Scale AI Clusters & HPC
Career Techniques
Career Techniques in New York seeks an experienced infrastructure engineer to design, deploy, and scale large-scale GPU clusters for AI research. You will work across compute, storage, OS, and automation to support hundreds of petabytes and thousands of nodes. You will profile GPU workloads, remove bottlenecks, and collaborate with researchers to translate findings into speedups. Expect to own end-to-end infrastructure projects from design through long-term support and vendor engagement. #J-18808-Ljbffr Career Techniques
- ...compute, storage, and automation behind large-scale trading workloads. You will join engineers responsible for hundreds of petabytes of storage and thousands of GPU-accelerated nodes, shaping architectures for AI clusters and performance profiling. Collaboration with researchers...Suggested
- ...a skilled professional in New York to design and operate large-scale GPU infrastructure for model inference and reinforcement learning. The role demands several years of experience in deploying GPU systems, optimizing model performance, and working with frameworks like...Senior
- A growing infrastructure company is seeking a Senior Systems Engineer to support the Department of Energy. This role involves guiding national... ...Strong technical presentation skills and understanding of HPC and AI/ML are crucial. Join us for this pivotal opportunity at...Senior
$295k
...first enterprise AI company. We... ...are building AI systems. We believe that... ...of researchers, engineers, designers, and... ...looking for a senior engineer to help... ...our frontier-scale language models... ...distributed systems, and HPC infrastructure.... ...on multi-node clusters (e.g., GB200/30...SeniorFull timeWork at officeLocal areaRemote workHome office$153k - $204k
...Essential Cloud for AI™. Built for... ...to build and scale AI with confidence... ...You'll Do The Systems Engineering team owns the... ...of the largest GPU fleets in the world... ...framework into HPC verification,... ...The Role As a Senior Software Engineer... .... HPC or large-cluster experience —...SeniorPermanent employmentFull timeTemporary workCasual workLive inWork at officeFlexible hours$108k - $172.5k
...unlimited potential of AI to define the... ...era in which our GPU acts as the... ...highly motivated Senior HPC Support Engineer focussing on InfiniBand... ...and supporting systems using Linux Operating... ...on large-scale networking and AI... ...and GPU Technology.Clustering or HPC Data-Center...SeniorFull timeWork experience placement- ...Senior GPU Systems / AI Infrastructure Engineer (NYC) Location: New York City (Hybrid / On-site... ...infrastructure powering large-scale model training and... ...training (multi-node, multi-GPU clusters) Improve memory... ...in systems engineering, HPC, GPU computing, or AI infrastructure...Full time
- ...the future of AI and HPC networking with... ...We’re seeking engineers who are energized... ...distributed software systems, and who are... ...performance at scale. Cornelis... ...efficiency of GPU, CPU and... ...-based compute clusters at any scale. Our... ...an experienced Senior Software Engineer...SeniorRemote jobFull timeFlexible hours
$150k - $300k
Hudson River Trading (HRT) is looking for GPU Systems Engineers to help scale and evolve our exceptionally sophisticated HPC/AI research environment. Joining our Research and... ...-scale storage and massive CPU and GPU clusters in globally distributed data centers. As such...Work at officeLocal areaImmediate start$168k - $270.25k
...Experience (NVEX) Solutions Engineering team is looking for... ...of NVIDIA’s GPU accelerated... ...will apply the latest AI technologies to triage... ...datacenters of rack-scale platforms, solve... ...GPU platformsStrong system software (firmware,... ...workloadsClustering or HPC data center...SeniorFull timeWeekend work$125k - $250k
...Senior Account Executive- GPU/AI Infrastructure Senior Account Executive - GPU and AI Infrastructure... ...era. Operating some of the largest GPU clusters globally, we deliver high-performance... ...Enterprise GPU clusters for large-scale AI initiatives On-demand GPU compute...SeniorTemporary workRemote workFlexible hours$152k - $241.5k
...the unlimited potential of AI to define the next era of... ...about building reliable systems software for cloud-scale GPU infrastructure, we... ...team. We are looking for a Senior Software Engineer to join our DGX Cloud / Fleet... ...operating software in AI, HPC, cloud, or large-scale...SeniorFull timeLocal areaRemote work$200k - $300k
...leading trading firms on an interesting GPU Systems Engineer position. This is a hands-on infrastructure role working at scale, with large CPU and GPU clusters spanning thousands of nodes and... ...with large-scale Linux systems in HPC, AI or distributed infrastructure environments...- ...you will join the engineers responsible for the... ..., operating systems, and automation behind... ...that work at serious scale: hundreds of... ...and large CPU and GPU clusters spanning thousands... ...architecture of a new AI cluster, the next... ...Linux systems in HPC, AI, or distributed...Work experience placement
- ...entity. Responsibilities As a senior Machine Learning Systems Engineer on the Search Platform team, you... ...retrieval workflows. Partner with Rovo and AI platform teams to evolve search... ...quality, freshness, and relevance at scale.Operational Excellence & Cost DisciplineDrive...SeniorWork at officeLocal area
- ...a Principal Infrastructure Engineer to own AI cluster performance and validation across... .... You’ll define healthy-at-scale criteria, set the architecture, and partner with senior leaders to production-... ...design scalable validation systems, and optimize fabric, storage...Senior
- ...innovation. Join us!About the RoleWe’re looking for a Senior Design Systems Engineer II to shape and scale the visual and interaction foundations of Blink’s... ...JavaScript, and modern front-end tooling.Expert use of modern AI-assisted tools (e.g., Cursor, GitHub Copilot, ChatGPT...Senior
- ...scalable worker platform in NYC. We’re seeking engineers with 5+ years in software development and deep experience with distributed systems and performance tuning. You’ll work... ...from first principles, enjoy operating large-scale bare-metal fleets, and are excited #J-188...Senior
$85k - $140k
Nebius is seeking a Data Center Support Engineer to manage large-scale bare metal GPU infrastructure supporting AI workloads. The role involves troubleshooting physical servers, Linux systems, networking components, and maintaining operational excellence across the data...Senior- ...The engineering team at Chainalysis is inspired... ...build a flexible, AI-driven platform that... ...our customers scale their workflows as... ...mainstream. As a Senior Infrastructure Engineer... ...on the Government Systems team - the group... ...RKE2 Kubernetes clusters across development...SeniorFlexible hours
$160k - $220k
...Every Identity, from AI to HumanIdentity is... ...are too, let's talk.Senior Database Reliability Engineer (DBRE) Experience Level... ...in PostgreSQL at scale and solid experience... ...scale, mission-critical systems. You will work... ...available PostgreSQL clusters (physical replication...SeniorPermanent employmentWork at officeLocal areaWorldwideFlexible hours$152k - $241.5k
We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency... ...and operating large scale compute... ...experience with NSight Systems and NSight ComputeExperience... ...Lustre and GPFS for AI/HPC workloadsFamiliarity with...SeniorFull timeRemote work$184k - $287.5k
...computing platforms for AI and HPC. Because of our... ..., and engineers can push the boundaries... ...motivated Senior Solutions Architect... ...with a focus on GPU, NVLink, and... ...generation GPU-based clusters enabling the... ...designing large-scale distributed systems, AI clusters, or...SeniorFull timeRemote work$184k - $287.5k
...seeking outstanding AI Solutions... ...work across product, engineering, sales, developer... ...advisor for accelerated systems architecture, GPU and networking systems, cluster design,... ..., including large-scale clustersSupport infrastructure... ...customers on AI/HPC infrastructureYour...SeniorFull timeRemote work- Blink Health is seeking a Senior Design Systems Engineer II to shape and scale the visual and interaction foundations of Blink’s digital ecosystem. You’ll partner with Product Design, Software Engineering, and cross‑functional teams to evolve the shared frontend design...Senior
$224k - $356.5k
...Performance Computing and AI workloads across... ...with NVIDIA Engineering, Product, and... ...next-generation GPU architectures.... ...DL, recommender systems, GNN, monte-... ...ML/DL models at scale on on-prem or public... ...cloud computing clusters in... ...developers building AI, HPC, or data analytics...SeniorFull time$224k - $356.5k
...Performance Computing and AI workloads across... ...with NVIDIA Engineering, Product, and Sales... ...next-generation GPU architectures.Work... ...ML/DL, recommender systems, GNN, monte-carlo... ...deploying ML/DL models at scale on public cloud... ...and/or on-prem HPC clusters in productionWays...SeniorFull timeRemote work$152k - $241.5k
...changing the world of AI Networking with... ...for Data Center scale systems for AI training/inference... .../ Software engineering or equivalent experience... ....Familiarity with cluster orchestration... ...Knowledge in MPI and HPC, InfiniBand,... ...invention of the GPU - the engine of modern...SeniorFull timeRemote work$139k - $257.55k
...seeking an outstanding Senior Site Reliability Engineer (SRE) to support... ...learning, autonomous AI workflows, and cloud-... ...native, containerized systems as we build the framework... ..., and operate large-scale, distributed, fault-... ...infrastructure — model serving, GPU workloads, language...SeniorFull timeTemporary workLocal areaRemote workWorldwide- A dynamic technology firm in New York is seeking a talented Senior/Staff level Systems Engineer to develop and scale a dedicated cloud for CI workloads. The role offers an opportunity to solve complex systems problems and build a new CI cloud from the ground up. Candidates...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior GPU Systems Engineer: Scale AI Clusters & HPC. Be the first to apply!
- senior staff systems engineer New York, NY
- application system engineer New York, NY
- system engineer contract New York, NY
- operations support system engineer New York, NY
- sr systems engineer New York, NY
- visual systems engineer New York, NY
- active directory systems engineer New York, NY
- system performance engineer New York, NY
- software system engineer New York, NY
- data systems engineer New York, NY



