Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior GPU Systems Engineer: Scale AI Clusters & HPC

Career Techniques

Career Techniques in New York seeks an experienced infrastructure engineer to design, deploy, and scale large-scale GPU clusters for AI research. You will work across compute, storage, OS, and automation to support hundreds of petabytes and thousands of nodes. You will profile GPU workloads, remove bottlenecks, and collaborate with researchers to translate findings into speedups. Expect to own end-to-end infrastructure projects from design through long-term support and vendor engagement. #J-18808-Ljbffr Career Techniques

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Senior GPU Systems Engineer: Scale AI Clusters & HPC in New York, NY vacancy
  •  ...compute, storage, and automation behind large-scale trading workloads. You will join engineers responsible for hundreds of petabytes of storage and thousands of GPU-accelerated nodes, shaping architectures for AI clusters and performance profiling. Collaboration with researchers... 
    Suggested

    Tower Research Capital

    New York, NY
    10 hours ago
  •  ...a skilled professional in New York to design and operate large-scale GPU infrastructure for model inference and reinforcement learning. The role demands several years of experience in deploying GPU systems, optimizing model performance, and working with frameworks like... 
    Senior

    Reflection

    New York, NY
    2 days ago
  • A growing infrastructure company is seeking a Senior Systems Engineer to support the Department of Energy. This role involves guiding national...  ...Strong technical presentation skills and understanding of HPC and AI/ML are crucial. Join us for this pivotal opportunity at... 
    Senior

    VAST Data

    New York, NY
    3 days ago
  • $295k

     ...first enterprise AI company. We...  ...are building AI systems. We believe that...  ...of researchers, engineers, designers, and...  ...looking for a senior engineer to help...  ...our frontier-scale language models...  ...distributed systems, and HPC infrastructure....  ...on multi-node clusters (e.g., GB200/30... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    10 hours ago
  • $153k - $204k

     ...Essential Cloud for AI™. Built for...  ...to build and scale AI with confidence...  ...You'll Do The Systems Engineering team owns the...  ...of the largest GPU fleets in the world...  ...framework into HPC verification,...  ...The Role As a Senior Software Engineer...  .... HPC or large-cluster experience —... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Casual work
    Live in
    Work at office
    Flexible hours

    CoreWeave

    New York, NY
    3 days ago
  • $108k - $172.5k

     ...unlimited potential of AI to define the...  ...era in which our GPU acts as the...  ...highly motivated Senior HPC Support Engineer focussing on InfiniBand...  ...and supporting systems using Linux Operating...  ...on large-scale networking and AI...  ...and GPU Technology.Clustering or HPC Data-Center... 
    Senior
    Full time
    Work experience placement

    Nvidia

    New York, NY
    3 days ago
  •  ...Senior GPU Systems / AI Infrastructure Engineer (NYC) Location: New York City (Hybrid / On-site...  ...infrastructure powering large-scale model training and...  ...training (multi-node, multi-GPU clusters) Improve memory...  ...in systems engineering, HPC, GPU computing, or AI infrastructure... 
    Full time
    New York, NY
    more than 2 months ago
  •  ...the future of AI and HPC networking with...  ...We’re seeking engineers who are energized...  ...distributed software systems, and who are...  ...performance at scale. Cornelis...  ...efficiency of GPU, CPU and...  ...-based compute clusters at any scale. Our...  ...an experienced Senior Software Engineer... 
    Senior
    Remote job
    Full time
    Flexible hours

    Cornelis Networks

    New York, NY
    10 hours ago
  • $150k - $300k

    Hudson River Trading (HRT) is looking for GPU Systems Engineers to help scale and evolve our exceptionally sophisticated HPC/AI research environment. Joining our Research and...  ...-scale storage and massive CPU and GPU clusters in globally distributed data centers. As such... 
    Work at office
    Local area
    Immediate start

    Hudson River Trading

    New York, NY
    10 hours ago
  • $168k - $270.25k

     ...Experience (NVEX) Solutions Engineering team is looking for...  ...of NVIDIA’s GPU accelerated...  ...will apply the latest AI technologies to triage...  ...datacenters of rack-scale platforms, solve...  ...GPU platformsStrong system software (firmware,...  ...workloadsClustering or HPC data center... 
    Senior
    Full time
    Weekend work

    Nvidia

    New York, NY
    10 hours ago
  • $125k - $250k

     ...Senior Account Executive- GPU/AI Infrastructure Senior Account Executive - GPU and AI Infrastructure...  ...era. Operating some of the largest GPU clusters globally, we deliver high-performance...  ...Enterprise GPU clusters for large-scale AI initiatives On-demand GPU compute... 
    Senior
    Temporary work
    Remote work
    Flexible hours

    ESR Healthcare

    New York, NY
    3 days ago
  • $152k - $241.5k

     ...the unlimited potential of AI to define the next era of...  ...about building reliable systems software for cloud-scale GPU infrastructure, we...  ...team. We are looking for a Senior Software Engineer to join our DGX Cloud / Fleet...  ...operating software in AI, HPC, cloud, or large-scale... 
    Senior
    Full time
    Local area
    Remote work

    Nvidia

    New York, NY
    10 hours ago
  • $200k - $300k

     ...leading trading firms on an interesting GPU Systems Engineer position. This is a hands-on infrastructure role working at scale, with large CPU and GPU clusters spanning thousands of nodes and...  ...with large-scale Linux systems in HPC, AI or distributed infrastructure environments... 

    Iceberg

    New York, NY
    4 days ago
  •  ...you will join the engineers responsible for the...  ..., operating systems, and automation behind...  ...that work at serious scale: hundreds of...  ...and large CPU and GPU clusters spanning thousands...  ...architecture of a new AI cluster, the next...  ...Linux systems in HPC, AI, or distributed... 
    Work experience placement

    Career Techniques Inc

    New York, NY
    2 days ago
  •  ...entity. Responsibilities As a senior Machine Learning Systems Engineer on the Search Platform team, you...  ...retrieval workflows. Partner with Rovo and AI platform teams to evolve search...  ...quality, freshness, and relevance at scale.Operational Excellence & Cost DisciplineDrive... 
    Senior
    Work at office
    Local area

    Atlassian

    New York, NY
    1 day ago
  •  ...a Principal Infrastructure Engineer to own AI cluster performance and validation across...  .... You’ll define healthy-at-scale criteria, set the architecture, and partner with senior leaders to production-...  ...design scalable validation systems, and optimize fabric, storage... 
    Senior

    Nscale

    New York, NY
    2 days ago
  •  ...innovation. Join us!About the RoleWe’re looking for a Senior Design Systems Engineer II to shape and scale the visual and interaction foundations of Blink’s...  ...JavaScript, and modern front-end tooling.Expert use of modern AI-assisted tools (e.g., Cursor, GitHub Copilot, ChatGPT... 
    Senior

    Blink Health

    New York, NY
    2 days ago
  •  ...scalable worker platform in NYC. We’re seeking engineers with 5+ years in software development and deep experience with distributed systems and performance tuning. You’ll work...  ...from first principles, enjoy operating large-scale bare-metal fleets, and are excited #J-188... 
    Senior

    Blacksmith

    New York, NY
    4 days ago
  • $85k - $140k

    Nebius is seeking a Data Center Support Engineer to manage large-scale bare metal GPU infrastructure supporting AI workloads. The role involves troubleshooting physical servers, Linux systems, networking components, and maintaining operational excellence across the data... 
    Senior

    Nebius

    New York, NY
    2 days ago
  •  ...The engineering team at Chainalysis is inspired...  ...build a flexible, AI-driven platform that...  ...our customers scale their workflows as...  ...mainstream. As a Senior Infrastructure Engineer...  ...on the Government Systems team - the group...  ...RKE2 Kubernetes clusters across development... 
    Senior
    Flexible hours

    Chainalysis Inc.

    New York, NY
    2 days ago
  • $160k - $220k

     ...Every Identity, from AI to HumanIdentity is...  ...are too, let's talk.Senior Database Reliability Engineer (DBRE) Experience Level...  ...in PostgreSQL at scale and solid experience...  ...scale, mission-critical systems. You will work...  ...available PostgreSQL clusters (physical replication... 
    Senior
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    New York, NY
    2 days ago
  • $152k - $241.5k

    We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency...  ...and operating large scale compute...  ...experience with NSight Systems and NSight ComputeExperience...  ...Lustre and GPFS for AI/HPC workloadsFamiliarity with... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    10 hours ago
  • $184k - $287.5k

     ...computing platforms for AI and HPC. Because of our...  ..., and engineers can push the boundaries...  ...motivated Senior Solutions Architect...  ...with a focus on GPU, NVLink, and...  ...generation GPU-based clusters enabling the...  ...designing large-scale distributed systems, AI clusters, or... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    4 days ago
  • $184k - $287.5k

     ...seeking outstanding AI Solutions...  ...work across product, engineering, sales, developer...  ...advisor for accelerated systems architecture, GPU and networking systems, cluster design,...  ..., including large-scale clustersSupport infrastructure...  ...customers on AI/HPC infrastructureYour... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    3 days ago
  • Blink Health is seeking a Senior Design Systems Engineer II to shape and scale the visual and interaction foundations of Blink’s digital ecosystem. You’ll partner with Product Design, Software Engineering, and cross‑functional teams to evolve the shared frontend design... 
    Senior

    BlinkRx

    New York, NY
    2 days ago
  • $224k - $356.5k

     ...Performance Computing and AI workloads across...  ...with NVIDIA Engineering, Product, and...  ...next-generation GPU architectures....  ...DL, recommender systems, GNN, monte-...  ...ML/DL models at scale on on-prem or public...  ...cloud computing clusters in...  ...developers building AI, HPC, or data analytics... 
    Senior
    Full time

    Nvidia

    New York, NY
    10 hours ago
  • $224k - $356.5k

     ...Performance Computing and AI workloads across...  ...with NVIDIA Engineering, Product, and Sales...  ...next-generation GPU architectures.Work...  ...ML/DL, recommender systems, GNN, monte-carlo...  ...deploying ML/DL models at scale on public cloud...  ...and/or on-prem HPC clusters in productionWays... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    10 hours ago
  • $152k - $241.5k

     ...changing the world of AI Networking with...  ...for Data Center scale systems for AI training/inference...  .../ Software engineering or equivalent experience...  ....Familiarity with cluster orchestration...  ...Knowledge in MPI and HPC, InfiniBand,...  ...invention of the GPU - the engine of modern... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    4 days ago
  • $139k - $257.55k

     ...seeking an outstanding Senior Site Reliability Engineer (SRE) to support...  ...learning, autonomous AI workflows, and cloud-...  ...native, containerized systems as we build the framework...  ..., and operate large-scale, distributed, fault-...  ...infrastructure — model serving, GPU workloads, language... 
    Senior
    Full time
    Temporary work
    Local area
    Remote work
    Worldwide

    Adobe Systems

    New York, NY
    2 days ago
  • A dynamic technology firm in New York is seeking a talented Senior/Staff level Systems Engineer to develop and scale a dedicated cloud for CI workloads. The role offers an opportunity to solve complex systems problems and build a new CI cloud from the ground up. Candidates... 
    Senior

    Crossing Hurdles

    New York, NY
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior GPU Systems Engineer: Scale AI Clusters & HPC. Be the first to apply!