Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior GPU Systems Engineer: Scale AI Clusters & HPC

Career Techniques

Career Techniques in New York seeks an experienced infrastructure engineer to design, deploy, and scale large-scale GPU clusters for AI research. You will work across compute, storage, OS, and automation to support hundreds of petabytes and thousands of nodes. You will profile GPU workloads, remove bottlenecks, and collaborate with researchers to translate findings into speedups. Expect to own end-to-end infrastructure projects from design through long-term support and vendor engagement. #J-18808-Ljbffr Career Techniques

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior GPU Systems Engineer: Scale AI Clusters & HPC in New York, NY vacancy
  •  ...compute, storage, and automation behind large-scale trading workloads. You will join engineers responsible for hundreds of petabytes of storage and thousands of GPU-accelerated nodes, shaping architectures for AI clusters and performance profiling. Collaboration with researchers... 
    Suggested

    Tower Research Capital

    New York, NY
    4 days ago
  •  ...a skilled professional in New York to design and operate large-scale GPU infrastructure for model inference and reinforcement learning. The role demands several years of experience in deploying GPU systems, optimizing model performance, and working with frameworks like... 
    Senior

    Reflection

    New York, NY
    1 day ago
  • Tower Research Capital seeks an accomplished engineer to design, deploy, and scale distributed GPU clusters, building out hardware selection, production operations, and monitoring across thousands of nodes. You will diagnose bottlenecks across compute, storage, and network... 
    Suggested

    Socket.dev

    New York, NY
    4 days ago
  • $153k - $204k

     ...Essential Cloud for AI. Built for...  ...innovators to build and scale AI with...  ...You'll Do:The Systems Engineering team owns the host...  ...of the largest GPU fleets in the...  ...framework into HPC verification,...  ...the role:As a Senior Software Engineer...  ...experience.HPC or large-cluster experience —... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Casual work
    Live in
    Work at office
    Flexible hours

    CoreWeave

    New York, NY
    4 days ago
  •  ...first enterprise AI company. We...  ...are building AI systems. We believe that...  ...of researchers, engineers, designers, and...  ...looking for a senior engineer to help...  ...our frontier-scale language models...  ...distributed systems, and HPC infrastructure....  ...on multi-node clusters (e.g., GB200/30... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    4 days ago
  •  ...Senior GPU Systems / AI Infrastructure Engineer (NYC) Location: New York City (Hybrid / On-site...  ...infrastructure powering large-scale model training and...  ...training (multi-node, multi-GPU clusters) Improve memory...  ...in systems engineering, HPC, GPU computing, or AI infrastructure... 
    Full time
    New York, NY
    more than 2 months ago
  •  ...SR. HIGH PERFORMANCE COMPUTING (HPC) SYSTEMS ENGINEER SpaceX is looking for an HPC Systems...  ...: Administer and manage HPC clusters, storage systems, and high-speed...  ...CFD, FEA) Familiarity with large scale AI training Familiarity with GPU usage in a compute cluster and Cuda... 
    Senior
    Permanent employment
    Flexible hours
    Weekend work

    SpaceX

    New York, NY
    4 days ago
  • $108k - $172.5k

     ...unlimited potential of AI to define the...  ...era in which our GPU acts as the...  ...highly motivated Senior HPC Support Engineer focussing on InfiniBand...  ...and supporting systems using Linux Operating...  ...on large-scale networking and AI...  ...and GPU Technology.Clustering or HPC Data-Center... 
    Senior
    Full time
    Work experience placement

    Nvidia

    New York, NY
    1 day ago
  • Scale AI is hiring for a senior software engineer to design, build, and scale full‑stack systems powering our GenAI data engine. You will work across front-end, back-end, and infrastructure, using React, TypeScript, Node.js, Python and data stores like MongoDB, Elasticsearch... 
    Senior

    Scale

    New York, NY
    22 hours ago
  • $150k - $300k

    Hudson River Trading (HRT) is looking for GPU Systems Engineers to help scale and evolve our exceptionally sophisticated HPC/AI research environment. Joining our Research and...  ...-scale storage and massive CPU and GPU clusters in globally distributed data centers. As such... 
    Work at office
    Local area
    Immediate start

    Hudson River Trading

    New York, NY
    4 days ago
  • $200k - $300k

     ...systematic trading and engineering talent. We empower...  ...the economies of scale that come from a...  ..., operating systems, and automation behind...  ...and large CPU and GPU clusters spanning thousands...  ...architecture of a new AI cluster, the next...  ...Linux systems in HPC, AI, or... 
    Work experience placement
    Casual work
    Work at office

    Socket

    New York, NY
    3 days ago
  • $125k - $250k

     ...Senior Account Executive- GPU/AI Infrastructure Senior Account Executive - GPU and AI Infrastructure...  ...era. Operating some of the largest GPU clusters globally, we deliver high-performance...  ...Enterprise GPU clusters for large-scale AI initiatives On-demand GPU compute... 
    Senior
    Temporary work
    Remote work
    Flexible hours

    ESR Healthcare

    New York, NY
    2 days ago
  • $182k - $242k

     ...Essential Cloud for AI. Built for...  ...innovators to build and scale AI with confidence...  ...What You'll Do:The Systems Engineering team owns the...  ...one of the largest GPU fleets in the world...  ...About the role:As a Senior Software Engineer...  ...hardware, platform, HPC, and Fleet teams... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    New York, NY
    4 days ago
  • Principal Senior Systems Engineer - HPE Networking (Mid-Telco Accounts)This role has been designated as...  ...IP, transport, data center, and AI-ready networking environments.The ideal...  ...expertise designing and supporting large-scale service provider networks including:BGPIS... 
    Senior
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    Remote work
    Work from home

    Hewlett Packard Enterprise

    New York, NY
    4 days ago
  •  ...entity. Responsibilities As a senior Machine Learning Systems Engineer on the Search Platform team, you...  ...retrieval workflows. Partner with Rovo and AI platform teams to evolve search...  ...quality, freshness, and relevance at scale.Operational Excellence & Cost DisciplineDrive... 
    Senior
    Work at office
    Local area

    Atlassian

    New York, NY
    22 hours ago
  •  ...innovation. Join us!About the RoleWe’re looking for a Senior Design Systems Engineer II to shape and scale the visual and interaction foundations of Blink’s...  ...JavaScript, and modern front-end tooling.Expert use of modern AI-assisted tools (e.g., Cursor, GitHub Copilot, ChatGPT... 
    Senior

    Blink Health

    New York, NY
    1 day ago
  • $112.9k - $155.24k

     .... SCOPE OF POSITION   The Senior Systems Engineer is responsible for working directly...  ...team to shape the future of AI infrastructure in hyperscale data...  ...AI/ML fabric architectures (scale-up / scale-out), including GPU/XPU cluster designs and optical I/O requirements... 
    Senior
    Full time
    Work experience placement

    Corning

    New York, NY
    26 days ago
  •  ...scalable worker platform in NYC. We’re seeking engineers with 5+ years in software development and deep experience with distributed systems and performance tuning. You’ll work...  ...from first principles, enjoy operating large-scale bare-metal fleets, and are excited #J-188... 
    Senior

    Blacksmith

    New York, NY
    3 days ago
  • You are one of the first systems engineers at a founder‑led AI firm building the execution layer for enterprise...  ...on tomorrow. You are coming in at a senior level, which means you are making...  ...written the runbook. When something scales, it is because you planned for it. Day... 
    Senior

    Saragossa

    New York, NY
    3 days ago
  • $85k - $140k

    Nebius is seeking a Data Center Support Engineer to manage large-scale bare metal GPU infrastructure supporting AI workloads. The role involves troubleshooting physical servers, Linux systems, networking components, and maintaining operational excellence across the data... 
    Senior

    Nebius

    New York, NY
    1 day ago
  • $160k - $220k

     ...Every Identity, from AI to HumanIdentity is...  ...are too, let's talk.Senior Database Reliability Engineer (DBRE) Experience Level...  ...in PostgreSQL at scale and solid experience...  ...scale, mission-critical systems. You will work...  ...available PostgreSQL clusters (physical replication... 
    Senior
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    New York, NY
    1 day ago
  • $152k - $241.5k

    We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency...  ...and operating large scale compute...  ...experience with NSight Systems and NSight ComputeExperience...  ...Lustre and GPFS for AI/HPC workloadsFamiliarity with... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    4 days ago
  • $184k - $287.5k

     ...computing platforms for AI and HPC. Because of our...  ..., and engineers can push the boundaries...  ...motivated Senior Solutions Architect...  ...with a focus on GPU, NVLink, and...  ...generation GPU-based clusters enabling the...  ...designing large-scale distributed systems, AI clusters, or... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    3 days ago
  • $150k - $250k

     ...systematic trading and engineering talent. We empower portfolio...  ...the economies of scale that come from a large,...  ...and optimize Kubernetes clusters in cloud environments Automate...  ...to-cloud migrations of HPC and distributed...  ...Infrastructure as Code, Linux systems, cloud networking, and... 
    Senior
    Casual work
    Work at office

    Tower Research Capital

    New York, NY
    4 days ago
  • $184k - $287.5k

     ...seeking outstanding AI Solutions...  ...work across product, engineering, sales, developer...  ...advisor for accelerated systems architecture, GPU and networking systems, cluster design,...  ..., including large-scale clustersSupport infrastructure...  ...customers on AI/HPC infrastructureYour... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    2 days ago
  • $150k - $250k

     ...This opportunity is ideal for an engineer looking to build and scale cloud platforms supporting high-...  ...manage and optimize Kubernetes clusters and Linux-based platforms. You...  ...high-performance computing (HPC) or distributed systems. You should have knowledge of... 
    Full time
    New York, NY
    18 days ago
  • $112.9k - $155.24k

    Senior Systems Engineer Co-Packaged Optics Locations: Santa Clara and Bay Area (West Coast...  ...team to shape the future of AI infrastructure in hyperscale data...  ...with AI/ML fabric architectures (scale-up / scale-out), including GPU/XPU cluster designs and optical I/O requirements... 
    Senior
    Full time
    Work experience placement

    Corning Incorporated

    New York, NY
    3 days ago
  • Viridien is seeking a highly experienced HPC Data Center Senior Linux IT Specialist to join the IT team, contributing to global HPC and Digital...  .... The role focuses on Linux administration, storage systems, and IT service management within a high-availability environment... 
    Senior

    SwiftCruit

    New York, NY
    4 days ago
  • $224k - $356.5k

     ...Performance Computing and AI workloads across...  ...with NVIDIA Engineering, Product, and...  ...next-generation GPU architectures....  ...DL, recommender systems, GNN, monte-...  ...ML/DL models at scale on on-prem or public...  ...cloud computing clusters in...  ...developers building AI, HPC, or data analytics... 
    Senior
    Full time

    Nvidia

    New York, NY
    4 days ago
  • $224k - $356.5k

     ...Performance Computing and AI workloads across...  ...with NVIDIA Engineering, Product, and Sales...  ...next-generation GPU architectures.Work...  ...ML/DL, recommender systems, GNN, monte-carlo...  ...deploying ML/DL models at scale on public cloud...  ...and/or on-prem HPC clusters in productionWays... 
    Senior
    Full time
    Remote work

    Nvidia

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior GPU Systems Engineer: Scale AI Clusters & HPC. Be the first to apply!