Lead GPU Systems Engineer - HPC & AI Infrastructure
Socket.dev
Tower Research Capital seeks an accomplished engineer to design, deploy, and scale distributed GPU clusters, building out hardware selection, production operations, and monitoring across thousands of nodes. You will diagnose bottlenecks across compute, storage, and network layers, collaborating with researchers to benchmark workloads and translate results into speedups. Strong Linux, CUDA/C++, and Python skills are essential. #J-18808-Ljbffr Socket.dev
- ...Research Capital is a leading quantitative trading firm... ...-performance compute infrastructure for systematic trading.... ...workloads. You will join engineers responsible for... ...storage and thousands of GPU-accelerated nodes, shaping architectures for AI clusters and performance...Suggested
- ...Senior GPU Systems / AI Infrastructure Engineer (NYC) Location: New York City (Hybrid / On-site preferred) Comp: Competitive + equity (Series A-C... ...match if you have: ~5-10+ years in systems engineering, HPC, GPU computing, or AI infrastructure ~ Deep experience...SuggestedFull time
$150k - $300k
Hudson River Trading (HRT) is looking for GPU Systems Engineers to help scale and evolve our exceptionally sophisticated HPC/AI research environment. Joining our Research... ...across the globe. We design, grow, and operate infrastructure at a large scale, including triple-digit...SuggestedWork at officeLocal areaImmediate start- ...Business Development Representative (BDR) for Lead Generation to identify and engage new enterprise customers in the AI and infrastructure sector. You will help generate qualified... ...or AI solutions, a strong understanding of GPU technology, and be self-motivated with excellent...SuggestedRemote job
- Obsidian is seeking an MLOps Engineer to join their AI lab's cutting-edge GenAI team in New York. This full-time role involves... ...candidate should have 2+ years of experience in ML infrastructure, with skills in custom GPU kernel optimization. This position offers a chance...SuggestedFull time
$200k - $300k
...Research Capital is a leading quantitative... ...trading and engineering talent. We... ...electronic trading infrastructure at a world class... ..., operating systems, and automation... ...large CPU and GPU clusters spanning... ...architecture of a new AI cluster, the... ...systems in HPC, AI, or distributed...Work experience placementCasual workWork at office$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead... ...high-quality, resilient systems at scale. The ideal candidate... ...—operating as an infrastructure-focused engineer, not a tool... ...output variance, and token/GPU cost anomalies. Translate...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week- Get AI‑powered advice on this job and more exclusive features. Job Title: Backup Administration Engineer Location: Greensboro, NC 27401 (Remote work allowed) Must Have Skills Design & stand‑up the Veeam infrastructure Develop a structured migration plan from Rubrik...Contract workRemote work
- ...are we?Cohere is the leading security-first enterprise AI company. We build... ...who are building AI systems. We believe that... ...team of researchers, engineers, designers, and more... ...and influence the Infrastructure team’s roadmap based... ...Kubernetes, and GPU workloads on those...Full timeWork experience placementWork at officeLocal areaRemote workHome office
- System Administrator / Infrastructure Engineer Location: Remote, UK / Europe About the role We are looking for a System... ...Technical Director, Technical Leads, and COO to develop and implement a... ...and ways of working, such as using AI to help solve problems. You have great...Local areaRemote workWork from homeFlexible hours
$144.9k - $283.5k
...industry. Here, you lead with innovative thinking... ...Development Leader, AI Native Ecosystems &... ...place within $100M+ GPU contracts across Neoclouds... ...disruptive AI engineering and global cloud infrastructure, establishing Everpure... ...Everpure’s internal AI/HPC product teams, ensuring...Flexible hoursShift work$153k - $204k
...Essential Cloud for AI. Built for... ...confidence. Trusted by leading AI labs, startups,... ...combines superior infrastructure performance with deep... ...You'll Do:The Systems Engineering team owns the host... ...one of the largest GPU fleets in the world... ...that framework into HPC verification,...Permanent employmentFull timeTemporary workCasual workLive inWork at officeFlexible hours$150k - $250k
...in your career? Join a leading quantitative trading firm... ...opportunity is ideal for an engineer looking to build and scale... ...automating, and optimizing infrastructure that powers compute-... ...high-performance computing (HPC) or distributed systems. You should have knowledge...Full time- ...Radix Trading is seeking a seasoned HPC Systems Engineer with a passion for Linux, HPC systems and research infrastructure to join us. This role requires the ability to quickly... ...is a self-starter that has experience leading projects end-to-end, fluency in scripting languages...
- A fast-growing infrastructure company in New York is seeking a Senior Systems Engineer. The role involves assisting clients with evaluations and installations, collaborating with R&D for product requirements, and demonstrating technical expertise in storage products. Candidates...
- Who are we?Cohere is the leading security-first enterprise AI company. We build... ...enterprises who are building AI systems. We believe that our... ...a team of researchers, engineers, designers, and more, who... ...distributed systems, and HPC infrastructure. You will design and maintain...Full timeWork at officeLocal areaRemote workHome office
- ...—until generative AI. Today, AI is the... ...to the New using leading-edge technologies... ...reliability of deployed systems.The Work:•... ...availability, system and infrastructure requirements -... ...including edge device and HPC• Engage in... ...business owners, engineers, architects, and UI...Full timeWork experience placementLive inWork at officeLocal areaShift work
- ...The engineering team at Chainalysis is inspired by solving... ...to build a flexible, AI-driven platform that... ...mainstream. As a Senior Infrastructure Engineer in the... ...leader on the Government Systems team - the group responsible... ...role, you'll: Lead the design and...Flexible hours
- ...human life on Mars. SR. HIGH PERFORMANCE COMPUTING (HPC) SYSTEMS ENGINEER SpaceX is looking for an HPC Systems Engineer with strong... ...computing (CFD, FEA) Familiarity with large scale AI training Familiarity with GPU usage in a compute cluster and Cuda Experience with...Permanent employmentFlexible hoursWeekend work
- ...skilled professional in New York to design and operate large-scale GPU infrastructure for model inference and reinforcement learning. The role demands several years of experience in deploying GPU systems, optimizing model performance, and working with frameworks like...
$110 per hour
...technical talent with leading AI research labs. Headquartered... ...Guide research and engineering teams to close... ...in MLOps , training infrastructure, and ML framework-level... ...solutions to MLOps and ML systems problems .... ...or optimizing custom GPU kernels using Triton...Contract workSummer workRemote workWeekday work- Basis is a nonprofit applied AI research organization with two mutually... ...value. About the Role ML Systems Engineers at Basis ensure training and evaluation infrastructure is fast, reliable, and scalable... .... You will manage GPU clusters, optimize cloud spending...Full timeWork at office
$182k - $242k
...Essential Cloud for AI. Built for... ...confidence. Trusted by leading AI labs, startups,... ...combines superior infrastructure performance with deep... ...You'll Do:The Systems Engineering team owns the Linux... ...one of the largest GPU fleets in the world... ..., platform, HPC, and Fleet teams to...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$200k - $300k
...critical software, analytics, and AI to life and is the company behind Astro, the industry-leading unified DataOps platform powered... ...Role Senior and Staff Software Engineers on our Infrastructure team own the architecture of the systems that stand up, configure, and...Full time$165k - $242k
...Essential Cloud for AI. Built for... ...confidence. Trusted by leading AI labs, startups... ...superior infrastructure performance with... ...skilled and motivated Systems Kernel Engineer to join our HAVOCK... ...virtualization, GPU/DPU enablement).Stack... ..., nydus, kubelet)HPC/AI workloads (...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours- ...interconnection projects and securing favorable energy terms for pioneering AI data centers in the US. You will own utility relationships,... ...role demands deep market awareness and coordination across development, legal, engineering, and finance teams. #J-18808-Ljbffr Fluidstack
$150k - $300k
...HRT) is looking for Systems Engineers to join our... ...incredibly large GPU and CPU compute clusters... ...focused on industry leading compute, network,... ...and keeping our HPC environment running... ...range of technical infrastructure Troubleshoot... ...be advised: Use of AI tools during interviews...Full timeWork at officeLocal areaImmediate startRemote workWorldwide- Fluidstack is seeking a Senior Legal Counsel in Seattle to review, draft, and negotiate a broad range of commercial agreements tied to a large-scale data center buildout. You will build scalable contracting and compliance frameworks, partnering with legal, finance, and ...
$131.1k - $196.7k
## Lead Compensation Program Manager - Market Strategy... ...projects, and build AI-leveraged tools necessary to scale our infrastructure. You will be a key partner... ...People Tech for seamless system integration.* Process Innovation... ...world’s leading game engine, powering play for more...Remote jobWork at officeImmediate startWorldwideRelocation packageShift work- ...looking for a Distributed Systems Engineer in a remote role... ...ThisWay Global, Inc. is an AI-first technology... ...intelligence, data center infrastructure, and workforce... ...and NVIDIA NVL72/GB300 GPU clusters. Amalgamy.ai... ...fault tolerance, and HPC-grade reliability. Location...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Lead GPU Systems Engineer - HPC & AI Infrastructure. Be the first to apply!
- lead algorithm engineer New York, NY
- lead system engineer New York, NY
- lead engineer New York, NY
- lead network engineer New York, NY
- lead industrial engineer New York, NY
- lead operating engineer New York, NY
- lead backend developer New York, NY
- lead solutions engineer New York, NY
- lead app. developer New York, NY
- lead security engineer New York, NY


