HPC Developer
Autonomai Recruitment
Role: HPC Data Center Developer
Location: New York City
Type: Full-time
About the role
We are looking for an HPC Data Center Developer to build the automation and tooling that supports large-scale high-performance computing infrastructure.
This is a development-focused role for an engineer who understands both software and data-centre hardware. You will automate the onboarding, provisioning, monitoring and lifecycle management of servers, network equipment, power systems, cooling infrastructure and other critical data-centre assets.
The ideal candidate will have strong, hands-on experience with Golang and a proven track record of automating physical hardware and infrastructure at scale.
Responsibilities
- Design, develop and maintain production-grade automation in Golang.
- Automate the full lifecycle of HPC data-centre hardware, from initial discovery and provisioning through to deployment, monitoring, maintenance and decommissioning.
- Build workflows for servers, GPUs, network switches, rack PDUs, CDUs, environmental sensors and related infrastructure.
- Develop end-to-end processes that take hardware from racked and cabled through configuration, validation and production readiness with minimal manual intervention.
- Integrate with hardware management interfaces, APIs and protocols such as Redfish, IPMI, SNMP and vendor-specific platforms.
- Build tooling for hardware health monitoring, diagnostics, alerting and automated recovery.
- Integrate telemetry from compute, networking, power, cooling and environmental systems into centralised monitoring platforms.
- Support capacity planning across power, cooling, rack space, networking and compute resources.
- Develop operational tools for inventory management, spares, change management, troubleshooting and hardware lifecycle tracking.
- Work closely with HPC engineering, data-centre operations, network, systems and infrastructure teams.
- Translate manual operational processes and pain points into reliable, maintainable automation.
- Own the reliability, documentation and ongoing improvement of the systems and tools you develop.
- Support large-scale infrastructure deployments, maintenance activities and production incidents when required.
Required experience
- Strong professional experience developing software in Golang.
- Demonstrable experience automating physical hardware or data-centre infrastructure.
- Experience building production automation for server, GPU, network, storage, power or cooling systems.
- Strong Linux and systems engineering knowledge.
- Understanding of server components, rack-scale infrastructure and data-centre operations.
- Experience developing reliable, scalable and observable production systems.
- Ability to work effectively with infrastructure, hardware and operations teams.
- Strong troubleshooting, debugging and problem-solving skills.
Desirable experience
- Experience supporting HPC, AI/ML, GPU or high-density compute environments.
- Experience with data-centre hardware such as GPUs, servers, switches, rack PDUs, CDUs and environmental monitoring systems.
- Knowledge of Kubernetes, Slurm or other cluster-management and workload-orchestration technologies.
- Experience with Python, Bash or another systems programming language.
- Familiarity with configuration management, infrastructure as code and CI/CD.
- Experience with Prometheus, Grafana or other observability platforms.
- Exposure to power, cooling and data-centre capacity planning.
- Experience in a high-availability, trading, cloud, hyperscale or other performance-critical environment.
What we are looking for
This role would suit a software engineer, infrastructure developer, systems engineer or data-centre automation engineer who enjoys working close to the hardware.
- ...Citis nextgeneration risk platformCiti is seeking a visionary Cloud HPC Engineer to lead the development and operation of our global... ...massivescale computation You will take the sophisticated pricing models developed by our top quants and operationalize them on a colossal grid...SuggestedFull timeLocal area
- ...governed, and reproducible AI workflows.2+ years of experience developing APIs, integration services, automation workflows, or platform services... ...managing the deployment of 1,000+ GPU clusters for AI, HPC, and agentic AI workloads with infrastructure services enabled....SuggestedFull timeWork experience placementLive inWork at officeLocal area
$225k - $275k
...join our team, you’ll be contributing to building the technology that powers the future. About the Role We’re hiring a Staff HPC Systems Software Engineer to define the technical direction and evolution of a core HPC platform domain at Nscale. In this role,...SuggestedTemporary workFlexible hours- engineeringjobs.net, Inc. is seeking engineers to design and improve workload scheduling, fleet management, and clustered file systems for large-scale compute environments. You will investigate kernel and network performance, build metrics tooling, and collaborate with ...Suggested
- Tower Research Capital is seeking a Senior Platform Engineer to build a multi-tenant, scalable research compute platform across on-premises and cloud environments in NYC. You will collaborate with Quant Researchers, Portfolio Managers, and Infrastructure to deliver elastic...Suggested
- NVIDIA Corporation in Memphis, TN is seeking a Senior HPC Support Engineer - Ethernet / AI Infrastructure to own customer interactions onsite and remotely, addressing complex installations and operations across NVIDIA Ethernet Switching technologies and Spectrum-X. You...Remote job
- To support NASA's human spaceflight programs, the full-time HPC Linux System Administrator will manage and improve a high-performance computing cluster, administering job schedulers and parallel filesystems while collaborating with scientists and engineers, with remote...Permanent employmentFull timeRemote work
$200k - $300k
Hudson River Trading’s High Performance Computing (HPC) Network Engineering team designs and engineers the low-latency communications infrastructure that underpins our incredibly large GPU and CPU compute clusters. Our mandate is to architect, optimize, and scale the high...Work at officeLocal areaImmediate startWorldwide$175k - $250k
...work across Windows, Linux, cloud, storage, virtualization and HPC environments, building and automating the platform that underpins... ...infrastructure, scaling storage, supporting AI workloads or improving the developer experience, you'll be building the platform rather than simply...Full time$168k - $270.25k
...root causing sophisticated customer issuesWork with R&D teams to develop bug fixes, workarounds, and solutions for critical customers... ...stand out from the crowd:Background with AI infrastructure and HPC networkingExperience programming switch and NIC ASICs and SDKsExperience...Full timeWeekend work- ...HPC Data Center Developer | Global Trading Firm Chicago or New York | We’re partnering with a global trading firm to hire an HPC Data Center Developer to build the software and automation behind its large-scale compute infrastructure. This is a development...Full timeWeekend workAfternoon shift
- ...leverages cutting-edge engineering, high-performance computing (HPC), artificial intelligence (AI), machine learning (ML), and... ...skilled Forward Deployed Engineer to join our team in designing, developing, and deploying custom applications for our diverse clientele. The...Full timeWork experience placementLive inWork at officeLocal area
- ...HPC Network EngineerMirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable... ...engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the...
$145k - $175k
...processes for ticketing and on-call procedures. Knowledge Sharing: Develop onboarding/training materials, knowledge base documentation, and... ..., including the ability to prioritize competing escalations. HPC Knowledge: Understanding of HPC technologies such as Infiniband...Full timeTemporary work$168k - $270.25k
...transform how new chemicals and materials are discovered. We are developing the software tools to make that possible! NVIDIA ALCHEMI is a... ...applicationsExperience in running material science simulations on large scale HPC systemsAbility to work independently and as part of a globally...Full timeRemote work$175k - $250k
...effectively in the most time‑critical markets. As a Senior HFT Developer on SPEED, you will design and build core low‑latency components... ...understanding of the HFT quantitative research pipeline. Experience with HPC grids (scheduling, storage, job management) for research and...$190k - $325k
...experience with cloud object storage such as S3 as well as POSIX-style filesystems.It’s a bonus if you have:Experience with parallel or HPC filesystems such as Weka, VAST, or Lustre.Familiarity with the data-loading and checkpointing patterns used in large-scale model...Full timeWork at officeLocal areaRemote workHome office$168k - $270.25k
...solution engineering workflows including solving customer cases and developing software, both products and internal tools.What you'll be doing... ...software performance of distributed workloadsClustering or HPC data center technologies including Upper Layer Protocols (i.e.,...Full timeWeekend work$227k - $303k
...(Nasdaq: CRWV) in March 2025. Learn more at .About the TeamThe Developer Experience team owns the services, systems, and developer-facing... ...Prior experience at a cloud infrastructure company, hyperscaler, or HPC environment where engineering scale and compute heterogeneity...Permanent employmentFull timeTemporary workCasual workWork at officeRemote workFlexible hours$150k - $170k
...fashion Qualifications Required: ~5+ years of development experience using C# & .NET framework 3.5 / 4.0 as full stack developer. ~ Experience building C# Console and windows applications for data processing. ~ Experience in Microsoft SQL Server and design...Full timeWork experience placement$160k - $200k
...experienced senior software engineer to join a growing team of developers within CCB and the other centers across the Flatiron Institute... ...major cloud platforms. Experience with high-performance computing (HPC) environments. REQUIRED APPLICATION MATERIALS Please submit a r...Full timeLocal area$155k - $200k
...Bring to the Team: ~2+ years of hands-on experience building, deploying, or operating cloud infrastructure, with exposure to AI/ML, HPC, or GPU workloads ~ Hands-on proficiency with at least one major cloud provider (AWS, GCP, or Azure), deploying and managing...Full timeTemporary workWork at officeMonday to Friday$800 - $1,000 per month
Our client is looking for a C# Developer to join their team in NYC. Pay: $800-1,000/dayQualificationsA Bachelor's Degree, Graduate Degree preferred 5 years of hands on programming experience with an understanding of Object Oriented Programming, Design patterns, Service...- We are seeking an experienced Algorithmic Trading Developer to design, build, and support sophisticated electronic trading solutions for equities and futures markets. This individual will work closely with trading, technology, and infrastructure teams to develop high-performance...
$145k - $165k
...Cloud Computing, Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing... ...of a growing, dynamic, highly-focused team responsible for developing new or existing opportunities in the global IT market. The FAE...Worldwide- ...functional infrastructure initiatives. Demonstrated leadership experience, including mentoring junior engineers. Experience with HPC or GPU cluster infrastructure, including Slurm.. Experience building or operating AI agents or agentic infrastructure....Work at officeLocal areaShift work
- ...compute farm expands Required qualifications BS or MS in Computer Science, Computer Engineering, or equivalent experience 8+ years in HPC or large-scale batch compute, with 5+ years specifically on IBM Spectrum LSF Proven expertise in LSF internals and debugging beyond...Full timeRemote work
$122k - $163k
...harness the full potential of our advanced Kubernetes-powered HPC cloud infrastructure. You'll be hands-on, collaborating with engineers... ...a Kind.In this role, you will:Guide and mentor team members in developing their technical skills and troubleshooting capabilities across...Full timeWork experience placementCasual workWork at officeRemote workWorldwideShift work$190k - $270k
...missions.The Databricks AI Research organization enables companies to develop AI models and agents using their own data, with technologies... ...‑scale experiments, data processing, and model training (e.g., HPC clusters, GPU fleets, or cloud‑based systems)Enable researchers...Local areaWorldwide$108k - $172.5k
...lasting impact on the world.We are seeking a highly motivated Senior HPC Support Engineer focussing on InfiniBand and NVLink technology,... ...customer experience, support tools, etc.As a technical resource develop, re-define and document standard methodologies to provide to...Full timeWork experience placement
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to HPC Developer. Be the first to apply!
- senior mulesoft developer New York, NY
- senior mainframe developer New York, NY
- peoplesoft developer New York, NY
- gis developer remote New York, NY
- entry developer New York, NY
- developer internship New York, NY
- senior tableau developer New York, NY
- program developer New York, NY
- cassandra developer New York, NY
- rpa developer New York, NY




