Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

HPC Developer

Full-time

Autonomai Recruitment

Role: HPC Data Center Developer

Location: New York City

Type: Full-time

About the role

We are looking for an HPC Data Center Developer to build the automation and tooling that supports large-scale high-performance computing infrastructure.

This is a development-focused role for an engineer who understands both software and data-centre hardware. You will automate the onboarding, provisioning, monitoring and lifecycle management of servers, network equipment, power systems, cooling infrastructure and other critical data-centre assets.

The ideal candidate will have strong, hands-on experience with Golang and a proven track record of automating physical hardware and infrastructure at scale.

Responsibilities

  • Design, develop and maintain production-grade automation in Golang.
  • Automate the full lifecycle of HPC data-centre hardware, from initial discovery and provisioning through to deployment, monitoring, maintenance and decommissioning.
  • Build workflows for servers, GPUs, network switches, rack PDUs, CDUs, environmental sensors and related infrastructure.
  • Develop end-to-end processes that take hardware from racked and cabled through configuration, validation and production readiness with minimal manual intervention.
  • Integrate with hardware management interfaces, APIs and protocols such as Redfish, IPMI, SNMP and vendor-specific platforms.
  • Build tooling for hardware health monitoring, diagnostics, alerting and automated recovery.
  • Integrate telemetry from compute, networking, power, cooling and environmental systems into centralised monitoring platforms.
  • Support capacity planning across power, cooling, rack space, networking and compute resources.
  • Develop operational tools for inventory management, spares, change management, troubleshooting and hardware lifecycle tracking.
  • Work closely with HPC engineering, data-centre operations, network, systems and infrastructure teams.
  • Translate manual operational processes and pain points into reliable, maintainable automation.
  • Own the reliability, documentation and ongoing improvement of the systems and tools you develop.
  • Support large-scale infrastructure deployments, maintenance activities and production incidents when required.

Required experience

  • Strong professional experience developing software in Golang.
  • Demonstrable experience automating physical hardware or data-centre infrastructure.
  • Experience building production automation for server, GPU, network, storage, power or cooling systems.
  • Strong Linux and systems engineering knowledge.
  • Understanding of server components, rack-scale infrastructure and data-centre operations.
  • Experience developing reliable, scalable and observable production systems.
  • Ability to work effectively with infrastructure, hardware and operations teams.
  • Strong troubleshooting, debugging and problem-solving skills.

Desirable experience

  • Experience supporting HPC, AI/ML, GPU or high-density compute environments.
  • Experience with data-centre hardware such as GPUs, servers, switches, rack PDUs, CDUs and environmental monitoring systems.
  • Knowledge of Kubernetes, Slurm or other cluster-management and workload-orchestration technologies.
  • Experience with Python, Bash or another systems programming language.
  • Familiarity with configuration management, infrastructure as code and CI/CD.
  • Experience with Prometheus, Grafana or other observability platforms.
  • Exposure to power, cooling and data-centre capacity planning.
  • Experience in a high-availability, trading, cloud, hyperscale or other performance-critical environment.

What we are looking for

This role would suit a software engineer, infrastructure developer, systems engineer or data-centre automation engineer who enjoys working close to the hardware.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the HPC Developer in New York, NY vacancy
  •  ...Citis nextgeneration risk platformCiti is seeking a visionary Cloud HPC Engineer to lead the development and operation of our global...  ...massivescale computation You will take the sophisticated pricing models developed by our top quants and operationalize them on a colossal grid... 
    Suggested
    Full time
    Local area

    Capgemini

    New York, NY
    17 days ago
  •  ...governed, and reproducible AI workflows.2+ years of experience developing APIs, integration services, automation workflows, or platform services...  ...managing the deployment of 1,000+ GPU clusters for AI, HPC, and agentic AI workloads with infrastructure services enabled.... 
    Suggested
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    New York, NY
    3 days ago
  • $225k - $275k

     ...join our team, you’ll be contributing to building the technology that powers the future. About the Role We’re hiring a Staff HPC Systems Software Engineer to define the technical direction and evolution of a core HPC platform domain at Nscale. In this role,... 
    Suggested
    Temporary work
    Flexible hours

    Jobleads-US

    New York, NY
    6 days ago
  • engineeringjobs.net, Inc. is seeking engineers to design and improve workload scheduling, fleet management, and clustered file systems for large-scale compute environments. You will investigate kernel and network performance, build metrics tooling, and collaborate with ...
    Suggested

    engineeringjobs.net, Inc.

    New York, NY
    3 days ago
  • Tower Research Capital is seeking a Senior Platform Engineer to build a multi-tenant, scalable research compute platform across on-premises and cloud environments in NYC. You will collaborate with Quant Researchers, Portfolio Managers, and Infrastructure to deliver elastic...
    Suggested

    Jobleads-US

    New York, NY
    4 days ago
  • NVIDIA Corporation in Memphis, TN is seeking a Senior HPC Support Engineer - Ethernet / AI Infrastructure to own customer interactions onsite and remotely, addressing complex installations and operations across NVIDIA Ethernet Switching technologies and Spectrum-X. You... 
    Remote job

    NVIDIA

    New York, NY
    4 days ago
  • To support NASA's human spaceflight programs, the full-time HPC Linux System Administrator will manage and improve a high-performance computing cluster, administering job schedulers and parallel filesystems while collaborating with scientists and engineers, with remote... 
    Permanent employment
    Full time
    Remote work

    Virtual Vocations Inc

    New York, NY
    3 days ago
  • $200k - $300k

    Hudson River Trading’s High Performance Computing (HPC) Network Engineering team designs and engineers the low-latency communications infrastructure that underpins our incredibly large GPU and CPU compute clusters. Our mandate is to architect, optimize, and scale the high... 
    Work at office
    Local area
    Immediate start
    Worldwide

    Hudson River Trading

    New York, NY
    1 day ago
  • $175k - $250k

     ...work across Windows, Linux, cloud, storage, virtualization and HPC environments, building and automating the platform that underpins...  ...infrastructure, scaling storage, supporting AI workloads or improving the developer experience, you'll be building the platform rather than simply... 
    Full time

    Saragossa

    New York, NY
    4 days ago
  • $168k - $270.25k

     ...root causing sophisticated customer issuesWork with R&D teams to develop bug fixes, workarounds, and solutions for critical customers...  ...stand out from the crowd:Background with AI infrastructure and HPC networkingExperience programming switch and NIC ASICs and SDKsExperience... 
    Full time
    Weekend work

    Nvidia

    New York, NY
    1 day ago
  •  ...HPC Data Center Developer | Global Trading Firm Chicago or New York | We’re partnering with a global trading firm to hire an HPC Data Center Developer to build the software and automation behind its large-scale compute infrastructure. This is a development... 
    Full time
    Weekend work
    Afternoon shift

    Autonomai Recruitment

    New York, NY
    4 days ago
  •  ...leverages cutting-edge engineering, high-performance computing (HPC), artificial intelligence (AI), machine learning (ML), and...  ...skilled Forward Deployed Engineer to join our team in designing, developing, and deploying custom applications for our diverse clientele. The... 
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    New York, NY
    3 days ago
  •  ...HPC Network EngineerMirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable...  ...engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the... 

    Mirantis

    New York, NY
    9 hours ago
  • $145k - $175k

     ...processes for ticketing and on-call procedures. Knowledge Sharing: Develop onboarding/training materials, knowledge base documentation, and...  ..., including the ability to prioritize competing escalations. HPC Knowledge: Understanding of HPC technologies such as Infiniband... 
    Full time
    Temporary work

    AI Chopping Block

    New York, NY
    4 days ago
  • $168k - $270.25k

     ...transform how new chemicals and materials are discovered. We are developing the software tools to make that possible! NVIDIA ALCHEMI is a...  ...applicationsExperience in running material science simulations on large scale HPC systemsAbility to work independently and as part of a globally... 
    Full time
    Remote work

    Nvidia

    New York, NY
    3 days ago
  • $175k - $250k

     ...effectively in the most time‑critical markets. As a Senior HFT Developer on SPEED, you will design and build core low‑latency components...  ...understanding of the HFT quantitative research pipeline. Experience with HPC grids (scheduling, storage, job management) for research and... 

    Millennium Management

    New York, NY
    1 day ago
  • $190k - $325k

     ...experience with cloud object storage such as S3 as well as POSIX-style filesystems.It’s a bonus if you have:Experience with parallel or HPC filesystems such as Weka, VAST, or Lustre.Familiarity with the data-loading and checkpointing patterns used in large-scale model... 
    Full time
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    9 hours ago
  • $168k - $270.25k

     ...solution engineering workflows including solving customer cases and developing software, both products and internal tools.What you'll be doing...  ...software performance of distributed workloadsClustering or HPC data center technologies including Upper Layer Protocols (i.e.,... 
    Full time
    Weekend work

    Nvidia

    New York, NY
    1 day ago
  • $227k - $303k

     ...(Nasdaq: CRWV) in March 2025. Learn more at .About the TeamThe Developer Experience team owns the services, systems, and developer-facing...  ...Prior experience at a cloud infrastructure company, hyperscaler, or HPC environment where engineering scale and compute heterogeneity... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    CoreWeave

    New York, NY
    9 hours ago
  • $150k - $170k

     ...fashion Qualifications Required: ~5+ years of development experience using C# & .NET framework 3.5 / 4.0 as full stack developer. ~ Experience building C# Console and windows applications for data processing. ~ Experience in Microsoft SQL Server and design... 
    Full time
    Work experience placement

    Jefferies

    New York, NY
    3 days ago
  • $160k - $200k

     ...experienced senior software engineer to join a growing team of developers within CCB and the other centers across the Flatiron Institute...  ...major cloud platforms. Experience with high-performance computing (HPC) environments. REQUIRED APPLICATION MATERIALS Please submit a r... 
    Full time
    Local area

    Simons Foundation

    New York, NY
    1 day ago
  • $155k - $200k

     ...Bring to the Team: ~2+ years of hands-on experience building, deploying, or operating cloud infrastructure, with exposure to AI/ML, HPC, or GPU workloads ~ Hands-on proficiency with at least one major cloud provider (AWS, GCP, or Azure), deploying and managing... 
    Full time
    Temporary work
    Work at office
    Monday to Friday

    Crusoe

    New York, NY
    2 days ago
  • $800 - $1,000 per month

    Our client is looking for a C# Developer to join their team in NYC. Pay: $800-1,000/dayQualificationsA Bachelor's Degree, Graduate Degree preferred 5 years of hands on programming experience with an understanding of Object Oriented Programming, Design patterns, Service... 

    Open Systems Technologies

    New York, NY
    1 day ago
  • We are seeking an experienced Algorithmic Trading Developer to design, build, and support sophisticated electronic trading solutions for equities and futures markets. This individual will work closely with trading, technology, and infrastructure teams to develop high-performance... 

    Huxley Associates

    New York, NY
    1 day ago
  • $145k - $165k

     ...Cloud Computing, Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing...  ...of a growing, dynamic, highly-focused team responsible for developing new or existing opportunities in the global IT market. The FAE... 
    Worldwide

    Support Revolution

    New York, NY
    4 days ago
  •  ...functional infrastructure initiatives. Demonstrated leadership experience, including mentoring junior engineers. Experience with HPC or GPU cluster infrastructure, including Slurm.. Experience building or operating AI agents or agentic infrastructure.... 
    Work at office
    Local area
    Shift work

    Cacheflow

    New York, NY
    4 days ago
  •  ...compute farm expands Required qualifications BS or MS in Computer Science, Computer Engineering, or equivalent experience 8+ years in HPC or large-scale batch compute, with 5+ years specifically on IBM Spectrum LSF Proven expertise in LSF internals and debugging beyond... 
    Full time
    Remote work

    Virtual Vocations Inc

    New York, NY
    3 days ago
  • $122k - $163k

     ...harness the full potential of our advanced Kubernetes-powered HPC cloud infrastructure. You'll be hands-on, collaborating with engineers...  ...a Kind.In this role, you will:Guide and mentor team members in developing their technical skills and troubleshooting capabilities across... 
    Full time
    Work experience placement
    Casual work
    Work at office
    Remote work
    Worldwide
    Shift work

    CoreWeave

    New York, NY
    3 days ago
  • $190k - $270k

     ...missions.The Databricks AI Research organization enables companies to develop AI models and agents using their own data, with technologies...  ...‑scale experiments, data processing, and model training (e.g., HPC clusters, GPU fleets, or cloud‑based systems)Enable researchers... 
    Local area
    Worldwide

    DataBricks

    New York, NY
    9 hours ago
  • $108k - $172.5k

     ...lasting impact on the world.We are seeking a highly motivated Senior HPC Support Engineer focussing on InfiniBand and NVLink technology,...  ...customer experience, support tools, etc.As a technical resource develop, re-define and document standard methodologies to provide to... 
    Full time
    Work experience placement

    Nvidia

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to HPC Developer. Be the first to apply!