Head of Infrastructure Support (New York)
Nscale
About Nscale
Nscale is the vertically integrated AI cloud engineered for AI. We own and operate the full stack — energy, data centres, GPU superclusters, orchestration, and AI services — delivering high-performance infrastructure to AI-native companies, enterprises, and governments across Europe and the US. We are deploying GPU capacity at hyperscale, operating some of the densest, most advanced AI infrastructure in the world.
About The Role
The Head of Infrastructure Support owns Infrastructure Support for their region — the team, the function, and its impact on customers. Reporting directly to the VP of Support and operating alongside counterpart Heads of Infrastructure Support across EMEA, the US, and APAC, you are accountable for the success of regional support outcomes: service performance, escalation quality, customer experience, and the health of the GPU estates your team supports.
The regional Infrastructure Support engineers report directly to you, and you own their management end to end — hiring, 1:1s, performance reviews, development planning, and documented performance management through to outcome. Their performance, growth, and results are your responsibility.
You will grow your regional Infrastructure Support team during a period of rapid company scaling, embed a consistent operating model with your counterparts in the other regions to deliver true follow-the-sun coverage, and act as the organisational accountability layer for your region — ensuring that strategic and tactical work spanning Support and Operations lands with clear owners and gets driven to completion. As the function maturing, you will own a global capability area on behalf of all regions and shape your team's structure — including developing team leads — as headcount grows.
You remain technically credible: close enough to GPU infrastructure, high-performance fabrics, and Linux operations to lead complex incident response, challenge technical decisions on their merits, and earn the respect of Senior Engineers — while spending the majority of your time leading.
Experience Required
8+ years in infrastructure, operations, or support engineering in production environments, including 4+ years of direct line management of engineers in an operational support function, with demonstrable ownership of performance management. Significant exposure to GPU, HPC, or large-scale data centre estates.
What You'll Be Doing
Regional Ownership & Accountability
- Own the success of Infrastructure Support for your region: service outcomes, customer impact, and team performance sit with you.
- Own regional service performance against defined KPIs — SLA adherence, MTTR, first-response time, backlog health, and CSAT — with accurate reporting to the VP of Support and senior leadership.
- Identify regional risks — capacity, capability, coverage, or customer — early, and either resolve them or escalate them with a clear recommendation.
- Own regional capacity modelling and headcount planning: forecast support demand against fleet growth and customer onboarding, and make the business case for investment to the VP of Support.
- Act as the regional accountability layer during rapid growth: when cross-functional work spanning Support, DC Operations, deployment, firmware, and Engineering lacks a clear owner, make sure it gets one and gets done.
- Partner with the Heads of Infrastructure Support in the other regions — across EMEA, the US, and APAC — to run a single global function: consistent standards, processes, and quality, with true follow-the-sun handover between regions.
- Own a global capability area on behalf of all regions — such as escalation management standards, the knowledge and runbook system, or the tooling and automation roadmap — working with other Heads of Infrastructure Support defining the standard every regional Support team operates to.
People Leadership & Team Building
- Own day-to-day people management for your regional Infrastructure Support team: regular 1:1s, performance reviews, development planning, and documented performance management — including underperformance — through to outcome.
- Hire and grow the team: define role requirements, run structured interviews, and build a bench of engineers who meet Nscale's technical and communication bar.
- Design your team's structure as the region scales, appointing and developing team leads and building second-line management capability as headcount grows.
- Set and monitor individual and team objectives, driving accountability and continuous improvement.
- Design and own shift planning, rota coverage, and on-call scheduling for the region, ensuring sustainable 24/7 support in coordination with the global coverage model.
- Identify skills gaps and drive upskilling through training, mentoring, and knowledge sharing across teams.
- Ensure roles, responsibilities, and expectations are clearly understood and consistently applied.
Service & Operational Performance
- Own ticket queue health for the region: accurate prioritisation, timely resolution, and clean escalation flow from frontline triage into L2/L3.
- Monitor team productivity and workload trends, addressing bottlenecks before they become service risks.
- Ensure adherence to ITIL-aligned processes across incident, request, change, and problem management.
- Improve dashboards, alerting, and runbooks to reduce repeat incidents and drive right-first-time resolution.
- Maintain consistent standards, processes, and documentation across regional teams; ensure compliance with audit, security, and operational requirements.
Incident, Escalation & Stakeholder Leadership
- Act as the senior regional escalation point for complex or high-impact incidents, including customer-facing escalations, participating in regional on-call as required.
- Lead post-incident reviews, identify recurring patterns, and ensure follow-up actions are tracked and delivered — converting incidents into problem records and durable fixes.
- Represent Infrastructure Support to regional customers and internal senior stakeholders; communicate clearly, candidly, and concisely at every level from engineer to executive.
- Contribute to readiness and support planning for new services, data centre deployments, and customer onboarding in the region.
Technical Leadership & Contribution
- Work alongside Senior Engineers on complex incidents, technical improvements, and operational tooling — close enough to the work to lead it credibly.
- Maintain hands‑on fluency across GPU infrastructure (drivers, firmware, hardware fault isolation, RMA workflows), Linux at scale, and east‑west high-performance fabrics (InfiniBand/RoCE diagnostics and fault isolation).
- Guide investigation quality: evidence‑led diagnosis, structured troubleshooting, and handovers that stand up to scrutiny.
- Contribute to scripting and automation direction to reduce toil across the regional operation.
- Travel to Nscale or customer sites when needed to lead onsite support activity.
About You
- Leadership experience.
- 5+ years of direct line management of engineers in an operational support environment, with end-to-end ownership of performance management: reviews, development plans, and documented underperformance processes through to outcome. You can describe your management framework and point to engineers you've grown.
- Operational ownership. Experience owning team workload, prioritisation, and service delivery against SLAs, with accountability for the numbers — and experience explaining those numbers to senior leadership.
- Function building. Experience hiring, scaling, or standing up support/operations capability in a fast-moving environment, including capacity modelling and headcount planning against demand; comfortable operating where processes are still evolving and helping define them without slowing delivery.
- Communication. Excellent written and verbal communication — clear, specific, and concise at every level, from ticket notes to executive updates to difficult customer conversations. We treat communication quality as a core leadership skill and assess it directly.
- Decisiveness and accountability. A bias for decisive action and calculated risk in ambiguous situations; you take ownership of outcomes, speak candidly, disagree when appropriate, and commit fully once decisions are made.
- Technical foundation — 8+ years across: Linux systems engineering in production, GPU infrastructure, high-performance east-wester fabrics, HPC scheduling and orchestration, networking fundamentals, data centre operations, observability and incident response, automation, process literacy.
- Strong understanding of ITIL-aligned incident, problem, and change management, and of SRE practices — runbooks, toil reduction, and continuous improvement.
- Adaptability. Comfortable with out-of-hours escalations, regional on-call participation, and travel for onsite leadership.
Nice to Have
- Deeper GPU/HPC exposure: NCCL-based performance troubleshooting, NVLink/NVSwitch, Slurm-scheduled multi-GPU workloads, or rack-scale systems.
- High-performance storage: Exposure to VAST or comparable AI-optimised storage platforms, Ceph, or NFS at scale.
- OpenStack and fleet operations tooling: OpenStack operations, or fleet-scale provisioning and health tooling (MAAS, NetBox, Redfish-driven automation, or similar).
- Kubernetes: Operating or supporting clusters and GPU operator stacks. Helpful context for our platform, though not the core of this role.
$125k - $250k
...reputation management firm with offices in DC, New York, Miami, Los Angeles, London, Paris, and... ...firm with a close‑knit, supportive team of professionals. Our model blends... ...relationships without major business‑development infrastructure. Superior case‑management skills and...SuggestedFull timeWork at officeImmediate startWorldwideFlexible hours$425k
...urgency and intellectual honesty and expect new team members to match our velocity. We... ...build together. The role At Aaru, infrastructure sets the limits of the product. A simulation... .... You want to build in person, in New York, at high speed. Strong candidates may...SuggestedFull time- ...Aaru in New York is seeking a lead Infrastructure Engineer to own the end-to-end compute and infrastructure that powers our agent-based simulations. You will set technical direction, build the infra team, and continue to write production code. You will design systems...SuggestedFull time
$350k
...InforCapital, partnership in New York, is seeking a Managing Director to lead the development and delivery of renewable energy projects, including solar and wind. The role entails team management, engagement with stakeholders, and strategic project oversight. Candidates...SuggestedFull time- ...Houston; New York; San Francisco; Seattle About Nscale Nscale... ...-effective, high-performance infrastructure for AI start-ups and large enterprise... ...capabilities and directly supports strategic business outcomes,... ...Overview We are seeking a Head of Infrastructure Operations...SuggestedFull timeContract workFor contractors
- Head of Infrastructure Security | Global Alternative Investment Firm Head of Infrastructure... ...every stage of the product lifecycle Support security evaluations for new products, platforms, and services... ...of Infrastructure” roles. New York, United States $248,600.00-$300,00...Full timeWork at officeRemote work
- ...A leading technology firm in New York is seeking a Head of Infrastructure to lead their global IT strategy. This senior position focuses on delivering innovative solutions and improving operational efficiency. Responsibilities include defining infrastructure strategy,...Full time
$250k
...You're stepping into one of the most ambitious infrastructure buildouts in private markets right now. This is a firm that administers over... ...across Cloud, Infrastructure Operations, and Application Support. Around 30 direct reports sit under you, and you're building...Full time- ...stage startup that is building the trust infrastructure layer for the AI economy. Starting in... ...is the first infrastructure hire - a Head of Infrastructure who will own and build... ...the customer security review process and support enterprise onboarding Implement monitoring...Full time
- ...We are seeking a Head of Infrastructure to lead the strategy and execution of our global IT Infrastructure organization. As a senior leader,... ...infrastructure principles (networking, storage, cloud, DevOps, end‑user support). Strong analytical, problem‑solving, communication, and...Full timeWork at officeFlexible hours
$50 per hour
...About the Role: We are seeking an experienced and visionary Head of Infrastructure Engineering to lead and scale our infrastructure, DevOps,... ...scalable, secure, and high-availability infrastructure to support complex microservice-based applications. Develop and...Full timeRemote work- ...A leading software security firm in New York is seeking a Sr. Sales Engineer to help developers integrate security into their applications effectively. You will use your extensive experience in customer success to drive engagement and influence the software development...Full time
$207k - $301k
...Google Inc. is looking for a Software Engineering Manager in New York, NY, to lead multiple teams in delivering high-quality software solutions. In this role, you will set technical goals, mentor engineers, and define project strategies to optimize performance and availability...Full time- ...Figma is seeking an Engineering Manager for the Growth Platform in New York. You will lead the initial engineering team, define the platform's direction, and integrate AI-driven approaches to enhance product capabilities. The ideal candidate has at least 3 years in engineering...Full timeRemote work
$270k - $310k
...receive an alert: Power and Energy Infrastructure Project Finance, Managing Director (New York) Requisition ID: 49999 Business... .... The role will report to the Head of Utilities, Power and Energy teams... ...inspiration, challenge, and support, with ample opportunities for visibility...Temporary workWork at officeImmediate startShift work3 days per week- ...Capital One is seeking a Sr Manager, Software Engineer in New York to lead a portfolio of distributed microservices and full-stack systems. You will guide a seasoned team of developers to deliver cloud-based solutions that meet regulatory requirements while driving innovation...Full time
- ...Ripple is looking for an innovative Engineering Manager to lead the Liquidity Management team in New York. This pivotal role involves driving the technical direction and development of a world-class platform for liquidity management. You will lead a team tackling complex...Full time
- ...AI, automation, and human expertise to support enterprises in more than 300 languages,... ...certifications. Welocalize is headquartered in New York with offices worldwide. The Cloud... ..., and optimization of cloud-based infrastructure, CI/CD pipelines, integration and delivery...Full timeRemote workWorldwide
- ...Sentio Recruitment LLC is seeking a Principal Cloud Architect for a US fintech scale-up in New York, NY. You will define the target AWS landing zone, lead the migration, and set security, reliability, and cost governance standards across engineering teams. You will partner...Full time
- ...services into a coherent system. The ideal candidate will have substantial experience in developing APIs, designing distributed systems, and mentoring technical teams. Offering a competitive compensation package, this position is based in New York, NY. #J-18808-LjbffrFull time
$165k - $190k
...Rumble is seeking a Senior Software Architect based in New York City to define and evolve the architecture of its Self-Service Cloud Platform. This role combines architectural leadership with hands-on development in Python, focused on building reliable, scalable services...Full time$207k - $301k
...corporate_fare Google place New York, NY, USA Qualifications ~ Bachelor’s degree, or equivalent practical experience... ...Reliability Engineering (SRE), managing high-availability service infrastructure, and on-call support rotations. Job Overview Like Google's own...Full time- ...Google Public Sector in New York is hiring a Principal Architect to lead enterprise architecture strategy for State of New York agencies... ...and alignment with executive stakeholders. Remote eligibility supports candidates located in New York, USA, to deliver Google Cloud...Full timeRemote work
- ...Google New York, NY, USA is seeking a Staff Software Engineer for the Office of the CTO to work on scalable, high-impact software systems and ML/AI initiatives. You will engage across distributed computing, data storage, security, and UI concerns, contributing to rapidly...Full timeWork at office
$196k - $240k
...Ripple is seeking an innovative Engineering Manager in New York, NY, to drive technical direction and build a world-class team. You’ll lead diverse technology projects and collaborate with product managers to deliver robust solutions. The role demands strong leadership...Full time$195k - $257.5k
...Circle, based in New York, is seeking a Technical Leader to drive their engineering teams. In this role, you will be responsible for managing dynamic teams, designing APIs, and ensuring alignment with company objectives. The ideal candidate should have over 7 years of...Full timeFlexible hours- ...A leading hedge fund in New York is seeking a Director of Software Engineering to lead technology delivery for their residential investment business. This senior role involves overseeing multiple development teams and ensuring stability across critical systems. Ideal...Full time
$139.9k - $274.8k
...power Microsoft’s cloud‑scale AI and Azure infrastructure. As a Network Engineer at L65, you will... ..., and telemetry frameworks that support Azure’s AI and data center infrastructure... ...year. For the San Francisco Bay area and New York City metropolitan area, the base pay...Full timeWork at officeLocal area- ...Versant Media in New York is looking for a Principal Engineer to spearhead modernization and design efforts. You will lead stakeholders and teams while leveraging SaaS and cloud-native technologies for impactful project delivery. This role mandates over 8 years in software...Full time
- ...Rove is seeking a Director of Software Engineering in New York to lead the engineering strategy and execution for its shopping and dining products. This senior role requires extensive experience in building high-scale consumer products, focusing on managing the technical...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Head of Infrastructure Support (New York). Be the first to apply!
- head of infrastructure New York, NY
- infrastructure manager New York, NY
- infrastructure engineering manager New York, NY
- director of infrastructure New York, NY
- information technology support New York, NY
- onsite support New York, NY
- recovery support New York, NY
- purchasing support New York, NY
- linux support New York, NY
- remote support New York, NY






















