Senior Data Center Network Engineer - AI/HPC Infrastructure
$210k - $240kStratitech
Workplace Type: On-site Employment Type: FTE Salary: $210,000–$240,000 base salary, plus equity and benefits Note: No C2C arrangements will be considered. Any attempt to use personal or household contact information for solicitation, candidate submission, or vendor outreach is strictly prohibited and will be reported to LinkedIn. Build and help define the network foundation behind a high-performance AI platform supporting GPU infrastructure, distributed compute, real-time analytics, and mission-critical enterprise workloads. This is a network engineering role first . We are looking for a deeply hands-on senior engineer with Staff-level ownership who has designed and operated modern data center networks using BGP, EVPN/VXLAN, spine-leaf architectures, ECMP, WAN connectivity, and network automation. This role requires more than implementing an established architecture. You should be comfortable taking an incomplete infrastructure problem, determining the actual requirements, evaluating technical and cost tradeoffs, designing the solution, and carrying it through implementation and production operation. General SRE, DevOps, cloud infrastructure, or Kubernetes experience without significant production network architecture depth will not be sufficient for this position. What You’ll Own Architect and operate secure, high-performance data center and edge networks supporting AI and distributed compute workloads. Translate loosely defined infrastructure needs into technical requirements, architecture decisions, implementation plans, and operational standards. Design, configure, and troubleshoot spine-leaf fabrics using BGP, EVPN/VXLAN, and ECMP. Own WAN and LAN transit networks, backhaul connectivity, switches, routers, firewalls, load balancers, VPNs, carrier circuits, and external peering. Work directly with carriers and colocation providers on circuit provisioning, cross-connects, path diversity, testing, migration, and production acceptance. Make network architecture decisions involving resilience, performance, capacity, security, operational complexity, and cost. Evaluate and implement security policies, including segmentation, application firewall capabilities, packet inspection, and IPsec/VPN connectivity. Help transition network infrastructure toward secure directory-based authentication and individual administrative accountability. Automate network configuration, validation, deployment, and operational workflows. Manage network addressing, inventory, capacity, and configuration through IPAM/DCIM platforms. Troubleshoot complex performance and connectivity issues across network devices, Linux systems, compute infrastructure, and production services. Drive root-cause analysis and systemic remediation for recurring network issues rather than relying solely on monitoring and alerting. Improve observability, incident response, documentation, operational readiness, and network standards. Partner directly with infrastructure, engineering, data science, security, and other technical teams. Participate in on-call responsibilities and remain hands-on from physical infrastructure through network architecture and production troubleshooting. Required Network Expertise Eight or more years of network, data center infrastructure, or closely related engineering experience. At least five years of hands-on data center network engineering or administration. Demonstrated ownership of production network architecture , not solely implementation of designs created by another team. Experience taking infrastructure requirements from an incomplete or ambiguous starting point through design, implementation, and production operation. Production expertise with BGP and routing policy . Strong experience with modern spine-leaf network architecture. Hands-on experience with EVPN/VXLAN and ECMP . Strong Layer 2 and Layer 3 networking fundamentals. Strong experience configuring and troubleshooting switches, routers, firewalls, and load balancers. Experience with WAN circuits, carrier provisioning, backhaul connectivity, VPNs, and external peering. Knowledge of MPLS, QoS, IPsec, network security, capacity planning, and failure-domain design. Experience with NetBox, Infoblox, Nautobot, or another IPAM/DCIM or network source-of-truth platform. Network automation experience using Ansible, Nornir, Python, Terraform, Jinja2, APIs, or similar technologies. Strong Linux and network troubleshooting skills. Demonstrated ability to investigate recurring outages, determine root cause, identify architectural fragility, and implement lasting corrective action. Ability to operate autonomously in an environment where requirements may not already be fully defined. Strong judgment around when to solve independently and when to escalate or bring in another specialist. What Staff-Level Ownership Looks Like Here The strongest candidates will be able to point to examples where they personally: Took an ambiguous networking or infrastructure problem and established the requirements. Designed the architecture rather than only implementing someone else's design. Owned a significant network environment or problem space end to end. Made technical tradeoffs involving reliability, performance, security, scalability, and cost. Diagnosed repeated failures and corrected the underlying system or architecture. Established network standards, automation, tooling, or operational practices used by others. Worked across network, systems, security, cloud, storage, or compute boundaries when needed to solve the actual problem. Stayed hands-on after making the architecture decision. A strong Senior Engineer may still be considered, but this role is best suited to someone already demonstrating Staff-level scope, autonomy, and architecture ownership . These Skills Are a Plus InfiniBand networking for GPU or HPC environments RDMA or RoCEv2 networking, including PFC, ECN, congestion management, and performance validation GPU cluster, HPC, or AI/ML infrastructure experience NVIDIA networking technologies or other high-performance network fabrics Kubernetes networking, including CNI, ingress, service networking, and network policy High-throughput or distributed storage networking, including NFS, Lustre, Ceph, BeeGFS, NVMe-oF, or iSCSI Hybrid cloud networking across AWS, Azure, or GCP Prometheus, Grafana, Datadog, ELK, OpenTelemetry, or similar observability platforms Experience with colocation environments such as Equinix Experience building or expanding physical data center environments Experience in a startup or rapidly scaling technical environment Experience supporting SOC 2 or other high-compliance environments The Environment This is a startup environment where the right answer may not already be documented. You should be comfortable: Asking the questions necessary to establish requirements Making cost-conscious technical tradeoffs Moving between architecture and hands-on implementation Working through unfamiliar infrastructure problems Managing escalations without immediately handing problems off Collaborating across technical domains Learning new technologies quickly when the problem requires it We are looking for someone who sees an undefined infrastructure problem and thinks, “I'll figure out what we need and build it,” rather than waiting for someone else to provide the complete architecture. Employment and Work Authorization Candidates must be authorized to work in the United States without current or future employer sponsorship. This opportunity does not support C2C arrangements, agencies, or third-party candidate submissions. #J-18808-Ljbffr Stratitech
$225k
...Join a rapidly scaling AI cloud infrastructure provider building next... ...in high-performance networking and AI-ready data center architecture. The company... ...is looking for a Senior Network Engineer to design, deploy, and... ...optimised for AI and HPC workloads within a rapidly...SuggestedFull timeRemote work$250k - $320k
Gimlet Labs, Inc. is seeking a Network Engineer to design and build network infrastructure for AI workloads at scale. This role involves ensuring robust and reliable networking for production systems across distributed environments, focusing on performance and efficiency...Suggested$250k
...Ready to architect AI infrastructure that powers next-generation... ...stealth-mode hyperscale data center startup building a next-generation... ...chance to join as a Senior Inference Platform Engineer at an early stage and help... ...distributed systems (ML inference, HPC, or similar). ~...SuggestedFull time$250k
...rapidly scaling AI cloud infrastructure provider building... ...looking for a Senior Storage Engineer with experience... ...infrastructure, networking, and platform engineering... ...teams across data-intensive AI and... ...AI and HPC workloads Manage... ...large-scale data center environments...SuggestedFull timeRemote work$275k
...compute and inference infrastructure. Backed by leading venture... ...fully managed AI infrastructure and offers... ...is seeking a Founding Engineer to take ownership of core... ..., high-bandwidth networking, and inference platforms... ...operating large-scale AI or HPC infrastructure in...SuggestedFull timeRelocation- ...scale systems that span data centers, GPUs, networking, and more, ensuring... ..., and responsible AI deployment over... ...role As a software engineer on the Fleet High Performance... ...Computing (HPC) team, you will be responsible... ...our supercomputing infrastructure. Our team...Full time
$224k - $284k
...Atoms builds Physical AI — real-world robots... ...We are roboticists, engineers, operators, and builders... ...do We’re seeking a HPC Network Engineer to join our... ...Collaborate across the infrastructure team to solve cross-discipline... ...GitHub; awareness of data center power and cooling....Full timeWork at officeImmediate startFlexible hours- IREN, a vertically integrated AI Cloud provider, seeks a Senior Director to lead end-to-end capacity and S&OP for data center infrastructure. You will align demand forecasts with GPU/... ...plans, coordinating with Operations, Engineering, Finance, and Commercial teams...
- About the Team OpenAI’s Infrastructure Operations team is... ...the world’s largest AI infrastructure networks. The team owns day-... ...Compute's data centers, working with colocation... ...Infrastructure Operations Engineer to operate and... ...center, cloud, AI, or HPC networks and can move...Permanent employment
- Oracle seeks a senior Strategic OCI & AI Infrastructure Lead to own and scale high-value AI and cloud deployments across global markets... ...multi-million-dollar negotiations, and align engineering, data center operations, networking, support, and sales to deliver ambitious...
$102.3k - $209.5k
...Overview Oracle Cloud Infrastructure (OCI) is... ...through advanced AI infrastructure and... ...executive leadership, engineering, and field teams... ...is seeking a senior leader to serve as... ...compute, storage, networking) Anticipate constraints... ...engineering, data center operations,...Contract workTemporary workFlexible hours- StratITech is hiring a senior network engineer to architect and operate secure, high-performance data center and edge networks supporting AI and distributed compute workloads. You will translate loosely defined needs into concrete requirements and plans, and own the end...
- ...Business Development Manager focused on infrastructure to source, structure, and... ...service providers, silicon, and data center equipment. You will work with engineering, legal, finance, and operations to... ...resources supporting state-of-the-art AI systems and growth in a fast-...
$226k - $285k
...TeamOpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference... ..., mechanical, controls, network, hardware, construction,... ...centers, AI infrastructure, HPC environments, colocation, or...Work at officeLocal areaFlexible hoursShift work- ...liquid market for GPU offtake and operates AI infrastructure for enterprise customers. We are... ...processes as we grow and work closely with engineering to maintain fast, reliable support.... ...drive CSAT and SLA performance, and partner with data-center #J-18808-Ljbffr SF Compute
- ...performing team that delivers infrastructure and performance... ...Lead Infrastructure Engineer at JPMorganChase... ...Enterprise Technology Data Center Networking team, you apply deep... ...enterprise-authorized AI capabilities within the... ...Data Direct I/O) for HPC workloadsFEDERAL DEPOSIT...
$225k - $275k
...only vertically integrated AI infrastructure company built from the... ...energy, manufacturing, data center construction, and cloud... ...RoleCrusoe Cloud is seeking a Senior Staff Network Deployment Engineer to serve as the... ...high-performance compute (HPC) and GPU-based AI infrastructure...Temporary workRemote work- ...building the next generation of AI infrastructure: large-scale AI datacenters... ...Gimlet Labs is seeking a Network Engineer to design, build, and scale... ...knowledge of modern data center networking and is comfortable... ...Understand high‑performance AI/HPC networking concepts such as...
$156.86k - $191.72k
...Research Scientific Computing Center (NERSC) is seeking a System Infrastructure / Platform Engineer to help build and manage HPC systems and Linux-based... ...storage, high-speed networking, Slurm, and Kubernetes, balancing... ...more of the following: data center networking (TCP/IP...Permanent employmentFull timeRemote workFlexible hours$195k - $235k
...vertically integrated AI infrastructure company built from... ..., manufacturing, data center construction, and cloud... ...is seeking a Staff Network Production Operations Engineer to help own... ...technical guidance to Senior engineers and contribute... ...fabrics for GPU or HPC workloads,...Temporary workWorldwide$160k - $195k
...only vertically integrated AI infrastructure company built from the... ...energy, manufacturing, data center construction, and cloud... ...This RoleWe’re seeking a Senior Cloud Infrastructure Engineer to own the design, implementation... ...workloads (VMs, virtual networks, storage, and supporting...Temporary workWork at office$200k - $250k
...Fluidstack, we build the compute, data centers, and power that will fuel... ...to the world’s biggest AI Labs at industry-defining... ...is looking for a Software Engineer, Infrastructure Platform to build the foundational... ...for infrastructure assets, network topology, and configuration...Full timeLocal area$193k - $234k
...only vertically integrated AI infrastructure company built from the... ...across energy, manufacturing, data center construction, and cloud... ...energy, detail-oriented Staff Network Production Engineer to lead the physical and... ...high-performance compute (HPC) and GPU-based AI infrastructure...Temporary workRemote work$293k - $385k
...capacity across first-party data centers, strategic partners, and industrial infrastructure environments. We focus... ...can support frontier AI training and inference... ...delivery, power, networking, hardware deployment,... ...Partner with infrastructure engineering, hardware, networking,...Work at officeLocal areaFlexible hours- Gimlet Labs is seeking a Network Engineer to design, build, and scale the network infrastructure powering production‑scale AI and distributed systems. The role covers physical infrastructure... .... The ideal candidate will have deep data center networking knowledge, experience with...
- OpenAI is seeking an Infrastructure Operations Engineer to operate and improve large-scale Ethernet fabrics supporting GPU clusters, storage, and management... ..., observability, and incident response across a global AI network. You will partner with architecture, deployment, and...
$300 per month
...only vertically integrated AI infrastructure company built from the ground... ...energy, manufacturing, data center construction, and cloud services... ...GPU Fleet Operations Engineer to join Crusoe’s Fleet Operations... ..., interconnects, and networking hardware.Conduct post-repair...Temporary workWork at office$105.4k - $207.8k
...you'll do As a Senior Consultant, Strategy... ...security, infrastructure security, artificial... ..., and data protectionPerforming... ...Security, Engineering, or Information... ...SC-400, SC-500, AI-300, AI-103, or... ...Cisco Certified Network Professional (CCNP... ...Global Call Center (GCC) at USTalentCICInbox...Local areaVisa sponsorship- ...intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we... ...experts across energy, manufacturing, data center construction, and cloud services.If... ...us at Crusoe.About the Role:As an Engineering Manager on the Managed AI team at...Temporary workWork at office
$190k - $270k
Staff Software Engineer - AI Research InfrastructureP-1215At Databricks,... ...we are obsessed with enabling data teams to solve the world’s... ...Software Engineer, AI Research Infrastructure, you will be developing and running... ..., and model training (e.g., HPC clusters, GPU fleets, or...Local areaWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Data Center Network Engineer - AI/HPC Infrastructure. Be the first to apply!
- junior data engineer remote San Francisco, CA
- director data engineering San Francisco, CA
- hadoop big data developer San Francisco, CA
- data engineer intern San Francisco, CA
- senior data quality engineer San Francisco, CA
- senior cloud data engineer San Francisco, CA
- data engineer contract San Francisco, CA
- data infrastructure engineer San Francisco, CA
- entry level big data engineer San Francisco, CA
- data science developer San Francisco, CA


