Director of Infrastructure Engineering
$225k - $325kRunpod
Director Of Infrastructure Engineering
Runpod is the AI Developer Cloud. More than one million developers, from indie researchers to teams running frontier models in production, use Runpod to experiment, train, fine-tune, deploy, and scale AI on one platform. The platform has processed more than 20 billion inference requests. We're at an inflection point for AI infrastructure, and we're building the platform the next generation of developers will depend on. We're a small, remote-first team. We take ownership seriously, move fast, and ship work that more than a million developers rely on every day. We're looking for people who care deeply, build with urgency, and want to matter at scale.
We're looking for a Director of Infrastructure Engineering to lead and scale Runpod's core cloud and bare-metal environments. This role owns the critical foundational layers of our platform—Site Reliability Engineering (SRE), global networking, High-Performance Computing (HPC) networks, and distributed storage engines. You'll build the operating rhythm, culture, and technical direction that ensures Runpod remains highly available, performant, and capable of scaling to meet massive GPU computing demands.
You will partner tightly with Product Engineering, Product, and GTM leadership to support enterprise customers. Your focus is ensuring that our underlying infrastructure provides the absolute fastest, most reliable, and lowest-latency path for customers training and running large-scale AI workloads.
Responsibilities:
- Own Core Infrastructure & SRE: Lead multiple engineering teams responsible for Site Reliability Engineering, networking, and storage. Establish rigorous SRE practices, driving SLA/SLO definitions, incident response, observability, and automated remediation.
- Architect HPC & Global Networking: Oversee the design, scaling, and operation of Runpod's global network backbone, as well as ultra-low-latency HPC cluster networks. Drive the implementation and optimization of InfiniBand and RDMA over Converged Ethernet (RoCE) to support massive, multi-node GPU training workloads.
- Drive Storage Engine Innovation: Direct the architecture and performance tuning of highly scalable, distributed storage systems. Ensure our storage engines can deliver the massive IOPS and throughput required to keep high-end GPUs fed with data during deep learning tasks.
- Build a High-Output Org: Hire, mentor, and grow highly technical engineering managers and senior ICs (network architects, systems engineers, SREs). Create a culture of ownership, operational excellence, and craft in a remote-first environment.
- Translate Scale into Strategy: Partner with Program Management and Product to forecast capacity requirements, shape technical roadmaps, and convert massive scale challenges into clear technical scopes, milestones, and measurable outcomes.
- Continuously Improve Systems & Flow: Drive measurable improvements in infrastructure reliability and delivery metrics, such as deployment frequency, MTTR (Mean Time To Recovery), infrastructure as code (IaC) coverage, and system uptime.
- Architectural Stewardship: Provide architectural oversight for bare-metal provisioning, virtualization layers, network fabrics, and storage clusters, ensuring seamless scalability without becoming a bottleneck for your teams.
- Cross-Functional Partnership: Coordinate cleanly with product delivery and platform teams to ensure the infrastructure primitives they rely on are robust, well-documented, and highly available.
Requirements:
- Engineering Leadership Experience: 7+ years leading software, infrastructure, SRE, or networking teams, including managing managers and multiple squads, with a proven record of scaling high-availability cloud environments.
- Deep Infrastructure Expertise: 8+ years building and operating large-scale distributed systems, bare-metal infrastructure, or public/private cloud platforms.
- HPC & Advanced Networking: Proven hands-on background or strong architectural understanding of ultra-low latency networking. Deep familiarity with InfiniBand and/or RoCE, spine-leaf architectures, and global WAN routing protocols (BGP).
- Storage Systems Knowledge: Experience building, operating, or tuning high-performance distributed storage systems and parallel file systems (e.g., Ceph, Lustre, Weka, NVMe-oF) capable of handling heavy AI/ML I/O loads.
- SRE / DevOps Culture: Strong foundation in reliability engineering, infrastructure-as-code (Terraform, Ansible), container orchestration (Kubernetes), and modern observability stacks.
- Remote-First Operating Excellence: Experience building culture, accountability, and momentum across distributed technical teams.
- Communication & Collaboration: Clear written and verbal communication, strong stakeholder management, and calm, decisive leadership during high-stakes operational incidents.
- Background Check: Successful completion of a background check.
Preferred Qualifications:
- Direct experience architecting and operating infrastructure specifically optimized for massive GPU clusters and AI/ML workloads.
- Deep understanding of hardware architectures, GPU interconnects (NVLink), and datacenter topology.
- Track record of scaling infrastructure teams in hyper-growth startup environments.
- Open-source contributions or active recognition within the infrastructure, networking, or Kubernetes communities.
What You'll Receive:
- The competitive base pay for this position ranges from ($225,000 - $325,000). This salary range may be inclusive of several career levels at Runpod and will be narrowed during the interview process based on a number of factors, including the candidate's experience, qualifications, and location.
- Meaningful equity in a fast-growing company- everyone on the team receives stock options — your impact drives our growth, and you share in the upside.
- Generous medical, dental & vision plans.
- Flexible PTO- take the time you need to recharge.
- Most roles are remote work first with an inclusive, collaborative teams utilizing slack as the main form of internal communication.
- Join a passionate team on the cutting edge of AI infrastructure — where culture, learning, and ownership are at the heart of how we scale.
- $1,200 Home Office & Equipment Stipend- We set you up for success from day one with gear and support to create your ideal workspace.
Runpod is committed to maintaining a workplace free from discrimination and upholding the principles of equality and respect for all individuals. We believe that diversity in all its forms enhances our team. As an equal opportunity employer, Runpod is committed to creating an inclusive workforce at every level. We evaluate qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, marital status, protected veteran status, disability status, or any other characteristic protected by law. We welcome every qualified candidate eligible to work in the United States; however, we are currently unable to sponsor employment visas.
- ...from @Rippling.com addresses.About The RoleRippling Infrastructure operates as a scaled engineering organization responsible for the mission-critical systems... ...and internal engineering teams.We are looking for a Director of Engineering to lead our Data Infrastructure...Suggested
- ...with a company that's devoted to shaping the future of infrastructure in financial services. Let's collaborate to explore... ...and achieve extraordinary feats together.As a Senior Director of Infrastructure Engineering at JPMorganChase within the Core Engineering Solutions...SuggestedTemporary workShift work
- ...To build and scale the platform that powers production AI systems, the full-time remote Director of Infrastructure Engineering will own the strategy and execution of cloud infrastructure, developer platform, and reliability practices while leading a high-performing engineering...SuggestedFull timeRemote work
$225k - $325k
...Series A in June 2026. We're at an inflection point for AI infrastructure, and we're building the platform the next generation... ...CEO's funding announcement: We're looking for a Director of Infrastructure Engineering to lead and scale Runpod's core cloud and bare-metal...SuggestedRemote workHome officeVisa sponsorshipWork visaFlexible hours- Job TitleFor positions that will be based in CA, the annual salary range for this position is below. Actual salaries may vary based on numerous factors including, among other things, an individual applicant's experience and qualifications for the position. This range does...Suggested
$220k - $275k
...can navigate the future with curiosity, resilience, and a love of learning. The Role Outschool is hiring a Director of Infrastructure Engineering to lead our Infrastructure and Platform team, whose mission is to accelerate innovation and deliver operation stability...Full timeTemporary workWork at officeLocal areaRemote workHome officeFlexible hours$229k - $285k
...retail, and the public sector.Visit gomotive.com to learn more.About the Role: We are looking for a seasoned engineering leader who will lead the Core Infrastructure and Operations Engineering. You will be responsible for managing and mentoring a multi-functional team...Temporary workWork experience placement$272k - $340k
...custody operation, and every line of code at Ripple runs on the infrastructure our platform teams build. We operate a multi-cloud,... ...environment spanning multiple regions and regulatory jurisdictions, engineered for the availability and security standards our financial institution...Full timeWork at officeLocal area- ...lead in shaping the future of technology, unleashing your full potential, and making your mark on the industry. As a Director of Infrastructure Engineering - Global Head of Multimedia, Audio Visual & Digital Experience Engineering at JPMorgan Chase within the Corporate...
- ...implementation, and operation of the cloud infrastructure and developer platform that power... ...response, and mentoring a high-performing engineering team to keep systems secure, observable... ...Listed compensation for the Director of Infrastructure Engineering core team...Full timeRemote work
- ...the boundaries of what's possible together. As a Senior Director of Software Engineering at JPMorganChase within the Card Technology... ...standardized, repeatable, self-service provisioning through infrastructure as code and configuration automation.Coach and lead engineering...
- ...the future of financial services is waiting for you. Let’s push the boundaries of what's possible together. As a Senior Director of Software Engineering at JPMorganChase within the Card Technology organization, you lead multiple technical areas, manage the activities of...
- ...boundaries of what's possible together. As a Senior Director of Software Engineering at JPMorganChase within the Card Technology organization... ...patterns, and production readiness practices. Lead infrastructure-as-code adoption and portfolio standards using Terraform...
$216.2k - $324.3k
..., we continue to innovate and build new ways to foster friendship and connection. That’s where you come in!The Director of Cloud Infrastructure & Engineering leads Hasbro's critical infrastructure transformation: moving completely from on-premises data centers to AWS....Contract work- ...Hybrid up to three days per week from home in either Owing Mills, MD or Baltimore, MD Our client seeks a Director of Cloud Infrastructure and Engineering to lead a globally distributed team that designs, develops, and deploys cloud infrastructure and automation....Hourly payPermanent employmentFull timeLocal area3 days per week
$350k
...About this role The Role We're looking for a Director of Infrastructure Engineering to build and scale the platform that powers production AI systems. You'll own the strategy and execution behind our cloud infrastructure, developer platform, and reliability practices...Full timeLocal areaRemote work- In this high-impact role, you can step up as a tech leader and innovator by providing your knowledge and mentorship to infrastructure engineers. Lead teams towards excellence and demonstrate your skills as a leader.As a Senior Manager of Infrastructure Engineering at JPMorganChase...Shift work
$180k - $215k
..., collaboration, and excellence then we’d love to meet you. Infrastructure Core Services teams focus on delivering compute, network, data... ...orchestration in the cloud. We partner with product engineering teams to build performant, scalable, and secure services that...Permanent employmentH1bLocal areaRemote workFlexible hours$150k - $170k
Engineering Manager, InfrastructureLocation: Austin, TX - ON SITE ROLESalary: $150,000-$170,000/Year + Bonus and Relocation. We are seeking an experienced Engineering Manager, Infrastructure to lead a team responsible for our CI/CD ecosystem, build and release pipelines...Relocation$197k - $246k
THE POSITIONOur roster has an opening with your name on itAs the Infrastructure Engineering Director, you'll own FanDuel's foundational cloud and outpost-based infrastructure layer. You'll establish operational excellence, standardization, and reliability across compute...Temporary workLocal areaWorldwide$290k - $340k
...picture and our vision at Postman.The OpportunityPostman is seeking a strategic and results-driven engineering leader who is passionate about cloud agnostic infrastructure, operational excellence, and enabling engineering teams to operate autonomously and build with...Work at officeFlexible hours3 days per week$161k - $264.5k
...City, a test course for mobility; and Cloud & AI, the digital infrastructure powering our collaborative foundation. Business-critical... ...automotive software developers. That's why the Enterprise Technology Engineering Team (EnTec) builds solutions that enhance productivity, so...Temporary workFor contractorsWork at office- ...Architect professional to join our Workplace Technology Desktop Engineering team in New York City or Pittsburgh.The Chief Desktop... ...organizational resiliency, and reducing the likelihood and impact of infrastructure disruptions across the enterprise. At BNY, our culture...WorldwideFlexible hours
- Job Summary: The Manager Network Engineering directs the development and ongoing maintenance of critical IT platforms, processes, and network infrastructure. Leads high performing teams that deliver reliable, secure, and scalable solutions. Responsible for overseeing day...Local areaRelocationFlexible hours
- ...GAPosition SummaryNCR Voyix is seeking a hands-on Manager, Network Engineering to lead our global network engineering team while remaining... ...operation of our enterprise data center and core network infrastructure.This is a true player-coach role, ideal for a technical...Full timeWorldwideFlexible hours
- Change the world. Love your job. As the leader of the Network Engineering team at TI, you will play a pivotal role in shaping and... ...managing the architecture and roadmap for TI's global network infrastructure. You will partner with infrastructure, security, cloud, and...Local area
$120k - $140k
Position: IT Infrastructure Director / NW Engineer (Player Coach)Location: Mason City, IA / Ames, IA (Initially onsite, then hybrid)Salary: $120,000 - $140,000 base annual salary plus excellent benefits*** For immediate and confidential consideration, please APPLY and EMAIL...Immediate start- ...Broadridge is undertaking a multiyear transformation of the infrastructure foundations that support our products, clients, and technology... ...platforms. Key Responsibilities Lead infrastructure engineering strategy across hybrid infrastructure domains Define and...Local area
$250k - $309k
...Jose, California, United StatesProducts - Engineering /Fulltime /HybridOver 50,000 customers... ...Overview We are seeking a Senior Director of Engineering to lead the strategy,... ...generation Network OS for Switching and routing Infrastructure. This leader will drive innovation...Full time- ...Leading the architecture and engineering of enterprise network systems, the full-time remote Network Engineering Manager will manage... ...relevant work experience in network engineering, operations, or infrastructure management 5+ years of experience in engineering roles...Full timeWork experience placementRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Director of Infrastructure Engineering. Be the first to apply!
- information technology infrastructure manager United States
- cloud infrastructure manager United States
- infrastructure manager United States
- director of infrastructure United States
- head of infrastructure United States
- infrastructure engineering manager United States
- senior infrastructure manager United States
- engineering director United States
- project engineer assistant project manager United States
- principal engineer United States



