GPU Infrastructure Engineer — Scale & Automate Massive Compute
$175k - $300kFluidstack
Fluidstack is seeking a highly skilled individual to manage compute fleet health and automate GPU repair workflows in Seattle. The role involves overseeing the entire GPU system, developing pipelines for metrics and alerts, and driving automation efforts in a fast-paced AI environment. The ideal candidate should possess a strong intuition for hardware and experience with AI tooling, and be ready to tackle ambiguous challenges. The salary range is competitive, from $175,000 to $300,000 annually, including equity options and comprehensive benefits. #J-18808-Ljbffr Fluidstack
$180k - $200k
...software with cost-efficient, large‑scale compute. Teams get the tools they need for... ...For Lightning AI is seeking a GPU & Compute Infrastructure Engineer to join our Infrastructure Engineering... ...systems, and software—developing automation, improving reliability, and...SuggestedRemote workWork from homeFlexible hours$175k - $300k
Fluidstack in Seattle is seeking a production engineer to oversee the health and operation of their massive GPU fleet. This role focuses on building automation for repair processes and ensuring high reliability and scalability in production. The ideal candidate has a strong...Suggested$184k - $287.5k
...5 Job Category: Engineering Time Type: Full... ...leading accelerated computing for the world's... ...across the stack—from GPU operator and... ...world problems at scale. In this pivotal... ...challenge of scaling AI infrastructure while optimizing... ...innovative, automated tests that...SuggestedFull timeWorldwide$165k - $242k
...innovators to build and scale AI with confidence... ...combines superior infrastructure performance with... ...and turn compute into capability. Founded... ...The Role Senior engineers are area owners who... ...teams to evolve our GPU performance... ...and operators to automate infrastructure workflows...SuggestedPermanent employmentTemporary workCasual workWork at officeRemote workFlexible hours- ...innovators to build and scale AI with confidence... ...combines superior infrastructure performance with... ...and turn compute into capability. Founded... ...the role Senior Engineers are area owners who... ...teams to evolve our GPU performance... ...performance tests and automation workflows to expand...Suggested
- Palantir is seeking a Senior Software Engineer for its Substrate team, building and operating the world’s leading Kubernetes-based infrastructure across on‑prem, cloud, and edge environments. You will automate, scale, and secure dozens of K8s clusters with zero manual...
$177.69k - $341.73k
...global, intelligent network infrastructure to meet the requirements of... ...center networking on a massive scale. Responsibilities Responsible... ...'s global high performance computing (HPC) networks. Work with... ...Science, Information Science, Engineering, Mathematics, or a related...Temporary workWork experience placementLocal areaRemote work$157k - $213.8k
Databricks is looking for a Senior Software Engineer to join their Networking Infrastructure team in Bellevue, Washington. You will design and automate the networking foundations for large-scale compute clusters, ensuring secure and efficient connectivity across major...$152k - $241.5k
...been transforming computer graphics, PC gaming... ...of forward‑thinking engineers tackling some of the... ...stack, including GPU operators, device plugins... ...problems at large scale and help shape how AI infrastructure runs in production.... ...to design automated, at‑scale workload...Remote work$202.16k - $368.22k
Senior Software Engineer - Compute Infrastructure (Orchestration & Scheduling) Location:... ...infrastructure powers hundreds of large‑scale clusters globally, with... ...cost efficiency on a massive scale within our rapidly... ...resources (including CPU, GPU, power, etc.), directly impacting...Temporary workLocal areaOverseas- The University of Washington seeks a Linux Web Infrastructure Engineer to maintain integration and software solutions within its Shared Web Hosting... .... Required qualifications include a Bachelor's degree in Computer Science, four years of relevant experience, and proficiency...Full time2 days per week
$188k - $275k
...innovators to build and scale AI with... ...combines superior infrastructure performance with... ...breakthroughs and turn compute into capability.... ...’ll Do The Field Engineering organization at... ...lifecycle: leading new GPU cluster bring‑up... ...scripting and automation related to bare‑...Permanent employmentContract workTemporary workCasual workWork at officeFlexible hours$202.16k - $368.22k
...Regular About the Team The Compute Infrastructure - Orchestration &... ...that powers hundreds of large‑scale clusters globally, with millions... ...resource cost efficiency on a massive scale within our rapidly... ...Computer Science, Computer Engineering or a related area with 3+ years...Temporary workInternshipLocal area- Allenai, based in Seattle, is seeking a Senior Software Engineer to lead AI infrastructure projects. This role involves designing systems that guarantee... ...and ensuring operational excellence through high-level automation. The ideal candidate has over 8 years of experience in...
- Insight Global is seeking a Senior Network Engineer to deploy and automate one of the world’s largest networks. The role emphasizes autonomous... ...communication, and project management to design scalable, reliable infrastructure. You will tackle complex routing and interconnectivity...
$112.73k - $177.84k
...Responsibilities The Compute Infrastructure team uses... ...hundreds of large‑scale clusters globally... ...efficiency on a massive scale within our... ...talented software engineers excited to optimize... ...(including CPU, GPU, power, etc.), directly... ...updates, automating routine resource...Temporary workInternshipLocal areaOverseas$180k - $200k
...with cost-efficient, large-scale compute. Teams get the tools they need... ...AI is seeking a Storage Infrastructure Engineer to join our Infrastructure... ...and operations—developing automation, improving reliability, and... ...technologies (e.g., RDMA, GPU Direct Storage) Experience...Remote workWork from homeFlexible hours$130k - $200k
...intent into action, automating website updates, design... ...The Role As a Senior Infrastructure Engineer at Gradial, you will... ...thrives in startup-to-scale up environments and... ...ownership of real-time, compute-intensive services:... ...infrastructure, including GPU provisioning and...Full time$139.8k - $277.8k
...performant, and reliable, 24/7. As an Infrastructure engineer, you'll be at the heart of making this... ...generative AI. You'll build resilient systems, automate across the stack, and champion... ...on building and operating large-scale, high-availability production systems...Full timeWork at officeLocal areaFlexible hoursShift work$200k - $275k
...’s High Performance Computing (HPC) Network Engineering team designs and engineers... ...communications infrastructure that underpins our incredibly large GPU and CPU compute... ...architect, optimize, and scale the high-performance... ...the cutting edge of automation in every part of our...Work at officeImmediate startWorldwide- CoreWeave in Bellevue is seeking a Senior Software Engineer to enhance the control plane for GPU data centers. The ideal candidate will design and operate... ...develop reliable APIs to manage the lifecycle of infrastructure. This role demands collaboration with hardware and...Flexible hours
$262k - $365k
Senior Staff Software Engineer, Google Cloud Compute, Performance Google — Seattle, WA, USA Benefits... ...experience building and developing large-scale infrastructure, distributed systems or networks,... ..., benchmark, and optimize their massive-scale workloads to ensure...Temporary work$130.05k - $175.95k
Compute Platform Engineer II We are a full‑stack shop consisting of product and... ..., data engineering, infrastructure and DevOps, data / metadata... ...generation, metadata‑ and automation‑driven data experience for... ...Aggressively engineering our data at scale, as one unified asset, to...Local area$188k - $275k
A leading AI cloud provider in Bellevue, WA, is hiring a Staff Software Engineer to lead the development of secure, scalable GPU infrastructure. The role requires significant engineering experience and expertise in cloud technologies, particularly with Go and Kubernetes...$272k - $431.25k
...has been transforming computer graphics, PC gaming,... ...Principal Software Engineer to join our DGX... ...’s high-performance GPU infrastructure. You will play a meaningful... ...crafting scalable automation solutions,... ...lifecycle operations at a massive scale. Drive technical alignment...$184k - $356.5k
...is seeking a Senior Software Engineer for the Accelerated Kubernetes... ...Runtime Team. You will design automation systems for seamless installation... ...optimizing them for NVIDIA's GPU architecture. Ideal candidates... ...implement automation in large-scale environments. NVIDIA offers a...$186.07k - $218.9k
As a Senior Software Engineer on the Compute Platform team within the Platform... ...compute orchestration infrastructure that every service at Coinbase... ...design and ship tooling, automation, and new capabilities that... ...and self‑healing at scale. Build developer‑facing tooling...Local area- Nscale in Seattle is seeking a Principal Observability Platform Engineer to own the technical direction of its observability platform. You... ...and architecture for the systems ensuring deep visibility into GPU clusters and AI workloads. The ideal candidate will possess over...
$202.16k - $368.22k
ByteDance is looking for a Senior Software Engineer specializing in AI Infra Compute to join our team in Seattle. This role involves designing and implementing... ...development, and a deep understanding of cloud infrastructure and AI technologies. We offer comprehensive benefits...$180.5k - $225.6k
...guide evolution, and scale the team. Ensure strong... ...varied use cases such as GPU onboarding, UDF... ...scale. Unify a fragmented compute surface by converging... ...Qualifications 5+ years managing engineers building and operating... ...technical fluency in infrastructure systems; ability to...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to GPU Infrastructure Engineer — Scale & Automate Massive Compute. Be the first to apply!
- principal infrastructure engineer Seattle, WA
- infrastructure engineering manager Seattle, WA
- infrastructure automation engineer Seattle, WA
- data infrastructure engineer Seattle, WA
- infrastructure developer Seattle, WA
- infrastructure engineer Seattle, WA
- entry level infrastructure engineer Seattle, WA
- remote infrastructure engineer Seattle, WA
- lead infrastructure engineer Seattle, WA
- security infrastructure engineer Seattle, WA

