Get new jobs by email
$148k - $216k
...utilization, data drift, and concept drift. Infrastructure Management: Provision and optimize cloud-based ML infrastructure (including GPU/CPU computing clusters) utilizing Infrastructure as Code (IaC) paradigms. Cross-Functional Collaboration: Work intimately with...SuggestedRemote workFlexible hours$132k - $191k
...Design and support cloud architectures for AI/ML workloads, including model training, inference, and high-performance compute (e.g., GPU/EDA burst capacity). Enable secure data pipelines, scalable compute environments, and integration of AI services while ensuring...SuggestedPermanent employment- ...infra /Kubernetes (Remote) Are you passionate about building scalable AI infrastructure and helping customers succeed with cutting-edge GPU platforms? We're looking for a Solutions Architect to join our team and work with enterprise customers deploying and optimizing AI/ML...SuggestedRemote work
- ...RESPONSIBILITIES: Own the discovery and definition of customer requirements for AI infrastructure use cases, including training, inference, GPU clusters, bare metal, managed orchestration, networking, and storage Work directly with strategic customers to understand their...SuggestedHourly payContract workLocal area
$175k - $220k
...on multi-functional teams to provide ethernet network expertise to server infrastructure builds, accelerated computing workloads and GPU enabled AI applications. Implementing tasks related to network configuration and validation for data centers. Create methods...SuggestedWorldwide- ...evaluate and guide the following areas: Future AI rack density and power consumption trends Impacts of next-generation GPU and AI chip architectures Optical networking and switching implications on infrastructure design AI workload impacts on utility...SuggestedWork at officeLocal areaWork visa
- ...as we shape the future of AI and beyond. Together, we advance your career. THE ROLE: AMD is looking for a Enterprise AI/HPC GPU architect to join our Datacenter System Architecture and Engineering team to develop world-class products around Instinct GPUs. In this...Suggested
- ...runbooks Enforce IAM least-privilege policies, secrets management, and FinOps cost controls Collaborate with AI/ML engineers on GPU workloads, model serving, and inference pipelines Own technical communication during clinet interaction, translating...Suggested
$123k - $163.5k
...workspaces for ML development. As we evolve, in this role, you’ll spearhead high-impact initiatives: designing multi-cloud setups to maximize GPU availability, driving deep-level model optimization, and building next-generation Agentic AI toolings. You will play a pivotal role...SuggestedFull timeWork at officeRemote work- ...engineering, product, and operations teams Define program scope, objectives, and success metrics for AI infrastructure initiatives, from GPU cluster buildouts to inference optimization projects Drive cross‑functional roadmap planning and prioritization, balancing immediate...SuggestedTemporary workWork at officeImmediate startFlexible hours
- ...to work across teams and manage up when needed NICE TO HAVE Experience with video‑first podcast formats Familiarity with AI/ML or GPU infrastructure topics Background in content strategy or editorial planning beyond pure production What We Offer Stock Options 100...SuggestedTemporary workWork at officeRemote workFlexible hours
- Overview Introl is seeking experienced Data Center Technicians to join our team as hourly, project-based W-2 employees supporting GPU infrastructure deployment projects across the U.S. This role is ideal for individuals who enjoy hands-on technical work in fast-paced environments...SuggestedHourly payLocal areaRelocationNight shiftWeekend work
- ...technology? We are looking for a talented Product Marketing Lead to join our dynamic team. Play a critical role in driving the success of our GPU cloud products by developing and executing innovative marketing strategies that resonate with our target audience. What You’ll Do...SuggestedTemporary workWork at officeRelocationFlexible hours
- ...engagement. You will strive to automate the delivery of existing and new Ubuntu products applied to all modern workloads from web servers to GPU‑aided AI for servers, VM’s and containers, and integrate our products with cloud native services. Come build a rewarding, meaningful...SuggestedWork at officeLocal areaRemote workWork from homeWorldwideFlexible hours
- ...keeps TensorWave running. This is a rare chance to shape IT at a company moving fast — working alongside engineering teams running GPU clusters across multiple data centers while establishing the systems and security controls that let the business scale without...SuggestedTemporary workWork at officeFlexible hoursShift work
- ...Experience in a customer-facing operations role at a cloud provider, managed services provider, or colocation facility ~ Exposure to GPU infrastructure, HPC clusters, or AI/ML compute environments. ~ Direct experience managing hardware replacement workflows through...Temporary workWork at officeLocal areaFlexible hours
- ...large-scale infrastructure platforms to support high-performance AI workloads across multiple data centers. Our environment includes GPU-intensive systems, high-throughput networking, and rapidly scaling compute clusters. We are looking for a Virtualization Operations...Temporary workWork at officeFlexible hours
- ...environments in real time across TensorWave data centers using monitoring and observability platforms Track key health indicators including GPU utilization, node availability, network performance, storage health, and Kubernetes cluster status Identify anomalies,...Temporary workWork at officeFlexible hoursShift workNight shiftRotating shiftDay shift
- ...Preferred Qualifications Experience with virtual cluster technologies (vcluster, Kamaji, or similar) Experience supporting GPU workloads in Kubernetes Familiarity with: NUMA-aware scheduling Topology-aware workloads Awareness of RDMA and high-throughput...Temporary workWork at officeFlexible hours
- ...code that makes it real. This is a hybrid architect-builder role. You'll move between high-level design sessions on our next-gen GPU fabric and the editor where you ship the Go, Python, or Rust that secures our APIs, orchestration plane, and firmware stack. If you believe...Temporary workWork at officeFlexible hoursShift work
- ...the Role We are hiring an AWS Cloud Engineer to design, provision, optimize, and support the AWS infrastructure powering our AMD GPU AI/HPC platform. This is a hands-on execution role — you'll work closely with Rust backend engineers, TypeScript developers, SREs, and...Temporary workWork at officeFlexible hours
- ...skills, particularly when engaging with both technical engineers and business stakeholders Preferred Experience Experience in GPU cloud, HPC, or AI/ML infrastructure operations Proficiency in Kubernetes administration, GPU workload scheduling, and multi-node...Temporary workWork at officeFlexible hoursNight shift
- ...high-performance infrastructure to power next-generation AI workloads. Our platform operates across multiple data centers and supports GPU-intensive environments with demanding requirements around performance, isolation, and scalability. We are looking for a Staff...Temporary workWork at officeLocal areaFlexible hours
$144k - $192k
...distributed training pipelines using frameworks such as PyTorch Distributed. Kernel Development : Design and maintain high-performance GPU kernels in Triton or CUDA for state-of-the-art ML workloads. Data Pipeline Engineering : Optimize robust data loading pipelines...Work at officeRemote work- ...’ll be responsible for building and operating the core systems that power large-scale ML training and inference across TensorWave’s GPU platform , working closely with cross-functional partners to support business objectives while upholding our standards for excellence...Temporary workWork at officeFlexible hours
- ...Support incident response and root cause analysis for storage-related issues Ensure storage platforms meet performance expectations for GPU and Kubernetes workloads Kubernetes Storage Operate and support Kubernetes-integrated storage - CSI drivers, StorageClasses,...Temporary workWork at officeFlexible hours
- ...related to services, networking, packages, storage, users, SSH, logs, and systemd. Support Ubuntu-based infrastructure systems and GPU node operating environments. Assist with baseline configuration, package management, service validation, and host-level...Temporary workWork at officeFlexible hours
$172k - $229k
...backend engineers to deploy distilled and RL models into production. Optimize for latency, throughput, and hardware efficiency across GPU/CPU clusters. Implement model versioning, A/B testing, and monitoring for performance regressions. Research and Integrate Agentic...Work at officeRemote work- ...commitments Partner with network, power, facilities, and infrastructure engineering teams to drive operational readiness for high-density GPU compute clusters Own post-incident retrospectives and corrective action tracking, driving lessons learned into durable process...Temporary workWork at officeFlexible hours
$84.9k - $209.5k
...directly with customers ranging from emerging AI startups to Fortune 500 enterprises, helping them architect and deploy complex HPC and GPU clusters, AI platforms, and intelligent agentic solutions across proof-of-concept and production environments. This highly visible...Temporary workFlexible hours
