Get new jobs by email
$132k - $191k
...Design and support cloud architectures for AI/ML workloads, including model training, inference, and high-performance compute (e.g., GPU/EDA burst capacity). Enable secure data pipelines, scalable compute environments, and integration of AI services while ensuring...SuggestedPermanent employment$148k - $216k
...utilization, data drift, and concept drift. Infrastructure Management: Provision and optimize cloud-based ML infrastructure (including GPU/CPU computing clusters) utilizing Infrastructure as Code (IaC) paradigms. Cross-Functional Collaboration: Work intimately with...SuggestedRemote workFlexible hours- ...Level Troubleshooting: Investigating and troubleshooting problems and hardware faults that our automation can't determine within our GPU platforms. This will involve taking data from system logs, kernel logs, BMC redfish APIs, and if the data is not there, working with...SuggestedLong term contractWork from home
- ...data handling, regulatory expectations, and third-party data use Secure AI infrastructure and supply chain: Harden AI platforms, GPU and container workloads, model registries, and artifact stores Assess risks in third-party models, libraries, embeddings, and...Suggested
- ...infra /Kubernetes (Remote) Are you passionate about building scalable AI infrastructure and helping customers succeed with cutting-edge GPU platforms? We're looking for a Solutions Architect to join our team and work with enterprise customers deploying and optimizing AI/ML...SuggestedRemote work
- ...Experience with distributed storage systems and understanding of one or more of object, block, and file storage paradigms. Hardware and GPU troubleshooting experience (nice to have). Exposure to OVN/OVS-based networking stack (nice to have). Strong communication skills....SuggestedNight shiftDay shift
- ...evaluate and guide the following areas: Future AI rack density and power consumption trends Impacts of next-generation GPU and AI chip architectures Optical networking and switching implications on infrastructure design AI workload impacts on utility...SuggestedWork at officeLocal areaWork visa
- ...RESPONSIBILITIES: Own the discovery and definition of customer requirements for AI infrastructure use cases, including training, inference, GPU clusters, bare metal, managed orchestration, networking, and storage Work directly with strategic customers to understand their...SuggestedHourly payContract workLocal area
- ...as we shape the future of AI and beyond. Together, we advance your career. THE ROLE: AMD is looking for a Enterprise AI/HPC GPU architect to join our Datacenter System Architecture and Engineering team to develop world-class products around Instinct GPUs. In this...Suggested
$175k - $220k
...on multi-functional teams to provide ethernet network expertise to server infrastructure builds, accelerated computing workloads and GPU enabled AI applications. Implementing tasks related to network configuration and validation for data centers. Create methods...SuggestedWorldwide- ...runbooks Enforce IAM least-privilege policies, secrets management, and FinOps cost controls Collaborate with AI/ML engineers on GPU workloads, model serving, and inference pipelines Own technical communication during clinet interaction, translating...Suggested
- ...technology services provider with 14 years of experience delivering innovative, business-critical solutions. We specialize in AI and GPU deployments, comprehensive data center support, IMAC services - desktop relocation, ensuring organizations transition seamlessly into...SuggestedFull timeFor subcontractorRelocation
$70k
...Experience with scripting languages like Python3 or MATLAB. Preferred Skills Experience with ROS/ROS2, real-time sensor fusion, and GPU-based processing. Knowledge of object detection, classification, and tracking algorithms. Familiarity with communication protocols...SuggestedFull timeLocal area- ...modeling to drug discovery, CHPC empowers researchers to solve complex problems through a massive ecosystem of 50,000+ CPU cores, 1,000+ GPU resources, and 50 PB+ of storage. CHPC seeks a network engineer to support connectivity both internally and externally to multiple...SuggestedFull timeCasual workWork at officeRemote workWork from homeRelocationHome office
- ...to drug discovery, CHPC empowers researchers to solve complex problems through a massive ecosystem of over 50,000 CPU cores, 1,000+ GPU resources, and 50PB+ of storage. CHPC seeks a network engineer to support connectivity internally and externally to multiple remote environments...SuggestedFull timeWork experience placementCasual workWork at officeRemote workWork from homeRelocationHome officeDay shift
$135k
...Responsibilities: * Design and development of hardware systems across RF/microwave, low-noise analog, power systems, FPGA hardware, and embedded GPU platforms * Evaluate tradeoffs between RF hardware and signal processing (detection, tracking, calibration, interference mitigation)...- ...3+ years working on AI/ML systems. ~ Strong understanding of modern AI infrastructure components, including distributed computing, GPU-accelerated systems, and large-scale storage. ~ Hands-on experience with cloud platforms such as AWS, Azure, or Google Cloud. ~ Proficiency...Local area
- ...technology services provider with 14 years of experience delivering innovative, business-critical solutions. We specialize in AI and GPU deployments, comprehensive data center support, IMAC services - desktop relocation, ensuring organizations transition seamlessly into...Work at officeRelocationMonday to Friday
$112.8k - $264.1k
...support for new HW platforms. We support both Bare Metal and Virtual machine instances across a diverse fleet of HW, including clustered GPU platforms. In addition, our customers demand high availability and security from our Cloud. We need a leader who can thrive in...Temporary workFlexible hours$135.2k - $306.4k
...support larger clusters, higher scale, improved operability, deeper OCI integrations, and increasingly demanding cloud native, AI, and GPU workloads. We are looking for a senior IC5 software engineer with deep Kubernetes expertise, required cloud infrastructure...Temporary workRemote workFlexible hours- ...inference optimization (quantization, SGLang/vLLM serving for audio, distillation), experience with at least one open‑source TTS family. GPU cost modeling. Proficiency in Python and familiarity with speech processing libraries and tools. Experience with cloud‑based...Work at officeRemote workFlexible hours
$114.6k - $234.6k
..., and maintain complex system software modules for managing, monitoring, and provisioning computer servers, storage, networking, and GPU subsystems in hyperscale environment. Utilizes advanced knowledge to design, develop, and deploy software layers and tooling to automate...Temporary workFlexible hoursShift work$95k - $171k
...activities involve coding, improving dashboards, enhancing alerts, and minimizing repetitive tasks. Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$126.2k - $264.1k
...area or are able to relocate permanently to the area. **** Preferred Qualifications Direct commissioning experience supporting GPU/high-density or liquid-cooled data halls (CDUs, leak detection, controls integration). Experience commissioning hyperscale or...Temporary workFor contractorsLive inLocal areaRelocationRelocation packageFlexible hours$146.4k - $263.6k
...other SREs, influence architecture decisions with product engineering teams, and shape SRE practices for AI inference workloads and GPU infrastructure at scale. As a Senior II Site Reliability Engineer, you will be responsible for: Taking ownership of observability...Work experience placementWork at office$89.2k - $209.5k
...team is responsible for deliver trusted, fast health determinations and customer‑initiated diagnostics that reduce false positives for GPU clusters, prevent unnecessary node returns, increase capacity for customers, protect revenue, and improve uptime—by providing an OCI‑...Temporary workFlexible hours$74.1k - $148.3k
...schedules. • Ensure that all work complies with OCI specifications, manufacturer warranty standards, and regional regulations. GPU Liquid-Cooled Rack Megaprojects • Serve as the technical and delivery lead for GPU-intensive data hall builds, managing low-voltage...Temporary workLive inLocal areaWorldwideRelocationRelocation packageFlexible hours- ...Google, Meta, Amazon, or similar). Deep familiarity with direct liquid cooling (DLC) and immersion cooling systems for high-density AI/GPU compute environments. Commissioning experience or ASHRAE/NEBB/AABC certification relevant to mission-critical mechanical systems....Contract workFor contractorsRemote work
