Get new jobs by email
$148k - $216k
...utilization, data drift, and concept drift. Infrastructure Management: Provision and optimize cloud-based ML infrastructure (including GPU/CPU computing clusters) utilizing Infrastructure as Code (IaC) paradigms. Cross-Functional Collaboration: Work intimately with...SuggestedRemote workFlexible hours$132k - $191k
...Design and support cloud architectures for AI/ML workloads, including model training, inference, and high-performance compute (e.g., GPU/EDA burst capacity). Enable secure data pipelines, scalable compute environments, and integration of AI services while ensuring...SuggestedPermanent employment- ...Experience with distributed storage systems and understanding of one or more of object, block, and file storage paradigms. Hardware and GPU troubleshooting experience (nice to have). Exposure to OVN/OVS-based networking stack (nice to have). Strong communication skills....SuggestedNight shiftDay shift
- ...data handling, regulatory expectations, and third-party data use Secure AI infrastructure and supply chain: Harden AI platforms, GPU and container workloads, model registries, and artifact stores Assess risks in third-party models, libraries, embeddings, and...Suggested
- ...Level Troubleshooting: Investigating and troubleshooting problems and hardware faults that our automation can't determine within our GPU platforms. This will involve taking data from system logs, kernel logs, BMC redfish APIs, and if the data is not there, working with...SuggestedLong term contractWork from home
- ...infra /Kubernetes (Remote) Are you passionate about building scalable AI infrastructure and helping customers succeed with cutting-edge GPU platforms? We're looking for a Solutions Architect to join our team and work with enterprise customers deploying and optimizing AI/ML...SuggestedRemote work
- ...RESPONSIBILITIES: Own the discovery and definition of customer requirements for AI infrastructure use cases, including training, inference, GPU clusters, bare metal, managed orchestration, networking, and storage Work directly with strategic customers to understand their...SuggestedHourly payContract workLocal area
- ...evaluate and guide the following areas: Future AI rack density and power consumption trends Impacts of next-generation GPU and AI chip architectures Optical networking and switching implications on infrastructure design AI workload impacts on utility...SuggestedWork at officeLocal areaWork visa
$175k - $220k
...on multi-functional teams to provide ethernet network expertise to server infrastructure builds, accelerated computing workloads and GPU enabled AI applications. Implementing tasks related to network configuration and validation for data centers. Create methods...SuggestedWorldwide- ...as we shape the future of AI and beyond. Together, we advance your career. THE ROLE: AMD is looking for a Enterprise AI/HPC GPU architect to join our Datacenter System Architecture and Engineering team to develop world-class products around Instinct GPUs. In this...Suggested
- ...runbooks Enforce IAM least-privilege policies, secrets management, and FinOps cost controls Collaborate with AI/ML engineers on GPU workloads, model serving, and inference pipelines Own technical communication during clinet interaction, translating...Suggested
- ...scalable and reliable services, and APIs, used by both internal teams and enterprise customers. Continuously optimize performance across GPU clusters, cloud infrastructure, and backend systems. Implement observability, reliability, and security best practices...SuggestedFull timeWork from homeFlexible hours1 day per week
$108.75k - $159.5k
...libraries (connectors, CMCs, ferrites, absorbers), and post-processing tools that emulate compliance receivers (RBW/VBW, detectors). HPC/GPU acceleration experience for EM workloads, and best practices for model version control, automation, and traceability What Makes...SuggestedFull timeTemporary workWork at officeImmediate startRemote workRelocationFlexible hours- ...similar language. Preferred Qualifications Experience supporting ML/AI workloads, including model serving infrastructure and GPU provisioning. GCP Professional DevOps Engineer or Cloud Architect certification. Familiarity with ArgoCD, Helm, or GitOps workflows...SuggestedFull timeContract work
$80.2k - $166.1k
...• Technical foundation that allows you to quickly absorb the complexities of cloud architecture, distributed systems, and emerging GPU/AI infrastructure. • Work across large organizations and bring people together around shared goals. • Embrace a growth mindset, learn...SuggestedTemporary workFlexible hours$95k - $171k
...with lab tools Have experience with standards, protocols and related technologies such as: Ethernet, TCP/IP, PCIe, NVMe, DDR and GPU architectures Have problem solving, documentation, and communication skills. This role will require your presence in our Cambridge...Work experience placementWork at office$126.2k - $264.1k
...area or are able to relocate permanently to the area. **** Preferred Qualifications Direct commissioning experience supporting GPU/high-density or liquid-cooled data halls (CDUs, leak detection, controls integration). Experience commissioning hyperscale or...Temporary workFor contractorsLive inLocal areaRelocationRelocation packageFlexible hours$193.6k - $414.4k
...planning and execution for the compute, networking, storage, data, inference/training at large scale, orchestration, observability, and GPU/HPC capacity required to support high-volume GenAI workloads. Drive disciplined capacity forecasting, utilization management, and...Temporary workFlexible hours$92.5k - $209.5k
...-agent coordination, policy enforcement, and evaluation. Develop distributed services optimized for low latency, high throughput, GPU efficiency, reliability, cost, operability, and secure multi-tenant operation. Define service boundaries, APIs, data models, state...Temporary workFlexible hours$74.1k - $148.3k
...schedules. • Ensure that all work complies with OCI specifications, manufacturer warranty standards, and regional regulations. GPU Liquid-Cooled Rack Megaprojects • Serve as the technical and delivery lead for GPU-intensive data hall builds, managing low-voltage...Temporary workLive inLocal areaWorldwideRelocationRelocation packageFlexible hours$102.3k - $209.5k
...construction audit practices, and formal quality management processes. Experience with AI infrastructure, high-density data halls, GPU deployments, liquid-cooled environments, or large-scale cloud infrastructure projects. Professional certifications such as CQM,...Temporary workFor contractorsFor subcontractorLive inRelocationRelocation packageFlexible hours

