Get new jobs by email
$132k - $191k
...Design and support cloud architectures for AI/ML workloads, including model training, inference, and high-performance compute (e.g., GPU/EDA burst capacity). Enable secure data pipelines, scalable compute environments, and integration of AI services while ensuring...SuggestedPermanent employment$148k - $216k
...utilization, data drift, and concept drift. Infrastructure Management: Provision and optimize cloud-based ML infrastructure (including GPU/CPU computing clusters) utilizing Infrastructure as Code (IaC) paradigms. Cross-Functional Collaboration: Work intimately with...SuggestedRemote workFlexible hours- ...data handling, regulatory expectations, and third-party data use Secure AI infrastructure and supply chain: Harden AI platforms, GPU and container workloads, model registries, and artifact stores Assess risks in third-party models, libraries, embeddings, and...Suggested
- ...infra /Kubernetes (Remote) Are you passionate about building scalable AI infrastructure and helping customers succeed with cutting-edge GPU platforms? We're looking for a Solutions Architect to join our team and work with enterprise customers deploying and optimizing AI/ML...SuggestedRemote work
- ...Experience with distributed storage systems and understanding of one or more of object, block, and file storage paradigms. Hardware and GPU troubleshooting experience (nice to have). Exposure to OVN/OVS-based networking stack (nice to have). Strong communication skills....SuggestedNight shiftDay shift
- ...Level Troubleshooting: Investigating and troubleshooting problems and hardware faults that our automation can't determine within our GPU platforms. This will involve taking data from system logs, kernel logs, BMC redfish APIs, and if the data is not there, working with...SuggestedLong term contractWork from home
$175k - $220k
...on multi-functional teams to provide ethernet network expertise to server infrastructure builds, accelerated computing workloads and GPU enabled AI applications. Implementing tasks related to network configuration and validation for data centers. Create methods...SuggestedWorldwide- ...RESPONSIBILITIES: Own the discovery and definition of customer requirements for AI infrastructure use cases, including training, inference, GPU clusters, bare metal, managed orchestration, networking, and storage Work directly with strategic customers to understand their...SuggestedHourly payContract workLocal area
- ...as we shape the future of AI and beyond. Together, we advance your career. THE ROLE: AMD is looking for a Enterprise AI/HPC GPU architect to join our Datacenter System Architecture and Engineering team to develop world-class products around Instinct GPUs. In this...Suggested
- ...evaluate and guide the following areas: Future AI rack density and power consumption trends Impacts of next-generation GPU and AI chip architectures Optical networking and switching implications on infrastructure design AI workload impacts on utility...SuggestedWork at officeLocal areaWork visa
- ...runbooks Enforce IAM least-privilege policies, secrets management, and FinOps cost controls Collaborate with AI/ML engineers on GPU workloads, model serving, and inference pipelines Own technical communication during clinet interaction, translating...Suggested
- Embedded Software Design Engineer (Microcontrollers Development, Xilinx FPGAs, RS422, RS485, Ethernet Switches, GPU Processors, Oscilloscopes) in Johnson City, TN Developing Microcontrollers, Embedded Software Engineering, Ethernet switches, Linux, power supplies, Python...SuggestedFull timeRelocation
- ...PREFERRED SKILLS AND EXPERIENCE Direct background in AI or hyperscale data center operations, including liquid cooling systems, high‑power GPU/accelerator environments, and on‑site power generation. Experience building or scaling fiber infrastructure for low‑latency, high‑...SuggestedFull timeFor contractorsWork at office
$30 - $32.5 per hour
...(WIRING AND HARDWARE) HEAVY LIFTING REQUIRED Job Description Hands‑on experience with server hardware (CPU, memory, power supply, GPU replacements, server builds) Experience with server swaps and deployments Experience replacing server components (CPUs, memory, power...SuggestedLong term contractContract workShift work- ...costs and central location for data center operations. The city is growing as a logistics and technology hub with increasing demand for GPU computing resources. About Introl Introl stands apart as a leader in GPU infrastructure deployments, specializing in large-scale GPU...SuggestedHourly payLocal areaRelocationNight shiftWeekend work
- ...detection. Perform fire hazard analyses, computational fluid dynamics (CFD) fire modeling, and risk assessments for high-heat-flux GPU clusters, battery storage, and natural gas turbine installations. Develop fire protection drawings, P&IDs, specifications,...Full timeFor contractorsLocal area
- ...installation, configuration, and maintenance of both Windows/Linux servers, both on‑premises and in the cloud, specifically configuring AWS GPU/CPU clusters for high performance computing. Develop and maintain custom toolchains to enhance our operations applying modern...Contract workWork at officeLocal areaRemote workRelocation packageFlexible hours
$80.2k - $166.1k
...forward. Technical foundation that allows you to quickly absorb the complexities of cloud architecture, distributed systems, and emerging GPU/AI infrastructure. Work across large organizations and bring people together around shared goals. Embrace a growth mindset, learn...Temporary workFlexible hours- ...Repair and Replacement: Perform repairs and replacements of faulty components, including but not limited to: Server hardware (e.g., GPU, CPU, Motherboard) Storage systems (e.g., Hard Drives, SSDs) Network devices (e.g., Switches, Routers, Firewalls) Ticket Management...Local areaFlexible hoursShift workWeekday work
$92.5k - $209.5k
...multi‑agent coordination, policy enforcement, and evaluation. Develop distributed services optimized for low latency, high throughput, GPU efficiency, reliability, cost, operability, and secure multi‑tenant operation. Define service boundaries, APIs, data models, state...Temporary workFlexible hours$96.8k - $251.6k
...Description OCI (Oracle Cloud Infrastructure) AI Infrastructure is at the forefront of building a cutting‑edge, ultra‑high‑performance GPU platforms designed to support AI/ML/HPC workloads. This is your chance to be part of the AI revolution, creating systems that allow...Temporary workFlexible hours- ...technology services provider with 14 years of experience delivering innovative, business‑critical solutions. We specialize in AI and GPU deployments, comprehensive data center support, IMAC services - desktop relocation, ensuring organizations transition seamlessly into...Relocation
- ...Willing to work shifts and be on-call Technical Ability Hands-on experience with server hardware (CPU, memory, power supply, GPU replacements, server builds) Experience with server swaps and deployments Experience replacing server components (CPUs, memory,...Contract workFlexible hoursShift work
- ...communication, stakeholder management, and leadership skills. Preferred Qualifications Experience operating large-scale AI, HPC, or GPU-intensive infrastructure environments. Knowledge of data center monitoring systems, BMS, DCIM, and environmental management platforms...
- ...vendor-recommended troubleshooting procedures Perform firmware upgrades, operating system installations, and system updates Support GPU and network interface configuration and testing Maintain equipment records, rack elevations, serial numbers, and runbooks...Contract work
- ...AI inference or distributed systems software, particularly in accelerator-rich environments. Strong technical skills in areas such as GPU programming, kernel optimization, and runtime design are essential. The ideal candidate will have a proven track record of delivering...
- ...workstations. Specify, configure, and tune engineering-class workstations for CAD, large-assembly, and FEA workloads. Maintain certified GPU driver, BIOS, and firmware standards and build repeatable deployment images. Own performance and reliability. Establish baselines...
- ...in a Linux environment Experience with surrogate modeling Experience with data analytics techniques Familiarity with C++ and GPU programming Familiarity with Python programming Experience with modern software development practices to ensure code quality...Work at officeRelocation packageFlexible hours
- ...Basic Network experience • Experience reading Linux logs • Linux cli experience Nice to have skills: • RoCE network experience • Nvidia GPU/NIC experience is a plus New Scope -• Provide day-to-day data center operational support • Escort and coordinate with dispatched...Contract workWork experience placementLocal areaFlexible hoursWeekend work
- ...integration. Infrastructure Optimization: Partner with engineering leaders to optimize cloud setups and manage dedicated compute/GPU resource allocations for localized model training. Enablement & Change Management: Drive tool adoption by upskilling non-technical...
