Get new jobs by email
  • $132k - $191k

     ...Design and support cloud architectures for AI/ML workloads, including model training, inference, and high-performance compute (e.g., GPU/EDA burst capacity). Enable secure data pipelines, scalable compute environments, and integration of AI services while ensuring... 
    Suggested
    Permanent employment

    Openkyber

    Vermont
    3 days ago
  • $148k - $216k

     ...utilization, data drift, and concept drift. Infrastructure Management: Provision and optimize cloud-based ML infrastructure (including GPU/CPU computing clusters) utilizing Infrastructure as Code (IaC) paradigms. Cross-Functional Collaboration: Work intimately with... 
    Suggested
    Remote work
    Flexible hours

    Openkyber

    Vermont
    4 days ago
  •  ...data handling, regulatory expectations, and third-party data use Secure AI infrastructure and supply chain: Harden AI platforms, GPU and container workloads, model registries, and artifact stores Assess risks in third-party models, libraries, embeddings, and... 
    Suggested

    Openkyber

    Vermont
    1 day ago
  •  ...Experience with distributed storage systems and understanding of one or more of object, block, and file storage paradigms. Hardware and GPU troubleshooting experience (nice to have). Exposure to OVN/OVS-based networking stack (nice to have). Strong communication skills.... 
    Suggested
    Night shift
    Day shift

    Openkyber

    Vermont
    1 day ago
  •  ...Level Troubleshooting: Investigating and troubleshooting problems and hardware faults that our automation can't determine within our GPU platforms. This will involve taking data from system logs, kernel logs, BMC redfish APIs, and if the data is not there, working with... 
    Suggested
    Long term contract
    Work from home

    Openkyber

    Vermont
    1 day ago
  •  ...infra /Kubernetes (Remote) Are you passionate about building scalable AI infrastructure and helping customers succeed with cutting-edge GPU platforms? We're looking for a Solutions Architect to join our team and work with enterprise customers deploying and optimizing AI/ML... 
    Suggested
    Remote work

    Openkyber

    Vermont
    3 days ago
  •  ...evaluate and guide the following areas: Future AI rack density and power consumption trends Impacts of next-generation GPU and AI chip architectures Optical networking and switching implications on infrastructure design AI workload impacts on utility... 
    Suggested
    Work at office
    Local area
    Work visa

    Openkyber

    Vermont
    4 days ago
  •  ...as we shape the future of AI and beyond. Together, we advance your career. THE ROLE: AMD is looking for a Enterprise AI/HPC GPU architect to join our Datacenter System Architecture and Engineering team to develop world-class products around Instinct GPUs. In this... 
    Suggested

    Openkyber

    Vermont
    5 days ago
  •  ...RESPONSIBILITIES: Own the discovery and definition of customer requirements for AI infrastructure use cases, including training, inference, GPU clusters, bare metal, managed orchestration, networking, and storage Work directly with strategic customers to understand their... 
    Suggested
    Hourly pay
    Contract work
    Local area

    Openkyber

    Vermont
    4 days ago
  • $175k - $220k

     ...on multi-functional teams to provide ethernet network expertise to server infrastructure builds, accelerated computing workloads and GPU enabled AI applications. Implementing tasks related to network configuration and validation for data centers. Create methods... 
    Suggested
    Worldwide

    Openkyber

    Vermont
    5 days ago
  •  ...runbooks Enforce IAM least-privilege policies, secrets management, and FinOps cost controls Collaborate with AI/ML engineers on GPU workloads, model serving, and inference pipelines Own technical communication during clinet interaction, translating... 
    Suggested

    Openkyber

    Vermont
    8 days ago
  • $146.4k - $263.6k

     ...other SREs, influence architecture decisions with product engineering teams, and shape SRE practices for AI inference workloads and GPU infrastructure at scale. As a Senior II Site Reliability Engineer, you will be responsible for: Taking ownership of observability... 
    Suggested
    Work experience placement
    Work at office

    Akamai

    East Montpelier, VT
    4 days ago
  • $89.2k - $209.5k

     ...team is responsible for deliver trusted, fast health determinations and customer‑initiated diagnostics that reduce false positives for GPU clusters, prevent unnecessary node returns, increase capacity for customers, protect revenue, and improve uptime—by providing an OCI‑... 
    Suggested
    Temporary work
    Flexible hours

    Oracle

    East Montpelier, VT
    3 days ago
  • $74.1k - $148.3k

     ...schedules. • Ensure that all work complies with OCI specifications, manufacturer warranty standards, and regional regulations. GPU Liquid-Cooled Rack Megaprojects • Serve as the technical and delivery lead for GPU-intensive data hall builds, managing low-voltage... 
    Suggested
    Temporary work
    Live in
    Local area
    Worldwide
    Relocation
    Relocation package
    Flexible hours

    Oracle

    East Montpelier, VT
    23 hours ago