Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Devops Infrastructure Engineer/GPU Infrastructure Engineer

Maxonic

Job Description

Job Description

Maxonic maintains a close and long-term relationship with our direct client. In support of their needs, we are looking for an  AI Devops Infrastructure Engineer/GPU Infrastructure Engineer.

Job Title:  AI Devops Infrastructure Engineer/GPU Infrastructure Engineer

Job Type:  Contract / Contract to Hire 

Job Location: San Jose, CA

Work Schedule:  Onsite

Pay Rate: $100 on W2 & $100-124 on c2c

Description:

We are looking for an AI Infrastructure Engineer to help build and operationalize infrastructure supporting AI development, experimentation, training, inference, and AI-enabled developer productivity. This role will focus on building scalable, highly resilient GPU-based infrastructure and enabling engineering teams to access AI resources through reliable, secure, and automated infrastructure services.

The ideal candidate has hands-on experience building and operating AI/ML infrastructure in real-world environments, with a strong foundation in infrastructure engineering, DevOps/SRE, platform engineering, or similar disciplines.

Responsibilities

  • Build and operate AI infrastructure supporting development, experimentation, training, and inference.
  • Design and manage infrastructure across GPU and accelerator environments, including NVIDIA and/or AMD.
  • Build scalable infrastructure for GPU capacity planning, utilization, forecasting, workload scheduling, resource allocation, and performance optimization.
  • Develop self-service AI infrastructure capabilities that allow engineering teams to request, provision, manage, extend, and release GPU, CPU, memory, and storage resources.
  • Build and maintain infrastructure across Kubernetes and containerized environments, including compute, networking, storage, and accelerator scheduling.
  • Automate infrastructure provisioning, configuration, scaling, patching, driver/firmware lifecycle management, and decommissioning.
  • Use Infrastructure-as-Code, APIs, automation, and scripting to improve infrastructure reliability and operational efficiency.
  • Establish monitoring, observability, alerting, capacity management, and operational processes for AI infrastructure.
  • Help define and implement best practices for scalable, feasible, and operationally sustainable AI infrastructure.
  • Integrate AI infrastructure with Developer Productivity tooling, CI/CD pipelines, source control, artifact management, build infrastructure, and developer tooling.
  • Partner with engineering, security, IT, and AI/ML teams to make AI resources accessible with minimal operational friction.
  • Troubleshoot complex infrastructure issues and perform root-cause analysis.
  • Help establish secure infrastructure standards covering network segmentation, access controls, authentication, authorization, secrets management, and governance.
  • Evaluate emerging AI infrastructure technologies and improve scalability, reliability, performance, automation, and developer experience.

Qualifications:

  • Strong experience in Infrastructure Engineering, Platform Engineering, DevOps, SRE, Cloud Engineering, or AI/ML Infrastructure.
  • Hands-on experience building and operating AI infrastructure, not just exposure to AI/ML concepts.
  • Strong experience with GPU infrastructure and GPU scalability.
  • Experience with Kubernetes and containerized production environments.
  • Strong understanding of compute, networking, storage, workload scheduling, and resource allocation.
  • Experience with GPU capacity planning, utilization, forecasting, workload scheduling, and performance optimization.
  • Experience building or operating self-service infrastructure/provisioning platforms.
  • Strong experience with Infrastructure-as-Code and automation, such as Terraform, Ansible, Python, APIs, or similar technologies.
  • Experience with CI/CD pipelines and infrastructure operations.
  • Experience with monitoring and observability tools such as Grafana, Prometheus, OpenTelemetry, Kibana, or similar platforms.
  • Experience working across bare-metal, private cloud, and/or public cloud environments.
  • Familiarity with AI/ML infrastructure, inference platforms, and distributed AI workloads.
  • Strong troubleshooting, root-cause analysis, communication, and cross-functional collaboration skills.

About Maxonic:

Since 2002 Maxonic has been at the forefront of connecting candidate strengths to client challenges. Our award winning, dedicated team of recruiting professionals are specialized by technology, are great listeners, and will seek to find a position that meets the long-term career needs of our candidates. We take pride in the over 10,000 candidates that we have placed, and the repeat business that we earn from our satisfied clients.

Interested in Applying?                    

Please apply with your most current resume. Feel free to contact Pramod Kumar for more details.

\nCompany Description

Maxonic, Inc. is a certified minority-owned technology consulting and staffing firm founded in 2002. We provide technology consulting and staffing services to Fortune 500 companies across the country. We collaborate with clients to design, develop and implement business-driven technology solutions. We have deep domain expertise with multiple big data platforms and have strong partnerships with industry leaders in the space. The founders have over 40 years of experience helping clients and consultants succeed. Our commitment is to provide customers with an honest, personalized, face-to-face, long-term relationship and provide consultants with ethical, respectful, and stress-free care. We believe in the power of building successful careers for everyone.

Specialties:
Staff Augmentation
Professional Services - Artificial Intelligence, Big Data Analytics, Mobile Solutions, and Cloud Solutions
Outsourced Services, Administration and Support
Contingent Workforce Services

Company Description

Maxonic, Inc. is a certified minority-owned technology consulting and staffing firm founded in 2002. We provide technology consulting and staffing services to Fortune 500 companies across the country. We collaborate with clients to design, develop and implement business-driven technology solutions. We have deep domain expertise with multiple big data platforms and have strong partnerships with industry leaders in the space. The founders have over 40 years of experience helping clients and consultants succeed. Our commitment is to provide customers with an honest, personalized, face-to-face, long-term relationship and provide consultants with ethical, respectful, and stress-free care. We believe in the power of building successful careers for everyone. \r\n\r\nSpecialties: \r\nStaff Augmentation \r\nProfessional Services - Artificial Intelligence, Big Data Analytics, Mobile Solutions, and Cloud Solutions\r\nOutsourced Services, Administration and Support \r\nContingent Workforce Services

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI Devops Infrastructure Engineer/GPU Infrastructure Engineer in San Jose, CA vacancy
  •  ...scientific discovery to powering AI and the technologies...  ...a hands-on Platform Engineer to build and operate the infrastructure foundations that enable engineering...  ...closely with software, DevOps, security, and AI/ML...  ...including model-serving, GPU-enabled compute, data-... 
    Suggested

    AMD

    San Jose, CA
    4 days ago
  • $184k - $287.5k

    We are seeking a Senior DevOps / Cloud Simulation Infrastructure Engineer to own the complete end-to-end cloud execution pipeline for...  ...for deploying a robust, multi-GPU pipeline that supports structural validation, AI-driven runtime behavioral testing, and automated... 
    Suggested
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    4 days ago
  • $224k - $356.5k

     ...into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of...  ...recipes, and model bring-up infrastructure — that lets developers run groundbreaking...  ...Computer Science, Computer Engineering, Electrical Engineering, or... 
    Suggested
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    2 days ago
  • $137k - $156k

     ...talented, passionate, and committed engineers, technologists, and business...  ...Senior Systems Engineer / GPU Platforms to support the...  ...-GPU server systems used for AI, HPC, enterprise computing, and...  ...technical enablement, HPC, AI infrastructure, or a related field.Strong... 
    Suggested
    Worldwide

    Super Micro Computer

    San Jose, CA
    2 days ago
  • Job Title: GPU Network Engineer Job Location: Sunnyvale, CA (hybrid)Job Salary: 200k-250k + BenefitsRequirements...  ..., CA, we are an innovative energy infrastructure company that develops cutting-edge...  ..., with direct experience in GPU or AI cluster environments.Hands-on... 
    Suggested
    Local area
    Relocation

    CyberCoders

    San Jose, CA
    4 days ago
  •  ...scientific discovery to powering AI and the technologies people...  ...are seeking a Senior Network Engineer to join the AMD IT System...  ...networks supporting large-scale AMD GPU clusters. The engineer will...  ...operating backend network infrastructure for GPU clusters with approximately... 

    AMD

    San Jose, CA
    3 days ago
  • $186k - $282k

     ...Why this role exists FloQast's AI products have outgrown the infrastructure patterns the rest of the platform runs on. Transform, AI Matching,...  ...ways a 500-rate dashboard never catches. Today, DevOps engineers carry this work alongside the wider fleet. We are making... 

    FloQast

    San Jose, CA
    1 day ago
  •  ...generation computing experiences—from AI and data centers, to PCs,...  ....As a Technical Marketing Engineer (TME) within the Software...  ...organization for AMD’s Data Center GPU Business Unit, you will play...  ...role focused on datacenter infrastructure or AI platforms.Strong... 

    AMD

    Santa Clara, CA
    14 hours ago
  • $151k - $226.6k

     ...ResponsibilitiesAs a Staff GPU Front-End Implementation Methodology Engineer, you will shape and...  ...support for an agentic AI workflow, and contribute...  ...issuesSolid background in DevOps practices, build automation...  ...implementation, and the infrastructure needed to deliver... 
    Hourly pay
    Full time
    Worldwide
    Relocation

    Samsung Semiconductor

    San Jose, CA
    14 hours ago
  • $136k - $218.5k

     ...re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and...  ...seeking a Senior ASIC Compose Verification Infrastructure and Tools Engineer to advance the systems that qualify,... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...world-leading technology company for AI and Bitcoin mining infrastructure. Bitdeer is committed to...  ...skilled and motivated Cloud Senior DevOps Engineer to join our AI Cloud team. In this...  ...specialized computing resources (e.g., GPU clusters) to support high-performance... 
    Remote job
    Full time
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    24 days ago
  • $152k - $241.5k

     ...Computing and Visualization. The GPU, our invention, serves as the...  ...for a motivated Performance engineer to influence the roadmap of...  ...data; build tools and infrastructure to visualize and analyze the...  ...existing vacancy. NVIDIA uses AI tools in its recruiting processes... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...NVIDIA, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self...  ...searching for outstanding senior system software engineer to join the NVIDIA's GPU Diagnostics SW team. Our... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    NVIDIA is searching for a highly motivated, creative engineer to join the GPU Software team. As a GPU/SOC system software engineer, you will work...  ..., 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $152k - $241.5k

     ...Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self...  ...is searching for a creative and highly motivated engineer with expertise in boot software to join the SOC... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    NVIDIA is seeking a Senior System SW Engineer with expertise in systems software, GPU drivers, embedded software, and platform security. You will provide technical...  ....This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of...  ...We are looking for a dedicated engineer for the Senior Systems Software...  ...practices in large-scale GPU infrastructure, delivering powerful tools, methodologies... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $224k - $356.5k

     ...our group of highly skilled and motivated engineers who bring GeForce NOW to life! As a member...  ...performance critical software, tools, CPU/GPU scheduler.Passionate about software design...  ...posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $224k - $356.5k

     ...searching for highly motivated, creative engineers to join the Platform Software team. You will...  ...Computing and Visualization. The GPU, our invention, serves as the visual cortex...  ...posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

    Senior Systems Software Engineer, Observability and Telemetry Platform...  ...internal and external facing GPU cloud services run maximum...  ...experience5+ years of experience with Infrastructure automation, distributed...  ...vacancy. NVIDIA uses AI tools in its recruiting processes... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $272k - $431.25k

     ...highly motivated Principal System Software Engineer to drive next-generation innovations in...  ...closely with hardware, architecture, kernel, AI, middleware, and platform teams to deliver...  ...optimization initiatives across CPU, GPU, memory, storage, networking, and platform... 
    Full time

    Nvidia

    Santa Clara, CA
    14 hours ago
  • $109k - $160k

     ...CoreWeave is The Essential Cloud for AI™. Built for pioneers by...  ...CoreWeave combines superior infrastructure performance with deep...  ...About the role A Software Engineer contributes to the design, implementation...  ...teams to evolve our GPU performance testing platform... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    Core Weave

    Sunnyvale, CA
    more than 2 months ago
  • $152k - $287.5k

     ...that you helped to craft as a member of the GPU Foundations Developer Tools team! Innovate...  ...datacenter scale. As a system software engineer in the Developer Tools group, you will be...  ...is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA... 
    Full time

    NVIDIA

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...NVIDIA, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self...  ...impact on the world.As a Developer Technology Engineer, you will be at the forefront of innovation,... 
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    2 days ago
  • $224k - $356.5k

     ...looking for a Senior System Software Engineer. As part of our team, you will work...  ..., embedded software, and cloud infrastructure.What you'll be doing:Drive AI-based automotive systems from design...  ...applications.Familiarity with CPU/GPU architectures, memory subsystems, and... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $152k - $241.5k

     ...performance computing platforms are powering the AI revolution across many applications and...  ...abstractions for Tensor Core and related GPU hardware features in MLIR, Python, and...  ...PhD degree in Computer Science, Computer Engineering, or related field (or equivalent... 
    Full time

    Nvidia

    Santa Clara, CA
    14 hours ago
  • $152k - $241.5k

    NVIDIA's invention of the GPU 1999 sparked the growth of the PC gaming market, redefined...  ...recently, GPU deep learning ignited modern AI — the next era of computing — with the GPU...  ...for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for its... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    NVIDIA’s invention of the GPU in 1999 fueled the growth of the PC gaming market, redefined...  ...Today, we are increasingly known as “the AI computing company.” We're looking to grow...  ...We are seeking an excellent Senior System Engineer to work on bring up, integration, validation... 
    Full time
    Early shift

    Nvidia

    Santa Clara, CA
    14 hours ago
  • $184k - $287.5k

     ...Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self...  ...are looking for an experienced Senior Software Engineer for the Embedded Platform team. This is an... 
    Full time

    Nvidia

    Santa Clara, CA
    14 hours ago
  •  ...performance computing, cloud, and AI. Whether you’re designing...  ...across future CPU, GPU, AI, and adaptive computing platforms...  ...Platform Functional Modeling Engineer, you will play a key role in...  ...and delivering the framework, infrastructure, and methodologies that power... 

    AMD

    San Jose, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Devops Infrastructure Engineer/GPU Infrastructure Engineer. Be the first to apply!