AI & HPC Infrastructure Engineer
$80.4k - $266.3kAccenture
We Are:
The Global AI Infrastructure team is at the center of enabling infrastructure reinvention for the next era of digital solutions powered by AI, accelerated computing, and high-performance workloads. We bring together deep technical expertise across cloud, on-premises, and hybrid environments to design, build, and operate advanced infrastructure that powers AI platforms, GPU-accelerated workloads, large-scale models, simulations, and emerging agentic AI solutions at scale. Our solutions enable some of our most strategic and mission-critical clients to unlock new levels of performance, efficiency, governance, and innovation. Our remit spans the full lifecycle-from strategy and architecture through implementation and operations-driving modernization across the entire infrastructure stack. We collaborate across the ecosystem to harness emerging technologies, fuel growth, and transform industries. In this rapidly growing market, our team is leading the way in shaping how enterprises leverage AI infrastructure to drive breakthrough innovation and reimagine what is possible.
Key Responsibilities:
Design and implement AI infrastructure and accelerated computing solutions, aligning system architecture and deployment roadmaps to industry-specific performance, scalability, resiliency, and governance needs
Deploy, configure, and manage XPU-based clusters (GPU, DPU, LPU, CPU) across bare-metal and containerized environments using workload schedulers (Slurm, Run:ai), Kubernetes orchestration, and container platforms to deliver scalable AI infrastructure services including Bare-Metal-aaS, GPUaaS, AIaaS, Token-aaS, model serving, and agentic AI frameworks
Integrate AI infrastructure platforms with existing IT systems, data pipelines, security frameworks, model-serving endpoints, and enterprise governance controls
Design and implement agentic AI infrastructure by integrating platform services, model endpoints, tool and function calling, retrieval patterns, and workflow orchestration with observability, identity, and policy controls through secure, deterministic APIs to support governed enterprise use cases
Build and integrate MCP servers, tools, connectors, and adapters that allows agents to monitor, troubleshoot, and tune infrastructure to ensure high availability, low-latency networking, and workload resiliency
Architect and deploy with NVIDIA platform tools including Base Command Manager (BCM), NGC, NCCL, NVLink, and CUDA along with LLM inference engines (TensorRT-LLM), production serving frameworks (vLLM, SGLang), inference orchestration (Triton Inference Server, NVIDIA Dynamo, llm-d), and GPU benchmarking and validation tools (MLPerf, NCCL tests, fio, iperf) to deploy, tune, profile, and validate AI cluster performance across compute and networking layers including multi-node training and inference workloads
Develop and maintain documentation including architecture diagrams, configuration baselines, and operational runbooks
Provide technical guidance, troubleshooting, and optimization across AI workloads including large-scale training, inference, multi-node simulations, and agentic pipelines while leveraging digital twins to validate infrastructure and drive performance, scalability, energy efficiency, and token cost optimization
Travel may be required for this role. The amount of travel will vary from 25% to 100% depending on business need and client requirements.
Required Skills and Qualifications:
Minimum of 5+ years of experience designing, deploying, and managing AI infrastructure and accelerated computing environments across on-premises, cloud, and hybrid environments for hyperscaler, neocloud, large enterprise, Telco/Mobile, Financial Services, Life Sciences, Manufacturing, and/or Retail clients.
Minimum of 5+ years of hands-on experience with accelerated computing platforms, including GPUs, DPUs, LPUs, CPUs, high-speed interconnects such as InfiniBand or Ethernet, data center networking such as SONiC, and AI storage architectures including NVMe, NVMe-oF, parallel file systems, VAST, Weka, or DDN.
Minimum of 5+ years of experience with cluster management, workload scheduling, orchestration, observability, and infrastructure automation using platforms and tools such as Kubernetes, Slurm, Run:ai, AWS, Azure, GCP, VMware, Nutanix, Python, Terraform, and Ansible.
Bachelor's degree or equivalent (minimum 12 years) work experience. If Associate's Degree, must have minimum 6 years work experience.
Preferred Skills and Qualifications:
2+ years of experience implementing MLOps, LLMOps, agentic AI, and DevSecOps frameworks to enable secure, automated, governed, and reproducible AI workflows.
2+ years of experience developing APIs, integration services, automation workflows, or platform services using Python and modern API patterns such as REST, OpenAPI, JSON/YAML schemas, webhooks, and event-driven integrations.
Experience designing and implementing agentic AI infrastructure, including LLM inference, tool/function calling, retrieval-augmented generation (RAG), agent orchestration, secure API integration, policy-based governance, and deterministic platform APIs.
Experience building and integrating MCP servers, tools, connectors, and adapters that allow agents to monitor, troubleshoot, and tune infrastructure for high availability, low-latency networking, workload resiliency, and intelligent observability.
Experience using NVIDIA platform tools including Base Command Manager (BCM), NGC, NCCL, NVLink, CUDA, TensorRT-LLM, Triton Inference Server, NVIDIA Dynamo, llm-d, vLLM, SGLang, MLPerf, NCCL tests, fio, and iperf to deploy, tune, profile, and validate AI cluster performance.
Experience managing the deployment of 1,000+ GPU clusters for AI, HPC, and agentic AI workloads with infrastructure services enabled.
Design and build experience in AI Cloud platforms from CoreWeave, Nebius, and other specialty cloud providers.
Knowledge of machine learning and AI frameworks such as TensorFlow, PyTorch, JAX, Jupyter notebooks, and Google Colab environments.
Industry certifications in NVIDIA infrastructure, public cloud providers, data science, infrastructure automation, networking, or security are a plus.
Compensation at Accenture varies depending on a wide array of factors, which may include but are not limited to the specific office location, role, skill set, and level of experience. As required by local law, Accenture provides a reasonable range of compensation for roles that may be hired as set forth below.
We anticipate this job posting will be posted until 10/15/2026.
U.S. Employee Benefits | Accenture
Role Location Annual Salary Range
California $94,400 to $266,300
Cleveland $87,400 to $213,000
Colorado $94,400 to $230,000
District of Columbia $100,500 to $245,000
Illinois $87,400 to $230,000
Maine $80,400 to $196,000
Maryland $94,400 to $230,000
Massachusetts $94,400 to $245,000
Minnesota $94,400 to $230,000
New York $87,400 to $266,300
New Jersey $100,500 to $266,300
Virginia $87,400 to $245,000
Washington $100,500 to $245,000
About Accenture
Accenture is a leading global professional services company that helps the world’s leading businesses, governments and other organizations build their digital core, optimize their operations, accelerate revenue growth and enhance citizen services—creating tangible value at speed and scale. We are a talent- and innovation-led company with approximately 791,000 people serving clients in more than 120 countries. Technology is at the core of change today, and we are one of the world’s leaders in helping drive that change, with strong ecosystem relationships. We combine our strength in technology and leadership in cloud, data and AI with unmatched industry experience, functional expertise and global delivery capability. Our broad range of services, solutions and assets across Strategy & Consulting, Technology, Operations, Industry X and Song, together with our culture of shared success and commitment to creating 360° value, enable us to help our clients reinvent and build trusted, lasting relationships. We measure our success by the 360° value we create for our clients, each other, our shareholders, partners and communities.Visit us at
What We Believe
We have an unwavering commitment to diversity with the aim that every one of our people has a full sense of belonging within our organization. As a business imperative, every person at Accenture has the responsibility to create and sustain an inclusive environment.
Inclusion and diversity are fundamental to our culture and core values. Our rich diversity makes us more innovative and more creative, which helps us better serve our clients and our communities. Read more here
Requesting An Accommodation
Accenture is committed to providing equal employment opportunities for persons with disabilities or religious observances, including reasonable accommodation when needed. If you are hired by Accenture and require accommodation to perform the essential functions of your role, you will be asked to participate in our reasonable accommodation process. Accommodations made to facilitate the recruiting process are not a guarantee of future or continued accommodations once hired.
If you would like to be considered for employment opportunities with Accenture and have accommodation needs such as for a disability or religious observance, please call us toll free at View phone number on aiapply.co or send us an email or speak with your recruiter.
Equal Employment Opportunity Statement
We believe that no one should be discriminated against because of their differences. All employment decisions shall be made without regard to age, race, creed, color, religion, sex, national origin, ancestry, disability status, military veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by applicable law. Our rich diversity makes us more innovative, more competitive, and more creative, which helps us better serve our clients and our communities.
For details, view a copy of the Accenture Equal Opportunity Statement
Accenture is an EEO and Affirmative Action Employer of Veterans/Individuals with Disabilities.
Accenture is committed to providing veteran employment opportunities to our service men and women.
Other Employment Statements
Applicants for employment in the US must have work authorization that does not now or in the future require sponsorship of a visa for employment authorization in the United States.
Candidates who are currently employed by a client of Accenture or an affiliated Accenture business may not be eligible for consideration.
Job candidates will not be obligated to disclose sealed or expunged records of conviction or arrest as part of the hiring process. Further, at Accenture a criminal conviction history is not an absolute bar to employment.
The Company will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant. Additionally, employees who have access to the compensation information of other employees or applicants as a part of their essential job functions cannot disclose the pay of other employees or applicants to individuals who do not otherwise have access to compensation information, unless the disclosure is (a) in response to a formal complaint or charge, (b) in furtherance of an investigation, proceeding, hearing, or action, including an investigation conducted by the employer, or (c) consistent with the Company's legal duty to furnish information.
California requires additional notifications for applicants and employees. If you are a California resident, live in or plan to work from Los Angeles County upon being hired for this position, please click here for additional important information.
Please read Accenture’s Recruiting and Hiring Statement for more information on how we process your data during the Recruiting and Hiring process.
- Crusoe Cloud seeks a Senior Cloud Support Engineer to empower customers using sustainable GPU compute power. You’ll be the main technical... ...reliability. The role emphasizes customer success, Kubernetes and HPC knowledge, and collaboration on onboarding materials and SOPs. A...Suggested
$193.3k - $261.5k
We are seeking an experienced engineer to work on distributed AI/ML systems. This role involves working on... ...experience with high-speed networking or HPC interconnects is valued highly.If... ...critical building blocks for EC2 infrastructure. Every instance in EC2 is running some...SuggestedInternshipLocal areaFlexible hours- ...Systems builds the world's largest AI chip, 56 times larger than... ...to join our Cluster Engineering Team and help shape the front... ...network fabrics for AI/ML and HPC clusters, optimizing for high... ...configuration, and validation of network infrastructure using Python, including...Suggested
$255k - $340k
...Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers... ...day is currently Tuesday.Hardware Engineering at Lambda is responsible for building... ...lead for integrating OEM and white-label HPC AI/ML, general purpose compute, storage...SuggestedWork at officeLocal areaWork from homeFlexible hours$163k - $237k
...High Performance Computing (HPC) products in advanced process... ...Bachelor's degree in Electrical Engineering, Computer Engineering,... ...work to shape the future of AI/ML hardware acceleration. You... ...driven systems. As an SoC Test Infrastructure Engineer, you will drive the...SuggestedWorldwide$155.42k - $205.9k
...DescriptionAbout the Team: The AI Validation Platform team owns... ...We’re proud to serve as the infrastructure platform for teams developing... ...a Senior ML Infrastructure engineer to help build and scale robust... ...high performance computing (HPC). Experience working with or...Full timeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours- A leading technology cloud provider in Sunnyvale, California, is looking for a Systems Kernel Engineer to enhance their Linux-based infrastructure. The ideal candidate will have over 5 years of experience in kernel engineering and a solid understanding of systems-level...
$161k - $264.5k
...vehicles; Woven City, a test course for mobility; and Cloud & AI, the digital infrastructure powering our collaborative foundation. Business-critical... ...developers. That's why the Enterprise Technology Engineering Team (EnTec) builds solutions that enhance productivity,...Temporary workFor contractorsWork at office- black.ai is looking for a skilled platform engineer in Palo Alto to enhance our AWS infrastructure and support quantum simulations. This role requires strong experience in platform engineering, DevOps practices, and GPU workloads. As a platform engineer, you will improve...
$168k - $270.25k
...into the unlimited potential of AI to define the next era of... ...Experience (NVEX) Solutions Engineering team is looking for a senior... ...that link GPUs and AI compute infrastructure. Candidates must have a software... ...with AI infrastructure and HPC networkingExperience programming...Full timeWeekend work- Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a tedious... ...work. Today, ServiceNow is the AI control tower for business reinvention.... ...solutions to a composable, agent-native infrastructure foundation that agents and applications...Full timeWork at officeImmediate startRemote workFlexible hoursShift work
- ...for the testing and evaluation of current and next-generation HPE HPC products. Ensure development issues are resolved in a cost-... ...remote labs. Drive appropriate automated test execution to test engineers at various global locations. Provide training and guidance to...Local areaRemote work
$225k - $275k
...intelligence . As the only vertically integrated AI infrastructure company built from the ground up, we... ...a Senior Staff Network Deployment Engineer to serve as the technical owner of how... ...footprint of high-performance compute (HPC) and GPU-based AI infrastructure, you...Temporary workRemote work$150k - $170k
...most advanced supercomputers or AI systems will never reach.... ...cooling and standard fiber-optic infrastructure. In 2024, PsiQuantum... ...OverviewPsiQuantum's Applications Software Engineering Team (ASET) builds tools for... ..., including to leverage HPC/GPU resourcesMake deployments...Full timeShift work$140k - $210k
...researchers and veteran systems engineers who share a vision for... ...foundations of distributed computing. As AI workloads grow increasingly complex, traditional infrastructure struggles to meet the demands... ...Experience supporting AI/ML, HPC, storage, or GPU cluster infrastructure...- ...Woven by Toyota is seeking a Head of Infrastructure Engineering in our Palo Alto office to lead a global team responsible for corporate infrastructure systems including network, cloud, productivity and AI tools. You will cultivate an internationally distributed team...Work at office
$262k - $364k
...Lead and coach a distributed engineering team, fostering innovation while... ...stack to accelerate HPC, RDMA, and ML workloads.Strategic... ...bridging RDMA, storage, and AI/ML, staying ahead of training... ...Systems or Machine Learning Infrastructure.Google's software engineers develop...Remote workWorldwide$150k - $217k
...recommending appropriate solutions in partnership with other engineering teams.Lead and improve the entire life-cycle of a network... ...with our suite of applications, products and services.The AI and Infrastructure team is redefining what’s possible. We empower Google customers...Worldwide- Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry... ....About The RoleWe are seeking a highly skilled WAN Network Engineer to design, implement, manage, and optimize global connectivity....
$200k - $220k
...Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers... ...talented, passionate, and committed engineers, technologists, and business... ...is seeking an experienced AI Network Software Solution... ...development of next-generation network infrastructure solutions optimized for AI...Worldwide- ...Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers... ...and on-call rotation for Network Engineering teamYouHave 10+ years of experience in... ...like Terraform/Ansible/SaltHands-on with HPC/AI networking: RoCEv2 and/or InfiniBand...Work at officeLocal areaWork from homeFlexible hours
$262k - $365k
...qualifications:Master’s degree or PhD in Engineering, Computer Science, or a related... ...in High-Performance Computing (HPC) systems and networking.... ...innovations that power Google's AI and High-Performance Computing (HPC) infrastructure. Your focus will be on the software...Worldwide$300k
...research, nurture the next generation of AI builders, and drive transformative... ...class researchers, data scientists, and engineers, tackling the most fundamental and impactful... ...We're looking for a distributed ML infrastructure engineer to help extend and scale our training...Full timeFlexible hours$150k - $185k
...years of experience in Platform Engineering, DevOps, or SRE roles. We... ..., and working through real infrastructure complexity. We expect practical... ...collaboration around HPC and GPU resources. You will... ...Security Terraform AI GitHub GitLab Support...Full time- ...of a high-performing team that delivers infrastructure and performance excellence. Your role will... ...companies.As a Lead Infrastructure Engineer at JPMorganChase within the Cloud Foundational... ...actionsUses enterprise-authorized AI capabilities within the work environment...Permanent employment
$153.2k - $234.1k
...about accelerating the future of autonomous driving? Join the Embodied AI Infra Foundation team at General Motors, where we build the critical infrastructure that powers every machine learning engineer working on our cutting-edge Autonomous Driving models. From...Full timeWork at officeLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours$153.2k - $234.1k
...of autonomous driving? Join the Embodied AI team at General Motors. Our team is... ...real-world scenarios. As a Senior ML Infra Engineer, you will work on the core systems that... ...distributed systems, applications, or ML infrastructure. Experience designing robust services or...Full timeLocal areaRemote workWork from homeRelocation packageFlexible hours$207k - $300k
Lead a team of software engineers to design, build, deploy, and in some cases, operate the... ...manage the global data center networking infrastructure.Cultivate a high-performance culture of... ...resilient disaster recovery protocols.The AI and Infrastructure team is redefining...Worldwide$120k - $250k
What MatX Is BuildingMatX's mission is to make the world’s best AI models run as efficiently as allowed by physics, bringing the... ...availability.About the RoleThis is a senior-level, highly specialized engineering role focused on the design, qualification, and productization...Full timeWork experience placementLocal areaRemote workMonday to FridayFlexible hours$124k - $250k
...in the LifeAs a member of our software engineering infra team, you'll solve technical challenges... ...implementing state-of-the-art software infrastructure. The team builds a high-performance,... ...tools, including artificial intelligence (AI), to help identify and evaluate...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI & HPC Infrastructure Engineer. Be the first to apply!
- ai engineer Mountain View, CA
- machine learning ai engineer Mountain View, CA
- senior ai engineer Mountain View, CA
- ai ml engineer Mountain View, CA
- ai prompt engineer Mountain View, CA
- ai developer Mountain View, CA
- principal infrastructure engineer Mountain View, CA
- remote infrastructure engineer Mountain View, CA
- infrastructure engineering manager Mountain View, CA
- infrastructure developer Mountain View, CA




