AI Infrastructure Operations Engineer
Accenture
Accenture is a global professional services company with leading capabilities in digital, cloud and security. Combining unmatched experience and specialized skills across more than 40 industries, we offer Strategy and, Interactive, Technology, and Operations services, all powered by the world’s largest network of Advanced Technology and Intelligent Operations centers. Our 738,000 people deliver on the promise of technology and human ingenuity every day, serving clients in more than 120 countries. We embrace the power of change to create value and shared success for our clients, people, shareholders, partners, and communities. Visit us at The Global AI Infrastructure team enables resilient, high-performance compute environments for strategic clients across cloud, on-premises, and hybrid deployments. We design, build, and operate large-scale GPU and accelerated-computing infrastructure that supports demanding AI training and inference, simulation, and high-performance compute workloads. Our work spans strategy, architecture, modernization, operations, governance, and continuous improvement across the infrastructure stack. We build reusable operational tools, automation workflows, and platform capabilities that make repeatable infrastructure tasks safer, faster, and more scalable. We collaborate across the technology ecosystem to harness new capabilities, drive business transformation, and deliver dependable services at scale. Key Responsibilities: Design and implement accelerated-computing infrastructure solutions aligned to system architecture, deployment roadmaps, performance, scalability, resiliency, and governance requirements. Deploy, configure, and operate GPU-based clusters across bare-metal and containerized environments, using workload schedulers and Kubernetes orchestration to support AI training, inference, and high-performance compute workloads. Integrate infrastructure platforms with enterprise systems, data platforms, security frameworks, service-management processes, and governance controls. Design, build, and maintain reusable tools, scripts, self-service capabilities, and automation workflows for infrastructure operations, including provisioning, configuration management, validation, capacity planning, monitoring, incident management, reporting, and recurring remediation. Establish repeatable operational processes for cluster provisioning, configuration management, patching, capacity planning, monitoring, incident response, and lifecycle management. Perform and automate GPU, compute, storage, and network benchmarking and validation; diagnose performance issues across multi-node AI training, inference, and distributed compute workloads. Develop and maintain architecture diagrams, configuration baselines, operational runbooks, and support documentation. Provide technical guidance, troubleshooting, and optimization for GPU clusters supporting AI training, inference, high-performance computing, and multi-node simulation workloads, with emphasis on availability, resiliency, scalability, energy efficiency, and cost management. Travel may be required for this role. The amount of travel will vary from 25% to 60% depending on business need and client requirements. Required Skills and Qualifications: Minimum of 5+ years of experience designing, deploying, and managing accelerated-computing infrastructure across on-premises, cloud, and hybrid environments for hyperscaler, neocloud, large enterprise, telecommunications, financial services, manufacturing, and/or retail clients. Minimum of 5+ years of hands-on experience with accelerated-computing platforms, including GPUs, DPUs, and CPUs, high-bandwidth network fabrics, and AI based storage architectures such as parallel file systems, NVMe-oF, etc. Minimum of 5+ years of experience with cluster management, workload scheduling, orchestration, observability, and infrastructure automation, including building operational tools and automation workflows with platforms such as Kubernetes, Slurm, Run:ai Minimum 6 months hands-on experience with Claude Code, AI automation tools, Terraform, Ansible, Python, and Bash scripting. Bachelor's degree or equivalent (minimum 12 years) work experience. (If Associate’s Degree, must have minimum 6 years work experience) Preferred Skills and Qualifications: Experience building AI infrastructure automation and operations tools, AgenticOps practices that enable secure, automated, governed, and reproducible platform operations. Experience developing reusable infrastructure code leveraging Python, platform services, and automation workflows using REST APIs, OpenAPI, JSON/YAML schemas, webhooks, and event-driven integrations. Experience operating large-scale GPU clusters, including capacity management, reliability engineering, change management, and performance validation for AI training, inference, HPC, and enterprise compute workloads. Experience using NVIDIA platform tools and libraries including Base Command Manager (BCM), NGC, NCCL, CUDA-X, NVAIE, Dynamo, benchmarking tools to deploy, tune, profile, and validate cluster performance for training and inference workloads. Experience managing deployments of 1,000+ GPU clusters with infrastructure services enabled for AI training, inference, high-performance, and enterprise compute environments. Design and build experience in running LLMs across AI Cloud platforms from CoreWeave, Nebius, and other specialty providers Knowledge of model deployments, tuning, and troubleshooting performance on AI Infrastructure Industry certifications in accelerated-computing infrastructure, public cloud providers, infrastructure automation, networking, or security are a plus. Compensation at Accenture varies depending on a wide array of factors, which may include but are not limited to the specific office location, role, skill set, and level of experience. As required by local law, Accenture provides a reasonable range of compensation for roles that may be hired as set forth below.We anticipate this job posting will be posted until 10/13/2026.Accenture offers a market competitive suite of benefits including medical, dental, vision, life, and long-term disability coverage, a 401(k) plan, bonus opportunities, paid holidays, and paid time off. See more information on our benefits here: U.S. Employee Benefits | Accenture ( Role Location Annual Salary RangeCalifornia $94,400 to $266,300Cleveland $87,400 to $213,000Colorado $94,400 to $230,000District of Columbia $100,500 to $245,000Illinois $87,400 to $230,000Maine $80,400 to $196,000Maryland $94,400 to $230,000Massachusetts $94,400 to $245,000Minnesota $94,400 to $230,000New York $87,400 to $266,300New Jersey $100,500 to $266,300Virginia $87,400 to $245,000Washington $100,500 to $245,000 Requesting an Accommodation Accenture is committed to providing equal employment opportunities for persons with disabilities or religious observances, including reasonable accommodation when needed. If you are hired by Accenture and require accommodation to perform the essential functions of your role, you will be asked to participate in our reasonable accommodation process. Accommodations made to facilitate the recruiting process are not a guarantee of future or continued accommodations once hired. If you would like to be considered for employment opportunities with Accenture and have accommodation needs such as for a disability or religious observance, please call us toll free at View phone number on click.appcast.io or send us an email or speak with your recruiter. Equal Employment Opportunity Statement We believe that no one should be discriminated against because of their differences. All employment decisions shall be made without regard to age, race, creed, color, religion, sex, national origin, ancestry, disability status, veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by federal, state, or local law. Our rich diversity makes us more innovative, more competitive, and more creative, which helps us better serve our clients and our communities. For details, view a copy of the Accenture Equal Opportunity Statement ( Accenture is an EEO and Affirmative Action Employer of Veterans/Individuals with Disabilities. Accenture is committed to providing veteran employment opportunities to our service men and women. Other Employment Statements Applicants for employment in the US must have work authorization that does not now or in the future require sponsorship of a visa for employment authorization in the United States. Candidates who are currently employed by a client of Accenture or an affiliated Accenture business may not be eligible for consideration. Job candidates will not be obligated to disclose sealed or expunged records of conviction or arrest as part of the hiring process. Further, at Accenture a criminal conviction history is not an absolute bar to employment. The Company will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant. Additionally, employees who have access to the compensation information of other employees or applicants as a part of their essential job functions cannot disclose the pay of other employees or applicants to individuals who do not otherwise have access to compensation information, unless the disclosure is (a) in response to a formal complaint or charge, (b) in furtherance of an investigation, proceeding, hearing, or action, including an investigation conducted by the employer, or (c) consistent with the Company's legal duty to furnish information. California requires additional notifications for applicants and employees. If you are a California resident, live in or plan to work from Los Angeles County upon being hired for this position, please click here for additional important information. Please read Accenture’s Recruiting and Hiring Statement for more information on how we process your data during the Recruiting and Hiring process. Accenture
- ...Description Job Description Role: Platform Engineer – AI & Automation Platform Location:... ...role is responsible for building, operating, and advancing the shared platform... ...similar platforms ~ Experience with infrastructure automation, deployment scripting, environment...SuggestedContract workLocal areaRemote work
- ...Strategy and, Interactive, Technology, and Operations services, all powered by the world’s... ...communities. Visit us at . The Global AI Infrastructure team enables resilient, high-... ...including capacity management, reliability engineering, change management, and performance validation...SuggestedFull timeWork experience placementLive inWork at officeLocal area
- ...Job Description Job Description LDN Mission Operations Engineer Houston, TX About Intuitive Machines Intuitive Machines is an innovative and cutting-edge space company making cislunar space accessible to both public and private customers. Our mission is to...SuggestedShift work
- ...Job Description Job Description Network Operations Engineer Houston, TX About Intuitive Machines Intuitive Machines is an innovative and cutting-edge space company making cislunar space accessible to both public and private customers. Our mission is to further...SuggestedWork at officeShift workNight shift
$72.12 - $120.19 per hour
Infrastructure Deployment Engineer This is a full-time exempt (40+ hours/week), remote and travel-heavy contract... ..., safety standards, and operational requirements while partnering closely... ...certifications Experience in hyperscale or AI datacenter environments Familiarity...SuggestedHourly payFull timeContract workFor contractorsRemote work- Senior Infrastructure Engineer — Hyper-V Virtualization & Infrastructure Engineering We're looking for... ...influential companies. You will engineer and operate enterprise-scale Microsoft Hyper-V... ...and capacity planning teams. Leverage AI-assisted tooling for research, prototyping...
- ...Senior Network Engineer Houston; New York; San Francisco; Seattle... ...is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise... ..., deployment, and ongoing operation of all front-end networking services...
- ...Senior Network Engineer ON.energy is building the backbone of energy and AI infrastructure powering grid-safe data centers and mission-critical facilities. The company supplies and operates hyperscale power systems that solve the toughest resilience challenges, delivering...Local areaRemote workWorldwide
- ...Vice President of Infrastructure Engineering About the Company Rapid-growing manufacturer & distributor of electronic components Industry... ...implementing a long-term technology vision that encompasses AI infrastructure, hyperscale data centers, enterprise networking...
- Cloud Infrastructure Solutions Engineer Fugro's IT department provides the technology, cybersecurity, applications, data management, AI enablement, and technical support that allow employees, vessels, remote operations, and business systems to operate securely and efficiently...Remote workWorldwideFlexible hours
- Infrastructure Engineer III You belong to the top echelon of talent in your field. At JPMorganChase, infrastructure... ...deep storage expertise to a team that operates at global scale, where your... ...running at scale Use enterprise-authorized AI capabilities to accelerate...Shift work
$80.4k - $224.6k
...Strategy and, Interactive, Technology, and Operations services, all powered by the world’s... ...us at . A successful Network Engineer Architect combines deep technical... ...modernization programs. Experience designing infrastructure to support AI, analytics, data-intensive, and high...Full timeWork experience placementLive inWork at officeLocal areaRemote work- ...focused firm in the energy sector is seeking a Vice President of Engineering to lead software development for cloud-based solutions. The... ...background in cloud technologies, and expertise in integrating AI. This full-time position is based in Houston, Texas. #J-18808-Ljbffr...Full time
$92.7k - $213.5k
...Description: HPE's Private Cloud AI organization is looking for a Senior Cloud... ...AI productivity and streamline AI operations. In this role, you will have the opportunity... ...with primary experience in DevOps or infrastructure engineering are not a fit for this role....Work experience placementWork at officeLocal areaImmediate start$302k - $335k
...and scaling enterprise-grade AI platforms that power... ...complex organization? As AI Infrastructure Director , you'll design, manage... ...closely with Cloud Engineering, AI Engineering, and the Chief... ...organization responsible for the operational excellence of the firm's AI...Contract workWork at officeWorldwideFlexible hours$133k - $166k
...innovation forward. What You'll Do Are you passionate about building and operating secure, scalable AI platforms that power real-world innovation? As an AI Infrastructure Engineer , you'll play a key role in shaping and supporting a modern AI platform within...WorldwideFlexible hours- ...new strategies through IT and operations services and ensures the... ...environment. In this role, the Senior Engineer will create network... ...Partner with application, cloud, infrastructure, cyber security, risk, audit,... ...events. Lead automation and AI‑enabled operations improvements...Work at officeLocal areaImmediate startRemote workRelocation
$118k - $151k
...with teams across marketing, operations, and data, and see the... ...fast. Our Global Platform Engineering team is responsible for architecting... ..., and maintaining the core infrastructure, tooling, and services that... ...necessary, we sometimes use AI tools to help with parts of...Flexible hours- ...Job Description The Cloud Platform Engineer will support the space program located... ...providers. Integrating cloud managed AI and data services with other bespoke and... ...Grafana and SuperSet. Proficiency with Infrastructure-as-Code utilizing Terraform for infrastructure...Permanent employmentFull timeWork experience placementWork at officeFlexible hours
- ...Life We are seeking a Platform Engineer with a strong interest in... ...with development, security, and operations teams to support CI/CD pipelines, infrastructure automation, container platforms,... ...solutions using scripting languages or AI platforms to automate operational...Work at office
- ...Job Title : Lead Cloud Engineer Location : Houston, TX Job... ...Architects and manage the Azure cloud infrastructure for end user device management like hardware, Operating Systems, patching in adherence... .... • Familiar with AI/ML integration frameworks...Work experience placement
- ...Resource Groups, Key Vaults, VMs, VMSS, Storage resources, and infrastructure design for applications. Expert level experience Architecting... ...and managing CI/CD Pipelines. Experience Integrating AI technologies into Azure Infrastructure to enhance automation, Machine...
$123k - $163k
...per year Requirements: Over 5 years of experience in data engineering or platform engineering, with a minimum of 2-3 years specifically... ...ingestion, Delta conventions, CI/CD, observability, and AI/ML methodologies Technologies: AI Airflow CI/CD...Full time- Job details Position: Software Engineer- Platform Compensation: $ 00 annually... ...wants to shape the internal tools and infrastructure that help other engineering teams move faster... ..., incident response, and internal AI tooling. You'll join a lean team, influence...Local areaImmediate startRemote work
$48 - $50 per hour
...Description: Our client is currently seeking a QA Automation Engineer We are seeking a QA Automation Engineer with strong Python experience... ...processes, and overall software quality. Utilize modern AI tools such as Claude or GitHub Copilot to accelerate test...Hourly pay- ...experienced engagement delivery consultant for the AI Business Solutions consultant to... ...Collaboration Communications Systems Engineer Associate (Exam MS-721) Compensation... ...commitment. Lead with Integrity: We operate with trust, ethical decision-making, and...Full time
$70 - $72 per hour
...them Title: GCP Cloud Platform Engineer (Contract-to-Hire) Locations:... ...Team: IT Platform Engineering, Infrastructure & Operations Role summary We are bringing... ...closely with identity, security, FinOps and AI Governance teams. Key responsibilities...Contract work- ...reference architectures, and engineering best practices. Evaluate... ...scalability, security, and operational sustainability. Platform Engineering... ...services. Develop reusable Infrastructure as Code modules using... ...Experience building enterprise AI/ML or cloud platforms. Experience...Local area
$160k - $190k
...entire energy value chain—from project development and engineering to construction, operations, and maintenance. By integrating advanced... ...POSITION OVERVIEW We are seeking a Cloud & AI Infrastructure Engineer to design, build, operate, and optimize the...Local area- ...developers, Python/Java developers, data analysts/data scientists, data engineers, machine learning engineers for full time positions with clients... ...and REST API's experience. For data science/data analyst/AI/machine learning positions preferred skills include an associate...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Infrastructure Operations Engineer. Be the first to apply!
- ai engineer Houston, TX
- ai engineer remote Houston, TX
- ai prompt engineer Houston, TX
- ai developer Houston, TX
- senior ai engineer Houston, TX
- machine learning ai engineer Houston, TX
- ai ml engineer Houston, TX
- lead infrastructure engineer Houston, TX
- principal infrastructure engineer Houston, TX
- infrastructure engineer Houston, TX



