AI Infra Ops Engineer: GPU Clusters & Automation

Accenture

Carmel (IN)

Hybrid

USD 100,000 - 210,000

Full time

7 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Medical, dental, vision insurance
Life insurance
Long-term disability
401(k) plan
Bonus opportunities
Paid holidays
Paid time off

Job summary

Accenture is hiring for an AI infrastructure role focusing on accelerated computing across on‑premises, cloud, and hybrid deployments. You will design, deploy, and optimize GPU‑based clusters supporting AI training and inference, with emphasis on reliability and cost efficiency.

The role requires deep Kubernetes, scripting, and automation skills, and offers collaboration across diverse teams within Accenture. Travel may be involved, with clear opportunities for impact in large‑scale workloads.

Qualifications

  • 5+ years designing, deploying, and managing accelerated‑computing infra in on‑prem, cloud, and hybrid environments.
  • Hands-on with GPUs, DPUs, CPUs, NVMe-oF, and high‑bw networks.

Responsibilities

  • Design and implement accelerated‑computing infrastructure solutions aligned to architecture and governance.
  • Deploy and operate GPU clusters across bare‑metal and containerized environments using schedulers and Kubernetes.
  • Integrate platforms with enterprise systems, data platforms, security, and governance controls.
  • Build reusable tools and automation workflows for provisioning, configuration, monitoring, and remediation.
  • Establish repeatable processes for provisioning, patching, capacity planning, and lifecycle management.
  • Benchmark and validate GPU, compute, storage, and network performance across multi‑node workloads.
  • Create architecture diagrams, runbooks, and support documentation.
  • Provide guidance on troubleshooting and optimization for AI training, inference, and HPC workloads.

Skills

Kubernetes
Cluster management
Python
Bash scripting
AI infrastructure
Networking
GPU clusters
Cloud & on-prem

Education

Bachelor's degree or equivalent

Tools

Kubernetes
Slurm
Run:ai
Terraform
Ansible

Job description

Accenture is hiring for an AI infrastructure role focusing on accelerated computing across on‑premises, cloud, and hybrid deployments. You will design, deploy, and optimize GPU‑based clusters supporting AI training and inference, with emphasis on reliability and cost efficiency.

The role requires deep Kubernetes, scripting, and automation skills, and offers collaboration across diverse teams within Accenture. Travel may be involved, with clear opportunities for impact in large‑scale workloads.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Infra Operations Engineer - GPU Clusters and Automation
AI Infra Operations Engineer - GPU Clusters and Automation

Accenture • City of Albany (NY)

On-site
USD 87,000 - 266,000
Medical insurance
Dental insurance
Vision insurance
+6
AI Infrastructure Engineer: GPU Clusters & Automation
AI Infrastructure Engineer: GPU Clusters & Automation

Accenture • Irving (TX)

On-site
USD 90,000 - 260,000
Medical coverage
Dental coverage
Vision coverage
+3
AI Infrastructure Ops Engineer: GPU Clusters, Hybrid Cloud
AI Infrastructure Ops Engineer: GPU Clusters, Hybrid Cloud

Accenture • Miami (FL)

On-site
USD 87,000 - 266,000
AI Infrastructure Engineer: GPU Clusters & Automation
AI Infrastructure Engineer: GPU Clusters & Automation

Socket.dev • Missouri

On-site
USD 87,000 - 266,000
Medical, dental, vision
401(k)
Paid holidays & time off
AI Infra Engineer: GPU Clusters, Automation & HPC
AI Infra Engineer: GPU Clusters, Automation & HPC

Accenture • Detroit (MI)

On-site
USD 110,000 - 210,000
Medical benefits
Dental benefits
401(k) plan
AI Infrastructure & GPU Compute Engineer
AI Infrastructure & GPU Compute Engineer

Accenture • Seattle (WA)

On-site
USD 101,000 - 245,000
Senior AI Infrastructure & GPU Compute Engineer
Senior AI Infrastructure & GPU Compute Engineer

Accenture • Austin (TX)

On-site
USD 140,000 - 260,000
Medical benefits
401(k) program
Paid holidays & time off
AI Infrastructure Engineer: GPU Clusters & Hybrid Cloud
AI Infrastructure Engineer: GPU Clusters & Hybrid Cloud

Accenture • Redmond (WA)

On-site
USD 101,000 - 245,000
AI Infra Engineer: Kubernetes on Bare Metal GPUs (Remote)
AI Infra Engineer: Kubernetes on Bare Metal GPUs (Remote)

vCluster • New York (NY)

Hybrid
USD 150,000 - 200,000
Competitive Salary
Platinum-Level Insurance
Flexible Working Schedule
+1
AI Infra & Cluster Engineer — Scale GPU/CPU Orchestration
AI Infra & Cluster Engineer — Scale GPU/CPU Orchestration

Linuxcareers • San Francisco (CA)

On-site
USD 120,000 - 160,000