AI Infrastructure Engineer: GPU Clusters & Automation

Socket.dev

Missouri

On-site

USD 87,000 - 266,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, vision
401(k)
Paid holidays & time off

Job summary

Accenture’s Global AI Infrastructure team builds resilient, high‑performance compute environments for strategic clients across cloud, on‑premises and hybrid deployments. We design, build, and operate large‑scale GPU and accelerated‑computing infrastructure to support AI training, inference, and HPC workloads with scalable automation.

We partner across technology ecosystems to deliver dependable services, reusable tooling, and governance across the infrastructure stack.

Qualifications

  • Minimum 5+ years designing, deploying, and managing accelerated-computing infra across on-prem, cloud, and hybrid environments.
  • 5+ years hands‑on with GPUs, DPUs, CPUs; high‑bw networks and AI storage architectures.
  • 5+ years experience with cluster management, workload scheduling, orchestration, observability, and automation.
  • 6 months hands‑on with Claude Code, AI tools, Terraform, Ansible, Python, and Bash.
  • Bachelor’s degree or equivalent (12 years) work experience.

Responsibilities

  • Design accelerated‑computing infrastructure solutions aligned to architecture and governance requirements.
  • Deploy and operate GPU clusters across bare‑metal and containerized environments using Kubernetes and schedulers.
  • Integrate infrastructure platforms with enterprise systems, data platforms, security, and governance controls.
  • Build reusable tools, scripts, and automation workflows for provisioning, config management, validation, and monitoring.
  • Establish repeatable operational processes for cluster provisioning, patching, and capacity planning.
  • Benchmark and diagnose performance across multi-node AI training and inference workloads.
  • Develop architecture diagrams, runbooks, and support documentation.
  • Provide technical guidance for GPU clusters with emphasis on availability, resiliency, scalability, energy efficiency.

Skills

Cloud architecture
Automation
Cluster management
SRE practices
Python scripting

Education

Bachelor's degree or equivalent
Associate’s degree option (6+ years experience)

Tools

Kubernetes
Slurm
Run:ai
Terraform
Ansible
Python
Bash
Claude Code

Job description

Accenture’s Global AI Infrastructure team builds resilient, high‑performance compute environments for strategic clients across cloud, on‑premises and hybrid deployments. We design, build, and operate large‑scale GPU and accelerated‑computing infrastructure to support AI training, inference, and HPC workloads with scalable automation.

We partner across technology ecosystems to deliver dependable services, reusable tooling, and governance across the infrastructure stack.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Infrastructure Engineer: GPU Clusters & Automation
AI Infrastructure Engineer: GPU Clusters & Automation

Accenture • Irving (TX)

On-site
USD 90,000 - 260,000
Medical coverage
Dental coverage
Vision coverage
+3
AI Infra Operations Engineer - GPU Clusters and Automation
AI Infra Operations Engineer - GPU Clusters and Automation

Accenture • City of Albany (NY)

On-site
USD 87,000 - 266,000
Medical insurance
Dental insurance
Vision insurance
+6
AI Infra Ops Engineer: GPU Clusters & Automation
AI Infra Ops Engineer: GPU Clusters & Automation

Accenture • Carmel (IN)

Hybrid
USD 100,000 - 210,000
Medical, dental, vision insurance
Life insurance
Long-term disability
+4
AI Infrastructure & GPU Compute Engineer
AI Infrastructure & GPU Compute Engineer

Accenture • Seattle (WA)

On-site
USD 101,000 - 245,000
AI Infrastructure Ops Engineer: GPU Clusters, Hybrid Cloud
AI Infrastructure Ops Engineer: GPU Clusters, Hybrid Cloud

Accenture • Miami (FL)

On-site
USD 87,000 - 266,000
AI Infra Engineer: GPU Clusters, Automation & HPC
AI Infra Engineer: GPU Clusters, Automation & HPC

Accenture • Detroit (MI)

On-site
USD 110,000 - 210,000
Medical benefits
Dental benefits
401(k) plan
Senior AI Infrastructure & GPU Compute Engineer
Senior AI Infrastructure & GPU Compute Engineer

Accenture • Austin (TX)

On-site
USD 140,000 - 260,000
Medical benefits
401(k) program
Paid holidays & time off
AI Infrastructure Engineer: GPU Clusters & Hybrid Cloud
AI Infrastructure Engineer: GPU Clusters & Hybrid Cloud

Accenture • Redmond (WA)

On-site
USD 101,000 - 245,000
AI Infrastructure Architect — Scalable GPU Compute
AI Infrastructure Architect — Scalable GPU Compute

EngineersOfAI • Sunnyvale (CA)

On-site
USD 150,000 - 200,000
Senior AI Infrastructure Architect — Enterprise GPU Clusters
Senior AI Infrastructure Architect — Enterprise GPU Clusters

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equity
Benefits