Datacenter Operations Manager

Larsen & Toubro

Concord (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical insurance
Vision insurance
401(k)

Job summary

A leading engineering firm is seeking an experienced AI Data Center Operations Manager to lead operations for their AI data center in California. This role requires managing critical infrastructure like power, cooling, and network systems while ensuring the performance of NVIDIA GPU clusters. Candidates should have 7+ years of experience in data center operations, with expertise in HVAC and mechanical systems. Benefits include medical insurance and a 401(k). This is a full-time position with a mid-senior level seniority.

Qualifications

  • 7+ years of experience in data center operations, with 3+ years managing AI or HPC environments.
  • Hands-on experience with GPU clusters (NVIDIA B300 preferred) and Dell PowerEdge hardware.
  • Strong knowledge of networking, telco systems, and high-availability architectures.

Responsibilities

  • Oversee daily operations of power systems, mechanical/electrical infrastructure, and HVAC.
  • Ensure continuous uptime and operational efficiency for AI compute clusters.
  • Coordinate hardware upgrades and troubleshooting with engineering teams.

Skills

Leadership
Communication
Vendor management
HVAC knowledge
Leadership skills
Vendor management

Education

Bachelor's degree in electrical engineering, Mechanical Engineering, or Computer Science
Master's degree preferred

Tools

NVIDIA B300 GPU
Dell PowerEdge hardware
Kubernetes
Slurm

Job description

AI Data Center Operations Manager

We are seeking an experienced AI Data Center Operations Manager to lead operations for a cutting-edge AI data center. This role involves managing all critical infrastructure systems—including power, mechanical, electrical, cooling, HVAC, liquid cooling, network, telco, chillers, water treatment plant, and generators—while ensuring optimal performance of NVIDIA B300 GPU clusters on Dell PowerEdge hardware. The position requires strong collaboration with vendors and internal teams to maintain high availability, security, and compliance.

Key Responsibilities
Infrastructure Operations
  • Oversee daily operations of power systems, mechanical/electrical infrastructure, HVAC, liquid cooling, chillers, water treatment plant, and backup generators.
  • Ensure continuous uptime and operational efficiency for AI compute clusters.
AI Hardware & Compute Management
  • Coordinate hardware upgrades and troubleshooting with engineering teams.
  • Act as primary liaison with vendors including Servers, HVAC providers, electrical contractors, and other critical service partners.
  • Negotiate and manage service agreements, maintenance schedules, and procurement.
Network & Telco
  • Maintain robust connectivity and manage network infrastructure supporting AI workloads.
  • Work with telecom providers to ensure redundancy and high availability.
Physical Security & Compliance
  • Partner with the physical security team to enforce access control, surveillance, and compliance with security protocols.
  • Ensure adherence to industry standards and environmental regulations.
  • Implement advanced monitoring systems for power, cooling, and compute resources.
  • Lead incident response and root cause analysis for outages or failures.
Qualifications
  • Bachelor’s degree in electrical engineering, Mechanical Engineering, Computer Science, or related field (master’s preferred).
  • 7+ years of experience in data center operations, with 3+ years managing AI or HPC environments.
  • Expertise in power systems, HVAC, liquid cooling, and mechanical/electrical infrastructure.
  • Hands‑on experience with GPU clusters (NVIDIA B300 preferred) and Dell PowerEdge hardware.
  • Strong knowledge of networking, telco systems, and high‑availability architectures.
  • Excellent leadership, communication, and vendor management skills.
Preferred Skills
  • Familiarity with AI workload orchestration tools (Kubernetes, Slurm).
  • Certifications: CDCP/CDCS, ITIL, or Data Center Management.
  • Experience with environmental sustainability practices in data centers.
Benefits
  • Medical insurance
  • Vision insurance
  • 401(k)
Seniority level
  • Mid-Senior level
Employment type
  • Full-time
Job function
  • Other
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Manager/Director - AI Data Center Operations
Manager/Director - AI Data Center Operations

Intelletec Energy • South Bend (IN)

On-site
USD 120,000 - 180,000
Manager/Director - AI Data Center Operations
Manager/Director - AI Data Center Operations

Intelletec Energy • Baton Rouge (LA)

On-site
USD 110,000 - 150,000
Generous PTO
Equity
Comprehensive benefits
AI Data Center Ops Leader
AI Data Center Ops Leader

Larsen & Toubro • Concord (CA)

On-site
USD 120,000 - 150,000
Senior Solutions Architect, Data Center Infrastructure, Senior Solutions Architect, Data Center[...]
Senior Solutions Architect, Data Center Infrastructure, Senior Solutions Architect, Data Center[...]

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Comprehensive benefits package
AI Operations & Infrastructure Engineer
AI Operations & Infrastructure Engineer

Invictus International • Geraghty Village (MD)

On-site
USD 100,000 - 130,000
Data Center Technician - AI Infrastructure
Data Center Technician - AI Infrastructure

Hamilton Barnes Associates Limited • South Carolina

On-site
USD 70,000 - 85,000
Funded training and certifications
Stable full-time schedule: Monday–Friday, 9-5
Healthcare benefits
AI Operations & Infrastructure Engineer
AI Operations & Infrastructure Engineer

Invictus International Consulting, LLC • Fort Meade (MD)

On-site
USD 100,000 - 130,000
Data Center Operations and Maintenance Engineer
Data Center Operations and Maintenance Engineer

Designworks Talent • Bellevue (WA)

Hybrid
USD 130,000 - 190,000
Medical insurance
Dental and vision insurance
401(k) with company match
+1
Datacenter Site Manager, Ops
Datacenter Site Manager, Ops

Blue-Signal-Search • Buffalo (NY)

On-site
USD 90,000 - 120,000
Comprehensive benefits package
Generous PTO policy
Competitive base salary
Director Datacenter Operations
Director Datacenter Operations

asobbi • New Jersey

On-site
USD 150,000 - 200,000