Site Manager/Director - GPU Data Center Hardware Operations

Intelletec

Lubbock (TX)

On-site

USD 180,000 - 240,000

Full time

5 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Generous PTO

Job summary

Intelletec is seeking a Data Center Site Operations Leader (Hardware/Network Focus) at the Manager/Director level to oversee 24/7 mission-critical data center operations. You will build and guide high-performing teams across facilities, hardware, and network operations to ensure maximum uptime and reliability.

You will own incident response, change management, and operational excellence initiatives, driving time-to-repair metrics and safety programs such as LOTO as capacity scales across a

Qualifications

  • Experience leading Data Center operations with strict availability requirements.
  • Strong understanding of electrical and mechanical infrastructure within data centers or other mission-critical environments.
  • Proven success building and developing technical teams in high-performance environments.
  • Experience managing incident response, repair operations, and operational processes at scale.
  • Deep commitment to safety programs, including LOTO and hazardous energy control practices.
  • Comfortable operating in fast-paced environments where processes are being built alongside rapid growth.
  • Excellent written and verbal communication skills.

Responsibilities

  • Lead 24/7 mission-critical data center operations to ensure uptime across key systems.
  • Build and manage high-performing teams across facilities, hardware, and network operations.
  • Own site reliability, incident response, change management, and operational excellence initiatives.
  • Drive hardware and network repair operations including triage and time-to-repair metrics.
  • Partner with construction, commissioning, and infra teams as capacity scales.
  • Establish and scale operational playbooks, safety programs, and best practices for rapid expansion.
  • Serve as escalation point for critical incidents and lead root cause analysis for long-term reliability.

Skills

Data center operations
Incident response
Team leadership
Operational processes
Safety programs (LOTO)
Communication skills

Job description

Data Center Site Operations Leader (Hardware/Network Focus) (Manager/Director level)

We're partnering with one of the fastest-growing AI infrastructure companies building next-generation compute platforms. This is an opportunity to join a team operating at the forefront of AI and help define what world-class data center operations look like at massive scale.

What You'll Do
  • Lead 24/7 mission-critical data center operations, ensuring uptime across electrical, mechanical, cooling, and building systems.
  • Build and manage high-performing teams across facilities, hardware, and network operations.
  • Own site reliability, incident response, change management, and operational excellence initiatives.
  • Drive hardware and network repair operations, including triage, break-fix processes, and time-to-repair metrics.
  • Partner closely with construction, commissioning, and infrastructure teams as new capacity comes online.
  • Establish and scale operational playbooks, safety programs, and best practices for a rapidly expanding infrastructure footprint.
  • Serve as the escalation point for critical incidents and lead root cause analysis efforts that improve long-term reliability.
What We're Looking For
  • Experience leading Data Center operations with strict availability requirements.
  • Strong understanding of electrical and mechanical infrastructure within data centers or other mission-critical environments.
  • Proven success building and developing technical teams in high-performance environments.
  • Experience managing incident response, repair operations, and operational processes at scale.
  • Deep commitment to safety programs, including LOTO and hazardous energy control practices.
  • Comfortable operating in fast-paced environments where processes are being built alongside rapid growth.
  • Excellent written and verbal communication skills.
Nice to Have
  • Hyperscale or Tier III/IV data center experience.
  • Experience supporting GPU clusters, HPC environments, Liquid Cooling, or large-scale network infrastructure.
  • Background in commissioning, operational readiness, or CMMS/EAM systems.
  • Trade certifications or journeyman licenses in electrical, HVAC, or controls disciplines.
Why This Opportunity?
  • Work on some of the most challenging infrastructure problems in AI - Supporting to AI Labs.
  • Join a team that values ownership, speed, and first-principles thinking.
  • Help build and scale mission-critical operations from the ground up.
  • Competitive compensation package including salary, equity, comprehensive benefits, and generous PTO.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Datacenter Operations Manager
Datacenter Operations Manager

Larsen & Toubro • Concord (CA)

On-site
USD 120,000 - 150,000
Medical insurance
Vision insurance
401(k)
Data Center Operations and Maintenance Engineer
Data Center Operations and Maintenance Engineer

Designworks Talent • Bellevue (WA)

Hybrid
USD 130,000 - 190,000
Medical insurance
Dental and vision insurance
401(k) with company match
+1
Datacenter Site Manager, Ops
Datacenter Site Manager, Ops

Blue Signal Search • Houston (TX)

On-site
USD 120,000 - 180,000
Competitive base salary
Comprehensive benefits package
Generous PTO policy
Technical Program Leader - AI Infrastructure
Technical Program Leader - AI Infrastructure

Designworks Talent • Bellevue (KY)

Hybrid
USD 180,000 - 240,000
Site Manager - Data Center Operations
Site Manager - Data Center Operations

Coders Connect • Buffalo (NY)

On-site
USD 150,000 - 190,000
Competitive total compensation including equity
Retirement or pension plan
Comprehensive health, dental, and vision coverage
+1
SVP, Global Data Center Operations
SVP, Global Data Center Operations

FourFound • Dallas (TX)

On-site
USD 350,000 - 600,000
Data Center Operations and Maintenance Engineering Leader
Data Center Operations and Maintenance Engineering Leader

Designworks Talent • Bellevue (WA)

Hybrid
USD 180,000 - 240,000
Health insurance
Vision coverage
401(k) plan
Senior Technical Program Manager
Senior Technical Program Manager

nebius • United States

Remote
USD 115,000 - 275,000
Healthcare coverage
401(k) with company match
Parental leave (20 weeks primary, 12–?
+3
Head of AI Data Center Infrastructure Platforms and Software
Head of AI Data Center Infrastructure Platforms and Software

Summit Group Solutions, LLC • United States

On-site
USD 150,000 - 350,000
Data Center Compute Engineer
Data Center Compute Engineer

Blue Signal Search • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Competitive compensation
Equity opportunity
Comprehensive benefits
+2