Director, Data Center Operations

Togetherai

San Francisco (CA)

On-site

USD 250,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
Startup equity
Health insurance
Other competitive benefits

Job summary

Togetherai, based in San Francisco, is seeking a Director of Data Center Operations to oversee the operational foundation of its data center portfolio in the US and Asia. This role involves leading the design and commissioning of high-density GPU workloads and requires extensive technical knowledge of data center power and cooling systems.

The ideal candidate will build a break-fix team, manage multiple sites, and ensure operational standards, offering an attractive compensation package including competitive salary and equity.

Qualifications

  • Deep hands-on knowledge of data center power and cooling systems.
  • Experience operating data center infrastructure at meaningful scale.
  • Comfortable building teams and functions from scratch.

Responsibilities

  • Own the design and commissioning of white space sites across regions.
  • Build and lead a break-fix and smart hands team.
  • Manage a portfolio of 5+ sites across two regions.

Skills

Technical knowledge of data center power and cooling systems
Experience in designing and commissioning data center infrastructure
Leadership experience in technical operations teams
Vendor and contractor management skills
Experience with GPU or AI infrastructure

Job description

About the Role

Together AI is scaling its physical AI infrastructure rapidly — and we're looking for a Director of Data Center Operations to help us build it right. This is a ground-floor opportunity to own the operational foundation of Together's growing data center portfolio across the US and Asia.

You'll be responsible for designing and commissioning white space deployments — taking pre-built environments and fitting them out with the power distribution, cooling distribution, and systems infrastructure needed to run high-density GPU workloads at scale. At the same time, you'll be building the break-fix and smart hands team from scratch: hiring, defining the playbook, and standing up the function that keeps our sites running around the clock.

This is not a steady-state operations role. It's a builder role. You'll be joining a small but fast-moving team, with real ownership over outcomes and the autonomy to shape how Together AI operates its physical infrastructure for years to come. If you've scaled data center infrastructure through hypergrowth before and want to do it again with more ownership — this is that opportunity.

Responsibilities
  • Own the design, fit-out, and commissioning of white space sites across the US and Asia, with a focus on power distribution (PDUs), cooling distribution (CDUs), and IT-adjacent infrastructure
  • Build and lead a ~20-person break-fix and smart hands team from scratch — define the operating model, hire the initial team, and establish the processes and playbooks that keep sites running
  • Manage a portfolio of 5+ sites across two regions in various stages of deployment and live operation
  • Partner with vendors, contractors, and equipment suppliers to drive site deployments to schedule and quality
  • Establish operational standards, runbooks, and escalation processes for a nascent but rapidly growing infrastructure function
  • Serve as the technical authority on data center infrastructure decisions — from white space evaluation through to live production operation
  • Travel to Asia periodically to oversee regional site deployments and support the local team
Requirements
  • Deep, hands‑on technical knowledge of data center power and cooling systems at the IT‑adjacent layer — PDUs, CDUs, power distribution, and white space fit‑out
  • Proven experience designing and commissioning data center infrastructure, from evaluation through to live operation
  • Experience operating data center infrastructure at meaningful scale, supporting large production workloads
  • People leadership experience — you've hired, developed, and led technical operations teams
  • A builder's instinct — you're comfortable standing up teams and functions from scratch, writing the playbook where none exists, and operating with ambiguity
  • Strong vendor and contractor management skills — you know how to hold external partners accountable to timeline and quality
  • Nice to have: experience with GPU or AI infrastructure deployments, multi‑site or multi‑region portfolio management, or familiarity with Asian markets
About Together AI

Together AI is a research‑driven AI infrastructure company on a mission to dramatically lower the cost of modern AI by co‑designing software, hardware, algorithms, and models. We believe open and transparent AI systems create the best outcomes for society — and we're building the physical and computational foundation to make that real. Our team has been behind landmark advances including FlashAttention, Hyena, FlexGen, and RedPajama.

Compensation

We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full‑time position is: $250,000 - $300,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job‑related knowledge.

Equal Opportunity

Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.

Please see our privacy policy at https://www.together.ai/privacy

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Director, Data Center Operations
Director, Data Center Operations

Together AI • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive compensation
Startup equity
Health insurance
+1
Director, Data Center Operations
Director, Data Center Operations

Together • San Francisco (CA)

On-site
USD 250,000 - 300,000
Program Manager, Data Center Delivery
Program Manager, Data Center Delivery

SupportFinity™ • San Francisco (CA)

Hybrid
USD 170,000 - 210,000
Equity
Health insurance
Remote-friendly
Associate, Infrastructure Strategy & Operations
Associate, Infrastructure Strategy & Operations

Together AI • San Francisco (CA)

Remote
USD 140,000 - 170,000
Startup equity
Health insurance
Remote work flexibility
+1
Strategic Finance Senior Associate - Compute
Strategic Finance Senior Associate - Compute

Togetherai • San Francisco (CA)

On-site
USD 138,000 - 175,000
Infrastructure Design Engineer
Infrastructure Design Engineer

Together AI • San Francisco (CA)

On-site
USD 210,000 - 250,000
Health insurance
Startup equity
Flexible remote work
Strategic Finance Senior Associate - Compute
Strategic Finance Senior Associate - Compute

Together AI • San Francisco (CA)

On-site
USD 138,000 - 175,000
Health insurance
Equity
Competitive benefits
Strategic Finance Senior Associate - Compute
Strategic Finance Senior Associate - Compute

Socket.dev • San Francisco (CA)

On-site
USD 138,000 - 175,000
Startup equity
Health insurance
Benefits
Data Center Operations Coordinator
Data Center Operations Coordinator

Togetherai • San Francisco (CA)

On-site
USD 150,000 - 200,000
Startup equity
Health insurance
Competitive benefits
Senior Software Engineer - Together Cloud Infrastructure
Senior Software Engineer - Together Cloud Infrastructure

Togetherai • San Francisco (CA)

Hybrid
USD 160,000 - 230,000
Equity
Health insurance
Flexible remote work options