Data Center Operations Coordinator

Together AI

San Francisco (CA)

On-site

USD 150,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
Startup equity
Competitive benefits

Job summary

Together AI is seeking a detail-oriented Data Center Operations professional to oversee break/fix activities across multiple locations. This role involves coordinating hardware incidents, managing tickets, and tracking assets to ensure optimal uptime. Candidates should have experience in IT infrastructure and a strong understanding of server and networking hardware.

Offering a competitive salary of $150,000-200,000, equity, and benefits, Together AI is committed to fostering transparent AI systems and innovative solutions.

Qualifications

  • Experience working in data center operations, IT infrastructure, or hardware support.
  • Strong understanding of server, storage, and networking hardware.
  • Ability to manage multiple priorities across several sites simultaneously.

Responsibilities

  • Track and manage all break/fix incidents across multiple data centers.
  • Monitor ticket queues and ensure SLA compliance for incident response and resolution.
  • Coordinate with on-site technicians, remote hands teams, vendors, and engineering groups.
  • Provide daily/weekly operational status reports and incident summaries.

Skills

Data center operations experience
IT infrastructure knowledge
Ticketing systems expertise (ServiceNow, Jira, Remedy)
Excellent communication skills
Organizational skills
SLA management familiarity
Proficiency in Excel

Tools

Excel
Reporting dashboards
Inventory tracking tools

Job description

Overview

We’re looking for a detail-oriented Data Center Operations professional to manage and track all break/fix activities across multiple data center locations. This role acts as the central point of coordination for hardware incidents, vendor dispatches, ticket management, asset tracking, and operational reporting to ensure maximum uptime and fast issue resolution.

Responsibilities
  • Track and manage all break/fix incidents across multiple data centers
  • Monitor ticket queues and ensure SLA compliance for incident response and resolution
  • Coordinate with on-site technicians, remote hands teams, vendors, and engineering groups
  • Maintain accurate records of failed hardware, replacements, RMAs, and repair status
  • Escalate critical outages and recurring infrastructure issues to leadership and engineering teams
  • Schedule and oversee maintenance windows and emergency repair activities
  • Provide daily/weekly operational status reports and incident summaries
  • Ensure all work follows data center operational procedures and change management policies
  • Identify trends in hardware failures and recommend process improvements
Requirements
  • Experience working in data center operations, IT infrastructure, or hardware support
  • Strong understanding of server, storage, and networking hardware
  • Experience with ticketing systems such as ServiceNow, Jira, or Remedy
  • Ability to manage multiple priorities across several sites simultaneously
  • Excellent communication and organizational skills
  • Familiarity with SLA management and incident escalation processes
  • Proficiency with Excel, reporting dashboards, and inventory tracking tools
Preferred Qualifications
  • Experience supporting enterprise or hyperscale data centers
  • Knowledge of remote hands operations and vendor management
  • Understanding of ITIL processes and change management
  • CompTIA Server+, Network+, or similar certifications
About Together AI

Together AI is a research-driven AI infrastructure company on a mission to dramatically lower the cost of modern AI by co-designing software, hardware, algorithms, and models. We believe open and transparent AI systems create the best outcomes for society — and we\'re building the physical and computational foundation to make that real. Our team has been behind landmark advances including FlashAttention, Hyena, FlexGen, and RedPajama.

Compensation

We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $150,000-200,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.

Equal Opportunity

Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.

Please see our privacy policy at https://www.together.ai/privacy

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Center Operations Coordinator
Data Center Operations Coordinator

Togetherai • San Francisco (CA)

On-site
USD 150,000 - 200,000
Startup equity
Health insurance
Competitive benefits
Director, Data Center Operations
Director, Data Center Operations

Together • San Francisco (CA)

On-site
USD 250,000 - 300,000
Data Center Operations Coordinator — Incident SLAs & Equity
Data Center Operations Coordinator — Incident SLAs & Equity

Together AI • San Francisco (CA)

On-site
USD 150,000 - 200,000
Health insurance
Startup equity
Competitive benefits
Director, Data Center Operations
Director, Data Center Operations

Together AI • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive compensation
Startup equity
Health insurance
+1
Director, Data Center Operations
Director, Data Center Operations

Togetherai • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive compensation
Startup equity
Health insurance
+1
Program Manager, Data Center Delivery
Program Manager, Data Center Delivery

SupportFinity™ • San Francisco (CA)

Hybrid
USD 170,000 - 210,000
Equity
Health insurance
Remote-friendly
Associate, Infrastructure Strategy & Operations
Associate, Infrastructure Strategy & Operations

Together AI • San Francisco (CA)

Remote
USD 140,000 - 170,000
Startup equity
Health insurance
Remote work flexibility
+1
IT Engineer
IT Engineer

Togetherai • San Francisco (CA)

On-site
USD 140,000 - 220,000
Health insurance
Startup equity
Comprehensive benefits
Strategic Finance Senior Associate - Compute
Strategic Finance Senior Associate - Compute

Socket.dev • San Francisco (CA)

On-site
USD 138,000 - 175,000
Startup equity
Health insurance
Benefits
IT Engineer
IT Engineer

Together AI • San Francisco (CA)

Hybrid
USD 140,000 - 220,000
Equity option
Health insurance