Data Center Operations Lead

IREN

Prince George

On-site

CAD 70,000 - 90,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive hourly rate
RRSP matching program
Relocation assistance
Comprehensive health and dental coverage
Paid vacation
Professional development support
Company events and team-building activities

Job summary

A leading AI Cloud Service Provider is seeking a Data Center Operations Lead in Prince George. This role involves overseeing critical incidents, ensuring operational continuity, and managing GPU clusters in a mission-critical environment. The ideal candidate will have over 3 years of experience in data center operations and a technical degree. Join a team that values sustainability and innovation while contributing to projects that impact the future of technology.

Qualifications

  • 3+ years in mission-critical, 24/7 data center operations.
  • Expert in real-time incident management.
  • Skilled in monitoring GPU compute clusters and facility infrastructure.

Responsibilities

  • Serve as the primary decision-maker for active incidents while embedded onsite.
  • Monitor IOC dashboards for GPU cluster and networking alerts.
  • Validate ticket prioritization and direct follow-up tasks.

Skills

Leadership
Real-time incident management
Monitoring GPU compute clusters
ITSM ticketing
Coordinating on-site support
Operational continuity
Client obsession
Teamwork
Accountability
Curiosity
Cultural fit

Education

Bachelor's degree in a technical field or equivalent military/technical operations experience

Job description

Job Type: Full-time | Location: Prince George, BC | Department: Operations | Reporting to: IOC Manager | Work Location Type: #onsite

IREN is a leading AI Cloud Service Provider, delivering large-scale GPU clusters for AI training and inference. IREN’s vertically integrated platform is underpinned by its expansive portfolio of grid‑connected land and data centers in renewable‑rich regions across the U.S. and Canada.

With 100% renewable energy, we build, own and operate our data centers and take pride in being at the forefront of sustainable solutions for the ever‑evolving applications of high‑performance compute. We believe that human progress is invaluable, but it should be done in the right way – responsibly, sustainably and having a positive impact on the communities we operate in.

Job Description

As the Data Center Operations Lead, you’ll play a pivotal role in developing new data center projects. This is your chance to thrive in a high‑growth environment, driving impactful projects at the cutting edge of energy and technology infrastructure.

Job Requirements
  • 3+ years in mission‑critical, 24/7 data center operations, including leadership, mentoring, and shift‑lead roles for small teams.
  • Expert in real‑time incident management, ensuring accurate categorization, prioritization, triage, and stabilization of IT hardware, network, and facility issues.
  • Skilled in monitoring GPU compute clusters, network environments, and facility infrastructure, validating alerts and coordinating immediate response actions.
  • Proficient in ITSM ticketing: validation, prioritization, routing, and initial root‑cause assessment for engineering follow‑up.
  • Experienced in coordinating on‑site smart‑hands support, directing technicians and facilities staff based on incident priority and operational impact.
  • Strong commitment to operational continuity through shift handovers, accurate logging, risk communication, and clear inter‑team collaboration.
  • Bachelor’s degree in a technical field or equivalent military/technical operations experience.
  • Client Obsession: You take pride in delivering outstanding service.
  • Teamwork: You collaborate openly and support others’ success.
  • Accountability: You own your work from start to finish.
  • Curiosity: You ask questions, learn quickly, and seek to improve.
  • Cultural Fit: You bring positivity, reliability, and a growth mindset to every interaction.

This role requires availability for a rotational shift schedule including day, night, and weekend coverage. Candidates must be comfortable working overnight and maintaining productivity during off‑peak hours. A shift premium may apply.

Job Responsibilities
  • Serve as the primary real‑time decision‑maker for active incidents while embedded onsite with shift teams, ensuring accurate categorization, prioritization, timely triage, clear communication, and stabilization actions across IT hardware, network, GPU clusters, and facility systems.
  • Monitor IOC dashboards for GPU cluster, network, and facility alerts, validating signal quality, assessing operational impact, and coordinating appropriate response actions in real time alongside the shift team.
  • Validate ticket prioritization, ensure accurate routing, and identify emerging ticketing patterns or repeat issues, directing follow‑up tasks to IOC analysts, day‑shift leads, or on‑site technicians as needed.
  • Coordinate smart‑hands activities and support requests related to GPU clusters, facility systems, and network operations, providing guidance to technicians and ensuring safe, timely execution.
  • Escalate critical or high‑impact incidents to the IOC Manager or Tier 2/Engineering teams with clear context, documented evidence, and recommended next steps.
  • Perform initial root‑cause assessment (RCA) by collecting evidence, timelines, logs, and observations to establish category and priority, then coordinate the handoff of deeper investigation and full RCA tasks to IOC analysts and day‑shift leads.
  • Conduct structured shift handover briefings and maintain precise operational logs, ensuring operational continuity, situational awareness, and seamless transition between shifts.
  • Contribute to process improvements by identifying operational gaps, recurring issues, or workflow inefficiencies observed during the shift, and proposing actionable solutions.
  • Provide guidance and mentorship to shift analysts, fostering consistent application of incident management standards and best practices while embedded onsite.
Why Join Us?
  • Be part of a mission‑critical environment that supports high‑performance computing.
  • Collaborate with industry professionals who are passionate about technology, reliability, and efficiency.
  • Contribute to a team that values innovation, growth, and technical excellence.
Job Benefits

Compensation & Rewards

  • Competitive hourly rate, finalized based on experience and impact.
  • RRSP matching program to help you plan for your future.
  • Relocation assistance and support to get you settled.

Wellbeing & Benefits

  • Comprehensive extended health and dental coverage to keep you and your family supported.
  • Paid vacation to recharge, travel, or simply enjoy more life outside of work.

Growth & Development

  • Professional development to support certifications, continuing education, or role‑related training.

Community & Culture

  • Company events and team‑building activities.

We value diverse perspectives and believe that skills can be developed. If you’re passionate about this role, we want to hear from you — whether you meet every criteria or not. Your unique experiences might be exactly what we need!

Podtech Data Centers Inc., the employing entity and proud member of the IREN Group is an equal opportunity employer that is committed to creating an inclusive workplace. We evaluate qualified applicants without regard to race, colour, religion, age, sex, sexual orientation, gender identity, genetic information, national origin, disability, veteran status, and other legally protected characteristics.

By applying for this position and submitting your resume and application materials, you consent to the processing of your personal information in accordance with our Job Applicant Privacy Statement available on our website at www.iren.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Center Technician
Data Center Technician

IREN • Prince George

On-site
Competitive hourly rate
RRSP matching program
Comprehensive health and dental coverage
+2
Data Center Technician
Data Center Technician

IREN • Mackenzie

On-site
Relocation assistance
Night shift premium
RRSP with company match
+1
Data Center Site Manager
Data Center Site Manager

IREN • Mackenzie

On-site
CAD 155,000 - 190,000
Health insurance
Life insurance
Disability coverage
+3
Lead Data Center Technician
Lead Data Center Technician

IREN • Mackenzie

On-site
Medical, dental, and vision insurance
Company-paid life and disability insurance
RRSP with company match
Project Manager
Project Manager

IREN • Mackenzie

On-site
CAD 120,000 - 150,000
Vacation: 3 weeks annually
Company events
Training & development
+1
Technical Service Manager
Technical Service Manager

IREN • Mackenzie

On-site
CAD 100,000 - 140,000
Warehouse Coordinator
Warehouse Coordinator

IREN • Mackenzie

On-site
RRSP matching
Relocation assistance
Extended health and dental
+2
Executive Assistant
Executive Assistant

IREN • Vancouver

On-site
CAD 90,000 - 120,000
Medical, dental, and vision insurance
RRSP with company match
3 weeks vacation and paid holidays
+2
Director, People Relations & Business Enablement
Director, People Relations & Business Enablement

Socket.dev • Vancouver

Hybrid
CAD 162,000 - 198,000
RRSP with company match
3 weeks vacation and paid holidays
Flexible Work Arrangements
+2
Director, People Relations & Business Enablement
Director, People Relations & Business Enablement

IREN • Vancouver

Hybrid
CAD 162,000 - 198,000
Medical, dental, and vision coverage
Life and disability insurance
Employee Assistance Program
+3