Data Center Operations Lead

IREN

Prince George

On-site

CAD 70,000 - 90,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Competitive hourly rate
RRSP matching program
Relocation assistance
Comprehensive health and dental coverage
Paid vacation
Professional development support
Company events and team-building activities

Job summary

A leading AI Cloud Service Provider is seeking a Data Center Operations Lead in Prince George. This role involves overseeing critical incidents, ensuring operational continuity, and managing GPU clusters in a mission-critical environment. The ideal candidate will have over 3 years of experience in data center operations and a technical degree. Join a team that values sustainability and innovation while contributing to projects that impact the future of technology.

Qualifications

  • 3+ years in mission-critical, 24/7 data center operations.
  • Expert in real-time incident management.
  • Skilled in monitoring GPU compute clusters and facility infrastructure.

Responsibilities

  • Serve as the primary decision-maker for active incidents while embedded onsite.
  • Monitor IOC dashboards for GPU cluster and networking alerts.
  • Validate ticket prioritization and direct follow-up tasks.

Skills

Leadership
Real-time incident management
Monitoring GPU compute clusters
ITSM ticketing
Coordinating on-site support
Operational continuity
Client obsession
Teamwork
Accountability
Curiosity
Cultural fit

Education

Bachelor's degree in a technical field or equivalent military/technical operations experience

Job description

Job Type: Full-time | Location: Prince George, BC | Department: Operations | Reporting to: IOC Manager | Work Location Type: #onsite

IREN is a leading AI Cloud Service Provider, delivering large-scale GPU clusters for AI training and inference. IREN’s vertically integrated platform is underpinned by its expansive portfolio of grid‑connected land and data centers in renewable‑rich regions across the U.S. and Canada.

With 100% renewable energy, we build, own and operate our data centers and take pride in being at the forefront of sustainable solutions for the ever‑evolving applications of high‑performance compute. We believe that human progress is invaluable, but it should be done in the right way – responsibly, sustainably and having a positive impact on the communities we operate in.

Job Description

As the Data Center Operations Lead, you’ll play a pivotal role in developing new data center projects. This is your chance to thrive in a high‑growth environment, driving impactful projects at the cutting edge of energy and technology infrastructure.

Job Requirements
  • 3+ years in mission‑critical, 24/7 data center operations, including leadership, mentoring, and shift‑lead roles for small teams.
  • Expert in real‑time incident management, ensuring accurate categorization, prioritization, triage, and stabilization of IT hardware, network, and facility issues.
  • Skilled in monitoring GPU compute clusters, network environments, and facility infrastructure, validating alerts and coordinating immediate response actions.
  • Proficient in ITSM ticketing: validation, prioritization, routing, and initial root‑cause assessment for engineering follow‑up.
  • Experienced in coordinating on‑site smart‑hands support, directing technicians and facilities staff based on incident priority and operational impact.
  • Strong commitment to operational continuity through shift handovers, accurate logging, risk communication, and clear inter‑team collaboration.
  • Bachelor’s degree in a technical field or equivalent military/technical operations experience.
  • Client Obsession: You take pride in delivering outstanding service.
  • Teamwork: You collaborate openly and support others’ success.
  • Accountability: You own your work from start to finish.
  • Curiosity: You ask questions, learn quickly, and seek to improve.
  • Cultural Fit: You bring positivity, reliability, and a growth mindset to every interaction.

This role requires availability for a rotational shift schedule including day, night, and weekend coverage. Candidates must be comfortable working overnight and maintaining productivity during off‑peak hours. A shift premium may apply.

Job Responsibilities
  • Serve as the primary real‑time decision‑maker for active incidents while embedded onsite with shift teams, ensuring accurate categorization, prioritization, timely triage, clear communication, and stabilization actions across IT hardware, network, GPU clusters, and facility systems.
  • Monitor IOC dashboards for GPU cluster, network, and facility alerts, validating signal quality, assessing operational impact, and coordinating appropriate response actions in real time alongside the shift team.
  • Validate ticket prioritization, ensure accurate routing, and identify emerging ticketing patterns or repeat issues, directing follow‑up tasks to IOC analysts, day‑shift leads, or on‑site technicians as needed.
  • Coordinate smart‑hands activities and support requests related to GPU clusters, facility systems, and network operations, providing guidance to technicians and ensuring safe, timely execution.
  • Escalate critical or high‑impact incidents to the IOC Manager or Tier 2/Engineering teams with clear context, documented evidence, and recommended next steps.
  • Perform initial root‑cause assessment (RCA) by collecting evidence, timelines, logs, and observations to establish category and priority, then coordinate the handoff of deeper investigation and full RCA tasks to IOC analysts and day‑shift leads.
  • Conduct structured shift handover briefings and maintain precise operational logs, ensuring operational continuity, situational awareness, and seamless transition between shifts.
  • Contribute to process improvements by identifying operational gaps, recurring issues, or workflow inefficiencies observed during the shift, and proposing actionable solutions.
  • Provide guidance and mentorship to shift analysts, fostering consistent application of incident management standards and best practices while embedded onsite.
Why Join Us?
  • Be part of a mission‑critical environment that supports high‑performance computing.
  • Collaborate with industry professionals who are passionate about technology, reliability, and efficiency.
  • Contribute to a team that values innovation, growth, and technical excellence.
Job Benefits

Compensation & Rewards

  • Competitive hourly rate, finalized based on experience and impact.
  • RRSP matching program to help you plan for your future.
  • Relocation assistance and support to get you settled.

Wellbeing & Benefits

  • Comprehensive extended health and dental coverage to keep you and your family supported.
  • Paid vacation to recharge, travel, or simply enjoy more life outside of work.

Growth & Development

  • Professional development to support certifications, continuing education, or role‑related training.

Community & Culture

  • Company events and team‑building activities.

We value diverse perspectives and believe that skills can be developed. If you’re passionate about this role, we want to hear from you — whether you meet every criteria or not. Your unique experiences might be exactly what we need!

Podtech Data Centers Inc., the employing entity and proud member of the IREN Group is an equal opportunity employer that is committed to creating an inclusive workplace. We evaluate qualified applicants without regard to race, colour, religion, age, sex, sexual orientation, gender identity, genetic information, national origin, disability, veteran status, and other legally protected characteristics.

By applying for this position and submitting your resume and application materials, you consent to the processing of your personal information in accordance with our Job Applicant Privacy Statement available on our website at www.iren.com.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Center Technician
Data Center Technician

IREN • Prince George

On-site
CAD 37,195 - 49,593
Competitive hourly rate
RRSP matching program
Comprehensive health and dental coverage
+2
Data Center Technician
Data Center Technician

IREN • Mackenzie

On-site
CAD 37,000 - 50,000
Relocation assistance
RRSP with company match
Senior Software Engineer
Senior Software Engineer

IREN • Vancouver

On-site
CAD 135,000 - 155,000
3 weeks vacation
RRSP with company match
Medical, dental, and vision insurance
Senior Technical Infrastructure Program Manager (Sr. TIPM)
Senior Technical Infrastructure Program Manager (Sr. TIPM)

Iris Energy • Canada

On-site
CAD 120,000 - 185,000
Relocation assistance
Paid time off
Flexible work arrangements
+1
Recruitment Coordinator
Recruitment Coordinator

IREN • Vancouver

Hybrid
CAD 54,000 - 62,000
Medical, dental, and vision insurance
RRSP with company match
Voluntary TFSA
+2
Director, People Relations & Business Enablement
Director, People Relations & Business Enablement

Iris Energy • Vancouver

On-site
CAD 180,000 - 200,000
RRSP with company match
Equity / long-term incentive
3 weeks vacation and holidays
+2
Project Scheduler
Project Scheduler

IREN • Mackenzie

On-site
CAD 120,000 - 150,000
Heavy Duty Mechanic
Heavy Duty Mechanic

IREN • Mackenzie

On-site
CAD 145,419,000 - 153,443,000
Relocation assistance
Medical coverage
Life & disability insurance
+3
Maintenance Technician
Maintenance Technician

Iren Group • Prince George

On-site
CAD 50,000 - 52,000
Relocation assistance
RRSP with company match
Training and development opportunities
+1
General Foreman - Mechanical - Prince George
General Foreman - Mechanical - Prince George

Iren Group • Vancouver

On-site
CAD 62,000 - 83,000
Relocation assistance
Living Out Allowance (LOA)
Travel allowance