Senior TPM - Bare Metal GPU & AI Cloud Delivery

Iris Energy

San Francisco (CA)

Hybrid

USD 200.000 - 260.000

Vollzeit

14 Tage+
Bewerbungsgenerator

Bekomme eine Antwort von diesem Arbeitgeber — ein Lebenslauf und ein Anschreiben, die genau auf die Eigenschaften eingehen, die gesucht werden.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Health insurance
401(k) matching
Paid time off
Flexible work arrangements
Career growth opportunities
Company events

Zusammenfassung

IREN, a leading AI Cloud provider, seeks a Principal Technical Program Manager for Bare Metal GPU & AI Cloud Delivery in San Francisco. You will own end-to-end delivery from planning to production handoff, coordinating across hardware, data center, and cloud software teams to ensure scalable, production-ready GPU cloud infrastructure for AI workloads.

The role requires deep experience in GPU infrastructure, data center readiness, and cross-functional leadership, with a focus on risk management

Qualifikationen

  • Bachelor's degree in a technical discipline or equivalent experience.
  • 10+ years of experience in technical program management, cloud infrastructure, data centers, hardware infrastructure, or software engineering programs.
  • Proven experience owning end-to-end delivery of complex infrastructure programs from planning and requirements through deployment, production readiness, customer launch, and operational handoff.
  • Strong understanding of bare metal GPU infrastructure, GPU server platforms, high-performance networking, storage, cloud platforms, distributed systems, and data center dependencies such as power, cooling, rack layout, and network readiness.
  • Experience leading GPU cluster bring-up, capacity enablement, hardware deployment, network integration, validation, and production launch across internal teams, vendors, and external partners.
  • Ability to manage critical paths, risks, dependencies, schedules, decision forums, executive communications, and customer-facing delivery commitments in high-ambiguity environments.

Aufgaben

  • Own and drive end-to-end delivery processes for Bare Metal GPU and AI Cloud programs, from requirements definition, infrastructure design, capacity planning, procurement readiness, and deployment planning through commissioning, software deployment, customer handover, and operational steady state.
  • Define, standardize, and scale repeatable delivery processes across multiple data centers, ensuring each site follows clear playbooks, milestones, dependency tracking, readiness criteria, risk management, escalation paths, and handoff procedures.
  • Lead cross-functional execution across infrastructure design, data center operations, supply chain, logistics, infrastructure installation, cabling, networking, commissioning, cloud software deployment, security, customer engineering, and support teams.
  • Partner with infrastructure design teams to translate AI Cloud capacity, GPU cluster architecture, power, cooling, rack layout, network fabric, storage, and operational requirements into executable multi-site delivery plans.
  • Collaborate with supply chain and vendor partners to align GPU systems, network equipment, racks, optics, storage, firmware, spares, logistics, and site delivery schedules with program milestones and customer commitments.
  • Coordinate infrastructure installation and site readiness activities, including rack placement, power and cooling validation, network turn-up, cabling completion, hardware acceptance, and issue remediation across internal teams and external contractors.
  • Drive commissioning and production readiness reviews for each GPU cluster, ensuring hardware, firmware, networking, storage, automation, observability, security controls, support processes, and service acceptance criteria are fully validated before launch.
  • Lead software deployment readiness across provisioning, orchestration, monitoring, capacity management, customer onboarding workflows, and platform service enablement to ensure AI Cloud environments are production-ready.
  • Own customer handover planning and execution, including launch readiness, acceptance criteria, documentation, support transition, known-issue tracking, stakeholder communications, and post-launch stabilization.
  • Establish portfolio-level governance for concurrent data center and AI Cloud delivery programs, providing executive-level reporting on progress, risks, dependencies, escalations, site readiness, customer impact, and business outcomes.
  • Identify bottlenecks across process, tooling, vendor execution, installation workflows, commissioning, software deployment, and operational handoffs; drive durable improvements that increase deployment velocity, repeatability, quality, reliability, and cost efficiency.

Kenntnisse

End-to-end program management
Cloud infrastructure
Data centers
GPU infrastructure
High-performance networking
Executive communications
Risk management

Ausbildung

Bachelor's degree in a technical discipline

Tools

Kubernetes
IaC
Automation
Observability

Jobbeschreibung

IREN, a leading AI Cloud provider, seeks a Principal Technical Program Manager for Bare Metal GPU & AI Cloud Delivery in San Francisco. You will own end-to-end delivery from planning to production handoff, coordinating across hardware, data center, and cloud software teams to ensure scalable, production-ready GPU cloud infrastructure for AI workloads.

The role requires deep experience in GPU infrastructure, data center readiness, and cross-functional leadership, with a focus on risk management

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior TPM, AI Cloud & Bare-Metal GPU Infra
Senior TPM, AI Cloud & Bare-Metal GPU Infra

Iren Group • USA

Vor Ort
USD 200.000 - 260.000
Health insurance
401(k) matching
Paid time off
Remote TPM for AI/GPU Infrastructure Deployments
Remote TPM for AI/GPU Infrastructure Deployments

5C Group • USA

Vor Ort
USD 155.000 - 175.000
Senior TPM – GPU AI Servers, Global Infra (RSUs)
Senior TPM – GPU AI Servers, Global Infra (RSUs)

Amazon Inc. • Seattle (WA), Northern (KY)

Hybrid
USD 149.000 - 201.000
Senior TPM: GPU AI Infrastructure & ML Servers
Senior TPM: GPU AI Infrastructure & ML Servers

Amazon Web Services (AWS) • Seattle (WA)

Vor Ort
USD 149.000 - 201.000
Health insurance
RSUs
401(k) matching
Senior TPM - AI Infrastructure & GPU Deployments
Senior TPM - AI Infrastructure & GPU Deployments

Hamilton Barnes Associates Limited • New York (NY)

Vor Ort
USD 200.000 - 240.000
Bonus
Equity
Senior TPM: GPU AI Server Infra for Global ML
Senior TPM: GPU AI Server Infra for Global ML

Amazon Web Services (AWS) • Cupertino (CA)

Vor Ort
USD 171.000 - 231.000
Health insurance
401(k) matching
Paid time off
+2
Senior TPM: GPU AI Server Infra for Global ML
Senior TPM: GPU AI Server Infra for Global ML

Amazon Web Services (AWS) • Cupertino (CA)

Vor Ort
USD 171.000 - 231.000
Health insurance
401(k) matching
Paid time off
+2
Staff TPM: AI Infrastructure Deployments & GPU Cloud
Staff TPM: AI Infrastructure Deployments & GPU Cloud

Crusoe • Bellevue (WA)

Vor Ort
USD 200.000 - 240.000
Equity
Paid time off
Health insurance
+2
Principal AI Cloud Solutions Architect – GPU HPC Infra
Principal AI Cloud Solutions Architect – GPU HPC Infra

IREN • San Francisco (CA)

Vor Ort
USD 180.000 - 260.000
Medical, dental, and vision insurance
401(k) plan with company match
Paid time off
+2
Principal Technical Program Manager, AI Cloud Infrastructure
Principal Technical Program Manager, AI Cloud Infrastructure

Iris Energy • San Francisco (CA)

Hybrid
USD 200.000 - 260.000
Health insurance
401(k) matching
Paid time off
+3