Technical Program Manager

GMI Cloud

Mountain View (CA)

On-site

USD 150,000 - 230,000

Full time

48 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

GMI Cloud is seeking a Technical Program Manager to own and drive complex programs spanning GPU infrastructure, Kubernetes platforms, and AI services. You will lead cross-functional teams from design to production, ensuring on-time delivery with high quality.

The role requires a hands-on TPM mindset, strong communication, and the ability to balance product tradeoffs with technical rigor in a fast-moving environment.

Qualifications

  • Proven ability to drive cross-functional technical programs from design to production with strong ownership and execution discipline.
  • Strong judgment around API reliability, latency, scalability, rollout quality, and operational readiness for production AI services.
  • Experience with AI infrastructure, GPU-based inference, and distributed systems is highly valued.

Responsibilities

  • Own end-to-end delivery of complex technical programs across AI infrastructure, GPU platforms, and cloud services.
  • Define goals, milestones, dependencies, risks, and success metrics for multi-team initiatives.
  • Drive execution rigor: timelines, accountability, escalation, and delivery predictability.
  • Ensure programs reach production-ready quality, not just prototype or design completion.
  • Partner with engineering to translate designs into executable delivery plans and monitor production readiness.

Skills

Program management
Production sense
AI/LLM systems
Inference & serving
AI infrastructure
Communication

Education

Bachelor’s or Master’s in CS/Engineering

Job description

GMI Cloud is a fast-growing, AI-native infrastructure company delivering high-performance GPU compute, inference services, and infrastructure for AI agents. Following 8x ARR growth, GMI Cloud continues to scale rapidly across the U.S. and APAC. As a Reference Platform NVIDIA Cloud Partner (NCP) and a validated leading NCP across both markets, we power production AI for leading AI-native companies including Fireworks AI, Cartesia, Reflection, and OpenRouter. From large-scale compute to optimized inference and agentic workloads, GMI Cloud gives AI teams the infrastructure they need to build, deploy, and scale on one unified cloud. One cloud for compute, inference, and agents.

About this role

We are looking for a Technical Program Manager (TPM) who combines strong program ownership, production sense, and execution rigor with a solid technical foundation in AI infrastructure and distributed systems.

In this role, you will own and drive complex, cross-functional programs that span GPU infrastructure, Kubernetes platforms, inference/training systems, and customer-facing AI services. You will ensure that high-impact initiatives move from design → implementation → production → scale, on time and with high quality.

This is a hands-on TPM role for someone who can:

  • Speak fluently with engineers
  • Think like a product manager about tradeoffs and user impact
  • And execute like an owner in a fast-moving environment.
Key Responsibilities
Program Ownership & Delivery Excellence
  • Own end-to-end delivery of complex technical programs across AI infrastructure, GPU platforms, and cloud services.
  • Define clear goals, milestones, dependencies, risks, and success metrics for multi-team initiatives.
  • Drive execution rigor: timelines, accountability, escalation, and delivery predictability.
  • Ensure programs reach production-ready quality, not just prototype or design completion.
Technical & Production Sense
  • Partner closely with engineering to translate technical designs into executable delivery plans.
  • Apply strong production judgment around reliability, scalability, performance, and operational readiness.
  • Identify risks early (capacity, performance, security, operational complexity) and drive mitigation plans.
  • Ensure launches include proper monitoring, alerting, rollback strategies, and operational ownership.
Cross-Functional Leadership
  • Act as the connective tissue across Engineering, Product, Infrastructure, SRE, and Go-To-Market teams.
  • Align stakeholders on priorities, tradeoffs, and sequencing in a resource-constrained environment.
  • Communicate clearly and concisely to both technical and non-technical audiences, including leadership updates.
Platform & Infrastructure Programs
  • Drive programs related to:
  • GPU cluster expansion and lifecycle management
  • AI inference and training infrastructure
  • Internal developer platforms and tooling
  • Coordinate roadmap execution across multiple regions and environments.
Minimum Qualifications
  • Program Management: Proven ability to drive cross-functional technical programs from design to production with strong ownership and execution discipline.
  • Production Sense: Strong judgment around API reliability, latency, scalability, rollout quality, and operational readiness for production AI services.
  • AI / LLM Systems: Solid understanding of LLM and multimodal model inference workflows, including text, image, audio, or video APIs.
  • Inference & Serving: Familiarity with model serving concepts such as throughput, tail latency, batching, streaming, and cost-performance tradeoffs.
  • AI Infrastructure: General understanding of GPU-based inference systems and their impact on performance and scalability.
  • Communication: Clear, structured communication with engineering, product, and leadership stakeholders.
Preferred Qualifications
  • Experience managing technical programs related to AI products, model APIs, or inference platforms.
  • Hands-on exposure to LLM or multimodal inference systems in production environments.
  • Familiarity with AI model APIs, SDKs, or developer-facing AI platforms.
  • Strong technical foundation with the ability to reason about system tradeoffs and customer impact.
  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related technical field.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Program Manager – AI Infrastructure / GPU Clusters
Technical Program Manager – AI Infrastructure / GPU Clusters

GMI Cloud • United States

On-site
USD 140,000 - 210,000
AI Infra TPM: Scale GPU & Inference
AI Infra TPM: Scale GPU & Inference

GMI Cloud • Mountain View (CA)

On-site
USD 150,000 - 230,000
Technical Program Manager, Data Center Infrastructure Delivery
Technical Program Manager, Data Center Infrastructure Delivery

GMI Cloud • United States

On-site
USD 170,000 - 230,000
Technical Program Manager (TPM), Infrastructure
Technical Program Manager (TPM), Infrastructure

cursor • San Francisco (CA)

On-site
USD 140,000 - 180,000
AI Technical Program Manager
AI Technical Program Manager

BuzzClan LLC • Irving (TX)

Hybrid
USD 140,000 - 190,000
Technical Program Manager - Data Center / HPC Infrastructure and Operations
Technical Program Manager - Data Center / HPC Infrastructure and Operations

Cienet-International • Seattle (WA)

On-site
USD 140,000 - 180,000
Medical insurance
Dental insurance
Vision insurance
+3
Technical Program Manager - Data Center / HPC Infrastructure and Operations
Technical Program Manager - Data Center / HPC Infrastructure and Operations

CIeNET International • Seattle (WA)

On-site
USD 140,000 - 190,000
Medical, Dental, Vision, Life Ins.
401(k) Matching
PTO & Holidays
+2
Technical Program Manager III, Hardware NPI, Platforms Infrastructure
Technical Program Manager III, Hardware NPI, Platforms Infrastructure

Socket.dev • Sunnyvale (CA)

On-site
USD 163,000 - 236,000
Technical Program Manager, Infrastructure Software Systems
Technical Program Manager, Infrastructure Software Systems

Meta • Menlo Park (CA)

On-site
USD 180,000 - 240,000
Site Reliability Lead
Site Reliability Lead

GMI Cloud • United States

On-site
USD 120,000 - 180,000