Technical Program Manager - Infrastructure Engineering

UST

Bengaluru

On-site

INR 1,500,000 - 2,100,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

UST in Bengaluru leads complex infrastructure programs across GPU orchestration, cluster management, and cloud platforms. The role coordinates SRE, platform engineers, and architects to deliver multi-quarter initiatives with measurable KPIs, ensuring reliability, security, and cost targets are met.

The Infra Engineering TPM supports top roadmap priorities, manages dependencies, and drives execution through scalable governance practices in a fast-moving startup-like environment.

Qualifications

  • Bachelor’s degree in computer science, Engineering, or equivalent experience.
  • 5+ years of technical program management experience at an infrastructure-oriented organization.
  • Hands-on familiarity in at least one: GPU compute stacks, Kubernetes, cluster orchestration, or cloud infrastructure.
  • Experience coordinating across SRE, platform engineering, and architecture teams to deliver complex programs.
  • Strong written and verbal communication, able to summarize trade-offs for executives.
  • Strong organizational skills: roadmaps, dependency mapping, risk management, prioritization.
  • Data-driven: defining, monitoring, and improving KPIs and SLAs.

Responsibilities

  • Lead end-to-end technical programs across infrastructure domains (GPU orchestration, cluster management, networking, storage, telemetry, provisioning pipelines).
  • Align program scope, milestones, dependencies, success metrics, and delivery timelines.
  • Drive program rituals: kickoff, sprint alignment, risk reviews, and postmortems.
  • Support Head of Infrastructure to set priorities and drive execution across teams.
  • Coordinate within infrastructure: SRE, Platform Developers, Architects.
  • Surface and manage dependencies, blockers, and risks across teams.
  • Communicate program status, trade-offs, and timelines to stakeholders.
  • Manage roadmaps and release plans for multi-team initiatives; align with strategic goals.
  • Implement scalable program management practices: OKRs, milestones, RACI, risk logs, SLAs.
  • Ensure engineering deliverables meet performance, reliability, security, and cost targets.
  • Maintain technical grounding on GPU infrastructure and NVIDIA Cloud Platform references.
  • Facilitate architecture reviews and translate architecture into executable work.
  • Coordinate reliability initiatives (SLOs, incident playbooks, SOP).
  • Prioritize reliability vs feature work and plan runbooks and rollout strategies.
  • Define and track KPIs (uptime, MTTR, provisioning latency).
  • Use data to drive prioritization and continuous improvement.

Skills

GPU compute stacks
Kubernetes
Cluster orchestration
Cloud infrastructure
Stakeholder communication
Roadmaps & planning
Risk management
SRE coordination

Education

Bachelor's degree in computer science or engineering

Job description

Client Job Title: Technical Program Manager - Infrastructure Engineering
Who we are:

At UST, we help the world s best organizations grow and succeed through transformation. Bringing together the right talent, tools, and ideas, we work with our client to co-create lasting change. Together, with over 30,000 employees in over 25 countries, we build for boundless impact-touching billions of lives in the process. Visit us at UST.com.

The Opportunity:

As an Infra Engineering TPM you will be the connective tissue between SRE, Platform Developers, Architects, and the broader engineering organization. You will own planning, execution, and communication for infrastructure programs‑ensuring work aligns to product priorities and is delivered on time with high quality. You ll support the Head of Infrastructure to turn strategic priorities into actionable roadmaps and to unblock teams operating in a fast-moving start-up environment.

Responsibilities
Program leadership

Lead end-to-end technical programs across infrastructure domains (GPU orchestration, cluster management, networking, storage, telemetry, provisioning pipelines).

Align program scope, milestones, dependencies, success metrics, and delivery timelines.

Drive program rituals: kickoff, sprint alignment, risk reviews, and postmortems.

Cross-team coordination stakeholder management

Proactively support the Head of Infrastructure in setting priorities, sequencing work, meeting coordination, and driving execution across teams.

Serve as the coordinator within infrastructure organization: SRE, Platform Developers, Architects.

Proactively surface and manage dependencies, blockers, and risks across teams.

Communicate program status, trade‑offs, and timelines to the Head of Infrastructure and other stakeholders.

Delivery and process

Manage roadmaps and release plans for multi‑team infrastructure initiatives; work directly with the Head of Infrastructure to align execution with strategic goals and adjust priorities as needed.

Implement scalable program management practices (OKRs, milestones, RACI, risk logs, SLAs).

Ensure engineering deliverables meet performance, reliability, security, and cost targets.

Technical grounding decision support

Maintain a strong technical understanding of GPU infrastructure, ensure architecture and program decisions align with NVIDIA NCP (NVIDIA Cloud Platform) reference architecture and best practices where applicable - including hardware configurations, networking/topology, telemetry, and security recommendations.

Facilitate architecture reviews and help translate architecture into executable work for platform engineers and SREs.

Reliability operations enablement

Coordinate SRE‑driven reliability initiatives (SLOs, incident response playbooks, SOP).

Help prioritize reliability work vs. feature work and plan operational runbooks, testing, and rollout strategies.

Metrics continuous improvement

Define and track key KPIs (uptime, mean time to recovery, provisioning latency).

Use data to drive prioritization and continuous improvement cycles.

Requirements Minimum

Bachelor s degree in computer science, Engineering, or equivalent experience.

5+ years of technical program management experience at an infrastructure‑oriented organization (start-up experience strongly preferred).

Strong technical background with hands‑on familiarity in at least one of: GPU compute stacks, Kubernetes, cluster orchestration, or cloud infrastructure.

Demonstrated experience coordinating across SRE, platform engineering, and architecture teams to deliver complex multi‑quarter programs.

Proven track record of delivering projects with multiple engineering teams and external dependencies.

Excellent written and verbal communication‑able to summarize technical trade‑offs and program status for executive stakeholders.

Strong organizational skills: roadmaps, dependency mapping, risk management, and prioritization.

Data‑driven: experience defining, monitoring, and improving KPIs and SLAs.

Comfortable with ambiguity and rapid change in a start‑up environment.

Experience with incident management, postmortems, and on‑call processes.

Nice-to-haves

Direct experience with GPU orchestration tools (Device Plugin, MIG, NCCL, CUDA drivers), Kubernetes GPU scheduling frameworks, or custom scheduler implementations.

Experience with on-prem/cloud deployments.

Familiarity with infrastructure‑as‑code, CI/CD, and observability tooling (Prometheus, Grafana, OpenTelemetry).

Prior experience at a GPU‑focused company, ML infrastructure team, or HPC environment.

What we believe:

We re proud to embrace the same values that have shaped UST since the beginning. Since day one, we ve been building enduring relationships and a culture of integrity. And today, its those same values that are inspiring us to encourage innovation from everyone to champion diversity and inclusion and to place people at the centre of everything we do.

Humility:

We will listen, learn, be empathetic and help selflessly in our interactions with everyone.

Humanity:

Through business, we will better the lives of those less fortunate than ourselves.

Integrity:

We honour our commitments and act with responsibility in all our relationships.

Equal Employment Opportunity Statement

UST is an Equal Opportunity Employer. We believe that no one should be discriminated against because of their differences, such as age, disability, ethnicity, gender, gender identity and expression, religion, or sexual orientation. All employment decisions shall be made without regard to age, race, creed, colour, religion, sex, national origin, ancestry, disability status, veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by federal, state, or local law. UST reserves the right to periodically redefine your roles and responsibilities based on the requirements of the organization and/or your performance.

  • - To support and promote the values of UST.
  • - Comply with all Company policies and procedures

Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Staff Site Reliability Engineer
Senior Staff Site Reliability Engineer

NVIDIA AI • Bengaluru

On-site
INR 1,800,000 - 3,200,000
Project Manager II
Project Manager II

UST • Bengaluru

On-site
INR 2,500,000 - 5,000,000
Infra Network Engineer
Infra Network Engineer

UST • Bengaluru

On-site
INR 4,000,000 - 6,000,000
DevOps Support Engineer
DevOps Support Engineer

UST • Bengaluru

On-site
INR 1,400,000 - 2,100,000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Jobgether • India

Hybrid
INR 8,247,000 - 12,371,000
Competitive salary
Annual discretionary bonus
Remote or hybrid options
+3
Cloud Engineer
Cloud Engineer

UST • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Lead SRE
Lead SRE

UST • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Technical Program Manager – Data Center Infrastructure Deployments
Technical Program Manager – Data Center Infrastructure Deployments

Nebius B.V. • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Senior Technical Program Manager
Senior Technical Program Manager

Everpure • Bengaluru

On-site
INR 500,000 - 900,000
Flexible time off
Wellness resources
Team events
Platform Engineer
Platform Engineer

UST • Bengaluru

On-site
INR 2,400,000 - 4,200,000