Technical Product Manager - Ai Compute Platform

Jobgether

Málaga

Presencial

EUR 90.000 - 130.000

Jornada completa

Hace 4 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Una candidatura hecha para este puesto de trabajo: un currículum y una carta de presentación adaptados que responden directamente a la oferta.

Supera los filtros ATS

Descripción de la vacante

Jobgether partners with a leading tech team in Spain to recruit a Technical Product Manager for an AI compute platform. You will own key parts of a hyperscale infrastructure product, collaborating with engineers on GPU orchestration, distributed training, and cloud APIs.

You will translate complex customer needs into platform capabilities, drive metrics, and work with cross-functional teams across engineering, SRE, networking, and billing to deliver reliable, scalable AI compute solutions.

Formación

  • 6+ years of experience in Product Management or equivalent leadership roles in SRE/engineering.
  • Strong technical foundation in cloud infrastructure, distributed systems, or AI/ML platforms.
  • Experience with large-scale infrastructure such as GPU clusters or multi-tenant cloud environments.
  • Proven track record of shipping complex technical products with measurable customer impact.
  • Strong analytical skills with telemetry and data-driven decision-making.

Responsabilidades

  • Own end-to-end product strategy, roadmap, and execution for a slice of the AI compute platform.
  • Define APIs, system behaviors, and developer-facing interfaces at hyperscaler quality.
  • Lead cross-functional execution across engineering, SRE, networking, storage, and billing teams.
  • Drive structured product discovery through customer interviews and usage analytics.
  • Translate complex technical challenges into clear product requirements and metrics.
  • Collaborate with engineering to evaluate architecture decisions and platform design.

Conocimientos

APIs design
Distributed systems
Cloud infrastructure
GPU infrastructure
SRE practices
Stakeholder management
Communication
Data-driven decisions

Herramientas

Kubernetes
Slurm
NCCL

Descripción del empleo

Technical Product Manager - AI Compute Platform

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Technical Product Manager - AI Compute Platform based in Spain.

In this role, you will help shape and scale a next-generation AI cloud platform powering some of the most demanding machine learning workloads in the world. You will own critical parts of a hyperscale infrastructure product, working at the intersection of engineering, customers, and platform strategy. This is a deeply technical product management role where you will collaborate as a peer with senior engineers on topics such as GPU orchestration, distributed training, cluster operations, and cloud APIs. You will translate complex customer needs into robust platform capabilities that enable large-scale AI training and inference. The environment is fast-paced, highly technical, and mission-driven, with direct impact on how frontier AI systems are built and deployed. This role is ideal for someone who thrives in ambiguity, enjoys solving infrastructure-scale problems, and wants to define the future of AI compute platforms.

Accountabilities
  • Own end-to-end product strategy, roadmap, and execution for a critical slice of an AI compute platform, ensuring alignment with customer and business outcomes.
  • Define and evolve platform contracts such as APIs, system behaviors, lifecycle semantics, and developer-facing interfaces at hyperscaler quality.
  • Lead cross-functional execution across engineering, SRE, networking, storage, observability, IAM, billing, capacity planning, and customer-facing teams.
  • Drive structured product discovery through customer interviews, usage analytics, incident analysis, and support feedback loops.
  • Translate complex technical and operational challenges into clear product requirements and measurable success metrics.
  • Collaborate as a technical peer with engineering teams to evaluate architecture decisions, system trade-offs, and platform design quality.
  • Own adoption and performance of shipped features, ensuring continuous improvement based on real-world usage and telemetry.
  • Serve as the escalation point for customer-facing teams on product behavior, system reliability, and platform design decisions.
  • Define success metrics tied to customer impact, platform efficiency, and operational excellence rather than output-based delivery.
Requirements
  • 6+ years of experience in Product Management, Platform/Product Infrastructure roles, or equivalent experience in SRE or Engineering leadership with strong product ownership.
  • Strong technical foundation in cloud infrastructure, distributed systems, or AI/ML platforms, with the ability to reason about system design and architecture.
  • Experience working with or operating large-scale infrastructure such as GPU clusters, HPC systems, or multi-tenant cloud environments.
  • Proven track record of shipping complex technical products with measurable impact on customers or platform performance.
  • Strong analytical skills with experience defining metrics, working with telemetry, and driving data-informed product decisions.
  • Experience leading discovery processes, including customer interviews, usage analysis, and support-driven insights.
  • Ability to engage confidently with engineering teams on topics such as API design, system reliability, control planes, and distributed systems behavior.
  • Excellent communication and stakeholder management skills across engineering, product, operations, and executive teams.
  • High ownership mindset with a strong bias toward execution, iteration, and operational excellence.
  • Familiarity with GPU infrastructure, Kubernetes, Slurm, or HPC environments is highly desirable.
  • Experience with distributed ML training or inference workloads (e.g., multi-node training, NCCL, checkpointing, fault-tolerant systems) is a strong plus.
  • Exposure to cloud platforms at hyperscaler scale (AWS, GCP, Azure) and developer experience design (APIs, CLI tools, observability systems) is advantageous.
  • Understanding of reliability engineering practices, SRE principles, and operational metrics such as MTTR, MTBF, or system-level goodput is a plus.
  • Experience working with emerging AI workloads such as agentic systems, RL pipelines, or large-scale inference serving is considered a bonus.
  • Competitive compensation package.
  • Opportunity to shape foundational infrastructure for the global AI ecosystem.
  • High-impact role with ownership over critical components of a hyperscale AI platform.
  • Flexible, trust-based work environment with strong autonomy.
  • Exposure to cutting-edge AI, GPU, and distributed systems technologies.
  • Collaborative international environment with world-class engineering teams.
  • Opportunity to work on problems at the frontier of cloud infrastructure and AI compute.
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

AI Compute Platform Product Manager – Hyperscale Infra
AI Compute Platform Product Manager – Hyperscale Infra

Jobgether • Málaga

Presencial
EUR 90.000 - 130.000
Product Mgr/Technical Project Mgr Template
Product Mgr/Technical Project Mgr Template

Jobgether • España

Presencial
EUR 70.000 - 110.000
Competitive compensation
Flexible working environment
International collaboration
AI Infrastructure Engineer (GPU) - Remote EMEA
AI Infrastructure Engineer (GPU) - Remote EMEA

Pragmatike • Madrid

Presencial
EUR 60.000 - 80.000
Work from home flexibility
Inclusive recruitment process
Opportunity to influence core engineering decisions
Product Manager, Expert
Product Manager, Expert

Keysight Technologies SAles Spain SL. • Barcelona

Presencial
EUR 80.000 - 120.000
Product Manager - AI Platform
Product Manager - AI Platform

Keysight Technologies • Barcelona

Presencial
EUR 70.000 - 100.000
Applied Ai Engineer
Applied Ai Engineer

Jobgether • Málaga

Híbrido
EUR 60.000 - 100.000
Healthcare coverage
Remote-first / Hybrid-friendly culture
Unlimited PTO
+1
Technical Support Engineer
Technical Support Engineer

Jobgether • España

Presencial
EUR 55.000 - 75.000
Competitive salary
Equity package
Learning and career growth
+1
Principal Platform Engineer (Python)
Principal Platform Engineer (Python)

Intellias • España

Presencial
EUR 80.000 - 110.000
Senior AI Platform Engineer
Senior AI Platform Engineer

N-iX • Valencia

Híbrido
EUR 90.000 - 130.000
Flexible remote/office option
Competitive salary
Career growth
+2
AI Platform Engineer
AI Platform Engineer

Super • Madrid

Presencial
EUR 60.000 - 90.000
Medical / Health Insurance
Open Annual Leave
Employee Assistance Programme
+1