Technical Account Manager

Lambda Inc.

San Francisco, Northern (CA, KY)

Hybrid

USD 140,000 - 210,000

Full time

9 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health and vision coverage
Wellness stipends
401k with company match

Job summary

Lambda, The Superintelligence Cloud, is seeking a Technical Account Manager to own the technical health of post-sales relationships for public cloud accounts. You will ensure workloads run efficiently, architectures are solid, SLAs are defensible, and technical risk is retired before it impacts revenue.

This hands-on role requires deep understanding of training, fine-tuning, and inference workloads, plus the ability to lead joint POCs, defend architectures, and build self-serve data tools for

Qualifications

  • 5+ years in technical account management, solutions engineering or related roles with customer-facing scope.
  • Hands-on fluency with GPU infrastructure: compute, networking (InfiniBand, Ethernet), storage, and schedulers (Slurm, Kubernetes).
  • Experience mapping AI/ML workloads: training, fine-tuning, and inference, with ability to map customer stacks.
  • Track record leading structured technical engagements: POCs with defined success criteria and architecture designs.
  • Scripting, SQL, and dashboarding to turn operational data into reusable tools.

Responsibilities

  • Own the technical health of your accounts: take handoff from solutions engineering and own outcomes through steady state, expansion, and renewal.
  • Understand customer AI use cases end to end: train, fine-tune, and serve models; map frameworks, data paths, and performance baselines.
  • Lead joint customer POC sessions: define success criteria, orchestrate provisioning, run tests, drive a clear verdict.
  • Lead customer architecture designs across compute, networking, storage and schedulers; define what is managed.
  • Own SLA and reliability engineering; build canonical methods for uptime, downtime, credits; validate events at root cause level.
  • Build tooling for account health: dashboards, telemetry, uptime history, health scoring, early warning signals.
  • Direct technical escalations during high-severity events; coordinate with engineering, support and vendors.
  • Drive the technical voice of the customer; push product improvements via feature-request pipelines and POCs.
  • Know market landscape: track GPU roadmaps, competing clouds, and training/inference stacks.

Job description

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.

If you'd like to build the world's best AI cloud, join us.

*Note: This position requires presence in our San Francisco, San Jose, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.


The Technical Account Manager owns the technical health of the post-sales relationship for Lambda’s public cloud accounts, spanning AI-native startups, Enterprises, and Fortune 500 companies. Where the Customer Success Manager owns the commercial health of an account, you own its technical health: the customer’s workloads run well, the architecture is right, the SLA story is defensible, and technical risk is found and retired before it threatens revenue. Solutions Engineering carries the account through pre-sales and hypercare; at handoff, you take ownership of the technical relationship for the life of the contract.

This is a hands-on role, not a coordination role. You will understand what customers are actually building (training runs, fine-tuning pipelines, inference services) deeply enough to lead joint POC sessions, design and defend architectures, validate SLA events at the root-cause level, and build the tooling that makes account health measurable. You will be the customer’s most credible technical advocate inside Lambda and Lambda’s most trusted technical voice inside the account.

What You’ll Do
  • Own the technical health of your accounts. Take the technical handoff from Solutions Engineering at the end of hypercare and own the account’s technical outcomes through steady state, expansion, and renewal. Know the state of every cluster and workload you are accountable for, and keep your commercial counterparts ahead of technical risk.

  • Understand customer AI use cases end to end. Map what it means for each customer to train, fine-tune, and serve models on Lambda: frameworks, schedulers, parallelism strategy, data paths, and performance baselines. Build the customer user journey and convert it into value-add opportunities across the platform, documentation, and escalation routing.

  • Lead joint customer POC sessions. Define success criteria with the customer before a node is provisioned: acceptance thresholds, benchmarks, timelines. Own the execution plan, coordinate capacity and provisioning, run or oversee the tests, and drive the POC to a clear verdict: win the workload, close the gap through product, or qualify out.

  • Lead customer architecture designs. Produce and defend reference architectures spanning compute, networking, storage, connectivity, and scheduler integration (Slurm, Kubernetes). Make support boundaries explicit: what is managed and what is not. Own the design as it evolves after handoff, pulling in engineering domain experts with specific, well-framed questions.

  • Own SLA and reliability engineering. Build and own the canonical methodology for uptime, downtime, and credit calculation. Validate breach events at the technical level, down to the specific Ethernet or InfiniBand failure, and arm CSMs and leadership with defensible numbers. Partner with product to standardize SLA language and structure across 1CC, on-demand, and reserved offerings.

  • Build the tooling that makes accounts measurable. Own the technical data surfaces for customer health end to end: dashboards, telemetry and uptime history views, health scoring, and churn early-warning signals. Scope, build, and drive adoption. Replace “escalate to engineering to answer a basic question” with self-serve data for the whole GTM org.

  • Direct technical escalations and incidents. Serve as the technical lead during high-severity events on your accounts: drive root cause, hold the quality bar on RCAs, coordinate engineering, support, and vendors (NVIDIA, storage, networking), and give account teams a technically accurate narrative. Run proactive maintenance and known-issue communication so customers hear about problems from Lambda first.

  • Drive the technical voice of the customer. Run a structured feature-request pipeline into product with committed triage timelines. Audit the platform hands‑on by provisioning as a customer and testing known friction points. Lead product-led POCs (for example, NVIDIA NIM) that open new value for customers.

  • Know the market technology landscape. Track GPU roadmaps, competing clouds and neoclouds, and the evolving training and inference stacks. Brief customers on what is coming and internal teams on where Lambda stands, and let that context shape architecture and expansion recommendations.

You
  • 5+ years in technical account management, solutions engineering or architecture, ML engineering, technical program or product management, or infrastructure engineering with significant customer-facing scope, in cloud, HPC, or AI infrastructure.

  • Hands‑on fluency with GPU infrastructure: able to provision, benchmark, and debug across compute, networking (InfiniBand, Ethernet), storage, and schedulers (Slurm, Kubernetes), and to read results critically.

  • Working command of AI/ML workloads (training, fine‑tuning, inference) sufficient to map a customer’s stack, identify constraints, and lead technical conversations with their ML and infrastructure engineers.

  • Track record leading structured technical engagements: POCs with defined success criteria, architecture designs, benchmark programs, or high‑severity escalations.

  • A builder’s toolkit: scripting, SQL, and dashboarding, with a history of turning operational data into tools other people depend on.

  • Executive‑grade communication of deeply technical content, in writing and in the room.

  • Comfort with ambiguity and a track record of building methodology where none exists.

Nice to Have
  • Experience at an AI cloud, neocloud, or hyperscaler serving large-scale GPU or HPC customers.

  • Applied LLM experience (fine‑tuning, RAG systems, or inference serving) that mirrors the workloads Lambda customers run.

  • Depth in the NVIDIA ecosystem: NIM and NeMo, Base Command / BCM, the CUDA stack, DGX‑class systems.

  • Familiarity with SLA structures, service credits, enterprise contract mechanics, and retention metrics (NRR/GRR).

  • Product management or TPM background with experience converting customer evidence into roadmap decisions.

Salary Range Information

The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.

About Lambda
  • Founded in 2012, with 500+ employees, and growing fast

  • Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In‑Q‑Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove

  • We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG

  • Our values are publicly available: https://lambda.ai/careers

  • We offer generous cash & equity compensation

  • Health, dental, and vision coverage for you and your dependents

  • Wellness and commuter stipends for select roles

  • 401k Plan with 2% company match (USA employees)

  • Flexible paid time off plan that we all actually use

Equal Opportunity Employer

Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Account CTO
Account CTO

Neura Market • Northern (KY)

Hybrid
USD 300,000 - 700,000
Senior Solution Engineer
Senior Solution Engineer

Lambda Labs • United States

On-site
USD 180,000 - 260,000
Health, dental, and vision coverage
401k Plan with 2% company match
Wellness and commuter stipends
+1
Senior Solution Engineer
Senior Solution Engineer

Lambda • San Francisco (CA)

Hybrid
USD 170,000 - 210,000
Health, dental, and vision coverage
401k with 2% company match
Wellness and commuter stipends
+1
Technical Success Engineer
Technical Success Engineer

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Health/Dental/Vision
401k with match
Wellness stipend
+2
Technical Success Engineer
Technical Success Engineer

Lambda • San Jose (CA)

On-site
USD 180,000 - 240,000
Health, dental, vision coverage
Wellness stipend
Commuter stipend
+2
Technical Success Engineer
Technical Success Engineer

Lambda Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Health, dental, and vision coverage
Wellness stipend
Commuter stipend
+1
Technical Success Engineer
Technical Success Engineer

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Health coverage
Dental coverage
Vision coverage
+4
Technical Success Engineer
Technical Success Engineer

Socket.dev • San Francisco (CA)

On-site
USD 140,000 - 200,000
Health, dental, vision coverage
401(k) with 2% company match
Wellness and commuter stipends
+1
Senior Software Engineer - Core Cloud Platform
Senior Software Engineer - Core Cloud Platform

Lambda Inc. • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Health, dental, and vision coverage
401(k) plan with company match (USA)
Senior Software Engineer - Core Cloud Platform
Senior Software Engineer - Core Cloud Platform

Lambda Labs • United States

Hybrid
USD 180,000 - 240,000
Health insurance
Dental insurance
Vision insurance
+5