Customer Support Engineer – GPU/AI Infra

The San Francisco Compute Company

San Francisco (CA)

On-site

USD 90,000 - 120,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Generous equity grant
Visa sponsorships
Retirement matching
Medical, dental & vision
Time off
Parental leave
Daily lunch
Unlimited office book budget

Job summary

San Francisco Compute is seeking a customer-facing support engineer to help enterprise customers run GPU/AI workloads with reliability and scale.

You will own tickets, triage complex issues across compute, network, and platform layers, and communicate clearly under pressure while coordinating with timezones for 24/7 coverage.

Join a team known for strong technical leadership, equity, and a mission to scale compute responsibly in a market-driven model.

Qualifications

  • 3-5 years in technical support, customer support engineering, or a similar customer-facing technical role.
  • Comfortable troubleshooting infrastructure-level issues - Linux administration, shell or Python scripting, and hands-on use of monitoring/observability tools (e.g. Prometheus, Grafana, Datadog).
  • Can explain technical problems clearly to both technical customers and internal engineering teams.
  • Calm under pressure - you don't rattle when a customer is frustrated or a system is down.
  • Comfortable working shift-based hours as part of a 24/7 global coverage model, including occasional after-hours, weekend, or holiday coverage during incidents.
  • Genuinely care about getting the customer to a good outcome, not just closing the ticket.
  • Excited to help build process and coverage from scratch, not just operate inside an existing one.
  • Excellent written communication skills, to both customers and internal teams

Responsibilities

  • Own inbound support tickets and live escalations for enterprise customers running workloads on our GPU/AI infrastructure.
  • Triage and resolve technical issues across compute, networking, and platform layers, spanning both scheduling/orchestration problems and performance issues.
  • Communicate clearly with technical customers under pressure - set expectations, give real status updates, close the loop.
  • Hand off cleanly across timezones so customers never feel the seams of follow-the-sun coverage.
  • Use monitoring and alerting tools to diagnose issues before or as customers report them.
  • Escalate hardware, data-center, or facility-level issues to the right internal or external party with a clear handoff.
  • Serve as a first responder on incidents, working alongside engineering through to resolution.
  • Work from and help improve runbooks, SOPs, and the knowledge base - flag gaps, don't just work around them.
  • Use AI-assisted tooling to work faster without losing quality or judgment.
  • Track and care about your own CSAT, first-response, and time-to-resolve numbers

Skills

Linux admin
Shell scripting
Python scripting
Monitoring tools
Written communication

Tools

Prometheus
Grafana
Datadog

Job description

San Francisco Compute is seeking a customer-facing support engineer to help enterprise customers run GPU/AI workloads with reliability and scale.

You will own tickets, triage complex issues across compute, network, and platform layers, and communicate clearly under pressure while coordinating with timezones for 24/7 coverage.

Join a team known for strong technical leadership, equity, and a mission to scale compute responsibly in a market-driven model.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Strategic GPU Infra & Customer Success Engineer
Strategic GPU Infra & Customer Success Engineer

Together AI • San Francisco (CA)

Hybrid
USD 260,000 - 290,000
Health insurance
Startup equity
Flexible remote work
Head of Infra Support Engineering (GPU/AI)
Head of Infra Support Engineering (GPU/AI)

Coreweave • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
Senior GPU Data Center Engineer
Senior GPU Data Center Engineer

Prime Intellect AI • San Francisco (CA)

On-site
USD 150,000 - 300,000
Compute Platform Engineer - GPU & Multi-Cloud Infra
Compute Platform Engineer - GPU & Multi-Cloud Infra

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health benefits
Paid parental leave
+2
GPU Infrastructure Engineer
GPU Infrastructure Engineer

Nscale • Houston (TX), San Francisco (CA), Seattle (WA)

On-site
USD 100,000 - 140,000
Infrastructure Support Engineering Manager - GPU/HPC
Infrastructure Support Engineering Manager - GPU/HPC

CoreWeave • Seattle (WA)

On-site
USD 157,000 - 210,000
Medical, dental, and vision insurance
Company-paid Life Insurance
Tuition Reimbursement
+3
Senior GPU Infrastructure Engineer — HPC & Clusters
Senior GPU Infrastructure Engineer — HPC & Clusters

Prime Intellect AI • San Francisco (CA)

On-site
USD 150,000 - 300,000
Senior GPU Solutions Engineer – AI Compute Systems
Senior GPU Solutions Engineer – AI Compute Systems

NVIDIA • Santa Clara (CA)

On-site
USD 140,000 - 270,000
Equity
Benefits
Senior GPU Compute Solutions Engineer - AI Platforms
Senior GPU Compute Solutions Engineer - AI Platforms

NVIDIA • Durham (NC)

On-site
USD 140,000 - 270,000
Equity
Benefits
Senior HPC & GPU Cluster Architect
Senior HPC & GPU Cluster Architect

The Consensus • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Visa sponsorships
401(k) retirement matching
Medical, dental & vision insurance
+2