Founding Platform Engineer

General Compute

San Francisco (CA)

On-site

USD 180,000 - 260,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

General Compute is building the inference cloud—control plane, API, and serving layer that turn racks into a sellable product. You will own the core platform, with no legacy system to constrain you, and will define how a heterogeneous mix of ASICs/GPUs gets scheduled and served reliably.

You will be a founding technical voice, working closely with data center deployment and model bring-up teams, shaping reliability and scale from day one.

Qualifications

  • Strong systems engineering background on cloud control planes/serving infra at scale.
  • Comfort being a high-impact IC rather than a manager.
  • Track record building reliability from scratch.
  • Comfort with hardware heterogeneity/ambiguity.

Responsibilities

  • Build/own the control plane — routing, model placement, scheduling across a mixed ASIC/GPU pool.
  • Build the API and serving layer exposing rack capacity as a sellable product.
  • Build in reliability and observability from day one.
  • Scale the platform ahead of the demand curve.
  • Partner closely with data center deployment and model bring-up teams.
  • Be a founding technical voice on platform architecture.

Skills

Control planes
Systems engineering
Reliability by design
Hardware heterogeneity

Job description

About us

General Compute is the neocloud for alternative chips.

Inference is fragmenting: purpose-built silicon from SambaNova, Cerebras, Positron, d-Matrix, and others already beats GPUs on decode, and we productionize that hardware — we buy the racks, find the data center space, and run it for our customers. Each piece of hardware runs the workload it's actually built for: prefill stays on GPUs, decode moves to the chip built for it, and today that means generating tokens 5–7× faster than existing GPU-based competitors. Our customers are frontier labs, fast-growing AI application companies, and asset-light clouds.

We closed a $15M seed round in May 2026, and have since closed a $400M debt facility — $100M funded upfront by Upper90, with the balance available for drawdown — collateralized by our inference chips.

About the role

You will build the inference cloud itself — the control plane, API, and serving layer that turn racks into a sellable product. There's no existing platform team to inherit or manage, no legacy system to work around, and no established playbook to follow — just the platform itself to build, with reliability treated as core infrastructure from day one rather than something bolted on after the first outage.The technical problem is also genuinely unsolved elsewhere. The fleet is heterogeneous by design — GPUs for prefill, multiple ASIC vendors for decode — so there's no single-vendor playbook to lean on; you'll be defining how a mixed-hardware inference cloud gets scheduled, routed, and served reliably, in close partnership with the teams standing up the physical fleet.

What you'll do:
  • Build/own the control plane — routing, model placement, scheduling across a mixed ASIC/GPU pool
  • Build the API and serving layer exposing rack capacity as a sellable product
  • Build in reliability and observability from day one
  • Scale the platform ahead of the demand curve
  • Partner closely with data center deployment and model bring-up teams
  • Be a founding technical voice on platform architecture
What we need from you:
  • Strong systems engineering background on cloud control planes/serving infra at scale
  • Comfort being a high-impact IC rather than a manager
  • Track record building reliability from scratch
  • Comfort with hardware heterogeneity/ambiguity
  • Genuine interest in being an early hire at a ~6–7 person company.
Nice-to-haves:
  • LLM-serving infra experience (vLLM, TGI, Ray Serve, etc.)
  • Experience running non-NVIDIA accelerators (TPUs/ASICs) in production.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Founding Inference Engineer
Founding Inference Engineer

General Compute • San Francisco (CA)

On-site
USD 180,000 - 320,000
Head of Infrastructure
Head of Infrastructure

General Compute • San Francisco (CA)

On-site
USD 190,000 - 280,000
Founding Platform Engineer — AI Inference Cloud
Founding Platform Engineer — AI Inference Cloud

General Compute • San Francisco (CA)

On-site
USD 180,000 - 260,000
Inference Engineering and Product Lead
Inference Engineering and Product Lead

United States Digital Space LLC • San Francisco (CA)

On-site
USD 240,000 - 360,000
Member of Technical Staff, Inference
Member of Technical Staff, Inference

Mount Thor • San Francisco (CA)

On-site
USD 240,000 - 320,000
Head of Inference, Perimeter Compute
Head of Inference, Perimeter Compute

Montauk Capital • New York (NY)

On-site
USD 150,000 - 200,000
Competitive compensation + equity
Studio support from Montauk Capital’s network
Principal Product Engineer, Cloud Platform
Principal Product Engineer, Cloud Platform

Verdigris Technologies Inc • Palo Alto (CA)

On-site
USD 130,000 - 180,000
Infrastructure Engineer, Perimeter Compute
Infrastructure Engineer, Perimeter Compute

Montauk Capital • New York (NY)

On-site
USD 120,000 - 160,000
Competitive compensation
Equity options
Studio support from Montauk Capital
Founding Engineer - ML Infrastructure
Founding Engineer - ML Infrastructure

uRun • San Francisco (CA)

On-site
USD 120,000 - 160,000
Health, dental, and vision
401(k)
Paid time off
+2
Inference Engineer
Inference Engineer

Hyperbolic Labs • San Francisco (CA)

On-site
USD 150,000 - 210,000