Inference Cloud Architect: Build Scalable AI Infrastructure

techire.®

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

techire.® is building an early-stage AI infrastructure platform in San Francisco, focused on turning a heterogeneous fleet of AI compute into reliable, scalable infrastructure customers can consume. This role is hands-on from 0→1, with no mature platform yet, and involves making architectural decisions that shape cloud behavior from day one.

You’ll work on building the actual product, defining how scheduling, serving and reliability components interact, and exploring trade-offs in how the

Qualifications

  • Experience building distributed systems at scale and in production
  • Ability to design cloud infrastructure and control planes
  • Familiarity with scheduling, routing, or large-scale serving systems

Responsibilities

  • Designing the control plane across heterogeneous compute
  • Building scheduling, routing and workload/model placement systems
  • Creating customer-facing APIs and serving infrastructure
  • Designing reliability, observability and failure handling from the start
  • Scaling the platform as workloads and customer demand grow

Job description

techire.® is building an early-stage AI infrastructure platform in San Francisco, focused on turning a heterogeneous fleet of AI compute into reliable, scalable infrastructure customers can consume. This role is hands-on from 0→1, with no mature platform yet, and involves making architectural decisions that shape cloud behavior from day one.

You’ll work on building the actual product, defining how scheduling, serving and reliability components interact, and exploring trade-offs in how the

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Inference Engineer
Inference Engineer

techire.® • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Founding Platform Engineer — AI Inference Cloud
Founding Platform Engineer — AI Inference Cloud

General Compute • San Francisco (CA)

On-site
USD 180,000 - 260,000
Senior AI Inference Platform Engineer
Senior AI Inference Platform Engineer

Cloudflare • Austin (TX)

Hybrid
USD 180,000 - 240,000
AI Inference Infrastructure Engineer
AI Inference Infrastructure Engineer

Scaled Cognition • New York (NY)

On-site
USD 120,000 - 170,000
Staff Cloud Inference Launch Engineer for Scalable AI
Staff Cloud Inference Launch Engineer for Scalable AI

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
AI Platform Engineer — Scale Cloud Infra, Equity
AI Platform Engineer — Scale Cloud Infra, Equity

Harrison Clarke • San Francisco (CA)

On-site
USD 150,000 - 210,000
Cloud-Scale AI Inference Architect
Cloud-Scale AI Inference Architect

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation and benefits package
Daily lunch and dinner
Unlimited snacks and beverages
+2
AI Infrastructure Architect — Scalable GPU Compute
AI Infrastructure Architect — Scalable GPU Compute

EngineersOfAI • Sunnyvale (CA)

On-site
USD 150,000 - 200,000
Remote AI Inference Platform Architect
Remote AI Inference Platform Architect

Utilidata • United States

Remote
USD 170,000 - 210,000
Stock options
Health insurance
401k match
+1
Senior AI Inference Cloud Engineer - Kubernetes
Senior AI Inference Cloud Engineer - Kubernetes

Arm • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Relocation package with visa support