Senior Backend Engineer - AI GPU Infra & Scheduling

zettabyte-space

United States

Hybrid

USD 180,000 - 260,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Competitive salary and equity
Hybrid work model (3 days in office, 2
days WFH

Job summary

Zettabyte is hiring a Backend Engineer to design APIs that expose GPU orchestration and resource management for AI workloads. You will work on scheduling, multi-tenant isolation, and billing systems while collaborating with frontend engineers to create intuitive interfaces.

The role emphasizes efficient GPU utilization, SLA adherence, and handling high-value compute resources in a fast-paced startup environment located in Palo Alto.

Qualifications

  • 5+ years backend engineering experience with distributed systems.
  • Proficiency in Go, Python, or similar backend languages.
  • Experience with resource scheduling, orchestration, and API design (REST, GraphQL, gRPC).
  • Understanding of hardware constraints and system optimization.
  • Linux systems knowledge and containerization experience (Docker, Kubernetes).
  • Startup mindset—comfortable with ambiguity and rapid iteration.

Responsibilities

  • Design APIs that abstract complex GPU operations into simple developer experiences.
  • Build scheduling algorithms that maximize GPU utilization while ensuring SLA compliance.
  • Develop resource management systems for GPU lifecycle—provisioning, allocation, scheduling, and release.
  • Create usage tracking and billing systems for GPU-hours, memory usage, and compute utilization.
  • Implement monitoring for GPU-specific metrics, health checks, and automatic failure recovery.
  • Build multi-tenancy systems with resource isolation, quota management, and fair scheduling.

Skills

Backend engineering
Distributed systems
Go
Python
API design
Linux

Tools

Docker
Kubernetes

Job description

Zettabyte is hiring a Backend Engineer to design APIs that expose GPU orchestration and resource management for AI workloads. You will work on scheduling, multi-tenant isolation, and billing systems while collaborating with frontend engineers to create intuitive interfaces.

The role emphasizes efficient GPU utilization, SLA adherence, and handling high-value compute resources in a fast-paced startup environment located in Palo Alto.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Backend Engineer: AI GPU Orchestration & Scheduling
Senior Backend Engineer: AI GPU Orchestration & Scheduling

MetAntz • Palo Alto (CA)

Hybrid
USD 150,000 - 190,000
Hybrid work model
Equity
Competitive salary
Senior/Staff Backend Engineer - Distributed System
Senior/Staff Backend Engineer - Distributed System

zettabyte-space • United States

Hybrid
USD 180,000 - 260,000
Competitive salary and equity
Hybrid work model (3 days in office, 2
days WFH
Senior, Staff Backend Engineer - Distributed System
Senior, Staff Backend Engineer - Distributed System

MetAntz • Palo Alto (CA)

On-site
USD 150,000 - 190,000
Hybrid work model
Equity
Competitive salary
Senior/Staff Backend Engineer - Distributed System
Senior/Staff Backend Engineer - Distributed System

Zettabyte Inc • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Competitive salary
Equity based on experience
Hybrid work model
Senior, Staff Backend Engineer - Distributed System
Senior, Staff Backend Engineer - Distributed System

SproutsAI • Palo Alto (CA)

On-site
USD 190,000 - 270,000
Competitive salary
Equity based on experience
Senior Backend Engineer AI Infra & GPU Scheduling (Hybrid)
Senior Backend Engineer AI Infra & GPU Scheduling (Hybrid)

SproutsAI • Palo Alto (CA)

Hybrid
USD 190,000 - 270,000
Competitive salary
Equity based on experience
Senior Backend Engineer – AI Infra & GPU Scheduling
Senior Backend Engineer – AI Infra & GPU Scheduling

Zettabyte Inc • Palo Alto (CA)

Hybrid
USD 180,000 - 240,000
Competitive salary
Equity based on experience
Hybrid work model
New Grad Software Engineer - Build AI Infra & Apps
New Grad Software Engineer - Build AI Infra & Apps

SproutsAI • Palo Alto (CA), Northern (KY)

Hybrid
USD 120,000 - 150,000
Software Engineering Intern (Summer 2026)
Software Engineering Intern (Summer 2026)

zettabyte-space • United States

Hybrid
USD 28,000 - 48,000
Intern stipend
Mentorship from senior engineers
Potential full-time transition
+1
GPU Cloud Infrastructure Intern — AI, Kubernetes, Hybrid
GPU Cloud Infrastructure Intern — AI, Kubernetes, Hybrid

zettabyte-space • United States

Hybrid
USD 28,000 - 48,000
Intern stipend
Mentorship from senior engineers
Potential full-time transition
+1