Senior Backend Engineer: AI GPU Orchestration & Scheduling

MetAntz

Palo Alto (CA)

Hybrid

USD 150,000 - 190,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Hybrid work model
Equity
Competitive salary

Job summary

Zettabyte is looking for a Backend Engineer to build systems that orchestrate GPU clusters for AI workloads. You will design APIs to abstract complex GPU operations, develop scheduling to maximize utilization, and manage GPU lifecycles from provisioning to release.

You will work across GPU resource tracking, billing, monitoring, and multi-tenancy while collaborating with frontend teams. This hybrid role requires strong backend expertise and a startup mindset.

Qualifications

  • 5+ years backend engineering with distributed systems experience.
  • Strong proficiency in Go and Python or similar languages.
  • Experience designing APIs (REST, GraphQL, gRPC) and building scalable services.

Responsibilities

  • Design APIs that abstract GPU operations into simple developer experiences.
  • Build scheduling algorithms to maximize GPU utilization while meeting SLA.
  • Develop resource management for GPU lifecycle: provisioning, allocation, scheduling, release.
  • Create usage tracking and billing systems for GPU-hours and memory usage.
  • Implement monitoring for GPU metrics, health checks, and failure recovery.
  • Build multi-tenancy with resource isolation and fair scheduling.
  • Collaborate with frontend engineers to expose infrastructure through intuitive interfaces.
  • Leverage AI-assisted coding tools to boost productivity and code quality.

Skills

Go
Python
Distributed systems
APIs (REST/GraphQL/gRPC)
Linux

Tools

Docker
Kubernetes
Git

Job description

Zettabyte is looking for a Backend Engineer to build systems that orchestrate GPU clusters for AI workloads. You will design APIs to abstract complex GPU operations, develop scheduling to maximize utilization, and manage GPU lifecycles from provisioning to release.

You will work across GPU resource tracking, billing, monitoring, and multi-tenancy while collaborating with frontend teams. This hybrid role requires strong backend expertise and a startup mindset.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior/Staff Backend Engineer - Distributed System
Senior/Staff Backend Engineer - Distributed System

Zettabyte Inc • Palo Alto (CA)

Hybrid
USD 180,000 - 240,000
Competitive salary
Equity based on experience
Hybrid work model
Senior, Staff Backend Engineer - Distributed System
Senior, Staff Backend Engineer - Distributed System

MetAntz • Palo Alto (CA)

On-site
USD 150,000 - 190,000
Hybrid work model
Equity
Competitive salary
Senior, Staff Backend Engineer - Distributed System
Senior, Staff Backend Engineer - Distributed System

SproutsAI • Palo Alto (CA)

On-site
USD 190,000 - 270,000
Competitive salary
Equity based on experience
Senior Backend Engineer AI Infra & GPU Scheduling (Hybrid)
Senior Backend Engineer AI Infra & GPU Scheduling (Hybrid)

SproutsAI • Palo Alto (CA)

Hybrid
USD 190,000 - 270,000
Competitive salary
Equity based on experience
Senior Backend Engineer – AI Infra & GPU Scheduling
Senior Backend Engineer – AI Infra & GPU Scheduling

Zettabyte Inc • Palo Alto (CA)

Hybrid
USD 180,000 - 240,000
Competitive salary
Equity based on experience
Hybrid work model
Senior Backend Engineer, GPU Cloud Infra & Kubernetes
Senior Backend Engineer, GPU Cloud Infra & Kubernetes

Socket.dev • New York (NY)

Hybrid
USD 180,000 - 250,000
Health insurance
Equity
401(k) matching
+6
Senior GPU Cluster Infra Engineer | Remote
Senior GPU Cluster Infra Engineer | Remote

AISafety • Berkeley (CA)

Hybrid
USD 120,000 - 180,000
Health Insurance
401(k) match
PTO 25 days per year
+3
Senior GPU Cloud Infrastructure Engineer
Senior GPU Cloud Infrastructure Engineer

Hyperbolic • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior AI Scheduling & GPU Orchestration Engineer
Senior AI Scheduling & GPU Orchestration Engineer

Bitdeer (NASDAQ: BTDR) • San Jose (CA)

On-site
USD 180,000 - 240,000
Senior GPU Infrastructure Engineer — HPC & Clusters
Senior GPU Infrastructure Engineer — HPC & Clusters

Prime Intellect AI • San Francisco (CA)

On-site
USD 150,000 - 300,000