Senior Backend Engineer - GPU Cloud Platform

Lightning AI

San Francisco, Seattle, New York (CA, WA, NY)

Hybride

USD 180 000 - 250 000

Plein temps

14 jours+
Générateur de candidature

Une candidature conçue pour ce poste — un CV et une lettre de motivation personnalisés qui correspondent à l’offre.

Passez les filtres ATS

Avantages offerts par ce poste

Comprehensive Health Coverage
Meaningful Equity
401(k) matching
Unlimited PTO
Hybrid work model
In-Office meals

Résumé du poste

Lightning AI is seeking a Senior Backend Engineer for the Managed Services team to design, build, and operate backend services that power our GPU infrastructure platform. You will develop control planes and automation to provision and scale Kubernetes and Slurm clusters across a distributed environment.

This role focuses on distributed systems, cloud-native architecture, and platform operations, with a hybrid office setup in SF/NYC/Seattle and a minimum of two in-office days per week.

Qualifications

  • Significant professional experience designing, building, and operating production backend systems using Go or Python.
  • Deep hands-on experience with Kubernetes or Slurm, including operating large-scale production environments.
  • Strong understanding of distributed systems, cloud-native architectures, and production infrastructure.
  • Experience designing and building scalable backend services, APIs, and automation for infrastructure or platform operations.
  • Strong understanding of networking, storage, and cloud infrastructure fundamentals.
  • Familiarity with observability, CI/CD, testing, production operations, and incident response.
  • Ability to own complex technical projects while collaborating effectively across engineering teams.

Responsabilités

  • Design, build, and operate backend services in Go or Python that power Lightning AI's managed infrastructure platform.
  • Develop control plane services that provision, orchestrate, and manage Kubernetes and Slurm clusters across large-scale GPU infrastructure.
  • Build distributed systems that automate cluster lifecycle management, workload scheduling, infrastructure provisioning, and platform operations.
  • Develop platform capabilities using Kubernetes APIs, controllers, operators, and other cloud-native technologies.
  • Improve the reliability, scalability, security, and observability of our managed platform through automation and operational excellence.
  • Diagnose and resolve complex production issues across Kubernetes, distributed systems, networking, and cloud infrastructure.
  • Collaborate with infrastructure, AI, and platform engineering teams to shape the future of our cloud platform.

Connaissances

Go
Python
Kubernetes
Slurm
Distributed systems

Description du poste

Lightning AI is seeking a Senior Backend Engineer for the Managed Services team to design, build, and operate backend services that power our GPU infrastructure platform. You will develop control planes and automation to provision and scale Kubernetes and Slurm clusters across a distributed environment.

This role focuses on distributed systems, cloud-native architecture, and platform operations, with a hybrid office setup in SF/NYC/Seattle and a minimum of two in-office days per week.

Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

Senior Backend Engineer — Kubernetes & Slurm Infra (Go/Python)
Senior Backend Engineer — Kubernetes & Slurm Infra (Go/Python)

Lightning AI • New York (NY)

Hybride
USD 180 000 - 250 000
Health Coverage
Equity
401(k) matching
+8
Senior GPU Infra Lead: Slurm, Kubernetes & Platform
Senior GPU Infra Lead: Slurm, Kubernetes & Platform

Jobgether SRL • États-Unis

À distance
USD 170 000 - 250 000
Product Lead, GPU Cloud & AI Infrastructure
Product Lead, GPU Cloud & AI Infrastructure

Lightning AI • New York (NY), San Francisco (CA)

Hybride
USD 200 000 - 250 000
Health coverage
Equity opportunity
401(k) + pension
Senior Cloud Platform Engineer, GPU Core & Lifecycle
Senior Cloud Platform Engineer, GPU Core & Lifecycle

Lambda Labs • États-Unis

Hybride
USD 180 000 - 240 000
Health insurance
Dental insurance
Vision insurance
+5
Senior Backend Engineer - Microservices & Cloud Platform
Senior Backend Engineer - Microservices & Cloud Platform

Lightning AI • San Francisco (CA)

Hybride
USD 188 000 - 254 000
Comprehensive Health Coverage
Meaningful Equity
401(k) matching
+1
Senior Cloud Platform Engineer - AI GPU Infrastructure
Senior Cloud Platform Engineer - AI GPU Infrastructure

Showcify, Inc. • États-Unis

À distance
USD 180 000 - 260 000
Senior GPU Data Center Engineer
Senior GPU Data Center Engineer

Prime Intellect AI • San Francisco (CA)

Sur place
USD 150 000 - 300 000
Senior Platform Engineer — GPU Cloud Orchestration
Senior Platform Engineer — GPU Cloud Orchestration

STN Incorporated • États-Unis

Hybride
USD 140 000 - 180 000
Senior Backend Engineer, Managed Services
Senior Backend Engineer, Managed Services

Lightning AI • New York (NY)

Sur place
USD 180 000 - 250 000
Health Coverage
Equity
401(k) matching
+8
Senior GPU Cloud Infrastructure Engineer
Senior GPU Cloud Infrastructure Engineer

Hyperbolic • San Francisco (CA)

Sur place
USD 180 000 - 240 000