Senior Cloud Platform Engineer - GPU Infrastructure

Socket.dev

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health, dental, and vision coverage
Wellness stipend
Commuter stipend
401k with company match

Job summary

Lambda, The Superintelligence Cloud, is seeking a Senior Software Engineer for the Core Cloud Platform to build control-plane systems powering its GPU cloud. You will develop APIs, workflows, schedulers, and orchestration services that turn physical GPU infrastructure into reliable, customer-facing cloud capacity.

This role requires in-office presence in San Francisco/San Jose four days per week, with Tuesday as the designated work-from-home day.

Qualifications

  • Bachelor's degree or equivalent working experience.
  • 6+ years of professional software engineering experience building production backend or distributed systems.
  • Strong in Python, Go, or a similar backend/system language.
  • Experience designing and operating APIs, workflow engines, schedulers, orchestration services, or other distributed systems.
  • Understand reliability fundamentals: fault tolerance, idempotency, retries, state machines, failure handling, and production debugging.
  • Experience with cloud or cloud-like infrastructure primitives such as compute, networking, storage, capacity management, identity, or fleet operations.
  • Comfortable with Linux, containers, Kubernetes, infrastructure automation, and service deployment patterns.
  • Owned production services, participated in on-call, and improved systems based on operational learnings.
  • Care about testability, CI/CD, observability, metrics, logging, alerting, and supportable operations.
  • Can take ambiguous infrastructure problems and drive them to clear designs, implementation plans, and production outcomes.
  • Communicate clearly across engineering, product, support, infrastructure teams, and leadership.

Responsibilities

  • Build the control-plane systems that power Lambda’s GPU cloud.
  • Develop APIs, workflows, schedulers, orchestration services, and operational surfaces.
  • Ensure deployment readiness, reliability, and tooling for production environments.
  • Collaborate with cross-functional teams across engineering, product, and support.

Skills

Python
Go
Distributed systems
API design
Kubernetes
Linux
On-call experience

Education

Bachelor's degree or equivalent

Job description

Lambda, The Superintelligence Cloud, is seeking a Senior Software Engineer for the Core Cloud Platform to build control-plane systems powering its GPU cloud. You will develop APIs, workflows, schedulers, and orchestration services that turn physical GPU infrastructure into reliable, customer-facing cloud capacity.

This role requires in-office presence in San Francisco/San Jose four days per week, with Tuesday as the designated work-from-home day.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Platform Engineer for AI GPU Infra
Senior Cloud Platform Engineer for AI GPU Infra

Lambda • San Jose (CA)

On-site
USD 170,000 - 260,000
Health, dental, and vision coverage
Wellness and commuter stipends
401k Plan with company match
+1
Senior Cloud Platform Engineer: Core GPU Infrastructure
Senior Cloud Platform Engineer: Core GPU Infrastructure

Lambda • San Francisco (CA)

Hybrid
USD 296,000 - 346,000
Health insurance
Dental insurance
Vision insurance
+2
Senior Cloud Platform Engineer, GPU Core & Lifecycle
Senior Cloud Platform Engineer, GPU Core & Lifecycle

Lambda Labs • United States

Hybrid
USD 180,000 - 240,000
Health insurance
Dental insurance
Vision insurance
+5
Senior Cloud Platform Engineer - GPU Infra, Hybrid
Senior Cloud Platform Engineer - GPU Infra, Hybrid

Neura Market • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Wellness stipend
Commuter stipend
401k with company match
Senior Cloud Platform Engineer — GPU Compute Orchestration
Senior Cloud Platform Engineer — GPU Compute Orchestration

Lambda Inc. • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Health, dental, and vision coverage
401(k) plan with company match (USA)
Senior Cloud Infrastructure Engineer – GPU & DPU
Senior Cloud Infrastructure Engineer – GPU & DPU

Lambda • United States

Remote
USD 180,000 - 260,000
Senior HPC Systems Architect: Liquid-Cooled GPU Cloud
Senior HPC Systems Architect: Liquid-Cooled GPU Cloud

Socket.dev • San Jose (CA)

Hybrid
USD 180,000 - 260,000
Senior GPU Cloud Solutions Engineer
Senior GPU Cloud Solutions Engineer

Lambda • San Francisco (CA)

Hybrid
USD 170,000 - 210,000
Health, dental, and vision coverage
401k with 2% company match
Wellness and commuter stipends
+1
Senior GPU Cloud Solutions Architect
Senior GPU Cloud Solutions Architect

Lambda • United States

Hybrid
USD 180,000 - 280,000
Staff Compute Platform Engineer – AI Cloud (Remote)
Staff Compute Platform Engineer – AI Cloud (Remote)

Applied Methods Ltd • Bellevue (WA), Northern (KY)

Hybrid
USD 190,000 - 260,000
Health coverage
Dental coverage
Vision coverage
+2