Compute Infra Engineer — GPU/VM, Scale & Reliability

Lambda

Bellevue (WA)

On-site

USD 150,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health, dental, and vision coverage
401k Plan with 2% company match
Generous cash & equity compensation
Wellness and commuter stipends
Flexible paid time off

Job summary

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure with a mission to make compute ubiquitous. This Compute Software Engineering role focuses on provisioning for bare-metal and VM infrastructure, improving workflows, and accelerating delivery velocity.

You will work across teams to scale engineering productivity and reliability. Based in an office location in Bellevue or other listed Bay Area/Bellevue sites, you will design and implement high-performance compute software,

Qualifications

  • 3+ years of experience with Go or Python in production.
  • 3+ years of experience with bare metal & virtualization hardware management.
  • Comfortable in Linux environments and debugging at OS, hardware, and networking layers.
  • Able to troubleshoot complex systems and communicate across software, infrastructure, and vendor teams.

Responsibilities

  • Design, develop and maintain software for GPU/CPU compute infrastructure with focus on performance, scalability, and reliability.
  • Implement and develop services for baremetal and VM instancing.
  • Develop distributed systems for managing and orchestrating compute resources across various SKU’s.
  • Troubleshoot and debug complex issues in a production and development environment.
  • On-call and incident ownership
  • Collaborate across multiple teams and drive ambiguity in requirements or solutions on RFCs.

Skills

Go
Python
Bare metal
Virtualization
Linux
Troubleshooting

Tools

Slurm
Kubernetes
KVM
QEMU
Temporal

Job description

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure with a mission to make compute ubiquitous. This Compute Software Engineering role focuses on provisioning for bare-metal and VM infrastructure, improving workflows, and accelerating delivery velocity.

You will work across teams to scale engineering productivity and reliability. Based in an office location in Bellevue or other listed Bay Area/Bellevue sites, you will design and implement high-performance compute software,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Center Systems Engineer for GPU Infrastructure
Data Center Systems Engineer for GPU Infrastructure

Lambda • Quincy (WA)

On-site
USD 70,000 - 110,000
Health insurance
Dental coverage
Vision coverage
+2
GPU-First Cloud Compute Engineer (Hybrid)
GPU-First Cloud Compute Engineer (Hybrid)

Lambda • San Francisco (CA)

Hybrid
USD 266,000 - 395,000
Health, dental, and vision coverage
Wellness stipend
Commuter stipend
+1
Cloud Compute Engineer - GPU/VM Infra & Bare-Metal
Cloud Compute Engineer - GPU/VM Infra & Bare-Metal

Lambda Inc. • San Francisco (CA)

On-site
USD 150,000 - 230,000
Health, dental, and vision coverage
Equity compensation
401(k) plan with company match
+3
Software Engineer - Compute
Software Engineer - Compute

Lambda • San Francisco (CA)

Hybrid
USD 266,000 - 395,000
Health, dental, and vision coverage
Wellness stipend
Commuter stipend
+1
Software Engineer - Compute
Software Engineer - Compute

Lambda • Bellevue (WA)

On-site
USD 150,000 - 210,000
Health, dental, and vision coverage
401k Plan with 2% company match
Generous cash & equity compensation
+2
Senior Cloud Platform Engineer - GPU Infra, Hybrid
Senior Cloud Platform Engineer - GPU Infra, Hybrid

Neura Market • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Wellness stipend
Commuter stipend
401k with company match
Staff Compute Platform Engineer – AI Cloud (Remote)
Staff Compute Platform Engineer – AI Cloud (Remote)

Applied Methods Ltd • Bellevue (WA), Northern (KY)

Hybrid
USD 190,000 - 260,000
Health coverage
Dental coverage
Vision coverage
+2
Senior Cloud Platform Engineer for AI GPU Infra
Senior Cloud Platform Engineer for AI GPU Infra

Lambda • San Jose (CA)

On-site
USD 170,000 - 260,000
Health, dental, and vision coverage
Wellness and commuter stipends
401k Plan with company match
+1
Software Engineer - Compute
Software Engineer - Compute

Socket.dev • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Health, dental, vision
401k with match
Wellness stipends
+2
GPU/VM Compute Engineer (Hybrid: 4 days on-site)
GPU/VM Compute Engineer (Hybrid: 4 days on-site)

Lambda Labs • United States

Hybrid
USD 130,000 - 180,000
Health, dental and vision coverage
Wellness stipend
Commuter stipend
+3