Software Engineer, Distributed Systems - San Francisco, California

MissionHires

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Unlimited office book budget
Generous equity grant
Retirement matching
Medical, dental & vision insurance
Unlimited paid time off
Weekly lunch coverage
Visa sponsorship

Job summary

MissionHires in San Francisco is looking for a distributed systems software engineer to develop our in-house resource orchestration system for GPU compute nodes. This role entails designing resilient architectures for high availability and performance optimization.

The ideal candidate has experience in managing distributed systems, Linux virtualization, and appreciates solid documentation practices. We offer competitive salary, generous equity grants, and various benefits including unlimited time off and health insurance.

Qualifications

  • Built fault tolerant distributed systems managing hardware resources at scale.
  • Experience with Linux virtualization tools such as Cloud Hypervisor and QEMU.
  • Strong documentation skills and system reliability.

Responsibilities

  • Design distributed system architectures for high availability and fault tolerance.
  • Automate deployment and optimize performance of GPU virtual machines.
  • Develop multi-tier high performance network attached storage systems.

Skills

Distributed systems design
Fault tolerant systems
Linux virtualization
Documentation

Tools

Cloud Hypervisor
QEMU
libvirt
Rust
etcd

Job description

As a distributed systems software engineer, you’ll be working on our in-house resource orchestration system. This system coordinates state and access to hundreds (soon thousands) of GPU compute nodes in multi-tenant clusters spanning across multiple data centers.

Responsibilities
  • Design of distributed system architectures that enable high availability fault tolerant state management
  • Deployment automation and performance optimization of virtual machines running on bare metal that utilize GPU passthrough
  • Design and deployment of multi-tier high performance network attached storage systems
Requirements
  • You have built fault tolerant distributed systems before that can manage hardware resources at scale
  • You enjoy creating self-correcting systems that contribute to hardware health and reliability
  • You have experience with Linux virtualization (Cloud Hypervisor, QEMU, libvirt, virtiofs, sr-iov, PCIe passthrough)
  • You appreciate and value good documentation
Nice to Haves
  • Experience with Rust (our VM orchestrator is written in Rust)
  • Experience with etcd
  • Experience with high performance storage systems (WEKA, VAST, Ceph, etc.)
Benefits
  • Unlimited office book budget: You can buy as many books for the office as you want. You’re encouraged to spend time during the workday reading!
  • Generous equity grant: Team members are offered a competitive salary along with equity in the company
  • Retirement matching: We match 401(k) plans up to 4%
  • Medical, dental & vision: We offer competitive medical, dental, vision insurance for employees and dependents and cover 100% of premiums
  • Time off: We offer unlimited paid time off as well as 10+ observed holidays
  • Parental leave: We offer biological, adoptive, and foster parents paid time off to spend quality time with family
  • Daily lunch: We cover lunch daily for employees
  • Visa Sponsorships: Yes, we sponsor visas and work permits
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Infrastructure
Software Engineer, Infrastructure

Fal • San Francisco (CA)

On-site
USD 180,000 - 250,000
Health, dental, and vision insurance
Learning and growth opportunities
Visa sponsorship and relocation assistance
+1
HPC/ GPU Cluster Architect
HPC/ GPU Cluster Architect

Electric Capital • San Francisco (CA)

Hybrid
USD 220,000 - 300,000
Generous equity grant
401(k) matching
Comprehensive medical, dental, and vision insurance
+3
HPC/ GPU Hardware Engineer
HPC/ GPU Hardware Engineer

The San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 180,000 - 260,000
Generous equity grant
Retirement matching
Comprehensive medical, dental, and vision insurance
+5
Distributed Systems Engineer: GPU Orchestration & HA
Distributed Systems Engineer: GPU Orchestration & HA

MissionHires • San Francisco (CA)

On-site
USD 120,000 - 160,000
Unlimited office book budget
Generous equity grant
Retirement matching
+4
Software Engineer, Virtualization
Software Engineer, Virtualization

Kindredventures • San Francisco (CA)

On-site
USD 180,000 - 250,000
Health, dental, and vision insurance
Learning and growth opportunities
Regular team events
+1
HPC/ GPU Cluster Architect
HPC/ GPU Cluster Architect

The Consensus • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Visa sponsorships
401(k) retirement matching
Medical, dental & vision insurance
+2
Software Engineer (C++ Systems)
Software Engineer (C++ Systems)

SK HR Consultants.com • San Francisco (CA)

On-site
USD 100,000 - 140,000
Relocation packages available
High-performance work environment
HPC/ GPU Cluster Architect
HPC/ GPU Cluster Architect

San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Generous equity grant
Competitive salary
Visa sponsorship
+6
Backend Software Engineer (Distributed Systems / Python)
Backend Software Engineer (Distributed Systems / Python)

Glint Tech Solutions • San Francisco (CA)

Hybrid
USD 170,000 - 230,000
Competitive equity package
Comprehensive benefits
Member of Technical Staff - Distributed Systems
Member of Technical Staff - Distributed Systems

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000