Systems Engineer II, Compute

ProducePay

San Francisco (CA)

On-site

USD 137,000 - 161,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance package options
401(k) with 100% match up to 4%
Generous paid time off
Tuition reimbursement
Cell phone reimbursement

Job summary

ProducePay, located in San Francisco, CA, is seeking a passionate Senior/Staff Software Engineer specializing in Systems Applications. This role focuses on design and development for their AI compute platform, emphasizing the construction of virtualization applications. The ideal candidate will have experience with the Linux kernel, hardware integration, and a solid understanding of performance analysis. The position includes comprehensive benefits, competitive salary range of $137,000 - $161,000, and an opportunity to work in a transformative cloud environment.

Qualifications

  • Experience building applications on Linux kernels, specifically pertaining to virtualization and device drivers.
  • Solid understanding of hardware devices such as GPUs, CPUs, Infiniband and Ethernet NICs.
  • Strong grasp of distributed applications and highly-scalable systems design.
  • Strong experience building software applications at both higher and lower levels.
  • Ability to collaborate with teams across an organization effectively.
  • Capable of adapting quickly and eager to research new technology.
  • General knowledge of hypervisors and virtual machine lifecycles.
  • Understanding how to build CI/CD pipelines delivering bug-free code.

Responsibilities

  • Design highly reliable and performant Linux applications to manage virtualization stack.
  • Integrate applications with various hardware and software AI stacks.
  • Collaborate with kernel and hypervisor teams for seamless integration.
  • Analyze and enhance performance of the virtualization stack.
  • Troubleshoot complex issues across virtualization stack.
  • Conduct thorough code reviews to ensure software quality.
  • Collaborate with hardware design and AI/ML teams.
  • Provide technical guidance and mentorship to junior engineers.

Skills

Linux Systems Familiarity
Hardware Integration
Systems Design
Software Architecture
Excellent Communication Skills
Rapid and Agile Learner
Virtualization Concepts
CI/CD and Validation

Job description

Location

San Francisco, CA - US

Employment Type

Full time

Location Type

On-site

Department

Cloud Engineering

Crusoe's mission is to accelerate the abundance of energy and intelligence. We’re crafting the engine that powers a world where people can create ambitiously with AI — without sacrificing scale, speed, or sustainability.

Be a part of the AI revolution with sustainable technology at Crusoe. Here, you'll drive meaningful innovation, make a tangible impact, and join a team that’s setting the pace for responsible, transformative cloud infrastructure.

About This Role:

The Crusoe Cloud Software Development team is seeking a passionate and experienced Senior/Staff Software Engineer specializing in Systems Applications. This pivotal role is critical in the design and development of our compute platform, specifically focusing on building compute applications for virtualized AI-platforms. An understanding of the linux kernel, virtualization, hardware tuning, distributed systems, object oriented programming, and low-level systems programming are critical to this role. Excellent communication skills and a desire to work with a wide range of technologies across the linux stack are both a must. This is a full-time position.

What You’ll Be Working On:
  • Compute Application Development & Scaleout: Design highly reliable and performant Linux applications used to manage our virtualization stack across thousands of AI compute servers in multiple global datacenters.

  • AI Hardware Platform Integration: Integrate Crusoe applications with a wide variety of hardware and software AI chip-vendor stacks. Build solutions to optimize and monitor virtualized hardware (GPUs, Infiniband/ROCe NICs, Ephemeral Storage, etc.) in cutting-edge AI/HPC environments.

  • Kernel & Hypervisor Integration - Work side by side with our Linux Kernel and Hypervisor teams to ensure our Crusoe applications are seamlessly integrated with a variety of kernels and hypervisors.

  • Performance Analysis & Tuning: Analyze and enhance the performance of the entire virtualization stack, from the hypervisor to the virtualized guest OS, with a specific focus on optimizing AI/ML workloads. This includes profiling, bottleneck identification, and implementing low-level optimizations.

  • System-Level Troubleshooting: Diagnose and resolve complex system issues across our virtualization stack (drivers, kernel, hypervisor, guest OS, and crusoe applications). Work closely with kernel and hypervisor teams to debug and resolve integration challenges.

  • Code Review and Quality Assurance: Conduct thorough code reviews to ensure the highest level of software quality, reliability, and security within compute applications and virtualization stack.

  • Cross-Functional Collaboration: Collaborate with other engineering teams, including hardware design, OS development, and AI/ML application teams, to ensure cohesive and integrated product development.

  • Technical Leadership: Provide technical guidance and mentorship to junior engineers, fostering a culture of technical excellence and collaborative problem-solving within the compute applications team.

What You’ll Bring to the Team:
  • Linux Systems Familiarity: Experience building applications on Linux kernels, specifically pertaining to virtualization, device drivers, memory management, and process scheduling.

  • Hardware Integration: Solid understanding of hardware devices such as GPUs, CPUs, Infiniband and Ethernet NICs, Ephemeral Disks, and PCI Express.

  • Systems Design: Strong grasp of distributed applications and highly-scalable systems design. Specific focus around communications protocols (GRPC, REST, TCP/IP, etc.), databases (Postgres, Redis), and systems design applications (Pub/Sub, Kafka).

  • Software Architecture: Strong experience building software applications, both at the higher (Golang, Java, Python) and lower (C, C++, Rust) levels. Keen eye for clean, maintainable code, and a unit-test driven mindset.

  • Excellent Communication Skills: Ability to collaborate with teams across an organization, blocking out noise, and focusing on what needs to get done to get a project across the line.

  • Rapid and Agile Learner: Capable of adapting quickly, eager to research new technology and not get overwhelmed by unfamiliar tech stacks.

  • Virtualization Concepts: General knowledge of hypervisors, virtual machine lifecycles, and Linux KVM tooling.

  • CI/CD and Validation: Understanding of how to build Gitlab or Github CI/CD pipelines that deliver bug-free code across a multitude of compute platforms.

Bonus Points:
  • Experience with virtualization specifically for AI/ML workloads, including GPU virtualization.

  • Previous work debugging or contributing to kernel or hypervisor code, particularly around device management.

  • Experience with configuring thousands of live compute nodes in a bare-metal production environment.

Benefits:
  • Industry competitive pay

  • Restricted Stock Units in a fast growing, well-funded technology company

  • Health insurance package options that include HDHP and PPO, vision, and dental for you and your dependents

  • Employer contributions to HSA accounts

  • Paid Parental Leave

  • Paid life insurance, short-term and long-term disability

  • Teladoc

  • 401(k) with a 100% match up to 4% of salary

  • Generous paid time off and holiday schedule

  • Cell phone reimbursement

  • Tuition reimbursement

  • Subscription to the Calm app

  • MetLife Legal

  • Company paid Commuter FSA benefit of $300 per month

Compensation:

Compensation will be paid in the range of $137,000 - $161,000. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant’s education, experience, knowledge, skills, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Hardware Systems Engineer
Senior Hardware Systems Engineer

Crusoe • Sunnyvale (CA)

On-site
USD 170,000 - 205,000
Health insurance
401(k) with match
Employee stock options
+2
Staff Hardware Systems Engineer, Performance
Staff Hardware Systems Engineer, Performance

Crusoe • San Francisco (CA)

On-site
USD 215,000 - 260,000
RSUs
Health insurance
401(k) with match
+3
Senior Hardware Systems Engineer, Performance
Senior Hardware Systems Engineer, Performance

Crusoe • Sunnyvale (CA)

On-site
USD 170,000 - 205,000
Health insurance
RSUs
401(k) match
+2
Senior Production Engineer, Compute
Senior Production Engineer, Compute

Crusoe Energy Systems LLC • Sunnyvale (CA)

On-site
USD 170,000 - 205,000
Industry competitive pay
RSUs in a fast growing tech company
Health insurance options including HDP
+1
Senior Production Engineer, Compute
Senior Production Engineer, Compute

crusoe • Sunnyvale (CA)

On-site
USD 170,000 - 205,000
Health insurance
RSUs
401(k) match
+7
Staff Cloud Support Engineer
Staff Cloud Support Engineer

Epoch Biodesign • San Francisco (CA)

On-site
USD 180,000 - 220,000
Health insurance
401(k) with employer match
Paid Parental Leave
+2
Senior Staff Deployment Automation Engineer
Senior Staff Deployment Automation Engineer

Crusoe • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive compensation and equity
Paid time off
Health, dental & vision insurance
+3
Senior Staff Deployment Automation Engineer
Senior Staff Deployment Automation Engineer

ProducePay • United States

On-site
USD 250,000 - 300,000
Competitive compensation
Equity packages
Paid time off
+5
Staff Production Engineer, Compute
Staff Production Engineer, Compute

Crusoe • San Francisco (CA)

On-site
USD 209,000 - 253,000
RSUs
PTO & holidays
Health insurance
+2
Principal Systems Software Engineer
Principal Systems Software Engineer

Crusoe Energy Systems LLC • San Francisco (CA)

On-site
USD 260,000 - 340,000
Competitive compensation
Restricted Stock Units
Paid time off
+10