Senior Virtualization Validation Engineer

Crusoe

San Francisco (CA)

On-site

USD 172,500 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading AI infrastructure company in San Francisco is looking for a Senior Systems Performance Engineer to lead the hardware evaluation and scaling of their AI infrastructure. This role involves defining the performance roadmap for their next-generation cloud, focusing on optimizing system efficiency. Ideal candidates have over 5 years of experience with large-scale GPU systems, strong programming skills in Python and C++, and a deep understanding of hardware architectures. The compensation ranges from $172,500 to $210,000 annually.

Qualifications

  • 5+ years experience in end-to-end hardware evaluation, reliability, and scaling of AI infrastructure.
  • Experience with large-scale GPU infrastructure optimization.
  • Knowledge of performance modeling for secure environments.

Responsibilities

  • Lead the evaluation and establishment of New Product Introduction across hardware architectures.
  • Conduct deep-dive performance evaluations and workload characterizations.
  • Design and implement 0-to-1 performance methodologies for scaling.

Skills

Expert-level proficiency in Python
Expert-level proficiency in C++
Proven experience in building and optimizing AI application systems
Deep knowledge of x86 and ARM architectures
Ability to write and debug ARMv8 assembly

Tools

Lauterbach Trace32
ARM DS-5

Job description

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world’s most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem‑solving, opportunity‑finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high‑performing team that believes in each other, come build with us at Crusoe.

Senior Systems Performance Engineer

San Francisco, Sunnyvale (Onsite)

Role Mission

At Crusoe, we are pioneering the future of sustainable computing. We are seeking a Senior Performance Engineer to serve as a technical lead for the end‑to‑end hardware evaluation, reliability, and scaling of our AI infrastructure. You will be responsible for defining the performance roadmap of our next‑generation cloud, ensuring that our SOTA (State‑of‑the‑Art) AI models run with peak efficiency across diverse hardware architectures.

What You’ll Be Working On:
  • Architectural Strategy: Lead the evaluation and establishment of New Product Introduction (NPI) across varied hardware architectures, focusing on Bare Metal and VM environments.
  • Full‑Stack Optimization: Conduct deep‑dive performance evaluations and workload characterizations across compute, memory, storage, and networking.
  • Performance Modeling: Develop sophisticated multivariable projection models and frameworks to analyze system design options through KPI tradeoffs, such as Power and TCO (Total Cost of Ownership).
  • Hardware‑Software Co‑Design: Collaborate with external vendors to drive platform customization and optimize server/AI architectures for maximum performance‑per‑TCO.
  • Infrastructure Scaling: Design and implement 0‑to‑1 performance methodologies that allow the team to scale evaluation processes for large‑scale GPU/AI data centers.
  • Industry Leadership: Actively engage in industry research and contribute technical insights to consortiums and standards committees to influence future hardware roadmaps.
What You’ll Bring to the Team:
  • 5+ Years experience in end‑to‑end hardware evaluation, reliability, and scaling of our AI infrastructure
  • Large‑Scale Systems: Proven experience in building and optimizing AI application systems for large‑scale GPU infrastructure.
  • Architecture & Microarchitecture: Deep knowledge of x86 and ARM architectures, including competitive analysis of microarchitecture and performance‑based validation.
  • Programming & Tooling: Expert‑level proficiency in Python and C++. Experience with cycle‑accurate simulators and hardware debuggers like Lauterbach Trace32 or ARM DS‑5 is essential.
  • Low‑Level Systems: Ability to write and debug ARMv8 assembly, implement data synchronization protocols (MESI/MOESI), and analyze RTL via simulation waveforms.
  • Security & HPC: Experience with performance modeling for secure environments (e.g., Intel SGX, TDX, VM Encryption) and high‑performance computing benchmarks.
Compensation:

Compensation will be paid in the range of $172,500 - $210,000. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant’s education, experience, knowledge, skills, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Performance Engineer
Senior Performance Engineer

Crusoe Energy Systems LLC • San Francisco (CA)

On-site
USD 170,000 - 205,000
Equity packages
Paid time off
Health insurance
+3
Senior Performance Engineer
Senior Performance Engineer

Crusoe • San Francisco (CA)

On-site
USD 170,000 - 205,000
Competitive compensation package
Equity options
Health, dental & vision insurance
+3
Staff Hardware Systems Engineer, Performance
Staff Hardware Systems Engineer, Performance

Crusoe • San Francisco (CA)

On-site
USD 215,000 - 260,000
RSUs
Health insurance
401(k) with match
+3
Senior Hardware Systems Engineer
Senior Hardware Systems Engineer

Crusoe • Sunnyvale (CA)

On-site
USD 170,000 - 205,000
Health insurance
401(k) with match
Employee stock options
+2
Senior Hardware Systems Engineer, Performance
Senior Hardware Systems Engineer, Performance

Crusoe • Sunnyvale (CA)

On-site
USD 170,000 - 205,000
Health insurance
RSUs
401(k) match
+2
Staff Hardware Systems Engineer
Staff Hardware Systems Engineer

Crusoe • San Francisco (CA)

On-site
USD 215,000 - 260,000
Restricted Stock Units
Principal Systems Software Engineer
Principal Systems Software Engineer

Crusoe Energy Systems LLC • San Francisco (CA)

On-site
USD 260,000 - 340,000
Competitive compensation
Restricted Stock Units
Paid time off
+10
Senior Hardware Systems Engineer
Senior Hardware Systems Engineer

Crusoe Energy Systems LLC • Sunnyvale (CA)

On-site
USD 170,000 - 205,000
RSUs included
Staff Production Engineer, Compute
Staff Production Engineer, Compute

Crusoe • San Francisco (CA)

On-site
USD 209,000 - 253,000
RSUs
PTO & holidays
Health insurance
+2
Senior Production Engineer, Compute
Senior Production Engineer, Compute

crusoe • Sunnyvale (CA)

On-site
USD 170,000 - 205,000
Health insurance
RSUs
401(k) match
+7