Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Get past ATS filters
Benefits offered by this job
Comprehensive medical, dental, and vision coverage
Generous paid time off
Flexible work environment
Job summary
Lightning AI in San Francisco is seeking a GPU & Compute Infrastructure Engineer to manage GPU-enabled systems and ensure efficiency across their infrastructure. The ideal candidate will have over 5 years of experience, focusing on image management, system diagnostics, and developing automation tools. This position offers a flexible work location, competitive compensation of $180,000 to $200,000 annually, and a comprehensive benefits package to support employee well-being.
Qualifications
5+ years of experience in infrastructure engineering, systems engineering, or related roles.
Hands-on experience with GPU-enabled systems and tools like NVIDIA DCGM.
Familiarity with bare-metal provisioning and system bring-up workflows.
Responsibilities
Own and evolve systems for image management and validation.
Run and maintain test clusters for system validation.
Validate firmware, drivers, and OS images across systems.
Skills
Experience in infrastructure engineering
Hands-on experience with GPU-enabled systems
Proficiency in Python
Ability to debug complex issues
Tools
NVIDIA DCGM
Job description
Lightning AI in San Francisco is seeking a GPU & Compute Infrastructure Engineer to manage GPU-enabled systems and ensure efficiency across their infrastructure. The ideal candidate will have over 5 years of experience, focusing on image management, system diagnostics, and developing automation tools. This position offers a flexible work location, competitive compensation of $180,000 to $200,000 annually, and a comprehensive benefits package to support employee well-being.