Senior Engineering Manager, AI Inference (LLM Production)

Crusoe Energy Systems LLC

San Francisco, Northern (CA, KY)

Hybrid

USD 250,000 - 300,000

Full time

6 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Competitive compensation and equity
Restricted Stock Units
Paid time off, holidays & leave
Healthcare: medical, dental & vision
HSA contributions
Parental leave
Life insurance
Tuition reimbursement
Mental health support
Commuter benefits
Cell phone stipend
401(k) with company match
Volunteer time off
Global travel insurance
Daily meals allowance
Location-specific perks

Job summary

Crusoe Energy Systems LLC is seeking a Senior Engineering Manager in San Francisco to lead an ML inference optimization team. You will stay hands-on with the stack, profiling performance and deploying improvements in production for large language models and diverse workloads.

The role blends leadership with deep technical work, collaborating with customers to tailor deployments and delivering measurable performance gains in real-world environments.

Qualifications

  • BS/MS/PhD in CS/Engineering/Math or related field.
  • Hands-on production coding in Python or C++.
  • Experience shipping ML inference systems.
  • Familiarity with optimizing LLMs for throughput and latency.
  • Understanding GPUs and kernel behavior.
  • Strong communication with customers and teammates.

Responsibilities

  • Bring current inference techniques into production and refine them.
  • Design and optimize serving architectures for latency, throughput and cost.
  • Dive into serving stack from frameworks to CUDA kernels to identify performance issues.
  • Scale optimization methods across ML models, especially large language models.
  • Profile and tune deployments to meet real production targets.
  • Lead team delivery end-to-end and align with product and engineering leadership.

Skills

Team leadership
Python
C++
LLM optimization
GPU basics
Communication

Education

Bachelors, Masters, or PhD in CS/Engineering/Math

Tools

Docker
Kubernetes

Job description

Crusoe Energy Systems LLC is seeking a Senior Engineering Manager in San Francisco to lead an ML inference optimization team. You will stay hands-on with the stack, profiling performance and deploying improvements in production for large language models and diverse workloads.

The role blends leadership with deep technical work, collaborating with customers to tailor deployments and delivering measurable performance gains in real-world environments.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Inference Engineer: High-Throughput LLMs
Senior AI Inference Engineer: High-Throughput LLMs

Crusoe Energy Systems • San Francisco (CA)

On-site
USD 180,000 - 240,000
Health benefits
401(k) match
Paid time off
+1
Senior Engineering Manager — AI Inference in Production
Senior Engineering Manager — AI Inference in Production

crusoe • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive compensation
Equity
Paid time off
+2
Senior AI Inference Engineer — Production LLM Optimizer
Senior AI Inference Engineer — Production LLM Optimizer

Crusoe Energy Systems LLC • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 300,000
Competitive compensation
Equity packages
Health/dental/vision insurance
+1
Senior AI Inference Engineer - Production LLM Optimizer
Senior AI Inference Engineer - Production LLM Optimizer

Crusoe Energy Systems LLC • San Francisco (CA), Northern (KY)

Hybrid
USD 215,000 - 260,000
Equity
RSUs
Paid time off
+13
Engineering Manager, AI Platform & LLM Infra
Engineering Manager, AI Platform & LLM Infra

Crusoe • United States

On-site
USD 215,000 - 260,000
Competitive compensation and equity
Restricted Stock Units
Paid time off, holidays & leave
+7
Senior AI Inference Engineer — Production-Scale LLMs
Senior AI Inference Engineer — Production-Scale LLMs

AI Chopping Block, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 300,000
Competitive compensation and equity
RSUs
Paid time off
+6
Senior AI Inference Architect for Production LLMs
Senior AI Inference Architect for Production LLMs

Crusoe • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive compensation and equity
Restricted Stock Units
Paid time off & holidays
+13
LLM Inference Systems Engineer
LLM Inference Systems Engineer

Acceler8 Talent • San Francisco (CA)

On-site
USD 150,000 - 230,000
Production AI Inference Engineer for Fast LLMs
Production AI Inference Engineer for Fast LLMs

Crusoe • San Francisco (CA)

On-site
USD 215,000 - 260,000
Competitive compensation and equity
Restricted Stock Units
Paid time off & holidays
+6
Staff Engineer - Customer-Facing AI Inference Infra
Staff Engineer - Customer-Facing AI Inference Infra

Simplify • San Francisco (CA)

On-site
USD 200,000 - 300,000
Housing stipend
Uber/Waymo rides