Senior/Staff ML Engineer, Performance Optimization

comfy-org

San Francisco (CA)

On-site

USD 100,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

comfy-org is seeking an AI Optimization Engineer in San Francisco to enhance model inference performance for ComfyUI. The role entails building and optimizing the core inference engine, ensuring models run faster and use less memory.

Ideal candidates are passionate about model optimization, have production experience in PyTorch, and enjoy tackling technical challenges in the visual AI domain. This position offers a unique opportunity to influence cutting-edge technology.

Qualifications

  • Experience optimizing model inference and memory management.
  • Proficiency in production PyTorch code.
  • Passion for exploring how models function internally.

Responsibilities

  • Build and optimize the core inference engine for ComfyUI.
  • Enhance model performance and reduce memory usage.
  • Collaborate with the core team on feature development.
  • Address technical challenges in visual AI.
  • Influence the direction of the technology.

Skills

Model inference optimization
Torch optimizations
Memory management
Production PyTorch code
Visual AI problems

Job description

The Role

We're looking for someone who loves optimizing model inference to join us in building the core of ComfyUI – the most complex and bleeding‑edge part of our engine. You'll be working on making AI models run faster and more efficiently than anyone thought possible.

You are a good fit if this describes you
  • You geek out about model inference, torch optimizations, and memory management
  • You’ve written production PyTorch code that pushes performance boundaries
  • You love diving deep into how models actually work under the hood
  • You get excited about making insanely optimized code that just works
  • You think the current state of ML deployment could be way better
What you’ll do
  • Build and optimize the core inference engine that powers ComfyUI
  • Make massive models run faster and use less memory than anyone else
  • Work directly with our core team on architecting new features
  • Tackle the hardest technical problems in the visual AI space
  • Help shape where we take this technology next
Bonus

If you’ve worked with diffusion/LLM models before or built custom nodes for ComfyUI, that’s awesome.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Engineer — Ultra-Fast Inference & Memory
Senior ML Engineer — Ultra-Fast Inference & Memory

comfy-org • San Francisco (CA)

On-site
USD 100,000 - 140,000
Member of Technical Staff, ML Performance
Member of Technical Staff, ML Performance

Odyssey • Palo Alto (CA)

On-site
USD 130,000 - 160,000
Member of Technical Staff, Inference
Member of Technical Staff, Inference

Reactor • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive SF salary
Early equity
Visa sponsorship
+2
Member of Technical Staff, ML Performance
Member of Technical Staff, ML Performance

Odyssey • Santa Clara (CA)

On-site
USD 120,000 - 160,000
ML Inference Engineer San Francisco · Engineering · Full Time →
ML Inference Engineer San Francisco · Engineering · Full Time →

Reactor • San Francisco (CA)

On-site
USD 120,000 - 160,000
Visa sponsorship
Relocation support
Generous health, dental, and vision coverage
Founding Engineer, ML Inference
Founding Engineer, ML Inference

Reactor • San Francisco (CA)

On-site
USD 180,000 - 280,000
Competitive salary
Early equity
Health, dental, and vision coverage
+1
Senior Software Engineer - Model Performance
Senior Software Engineer - Model Performance

inference.net • San Francisco (CA)

Hybrid
USD 220,000 - 320,000
Equity in a high-growth startup
Comprehensive benefits
Member of Technical Staff, Performance Optimization
Member of Technical Staff, Performance Optimization

Fireworks AI • San Mateo (CA)

On-site
USD 180,000 - 260,000
Senior Software Engineer - Model Performance
Senior Software Engineer - Model Performance

Inference • San Francisco (CA)

On-site
USD 220,000 - 320,000
Competitive compensation
Equity in a high-growth startup
Comprehensive benefits
Senior ML Performance Engineer
Senior ML Performance Engineer

well-funded deeptech startup • California (MO)

On-site
USD 200,000 - 250,000