Senior ML Engineer — Ultra-Fast Inference & Memory

comfy-org

San Francisco (CA)

On-site

USD 100,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

comfy-org is seeking an AI Optimization Engineer in San Francisco to enhance model inference performance for ComfyUI. The role entails building and optimizing the core inference engine, ensuring models run faster and use less memory.

Ideal candidates are passionate about model optimization, have production experience in PyTorch, and enjoy tackling technical challenges in the visual AI domain. This position offers a unique opportunity to influence cutting-edge technology.

Qualifications

  • Experience optimizing model inference and memory management.
  • Proficiency in production PyTorch code.
  • Passion for exploring how models function internally.

Responsibilities

  • Build and optimize the core inference engine for ComfyUI.
  • Enhance model performance and reduce memory usage.
  • Collaborate with the core team on feature development.
  • Address technical challenges in visual AI.
  • Influence the direction of the technology.

Skills

Model inference optimization
Torch optimizations
Memory management
Production PyTorch code
Visual AI problems

Job description

comfy-org is seeking an AI Optimization Engineer in San Francisco to enhance model inference performance for ComfyUI. The role entails building and optimizing the core inference engine, ensuring models run faster and use less memory.

Ideal candidates are passionate about model optimization, have production experience in PyTorch, and enjoy tackling technical challenges in the visual AI domain. This position offers a unique opportunity to influence cutting-edge technology.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior/Staff ML Engineer, Performance Optimization
Senior/Staff ML Engineer, Performance Optimization

comfy-org • San Francisco (CA)

On-site
USD 100,000 - 140,000
Senior ML Performance Engineer — Ultra-Fast Inference + Equity
Senior ML Performance Engineer — Ultra-Fast Inference + Equity

well-funded deeptech startup • California (MO)

On-site
USD 200,000 - 250,000
Senior ML Engineer - Low-Latency Inference & Systems
Senior ML Engineer - Low-Latency Inference & Systems

Inworld • Germany (OH)

Hybrid
USD 120,000 - 160,000
High-Performance ML Inference Engineer
High-Performance ML Inference Engineer

Reactor • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive SF salary
Early equity
Visa sponsorship
+2
Inference Performance Engineer: Optimize Model Serving
Inference Performance Engineer: Optimize Model Serving

Adaption • San Francisco (CA)

On-site
USD 180,000 - 240,000
Lunch stipend
Travel stipend (Adaption Passport)
Well-being benefits
+1
ML Inference Engineer San Francisco · Engineering · Full Time →
ML Inference Engineer San Francisco · Engineering · Full Time →

Reactor • San Francisco (CA)

On-site
USD 120,000 - 160,000
Visa sponsorship
Relocation support
Generous health, dental, and vision coverage
Senior Inference Systems Engineer — Low-Latency ML Serving
Senior Inference Systems Engineer — Low-Latency ML Serving

Jobtailor • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff, Inference
Member of Technical Staff, Inference

Reactor • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive SF salary
Early equity
Visa sponsorship
+2
Founding ML Inference Engineer — Ultra-Low Latency AI
Founding ML Inference Engineer — Ultra-Low Latency AI

Reactor • San Francisco (CA)

On-site
USD 180,000 - 280,000
Competitive salary
Early equity
Health, dental, and vision coverage
+1
Staff ML Engineer: Build Ultra-Fast AI at Scale (Relocation)
Staff ML Engineer: Build Ultra-Fast AI at Scale (Relocation)

Inworld AI • Mountain View (CA)

On-site
USD 270,000 - 500,000