Senior AI Inference Engineer — Scale LLMs & PyTorch

Togetherai

San Francisco (CA)

On-site

USD 160,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
Startup equity
Competitive benefits

Job summary

Togetherai in San Francisco is seeking a skilled Machine Learning Engineer to enhance the performance of AI inference systems. The role involves collaboration with top AI researchers and engineers, with responsibilities including building production systems and optimizing large-scale applications.

The ideal candidate has strong experience with Python and PyTorch, along with a comprehensive understanding of performance-related aspects of systems. This position offers a competitive base salary range of $160,000 - $230,000 plus equity and benefits.

Qualifications

  • 3+ years of experience writing high-performance, well-tested, production-quality code.
  • Proficiency with Python and PyTorch.
  • Excellent understanding of low-level operating systems concepts.

Responsibilities

  • Design and build production systems for AI inference engine.
  • Develop and optimize runtime inference services.
  • Collaborate with researchers and engineers on new features.

Skills

Python
PyTorch
Multi-threading
Memory management
Networking
CUDA

Tools

Rust
Cython

Job description

Togetherai in San Francisco is seeking a skilled Machine Learning Engineer to enhance the performance of AI inference systems. The role involves collaboration with top AI researchers and engineers, with responsibilities including building production systems and optimizing large-scale applications.

The ideal candidate has strong experience with Python and PyTorch, along with a comprehensive understanding of performance-related aspects of systems. This position offers a competitive base salary range of $160,000 - $230,000 plus equity and benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Machine Learning Engineer - Inference
Machine Learning Engineer - Inference

Togetherai • San Francisco (CA)

On-site
USD 160,000 - 230,000
Health insurance
Startup equity
Competitive benefits
Distributed LLM Inference Engineer - Scale HighThroughput AI
Distributed LLM Inference Engineer - Scale HighThroughput AI

Cerebras • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Stock Options
Healthcare plans with 99% premium coverage
401k Retirement Plan
+6
LLM Inference Architect & Systems Optimizer
LLM Inference Architect & Systems Optimizer

Togetherai • San Francisco (CA)

On-site
USD 160,000 - 230,000
Health insurance
Startup equity
Competitive benefits
Senior Backend Engineer, AI Inference Platform
Senior Backend Engineer, AI Inference Platform

Togetherai • San Francisco (CA)

On-site
USD 160,000 - 250,000
Health insurance
Startup equity
Competitive benefits
Distributed LLM Inference Engineer - Scale AI at Speed
Distributed LLM Inference Engineer - Scale AI at Speed

Anyscale • San Francisco (CA)

On-site
USD 120,000 - 160,000
Stock Options
Healthcare plans
401k Retirement Plan
+6
Senior ML Inference Engineer: Scale AI in Production
Senior ML Inference Engineer: Scale AI in Production

Oscar • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Equity
401k matching
Medical coverage
Member of Technical Staff (AI Inference Engineer)
Member of Technical Staff (AI Inference Engineer)

Kindredventures • Palo Alto (CA)

On-site
USD 190,000 - 250,000
Comprehensive health insurance
Dental insurance
Vision insurance
+1
AI Systems Architect – Inference & RL at Scale
AI Systems Architect – Inference & RL at Scale

Togetherai • San Francisco (CA)

On-site
USD 200,000 - 280,000
Startup equity
Health insurance
Competitive benefits
Engineering Manager, ML Inference & Scale
Engineering Manager, ML Inference & Scale

Anthropic • San Francisco (CA)

Hybrid
USD 425,000 - 560,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave
LLM Inference Frameworks and Optimization Engineer
LLM Inference Frameworks and Optimization Engineer

Together AI • San Francisco (CA)

On-site
USD 160,000 - 230,000
Startup equity
Health insurance
Competitive benefits