Staff Engineer - Training & Inference for Distributed AI

Boson AI

Santa Clara (CA)

On-site

USD 150,000 - 600,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

A pioneering AI startup located in Santa Clara is looking for research scientists and engineers to join their team. The role involves optimizing model architectures and implementing distributed optimization algorithms. The ideal candidates should have a strong background in CUDA, PyTorch, and deep learning architectures, with at least a Master’s or Doctoral degree in computer science. Competitive salary ranges from $150,000 to $600,000 per year.

Qualifications

  • Experience in writing clean and efficient code.
  • Proficiency in at least one deep learning framework.
  • Participation in at least 1 research project related to distributed training.

Responsibilities

  • Optimize model architectures to handle images, video, text, speech, and audio data.
  • Implement and optimize kernels for efficient training on GPUs.
  • Conduct performance optimization and distributed training.

Skills

CUDA
PyTorch
Distributed optimization
Deep learning architectures
Performance optimization

Education

Master or Doctoral degree in computer science or equivalent

Tools

Triton
JAX

Job description

A pioneering AI startup located in Santa Clara is looking for research scientists and engineers to join their team. The role involves optimizing model architectures and implementing distributed optimization algorithms. The ideal candidates should have a strong background in CUDA, PyTorch, and deep learning architectures, with at least a Master’s or Doctoral degree in computer science. Competitive salary ranges from $150,000 to $600,000 per year.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Software Engineer — Scalable RL & Distributed Training
Research Software Engineer — Scalable RL & Distributed Training

Reflection AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health insurance
Fully paid parental leave
+2
Staff Engineer, Scalable AI Inference Infrastructure
Staff Engineer, Scalable AI Inference Infrastructure

Inferact • San Francisco (CA)

Hybrid
USD 200,000 - 400,000
Generous health, dental, and vision benefits
401(k) company match
Equity options
Senior ML Performance Engineer - Distributed Training
Senior ML Performance Engineer - Distributed Training

Odyssey • Santa Clara (CA)

On-site
USD 120,000 - 160,000
Research Software Engineer — Scalable RL & Distributed Training
Research Software Engineer — Scalable RL & Distributed Training

Reflection AI • New York (NY)

On-site
USD 120,000 - 180,000
Top-tier compensation
Comprehensive medical, dental, vision insurance
Fully paid parental leave
+2
Senior Model Inference Engineer for Production-Scale AI
Senior Model Inference Engineer for Production-Scale AI

OpenAI • San Francisco (CA)

On-site
USD 325,000 - 490,000
Distributed AI Training Research Engineer (Remote)
Distributed AI Training Research Engineer (Remote)

Prime Intellect • United States

Hybrid
USD 110,000 - 150,000
Competitive compensation including equity incentives
Flexible work arrangements
Visa sponsorship and relocation assistance
+2
Staff ML Systems Engineer — Distributed Training at Scale
Staff ML Systems Engineer — Distributed Training at Scale

RadixArk • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Comprehensive benefits
Flexible work arrangements
Staff ML Engineer: Build Ultra-Fast AI at Scale (Relocation)
Staff ML Engineer: Build Ultra-Fast AI at Scale (Relocation)

Inworld AI • Mountain View (CA)

On-site
USD 270,000 - 500,000
Relocation assistance
Equity options
Comprehensive benefits package
Senior AI Inference Infrastructure Engineer
Senior AI Inference Infrastructure Engineer

Modular • United States

Hybrid
USD 167,000 - 273,000
Competitive salary
Premier insurance plans
Flexible paid time off
+1
Staff GenAI Inference Engineer: Optimize LLM Serving Latency
Staff GenAI Inference Engineer: Optimize LLM Serving Latency

Menlo Ventures • San Francisco (CA)

On-site
USD 190,900 - 232,800
Annual performance bonus
Equity options
Comprehensive health benefits