Lead Engineer, Inference Platform & Scale

Cerebras

Sunnyvale (CA)

On-site

USD 150,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Job stability with startup vitality
Open-source AI research
Simple, non-corporate work culture

Job summary

A pioneering AI hardware company in Sunnyvale is looking for an engineering leader to oversee the Inference Service Platform. You will guide a team in scaling LLM inference and architecting low latency systems. Candidates should have substantial experience in distributed systems, strong technical leadership, and a history of optimizing performance. The position emphasizes collaboration across teams to deliver enterprise solutions and requires a hands-on approach to technology and team mentorship.

Qualifications

  • 6+ years in high-scale software engineering, with 3+ years leading distributed systems.
  • Proven track record scaling LLM inference with optimizations.
  • Deep experience with orchestration and large clusters.

Responsibilities

  • Own the technical vision for Cerebras Inference Platform.
  • Lead development of distributed inference systems.
  • Drive operational excellence, ensuring platform reliability.

Skills

Technical Leadership
Inference Expertise
ML Systems Knowledge
Frameworks & Tools
Infrastructure
Operations & Monitoring
Leadership & Collaboration

Tools

Kubernetes
TensorRT-LLM
PyTorch
Hugging Face

Job description

A pioneering AI hardware company in Sunnyvale is looking for an engineering leader to oversee the Inference Service Platform. You will guide a team in scaling LLM inference and architecting low latency systems. Candidates should have substantial experience in distributed systems, strong technical leadership, and a history of optimizing performance. The position emphasizes collaboration across teams to deliver enterprise solutions and requires a hands-on approach to technology and team mentorship.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Inference Infrastructure Engineer
Senior AI Inference Infrastructure Engineer

Modular • United States

Hybrid
USD 167,000 - 273,000
Staff Engineer, Inference Cloud Platform
Staff Engineer, Inference Cloud Platform

Cerebras Systems • Sunnyvale (CA)

On-site
USD 130,000 - 160,000
Opportunity to publish and open source AI research
Work on one of the fastest AI supercomputers
Non-corporate work culture
Staff AI Cloud Platform Engineer – Inference & Training
Staff AI Cloud Platform Engineer – Inference & Training

Cerebras • Sunnyvale (CA)

On-site
USD 120,000 - 150,000
Staff Engineer, Inference Runtime — Performance & Scale
Staff Engineer, Inference Runtime — Performance & Scale

Menlo Ventures • New York (NY)

Hybrid
USD 405,000 - 485,000
Senior AI Model Serving Engineer — Low-Latency Inference
Senior AI Model Serving Engineer — Low-Latency Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 166,000 - 225,000
Annual performance bonus
Equity options
Comprehensive benefits package
Staff Software Engineer, Inference Systems at Scale
Staff Software Engineer, Inference Systems at Scale

Anthropic • San Francisco (CA)

On-site
USD 300,000 - 485,000
Competitive salary
Flexible working hours
Generous vacation and parental leave
Director, AI Inference & GPU-Accelerated Pipelines
Director, AI Inference & GPU-Accelerated Pipelines

WEKA • United States

On-site
USD 150,000 - 200,000
Staff Engineer, LLM Inference & Infra
Staff Engineer, LLM Inference & Infra

Prime Intellect • United States

Hybrid
USD 120,000 - 150,000
Competitive compensation
Flexible work arrangement
Full visa sponsorship
+2
Principal AI Systems Architect — Inference & Hardware
Principal AI Systems Architect — Inference & Hardware

Conductor • San Jose (CA)

On-site
USD 219,000 - 351,000
4+ weeks of paid time off
Medical/Dental/Vision/401k
Flexible work environment
+2
Staff Engineer, Scalable AI Inference Infrastructure
Staff Engineer, Scalable AI Inference Infrastructure

Inferact • San Francisco (CA)

Hybrid
USD 200,000 - 400,000