Founding ML Infra Engineer — Production-Grade LLMs

Realmlabs

Sunnyvale (CA)

On-site

USD 210,000 - 350,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Market aligned compensation
Founding engineer equity
Medical, Dental, Vision, and Life insurance
401-K
In-office lunch

Job summary

An innovative AI startup is seeking a Founding ML Infrastructure Engineer to take charge of deploying and optimizing production-grade LLM systems. In this core role, you will be responsible for building and managing a full ML serving stack, working closely with product teams to ensure system performance and reliability. The ideal candidate will have extensive experience in ML infrastructure, particularly with LLMs, and will be proficient in relevant technologies such as PyTorch, TensorFlow, and Kubernetes. This position offers a unique opportunity to shape the future of AI within the company.

Qualifications

  • Extensive experience with GPU inference technologies.
  • Proficient in building production-grade ML infrastructure.
  • Keen awareness of latency and throughput optimization.

Responsibilities

  • Own the end-to-end LLM inference stack.
  • Design high-performance LLM serving systems.
  • Collaborate with product teams for deployment.

Skills

Deep understanding of LLM internals
GPU inference optimization
Software engineering fundamentals
Collaboration with cross-functional teams

Education

5+ years of professional experience in ML infrastructure

Tools

PyTorch
TensorFlow
TensorRT
Triton Inference Server
Kubernetes

Job description

An innovative AI startup is seeking a Founding ML Infrastructure Engineer to take charge of deploying and optimizing production-grade LLM systems. In this core role, you will be responsible for building and managing a full ML serving stack, working closely with product teams to ensure system performance and reliability. The ideal candidate will have extensive experience in ML infrastructure, particularly with LLMs, and will be proficient in relevant technologies such as PyTorch, TensorFlow, and Kubernetes. This position offers a unique opportunity to shape the future of AI within the company.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer, LLM Inference & Infra
Staff Engineer, LLM Inference & Infra

Prime Intellect • United States

Hybrid
USD 120,000 - 150,000
Competitive compensation
Flexible work arrangement
Full visa sponsorship
+2
Remote ML Engineering Manager: LLM Serving & Infra
Remote ML Engineering Manager: LLM Serving & Infra

Jobgether • United States

Remote
USD 176,000 - 252,000
Senior ML Engineer, AI Platform & Products (LLM)
Senior ML Engineer, AI Platform & Products (LLM)

BetterUp • New York (NY)

Hybrid
USD 200,000 - 275,000
ML Infrastructure Engineer for Scalable LLMs
ML Infrastructure Engineer for Scalable LLMs

ServiceNow • Mountain View (CA)

Hybrid
USD 130,000 - 160,000
Staff Engineer - LLM Inference & Serving at Scale
Staff Engineer - LLM Inference & Serving at Scale

Prime Intellect • San Francisco (CA)

Hybrid
USD 150,000 - 300,000
Principal ML Engineer — GenAI, LLMs & Production
Principal ML Engineer — GenAI, LLMs & Production

Alexander Chapman • Boulder (CO)

Hybrid
USD 120,000 - 160,000
ML Engineer: Production-Scale LLMs & AI Infra
ML Engineer: Production-Scale LLMs & AI Infra

Harrison Clarke • San Francisco (CA)

On-site
USD 150,000 - 190,000
ML Engineer, Integrity — Production LLMs & Equity
ML Engineer, Integrity — Production LLMs & Equity

OpenAI • Los Angeles (CA)

On-site
USD 266,000 - 555,000
ML Infra Engineer — Open-Source LLM Performance
ML Infra Engineer — Open-Source LLM Performance

well-funded deeptech startup • San Francisco (CA)

Hybrid
USD 250,000 - 500,000
Bonus
Medical
Dental
+4
Production ML Engineer: LLMs, RAG & MLOps on GCP
Production ML Engineer: LLMs, RAG & MLOps on GCP

DeepHow • United States

On-site
USD 140,000 - 210,000