Founding ML Infra Engineer — Production-Grade LLMs
Realmlabs
Sunnyvale (CA)
On-site
USD 210,000 - 350,000
Full time
14 days+
Application generator
Stand out for this role — generate a tailored resume and cover letter in about a minute.
Get past ATS filters
Benefits offered by this job
Market aligned compensation
Founding engineer equity
Medical, Dental, Vision, and Life insurance
401-K
In-office lunch
Job summary
An innovative AI startup is seeking a Founding ML Infrastructure Engineer to take charge of deploying and optimizing production-grade LLM systems. In this core role, you will be responsible for building and managing a full ML serving stack, working closely with product teams to ensure system performance and reliability. The ideal candidate will have extensive experience in ML infrastructure, particularly with LLMs, and will be proficient in relevant technologies such as PyTorch, TensorFlow, and Kubernetes. This position offers a unique opportunity to shape the future of AI within the company.
Qualifications
Extensive experience with GPU inference technologies.
Proficient in building production-grade ML infrastructure.
Keen awareness of latency and throughput optimization.
Responsibilities
Own the end-to-end LLM inference stack.
Design high-performance LLM serving systems.
Collaborate with product teams for deployment.
Skills
Deep understanding of LLM internals
GPU inference optimization
Software engineering fundamentals
Collaboration with cross-functional teams
Education
5+ years of professional experience in ML infrastructure
Tools
PyTorch
TensorFlow
TensorRT
Triton Inference Server
Kubernetes
Job description
An innovative AI startup is seeking a Founding ML Infrastructure Engineer to take charge of deploying and optimizing production-grade LLM systems. In this core role, you will be responsible for building and managing a full ML serving stack, working closely with product teams to ensure system performance and reliability. The ideal candidate will have extensive experience in ML infrastructure, particularly with LLMs, and will be proficient in relevant technologies such as PyTorch, TensorFlow, and Kubernetes. This position offers a unique opportunity to shape the future of AI within the company.