ML Infra Engineer: Scale Training & Inference (Hybrid)
Lattice, Inc.
San Francisco (CA)
Hybrid
USD 200,000 - 280,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Competitive salary
Premium health, dental, and vision insurance
Unlimited PTO
$5,000 annual learning & development budget
Conference attendance and speaking opportunities
Job summary
A leading technology company is looking for an ML Infrastructure Engineer in San Francisco. The successful candidate will build and maintain ML training pipelines and ensure low-latency model serving. Candidates should have over 4 years of experience in ML engineering, a strong proficiency in Python, and familiarity with Kubernetes. This role offers competitive salaries, premium health benefits, and a hybrid work model with office access and a $5,000 annual learning budget.
Qualifications
4+ years of experience in ML engineering or infrastructure.
Strong proficiency in Python and experience with PyTorch or JAX.
Experience with ML training frameworks and distributed training.
Responsibilities
Build and maintain ML training pipelines and infrastructure.
Design model serving systems for low-latency inference.
Implement monitoring and observability for ML systems.
Skills
Python
ML Engineering
Kubernetes
PyTorch or JAX
ML Training Frameworks
Tools
TensorRT
ONNX
vLLM
Job description
A leading technology company is looking for an ML Infrastructure Engineer in San Francisco. The successful candidate will build and maintain ML training pipelines and ensure low-latency model serving. Candidates should have over 4 years of experience in ML engineering, a strong proficiency in Python, and familiarity with Kubernetes. This role offers competitive salaries, premium health benefits, and a hybrid work model with office access and a $5,000 annual learning budget.