Senior Model Serving Engineer - Low-Latency AI Platform
Menlo Ventures
San Francisco (CA)
On-site
USD 192,000 - 260,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Comprehensive benefits
Eligibility for annual performance bonus
Equity opportunities
Job summary
A leading data and AI company in San Francisco is seeking a Staff Engineer to design and implement systems for their AI/ML Model Serving platform. You will collaborate with product, infrastructure, and research teams to ensure high-performance system delivery. The ideal candidate has over 10 years of experience in distributed systems and model serving. This role offers a competitive salary and a commitment to diversity and inclusion.
Qualifications
10+ years of experience building and operating large-scale distributed systems.
Deep expertise in model serving and related infrastructure.
Proven ability to deliver technically complex, high-impact initiatives.
Responsibilities
Design and implement systems and APIs for Databricks Model Serving.
Define the technical roadmap for serving workloads.
Lead initiatives to improve performance and cost-effectiveness.
Skills
Building and operating large-scale distributed systems
Model serving and inference systems
Algorithms and system design
Strong communication skills
Mentoring and technical guidance
Job description
A leading data and AI company in San Francisco is seeking a Staff Engineer to design and implement systems for their AI/ML Model Serving platform. You will collaborate with product, infrastructure, and research teams to ensure high-performance system delivery. The ideal candidate has over 10 years of experience in distributed systems and model serving. This role offers a competitive salary and a commitment to diversity and inclusion.