Senior AI Model Serving Engineer — Low-Latency Inference
Menlo Ventures
San Francisco (CA)
On-site
USD 166,000 - 225,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Annual performance bonus
Equity options
Comprehensive benefits package
Job summary
A leading data and AI company in San Francisco is seeking a Senior Engineer to enhance their Model Serving platform. This role requires expertise in building large-scale distributed systems and collaboration across teams to optimize performance and reliability. Ideal candidates will have a strong foundation in algorithms and system design, along with a passion for mentoring others. The position offers a competitive salary and generous benefits.
Qualifications
5+ years of experience building large-scale distributed systems.
Experience in model serving, inference systems, or related infrastructure.
Strong background in algorithms, data structures, and system design.
Responsibilities
Design core systems and APIs for Model Serving.
Drive architectural decisions for performance optimization.
Collaborate with cross-functional teams.
Skills
Building large-scale distributed systems
Model serving
System design
Collaborative communication
Customer-focused mindset
Mentoring engineers
Job description
A leading data and AI company in San Francisco is seeking a Senior Engineer to enhance their Model Serving platform. This role requires expertise in building large-scale distributed systems and collaboration across teams to optimize performance and reliability. Ideal candidates will have a strong foundation in algorithms and system design, along with a passion for mentoring others. The position offers a competitive salary and generous benefits.