Real-Time ML Inference Engineer for Scalable Serving
Yobi
New York (NY)
Hybrid
USD 100,000 - 150,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Competitive Base Salary
Meaningful equity
Annual performance bonus
Comprehensive health benefits
Unlimited PTO
401k with company match
Job summary
A Behavioral AI company is seeking a Machine Learning Engineer to design and optimize systems for bringing their models to life. The role involves ensuring ML models are efficient and reliable, requiring experience in model deployment and robust coding skills. Candidates should be familiar with low-latency techniques and operational maturity in ML systems. This position can be remote or hybrid from several hubs.
Qualifications
Deep expertise in model deployment and scaling production ML serving systems.
Understand low-latency model inference techniques.
Robust coding skills in Go, Rust, C++, or Java.
Experience in monitoring and observing models.
Responsibilities
Design, optimize, and operate systems for Behavioral AI models.
Package, version, and roll out ML models in production environments.
Ensure models are fast, accurate, and accountable.
Skills
Model deployment
Low-latency optimization
High-performance coding in Go, Rust, C++, Java
Monitoring model drift
Infrastructure design
Applied ML understanding
Job description
A Behavioral AI company is seeking a Machine Learning Engineer to design and optimize systems for bringing their models to life. The role involves ensuring ML models are efficient and reliable, requiring experience in model deployment and robust coding skills. Candidates should be familiar with low-latency techniques and operational maturity in ML systems. This position can be remote or hybrid from several hubs.