ML Infrastructure Engineer

Mixpeek

San Mateo (CA)

On-site

USD 150,000 - 230,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Full benefits (medical, dental, vision
401k
H-1B visa sponsorship and immigration”

Job summary

zaimler is seeking an engineer to own our inference and model-serving infrastructure end to end, enabling production-grade AI agents to run fast and reliably at scale. You’ll report to Sofus and collaborate with ML and infra teams on deployment and scalability.

You will set up and scale inference systems, manage GPU resources for concurrency, and optimize the engine that powers scalable agent orchestration, joining an early engineering team in a cutting-edge context.

Qualifications

  • Proven ability to build scalable ML/AI platforms end-to-end for production.
  • Deep understanding of the inference stack: vLLM, KV cache, and the optimization layers underneath model serving.
  • Experience building distributed AI/ML workloads at scale, connecting them to real product or vertical integrations.
  • 3+ years of relevant experience. We care about capability, not tenure.

Responsibilities

  • Set up and scale inference/Ray Serve for ML and LLM model serving.
  • Scale agent GPU infrastructure for concurrency and efficiency across multiple agent workloads.
  • Optimize and improve the engine builder and model server that power scalable agent orchestration.

Skills

Scalable ML platforms
Distributed systems
Model serving
Ray Serve

Tools

Ray Serve
vLLM
KV cache
Distributed systems tooling

Job description

About zaimler

AI agents can't reason over data they don't understand. Enterprise data today is fragmented across dozens of systems with no shared context, meaning, or structure, and that's why most enterprise AI is failing. The shift from copilots to autonomous agents is creating an entirely new infrastructure layer, and we're building it.

zaimler is the context infrastructure for the agentic era: a platform that automatically discovers domain knowledge, maps relationships, and gives AI agents the semantic understanding to operate with precision at scale. Imagine knowledge graphs that support real-time inference, built for systems that need to reason, not just retrieve.

zaimler was founded by Biswajit Das (ex-VP Engineering, Truera), a Data Infra veteran and former Chief Architect at Visa, and Sofus Macskassy (ex-Director of Engineering, LinkedIn), who built one of the largest knowledge graphs in production in the industry at LinkedIn. We're growing and deploying with major enterprises across insurance, travel, and technology. If you want to build infrastructure that the next decade of enterprise AI runs on, we'd love to talk.

About the Role

You’ll own our inference and model-serving infrastructure end to end. This isn't a research role. It's a build role: you're setting up and scaling the systems that let our agents actually run in production, fast and reliably, at increasing concurrency.

You report to Sofus and work closely with our ML and infra teams.

What You’ll Own
  • Set up and scale inference/Ray Serve for ML and LLM model serving, integrated with our data analysis and agent workflows
  • Scale agent GPU infrastructure for concurrency and efficiency across multiple agent workloads
  • Optimize and improve the engine builder and model server that power scalable agent orchestration
What You Need
  • Proven ability to build scalable ML/AI platforms from scratch, end-to-end, for production use cases. You've owned a zero-to-one build before, or can show you're capable of it
  • Deep understanding of the inference stack: vLLM, KV cache, and the optimization layers underneath model serving
  • Experience building distributed systems for AI/ML workloads at scale, connecting them to real product or vertical integrations
  • 3+ years of relevant experience. We care about capability, not tenure
Nice to Have
  • Ray / Ray Serve experience
  • Familiarity with AIBrix
Why Join
  • A rare chance to shape both company and product direction as an early team engineer
  • Work alongside engineers and researchers from LinkedIn, Visa, Meta, and Branch
  • Onsite culture in San Mateo, built for deep collaboration and high-velocity building
  • Full benefits (medical, dental, vision, 401k)
  • We sponsor H-1B visas and assist with immigration

We value builders over résumés. If this role excites you but you don't check every box, we still want to hear from you. zaimler is an equal opportunity employer.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Cloud Infrastructure Engineer
Cloud Infrastructure Engineer

Mixpeek • San Mateo (CA)

On-site
USD 180,000 - 280,000
Equity
Full benefits (Medical, Dental, Vision
401k
+2
Staff Applied Machine Learning Engineer
Staff Applied Machine Learning Engineer

Mixpeek • San Mateo (CA)

On-site
USD 180,000 - 240,000
Head of Product
Head of Product

Mixpeek • San Mateo (CA)

On-site
USD 180,000 - 240,000
On-site in San Mateo, CA
Competitive equity
Health, dental, and vision insurance
+2
Full Stack Engineer
Full Stack Engineer

Mixpeek • San Mateo (CA)

On-site
USD 140,000 - 210,000
Data Infrastructure Engineer (Query Engine)
Data Infrastructure Engineer (Query Engine)

zaimler, Inc • San Mateo (CA)

On-site
USD 150,000 - 210,000
Medical benefits
401k
Onsite in San Mateo
Forward Deployed Engineer
Forward Deployed Engineer

zaimler, Inc • Northern (KY), New York (NY)

Hybrid
USD 104,000 - 173,000
Scale ML Inference & Model-Serving Engineer
Scale ML Inference & Model-Serving Engineer

Mixpeek • San Mateo (CA)

On-site
USD 150,000 - 230,000
Full benefits (medical, dental, vision
401k
H-1B visa sponsorship and immigration”
Staff Applied ML Engineer, AI Infrastructure
Staff Applied ML Engineer, AI Infrastructure

Mixpeek • San Mateo (CA)

On-site
USD 180,000 - 240,000
Founding Engineer
Founding Engineer

Embedding VC • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Medical, Dental, Vision insurance
401(k)
Lunch & Dinner in the office
+3
Senior/Staff AI Engineer
Senior/Staff AI Engineer

AI Talent Now • San Mateo (CA)

On-site
USD 120,000 - 160,000