An application made for this job — a tailored resume and cover letter that speak straight to the posting.
LambdaQ labs Pvt. Ltd. in Maharashtra is seeking a seasoned ML infra engineer to build and optimize the serving stack for fast, reliable, and affordable Vikasit Inference at India scale.
You’ll own batching, caching, routing, autoscaling, and cost management across GPU-heavy model fleets, while ensuring the OpenAI-compatible API remains rock-solid for streaming and tool calls.
You’ll build and optimize the serving stack that makes Vikasit Inference fast, reliable, and affordable at India scale — kernels, batching, caching, and the OpenAI-compatible API surface developers love.
We hire for skill over credentials. Tell us why you're a fit — links and projects welcome.