Get more replies from employers
Send a job-specific resume in minutes.
AZX is seeking an ML Engineer to own the technical backbone of model serving at scale, shaping inference platforms, GPU scheduling, and autoscaling across cloud and customer-managed clusters for vLLM/SGLang models.
You’ll set standards, guide evaluation systems, and mentor engineers in a high-impact, architecture-focused role that balances cost, performance, and reliability for critical AI workloads.
AZX is seeking an ML Engineer to own the technical backbone of model serving at scale, shaping inference platforms, GPU scheduling, and autoscaling across cloud and customer-managed clusters for vLLM/SGLang models.
You’ll set standards, guide evaluation systems, and mentor engineers in a high-impact, architecture-focused role that balances cost, performance, and reliability for critical AI workloads.