Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Hire Rightt seeks an AI Infrastructure Engineer to design, deploy, and optimize production-grade LLM and Generative AI infrastructure. You will work on LLM inference, GPU optimization, Kubernetes orchestration, and cloud-based AI services.
The role covers observability, state management, and scalable AI workflows. You will deploy self-hosted LLMs, optimize performance across GPUs, and build CI/CD pipelines for AI services, while ensuring reliability and cost efficiency in a fast-paced
Dubai
The job holder will be responsible for building, deploying, optimizing, and operating production-grade LLM and Generative AI infrastructure. The role combines LLM inference, GPU optimization, Kubernetes, cloud infrastructure, observability, RAG, and AI platform engineering.