A complete application in a minute — tailored resume and cover letter, ready to send.
PhoenixAI in Mumbai, India is seeking an experienced ML infrastructure engineer to fine-tune, deploy, and scale LLMs and foundation models for design tasks. You will build scalable inference systems and RAG pipelines, aiming for sub-second latency in a cloud-native setting.
Strong Python and PyTorch expertise, experience with LangChain/LlamaIndex, vector search, and high-availability architectures are essential.
AI is reshaping every industry—and Physical AI is next. At PhoenixAI, you'll work at the intersection of cutting-edge ML and Robotics innovation, helping bring frontier models to production. From deploying optimized foundation models to building low-latency inference infrastructure, your work will define how engineers interact with intelligence. If you're excited to make research real and shape the future of design, this role is for you.
We're a fast-moving AI startup with a collaborative, high-trust culture. You'll work side-by-side with top engineers and scientists, translating cutting-edge research into real product impact. We value curiosity, speed, and precision—engineering systems that are elegant, pragmatic, and built to last. If you're passionate about bringing frontier Physical AI to the real world, you'll feel right at home.
This position is available in Mumbai, India. We believe a mindmelt between engineers and scientists leads to our professional growth and to great products; we have a hybrid schedule.