An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Axiom Global Technologies seeks a Multimodal AI Engineer to own streaming speech and real-time computer vision for identity verification and proctoring. You will work on vision-language models in production and optimize models for constrained hardware while maintaining rigorous evaluation practices.
The role requires strong PyTorch, Python, Linux, SQL skills, and autonomous work with clean pull requests in a Kubernetes-deployed service. Excellent English communication is essential.
Owns streaming speech in and out of the interview and the real-time computer vision for identity verification and proctoring, including vision-language models.