Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Virtusa is seeking an AI/ML-focused Data Engineer to build intelligent data pipelines for unstructured content and integrate with modern ML ecosystems. The role emphasizes PySpark and Python, with a focus on document classification, cleansing, quality metrics, and working with LLMs, vector databases, and RAG frameworks.
You will bridge data engineering and machine learning to enable AI-first applications, collaborating with AI architects and platform teams to design end-to-end AI data readiness
We are looking for an AI/ML-focused Data Engineer who brings deep expertise in building intelligent data pipelines for unstructured content and is experienced in integrating with modern machine learning ecosystems. The ideal candidate will have hands-on experience in PySpark and Python with a strong focus on document classification cleansing quality metrics and the ability to work with LLMs vector databases and Retrieval-Augmented Generation (RAG) frameworks. Candidates will play a critical role in bridging data engineering and machine learning enabling the development of AI-first applications across