Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Turing in Boston is seeking a software engineer to create and refine datasets and code examples for training and benchmarking large language models, with a focus on Python and full-stack deployment.
You'll build automated verification tools to evaluate AI-generated code for efficiency, scalability, and reliability, collaborating with ML teams to improve model evaluation workflows.
Create and refine high-quality datasets and code examples to train and benchmark large language models. Develop automated verification tools in Python to evaluate AI-generated code for efficiency, scalability, and reliability.
Requirements: Requires at least 3 years of software engineering experience with deep expertise in Python and full-stack application deployment. Strong understanding of software architecture and the ability to provide structured evaluation rationales is essential.
Key Skills: Python, JavaScript, ReactJS, C/C++, Java, Rust, Go, Software Architecture, Full-stack Development, ML Infrastructure, Data Pipelines, Code Review, Debugging, API Design, Automated Verification Tools, Large Language Models