About the projects
We are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems. Oneederland approach is to build verifiable SWE tasks based on public repository histories in a synthetic approach with human‑in‑the‑loop, expanding the dataset coverage to different types of tasks in terms of programming language, difficulty level, and more.
About the Role
We are looking for experienced software engineers (tech lead level) familiar with high‑quality public GitHub repositories who can contribute to this project. The role involves hands‑on software engineering work, including automating development environments, triaging issues, and evaluating test coverage and quality.
Why Join Us?
Turing is one of the world’s fastest‑growing AI companies accelerating the advancement and deployment of powerful AI systems. You’ll be at the forefront of evaluating how LLMs interact with real code, influencing the future of AI‑assisted software development. This is a unique opportunity to blend practical software engineering with AI research.
What does day‑to‑day look like
- Analyze and triage GitHub issues across trending open‑source libraries.
- Set up and configure code repositories, including Dockerization and environment setup.
- Evaluate unit test coverage and quality.
- Modify and run codebases locally to assess LLM performance in bug‑fixing scenarios.
- Collaborate with researchers to design and identify repositories and issues that are challenging for LLMs.
- Lead a team of junior engineers on collaborative projects.
Required Skills
- Minimum 3+ years of overall experience.
- Strong experience with at least one of the following languages: Python.
- Proficiency with Git, Docker, and basic software pipeline setup.
- Ability to understand and navigate complex codebases.
- Comfortable running, modifying, and testing real‑world projects locally.
- Experience contributing to or evaluating open‑source projects is a plus.
Nice to Have
- Previous participation in LLM research or evaluation projects.
- 相关经验 building or testing developer tools or automation agents.
Freelancing Perks
- Work in a fully remote environment.
- Opportunity to work on cutting‑edge AI projects with leading LLM companies.
Commitments Required
- At least 4 hours per day and minimum 20 hours per week with overlap of 4 hours with PST. (Options: 20 hrs/week, 30 hrs/week, or 40 hrs/week)
Employment type
- Contractor assignment (no medical/paid leave)
Duration of contract
- 3 month; expected start date is next week.
After applying, you will receive an email with a login link. Please use that link to access the portal and complete your profile.