Remote Part-Time JS Engineer — AI Benchmarking & LLM QA

Jointaro

United States

On-site

USD 74,000 - 105,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job description

Handshake is seeking motivated Software Engineers to evaluate Large Language Models (LLMs) in collaboration with top AI labs.

Handshake AI projects are remote and part-time opportunities. Must be located in the US and have proper work authorization (OPT and H-1B not supported).

Key responsibilities include developing and validating coding benchmarks by curating issues, solutions, and test suites from real-world repositories, ensuring comprehensive unit and integration tests for solution verification, and maintaining the consistency and scalability of benchmark task distribution. You'll also provide structured feedback on solution quality, debug and optimize benchmark code, and document processes for reproducibility.

Shape the future of AI while building skills and confidence to navigate an AI-driven job market. This role will leverage your engineering expertise in a flexible environment, but you must be able to commit 15+ hours/week. Pay starts at $65/hour, plus bonuses for task completion.

Handshake isn't offering any visa sponsorship support at this time.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

JavaScript Software Engineer
JavaScript Software Engineer

Jointaro • United States

Remote
USD 74,000 - 105,000
Remote Part-Time Python AI Engineer — LLM Benchmarking
Remote Part-Time Python AI Engineer — LLM Benchmarking

Jointaro • United States

On-site
USD 90,000 - 124,000
Python Software Engineer
Python Software Engineer

Jointaro • United States

Remote
USD 90,000 - 124,000
Remote Part-Time Java Engineer — LLM Benchmarking
Remote Part-Time Java Engineer — LLM Benchmarking

Jointaro • United States

On-site
Remote AI Research Fellow — LLM Prompts & Domain Expertise
Remote AI Research Fellow — LLM Prompts & Domain Expertise

Handshake • United States

On-site
AI QA Trainer - LLM Evaluation - Freelance Project
AI QA Trainer - LLM Evaluation - Freelance Project

Meridial • United States

Remote
Secure computer and high-speed internet required
Computer Science Expert
Computer Science Expert

Handshake • United States

On-site
Engineering Manager, AI Evaluation
Engineering Manager, AI Evaluation

Cacheflow • San Francisco (CA)

On-site
USD 180,000 - 240,000
Equity
401(k) match
Parental leave
+4
Remote AI Math Evaluator & LLM Benchmark Designer
Remote AI Math Evaluator & LLM Benchmark Designer

United States Digital Space LLC • United States

Remote
Fully remote
AI projects
Contract extension potential
STEM Professor at Handshake
STEM Professor at Handshake

aitrainer • Northern (KY)

Hybrid
USD 124,000 - 207,000