Remote Applied AI Research Scientist (LLM & Evaluation)
Rex.zone
United States
Remote
USD 41,328 - 68,880
Full time
14 days+
Application generator
Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Get past ATS filters
Job summary
A tech company specializing in AI research is seeking a mid-senior level researcher to manage applied AI research projects. The role involves end-to-end research cycles, building and evaluating LLM systems, and collaborating on dataset development. The ideal candidate should have strong Python skills, LLM evaluation experience, and strong written communication skills. This is a remote, full-time position with competitive hourly compensation ranging from $30 to $50.
Qualifications
Mid-Senior experience delivering applied ML research or productionized ML evaluation.
Hands-on LLM evaluation, prompt evaluation, or RLHF experience.
Familiarity with dataset development: data labeling, QA evaluation, and guideline compliance checks.
Responsibilities
Own end-to-end applied research cycles: problem framing, baselines, ablations, and reporting.
Build and evaluate LLM systems using offline metrics and human-in-the-loop evaluation.
Perform error analysis and model debugging to improve robustness, safety, and helpfulness.
Skills
Python
Applied ML research
LLM evaluation
Experiment design
Strong written communication
Tools
PyTorch
Job description
A tech company specializing in AI research is seeking a mid-senior level researcher to manage applied AI research projects. The role involves end-to-end research cycles, building and evaluating LLM systems, and collaborating on dataset development. The ideal candidate should have strong Python skills, LLM evaluation experience, and strong written communication skills. This is a remote, full-time position with competitive hourly compensation ranging from $30 to $50.