AI Evaluator - Java (Freelance Opportunity)

Biz Tech Consultants

New Delhi

On-site

INR 1,339,000 - 2,009,000

Part time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Biz Tech Consultants is seeking experienced software engineers to work on RLHF projects supporting AI coding agent training and evaluation. You will design technical tasks, review AI-generated responses for correctness, and provide evidence-based feedback across multiple languages and tools.

The role demands strong Java skills, production experience, and comfort with Docker, databases (SQL and NoSQL), and remote or international teams.

Qualifications

  • 4 to 7 years of experience building production systems.
  • Strong hands-on skill in Java.
  • Comfortable working with Docker and command line tools.
  • Comfortable working with one relational and one NoSQL database.
  • Can point to a clear before and after metric from past work.
  • Experience working with remote or international clients or teams.
  • Some exposure to secure data handling or compliance work.

Responsibilities

  • Design technical tasks across areas such as software development, debugging, data processing, security, and machine learning, and set up any supporting environment and automated tests needed to validate them, where the project calls for task authoring.
  • Review AI generated responses to technical tasks, assessing correctness, code quality, and adherence to instructions.
  • Provide clear, evidence based written feedback and comparative ratings across responses.
  • Validate task or evaluation difficulty against AI model outputs and revise based on review feedback.

Skills

Java
Python
JavaScript
Node.js
Go
Rust
CI/CD
System Design
AI Integration
LLM

Tools

Docker
Kubernetes
Kafka
RabbitMQ
gRPC
Redis
Git
pytest

Job description

Job Description

We are looking for experienced software engineers to work on RLHF projects that support the training and evaluation of AI coding agents. Depending on the project, the work can involve designing realistic technical tasks, or reviewing and rating AI generated responses for correctness, clarity, and reliability. It suits engineers who enjoy precise technical writing and careful quality review as much as building.


Key Responsibilities

Design technical tasks across areas such as software development, debugging, data processing, security, and machine learning, and set up any supporting environment and automated tests needed to validate them, where the project calls for task authoring.

Review AI generated responses to technical tasks, assessing correctness, code quality, and adherence to instructions.

Provide clear, evidence based written feedback and comparative ratings across responses.

Validate task or evaluation difficulty against AI model outputs and revise based on review feedback.


Desired Candidate Profile
Highly Required
  • 4 to 7 years of experience building production systems, not just internal tools.
  • Strong hands-on skill in Java
  • Comfortable working with Docker and command line tools.
  • Comfortable working with one relational and one NoSQL database.
  • Can point to a clear before and after metric from past work.
  • Experience working with remote or international clients or teams.
  • Some exposure to secure data handling or compliance work.
Preferred
  • Has integrated an AI or LLM API into a real, shipped feature.
  • Experience in AI training and evaluation work.
  • Prior experience debugging, reviewing code, and writing documentation or test cases.

Employment Type Contractual/Freelance

Key Skills
  • Python
  • Java
  • JavaScript
  • Node.js
  • Go
  • Rust
  • Docker
  • Git
  • pytest
  • AI Integration
  • LLM
  • NoSQL
  • CI/CD
  • System Design
Nice to Have
  • Kubernetes
  • Kafka
  • RabbitMQ
  • gRPC
  • Redis

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Evaluator - Freelance Opportunity (Python)
AI Evaluator - Freelance Opportunity (Python)

Biz Tech Analytics • New Delhi

On-site
INR 1,500,000 - 2,500,000
AI Agent Evaluation Analyst (Freelance)
AI Agent Evaluation Analyst (Freelance)

Mindrift • Pune District

Remote
Flexible remote work
Competitive pay up to $17/hour
Experience on advanced AI projects
AI/ML Evaluator - English
AI/ML Evaluator - English

Welocalize • Delhi

On-site
INR 977,000 - 1,391,000
AI Agent Evaluation Analyst (Freelance)
AI Agent Evaluation Analyst (Freelance)

Mindrift • Maharashtra

Remote
Flexible schedule
Competitive pay up to $12/hour
Gain experience in advanced AI projects
AI/ML Evaluator - English(India)
AI/ML Evaluator - English(India)

Welocalize • Delhi

On-site
INR 1,058,000 - 1,454,000
AI Agent Evaluation Analyst (Freelance)
AI Agent Evaluation Analyst (Freelance)

Mindrift • Ahmedabad District

Remote
Flexible project schedule
Competitive hourly rates
Experience in advanced AI projects
Python AI Developer
Python AI Developer

Nexon Software Solutions • Hyderabad, Chennai District, Bengaluru

Hybrid
INR 1,500,000 - 2,800,000
AI/ML Engineer
AI/ML Engineer

Vaiticka Solution • Dadri

On-site
INR 2,000,000 - 3,500,000
Freelance Agent Evaluation Analyst
Freelance Agent Evaluation Analyst

Mindrift • New Delhi

On-site
INR 1,250,090 - 1,500,108
Flexible work hours
Competitive pay up to $12/hour
Experience on advanced AI projects
AI Engineer
AI Engineer

Benchmarkit • Pune District

On-site
INR 600,000 - 1,200,000