Senior AI Evaluation Engineer - Gatekeeper of Model Quality

Singapore Telecommunications Limited

Singapore

On-site

SGD 120,000 - 180,000

Full time

8 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Singtel, a leader in AI & Data Analytics, seeks a senior technical evaluator to design, calibrate, and adjudicate evaluation suites for agent archetypes. You will lead gate reviews under the Team Lead and own the regression-pack methodology, ensuring quality and drift hygiene for AI/ML changes.

You will mentor teammates, produce monthly quality reports, and influence build teams on evaluation rigor. 6+ years in ML/data/software with hands-on AI evaluation is required.

Qualifications

  • Bachelor's or Master's in CS or related field.
  • Experience with ML/AI evaluation, quality assurance, or models.
  • Strong ability to design and run evaluation suites for AI systems.
  • Proficiency in Python and evaluation tooling.
  • Experience with LLM or agent evaluation.

Responsibilities

  • Design and maintain offline evaluation suites and scoring pipelines.
  • Conduct operability gate reviews with documented findings.
  • Own regression pack methodology and adjudicate against baselines.
  • Calibrate gate thresholds to production reality and manage drift.
  • Produce monthly quality reports with trends and defect clusters.
  • Mentor team members and ensure work quality.

Skills

LLM Evaluation Design
Statistical rigour
Python & tooling
Data analysis & metrics
Tracing/observability
Adversarial testing
Gate-review leadership
Technical writing

Education

Bachelor's or Master's in Computer Science or related field

Tools

promptfoo
DeepEval
Custom evaluation harnesses
Reproducibility tooling

Job description

Singtel, a leader in AI & Data Analytics, seeks a senior technical evaluator to design, calibrate, and adjudicate evaluation suites for agent archetypes. You will lead gate reviews under the Team Lead and own the regression-pack methodology, ensuring quality and drift hygiene for AI/ML changes.

You will mentor teammates, produce monthly quality reports, and influence build teams on evaluation rigor. 6+ years in ML/data/software with hands-on AI evaluation is required.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Evaluation Architect
Senior AI Evaluation Architect

Singtel Group • Singapore

On-site
SGD 120,000 - 180,000
Head of AI Agent Evaluation & Instrumentation
Head of AI Agent Evaluation & Instrumentation

Singapore Telecommunications Limited • Singapore

On-site
SGD 180,000 - 240,000
Senior AI Evaluation Engineer #AIDA
Senior AI Evaluation Engineer #AIDA

Singtel Group • Singapore

On-site
SGD 120,000 - 180,000
Senior AI Evaluation Engineer #AIDA
Senior AI Evaluation Engineer #AIDA

Singapore Telecommunications Limited • Singapore

On-site
SGD 120,000 - 180,000
Head of AI Agent Quality & Instrumentation
Head of AI Agent Quality & Instrumentation

Singtel • Singapore

On-site
Confidential
Head of AI Agent Evaluation & Instrumentation
Head of AI Agent Evaluation & Instrumentation

Singtel Group • Singapore

On-site
SGD 120,000 - 180,000
Head of AI Agent Evaluation & Instrumentation
Head of AI Agent Evaluation & Instrumentation

Singtel • Singapore

On-site
Confidential
Associate Director/Senior Manager, Agent Evaluation & Instrumentation (AIDA)
Associate Director/Senior Manager, Agent Evaluation & Instrumentation (AIDA)

Singtel • Singapore

On-site
Confidential
AI Evaluation Architect
AI Evaluation Architect

Allegis Group Singapore Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
Lead Agent Evaluation & Instrumentation #AIDA
Lead Agent Evaluation & Instrumentation #AIDA

Singtel Group • Singapore

On-site
SGD 120,000 - 180,000