Senior AI Evaluation Scientist

Oracle

Nashville (TN)

On-site

USD 115,000 - 235,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Medical, dental, vision insurance
401(k) with company match
Paid time off

Job summary

Oracle’s OCI AI Evaluation team seeks a Senior Applied Scientist to own complex evaluation work from problem definition to final recommendation. You will translate questions into measurable hypotheses, design benchmarks, and build evaluation pipelines.

You will work with researchers, engineers, product teams, and leadership to deliver defensible results. You will write robust Python code, validate data and metrics, and examine factors beyond aggregate scores such as robustness, data provenance,

Qualifications

  • PhD in Computer Science, ML, AI, Statistics or related quantitative field; or Master/Bachelor with equivalent industry experience.
  • Experience designing and executing ML experiments, defining hypotheses, datasets, metrics, baselines, and interpreting results.
  • Strong knowledge of modern ML, deep learning, NLP, and generative AI methods.
  • Hands-on experience evaluating LLMs, foundation models, AI agents, or related systems.
  • Proficiency in Python and production-oriented experiment code.
  • Experience with ML frameworks and data-science libraries (PyTorch, TensorFlow, HuggingFace, NumPy, pandas).
  • Ability to analyze complex datasets, perform statistical and error analyses, and identify data-quality issues.
  • Experience translating technical findings into clear recommendations for diverse stakeholders.
  • Ability to work independently while collaborating with cross-functional teams.
  • Strong written and verbal communication skills.

Responsibilities

  • Own end-to-end evaluations of foundation models and AI systems, from design to analysis and stakeholder review.
  • Translate customer and business needs into testable hypotheses, datasets, and metrics.
  • Design and maintain benchmarks and evaluation methods across ML capabilities.
  • Write production-oriented evaluation code and build reproducible pipelines and checks.
  • Evaluate model behavior across quality, cost, latency, reliability, safety, and domain fit.
  • Conduct statistical and failure-mode analyses to explain results and differences.
  • Develop automated evaluators, including LLM-as-a-judge methods, against human judgments.
  • Design human-evaluation workflows with rubrics and quality controls.

Skills

PhD in CS/ML
Experiment design
Statistical analysis
Technical writing
Independent work
Strong communication

Education

PhD in Computer Science/ML
Master/Bachelor with equivalent experience

Tools

PyTorch
TensorFlow
HuggingFace
NumPy
pandas

Job description

Oracle’s OCI AI Evaluation team seeks a Senior Applied Scientist to own complex evaluation work from problem definition to final recommendation. You will translate questions into measurable hypotheses, design benchmarks, and build evaluation pipelines.

You will work with researchers, engineers, product teams, and leadership to deliver defensible results. You will write robust Python code, validate data and metrics, and examine factors beyond aggregate scores such as robustness, data provenance,

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Applied Scientist: AI Evaluation & Benchmarks
Senior Applied Scientist: AI Evaluation & Benchmarks

Oracle • United States

On-site
USD 115,000 - 235,000
Medical, dental, and vision insurance
Short/Long-term disability
Life insurance & AD&D
+2
Senior AI Evaluation Scientist — Benchmarks & Systems
Senior AI Evaluation Scientist — Benchmarks & Systems

Oracle • Santa Clara (CA)

On-site
USD 115,000 - 235,000
Medical, dental, and vision insurance
401(k) Savings with company match
Paid time off and holidays
+1
Senior Applied Scientist - End-to-End AI Evaluations
Senior Applied Scientist - End-to-End AI Evaluations

Oracle • Austin (TX)

On-site
USD 115,000 - 234,000
Health insurance
401(k) plan
Paid time off
Senior Foundation-Model Evaluation Scientist
Senior Foundation-Model Evaluation Scientist

Oracle • Seattle (WA)

On-site
USD 115,000 - 235,000
401(k) matching
Employee Stock Purchase Plan
Paid time off
+2
Senior Principal AI Platform Scientist
Senior Principal AI Platform Scientist

Oracle • Nashville (TN)

On-site
USD 158,000 - 355,000
Health benefits
401(k) match
Paid time off
Senior Applied Scientist: Research & AI Strategy
Senior Applied Scientist: Research & AI Strategy

Oracle • Seattle (WA)

On-site
USD 158,000 - 355,000
Medical insurance
Dental insurance
Vision insurance
+18
Senior Principal AI Scientist - Enterprise Cloud Platform
Senior Principal AI Scientist - Enterprise Cloud Platform

Oracle • United States

On-site
USD 158,000 - 355,000
Medical insurance
401(k) with company match
Paid time off
Senior Applied Scientist
Senior Applied Scientist

Oracle • Santa Clara (CA)

On-site
USD 115,000 - 235,000
Medical, dental, and vision insurance
401(k) Savings with company match
Paid time off and holidays
+1
Senior Applied Scientist
Senior Applied Scientist

Oracle • Nashville (TN)

On-site
USD 115,000 - 235,000
Medical, dental, vision insurance
401(k) with company match
Paid time off
Senior Applied AI Scientist & Research Strategy Lead
Senior Applied AI Scientist & Research Strategy Lead

Oracle • United States

On-site
USD 158,000 - 355,000
Medical, dental, and vision insurance
401(k) Savings and Investment Plan
Paid time off
+2