Tech Lead for Agent Evaluation Platform

Servicenow

Mountain View (CA)

Hybrid

USD 180,000 - 260,000

Full time

11 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ServiceNow is seeking an experienced engineer to own the evaluation framework for AI agents within Moveworks. You will design rubrics and judges, orchestrate large-scale multi-turn scenarios, and build stateful simulators that enable reliable, reusable evals across enterprise environments.

You’ll work across distributed systems, orchestration runtimes, and observability, using Python or Go to implement scalable backend services, with a focus on measurable agent performance and safe, reproducible

Qualifications

  • 10+ years building production backend or infrastructure systems.
  • Experience in 3 of: distributed systems, orchestration/workflows, observability internals, async programming, data pipelines, gRPC/protobuf.
  • Strong in Python or Go (ideally both).

Responsibilities

  • Build the eval judgement layer for agent evaluation: rubrics, judges, calibration against human labels.
  • Orchestrate end-to-end multi-turn agent scenarios at scale.
  • Develop stateful simulators and isolated sandboxes for reproducibility.
  • Enable self-serve eval workflows with reliable harness and rollout safeguards.

Skills

Python
Go
Distributed systems
Orchestration & workflows
Observability
Async programming
Data pipelines
gRPC/protobuf

Tools

Temporal
Airflow
Argo

Job description

ServiceNow is seeking an experienced engineer to own the evaluation framework for AI agents within Moveworks. You will design rubrics and judges, orchestrate large-scale multi-turn scenarios, and build stateful simulators that enable reliable, reusable evals across enterprise environments.

You’ll work across distributed systems, orchestration runtimes, and observability, using Python or Go to implement scalable backend services, with a focus on measurable agent performance and safe, reproducible

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer — Scale AI Evaluation & Orchestration
Staff Software Engineer — Scale AI Evaluation & Orchestration

Servicenow • Mountain View (CA)

On-site
USD 180,000 - 240,000
Senior AI Systems Engineer, Evaluation & Orchestration
Senior AI Systems Engineer, Evaluation & Orchestration

Servicenow • Mountain View (CA)

On-site
USD 161,000 - 274,000
Health plans
401(k) with company match
Employee stock purchase plan (ESPP)
+2
Senior AI Evaluation Systems Engineer
Senior AI Evaluation Systems Engineer

Servicenow • Santa Clara (CA)

On-site
USD 180,000 - 320,000
Senior ML Engineer: Agentic Evaluation & Reward Modeling
Senior ML Engineer: Agentic Evaluation & Reward Modeling

Servicenow • Santa Clara (CA)

Hybrid
USD 150,000 - 230,000
Senior Distributed Systems Engineer — AI Agent Orchestration
Senior Distributed Systems Engineer — AI Agent Orchestration

Servicenow • Mountain View (CA)

Hybrid
USD 180,000 - 240,000
LLM Evaluation Engineer – AI Agent Platforms
LLM Evaluation Engineer – AI Agent Platforms

Servicenow • Santa Clara (CA)

On-site
USD 126,000 - 195,000
Equity
Health plans
401(k) Plan with company match
+3
AI Eval & Benchmarking Manager - 8-Engineer Team Lead
AI Eval & Benchmarking Manager - 8-Engineer Team Lead

ServiceNow • California (MO)

On-site
USD 166,500 - 291,400
Staff Software Engineer, Agent Eval Platform
Staff Software Engineer, Agent Eval Platform

Servicenow • Mountain View (CA)

On-site
USD 180,000 - 240,000
AI Build Agent Engineering Manager
AI Build Agent Engineering Manager

ServiceNow • Santa Clara (CA)

On-site
USD 166,500 - 291,400
Health plans
401(k) Plan with company match
ESPP
+3
Tech Lead, Agent Eval Platform
Tech Lead, Agent Eval Platform

Servicenow • Mountain View (CA)

On-site
USD 180,000 - 260,000