Get more replies from employers
Send a job-specific resume in minutes.
ServiceNow is seeking an expert in AI evaluation to advance its Build Agent team, the AI coding assistant for the platform. You will work on evaluation infrastructure, scoring, and failure analysis to ensure high-quality model performance across ServiceNow metadata types and workflows.
Responsibilities include large-scale eval orchestration and improving agent observability and tracing. A track record in LLM evaluation and model benchmarking is required; you will influence release quality and
ServiceNow is seeking an expert in AI evaluation to advance its Build Agent team, the AI coding assistant for the platform. You will work on evaluation infrastructure, scoring, and failure analysis to ensure high-quality model performance across ServiceNow metadata types and workflows.
Responsibilities include large-scale eval orchestration and improving agent observability and tracing. A track record in LLM evaluation and model benchmarking is required; you will influence release quality and