Software Engineering Manager - Build Agent

ServiceNow

California (MO)

On-site

USD 166,500 - 291,400

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ServiceNow is seeking a senior AI evaluation lead to own the Build Agent evaluation framework, roadmap, and telemetry, with responsibility for scoring, failures, and model benchmarking.

You will manage an eight-engineer team, align across teams, and drive fixes with cross-functional partners to ensure scalable, measurable improvements. This role offers a high-impact path in AI-native product development.

Qualifications

  • Direct experience building or leading evaluation systems for LLM-based products — scoring methodology, benchmark design, and failure analysis.
  • Experience driving cross-team accountability without direct authority — getting other teams to own and fix issues surfaced by data.
  • A working understanding of coding agent architecture: tool use, agentic loops, context management, and tradeoffs between generic and domain-specific scaffolding.
  • 6+ years of experience with technologies relevant to ServiceNow, including advanced coding skills and fluency in Java, C++, Ruby, Shell, or JavaScript.
  • Experience critically evaluating foundation models — distinguishing real capability differences from tuning or scaffold artifacts.
  • Ability to execute against ambiguous priorities, weighing context, risk, and outcomes.
  • People management experience is required; managing 8+ engineers is a plus.

Responsibilities

  • Own eval strategy and roadmap: prompt set design, scoring methodology, failure-mode taxonomy, telemetry.
  • Coordinate across teams to arbitrate quality standards and drive fixes, not just reports.
  • Lead model benchmarking efforts, recommending model support decisions.
  • Manage the daily activities of an 8-engineer team, including staffing and mentoring.
  • Solve cross-cutting problems where eval signal, model capability, and architecture intersect.
  • Represent eval results and model recommendations to cross-functional and executive stakeholders.

Skills

LLM evaluation
Cross-team leadership
Java
C++
Ruby
JavaScript
Shell
People management

Tools

Evaluation tooling
Agentic loops
Context management

Job description

It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. That moment inspired Fred to build a company that could do that for everyone—freeing people from busywork so they could focus on meaningful work. Today, ServiceNow is the AI control tower for business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any workflow— helping 85% of the Fortune 500® work smarter, faster, and better. We're building an AI-native culture where technology and talent are unstoppable together. And we're just getting started.

Join us to put AI to work for people.

The team

Build Agent is ServiceNow's AI coding assistant, purpose-built for the platform's metadata-driven substrate - operating natively across ServiceNow's scoped applications, tables, and metadata types. This role leads the team responsible for the Build Agent evaluation framework, model support, and telemetry.

The framework is critical to Build Agent success — it's already driven measurable wins: recovering Build Agent correctness, cutting latency and inference cost ahead of a major release, and catching high-severity defects that manual testing missed before they reached production. The team is 8 engineers and is currently the bottleneck on its own scale — this role exists to convert a high-performing initiative into a durable, scalable function.

Key areas of focus include:

  • Evaluation Infrastructure: Golden prompt sets, Pass@1 and functional scoring, failure categorization (plumbing vs. metadata-creation failures), and coverage across ServiceNow metadata types and UI workflows.
  • Model Support & Benchmarking: Structured evaluation of candidate foundation models against the production default, with failure-consistency analysis to separate scaffold/tuning issues from genuine capability gaps.
  • Telemetry: Token usage, cache efficiency, and inference cost tracked alongside correctness as first-class release signals.
  • Cross-Team Arbitration: Prioritizing eval coverage, defining what "good" means across surfaces owned by different contributing teams, and turning eval signal into committed fixes by the owning teams rather than open-ended findings.
What you get to do in this role
  • Own eval strategy and roadmap: prompt set design, scoring methodology, failure-mode taxonomy, and telemetry instrumentation.
  • Coordinate across the multiple teams contributing to Build Agent to arbitrate quality standards and drive eval findings to committed, owned fixes — not just reports.
  • Lead model benchmarking efforts, making data-backed recommendations on model support decisions (e.g., evaluating frontier coding models against the incumbent default).
  • Manage the daily activities of an 8-engineer team, including staffing, mentoring, and performance management, while building the scale (people, process, or automation) to remove the team as its own bottleneck.
  • Solve ambiguous, cross-cutting problems where eval signal, model capability, and platform architecture intersect — and where "correct" isn't obvious until you've defined the bar.
  • Represent eval results and model support recommendations to cross-functional and executive stakeholders, including in response to org-wide benchmark initiatives.

To be successful in this role you have:

  • Direct experience building or leading evaluation systems for LLM-based products — scoring methodology, benchmark design, and failure analysis, not just prompt engineering.
  • Experience driving cross-team accountability without direct authority — getting other teams to own and fix issues surfaced by your data.
  • A working understanding of coding agent architecture: tool use, agentic loops, context management, and the tradeoffs between generic and domain-specific scaffolding.
  • 6+ years of experience with technologies relevant to ServiceNow, including advanced coding skills and fluency in one or more of Java, C++, Ruby, Shell, or JavaScript.
  • Experience critically evaluating foundation models — distinguishing real capability differences from tuning or scaffold artifacts.
  • Ability to execute against ambiguous priorities, weighing context, risk, and desired outcomes — particularly where eval data is incomplete or contested.
  • People management experience is required; experience managing at scale (8+ engineers) or managing managers is a plus.

For positions in this location, we offer a base pay of $166,500 - $291,400, plus equity (when applicable), variable/incentive compensation and benefits. Sales positions generally offer a competitive On Target Earnings (OTE) incentive compensation structure. Please note that the base pay shown is a guideline, and individual total compensation will vary based on factors such as qualifications, skill level, competencies, and work location. We also offer health plans, including flexible spending accounts, a 401(k) Plan with company match, ESPP, matching donations, a flexible time away plan and family leave programs. Compensation is based on the geographic location in which the role is located and is subject to change based on work location.

Work Personas

We approach our distributed world of work with flexibility and trust. Work personas (flexible, remote, or required in office) are categories that are assigned to ServiceNow employees depending on the nature of their work and their assigned work location. Learn more here. To determine eligibility for a work persona, ServiceNow may confirm the distance between your primary residence and the closest ServiceNow office using a third-party service.

Equal Opportunity Employer

ServiceNow is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, national origin, age, disability, gender identity, veteran status, or any other category protected by law. In addition, all qualified applicants with arrest or conviction records will be considered for employment in accordance with legal requirements.

Accommodations

We strive to create an accessible and inclusive experience for all candidates. If you require a reasonable accommodation to complete any part of the application process, or are unable to use this online application and need an alternative method to apply, please contact globaltalentss@servicenow.com for assistance.

Export Control Regulations

For positions requiring access to controlled technology subject to export control regulations, including the U.S. Export Administration Regulations (EAR), ServiceNow may be required to obtain export control approval from government authorities for certain individuals. All employment is contingent upon ServiceNow obtaining any export license or other approval that may be required by relevant export control authorities.

From Fortune. ©2026 Fortune Media IP Limited. All rights reserved. Used under license.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineering Manager - Build Agent
Software Engineering Manager - Build Agent

ServiceNow • Santa Clara (CA)

On-site
USD 166,500 - 291,400
Health plans
401(k) Plan with company match
ESPP
+3
Software Engineering Manager - Build Agent
Software Engineering Manager - Build Agent

ServiceNow • Santa Clara (CA)

On-site
USD 166,000 - 292,000
Health plans
401(k) plan with company match
ESPP
+3
Senior Software Engineer - Agent Development
Senior Software Engineer - Agent Development

ServiceNow • California (MO)

On-site
USD 143,000 - 244,000
Equity
401(k) Plan with company match
Employee Stock Purchase Plan (ESPP)
+3
Senior Engineering Manager, Agentic & Generative AI Benchmarking and Evaluations
Senior Engineering Manager, Agentic & Generative AI Benchmarking and Evaluations

ServiceNow • Santa Clara (CA)

On-site
USD 201,000 - 353,000
Staff ServiceNow Developer
Staff ServiceNow Developer

ServiceNow • Illinois

On-site
USD 149,000 - 263,000
Health plans
401(k) Plan with company match
ESPP
+3
Staff AI Engineer - Conversational & Agentic AI
Staff AI Engineer - Conversational & Agentic AI

ServiceNow • Santa Clara (CA)

Hybrid
USD 176,000 - 309,000
Health plans
401(k) Plan with company match
Flexible time away plan
+1
Senior Staff, Product Manager- AI Evaluation & Quality
Senior Staff, Product Manager- AI Evaluation & Quality

ServiceNow • California (MO)

On-site
USD 190,000 - 335,000
Health plans
401(k) Plan with company match
ESPP
+3
Senior Software Engineer, Agent Eval Platform
Senior Software Engineer, Agent Eval Platform

Servicenow • Mountain View (CA)

On-site
USD 161,000 - 274,000
Health plans
401(k) with company match
Employee stock purchase plan (ESPP)
+2
Senior Software Engineer, Agent Eval Platform
Senior Software Engineer, Agent Eval Platform

Moveworks • Mountain View (CA)

Hybrid
USD 161,000 - 274,000
Senior Staff Applications Development Engineer
Senior Staff Applications Development Engineer

ServiceNow • Santa Clara (CA)

Hybrid
USD 191,000 - 334,000
Health plans
401(k) Plan with company match
ESPP
+3