Senior Director - AI Evaluation Platform at ServiceNow
Santa Clara, CA
$279,100–$488,400from the description
Jul 7, 2026
Santa Clara, CA
Jul 21, 2026
What this job asks for AI summary
A senior leadership role responsible for building and scaling an AI evaluation platform — covering benchmark design, automated scoring pipelines, and verifier systems — used both by internal research and engineering teams and by external customers. The position involves setting strategy, managing a large organization of researchers and engineers, and advancing evaluation methodologies for generative and agentic AI. It suits a technically deep leader with a background in AI quality assurance or evaluation infrastructure at scale.
Principal level · San Jose-Sunnyvale-Santa Clara, CA · Full-time
Advertised as Senior, but the requirements read as Principal.
“or” means any one of them counts — you don't need all of them.
We read this from the posting text with AI. Skim the description below before ruling yourself out.
How this req sits in the market our data
Rare in this occupation — lead with these, and say what you built with them.
What the occupation pays Median $298,074 (middle half $224,695–$326,567). This posting is about at that midpoint.
Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.
Why we read it this way (8)
This is a people-management role (Senior Director) explicitly requiring leadership of organizations of 50+ engineers and researchers — mapped to 11-3021 (Computer and Information Systems Managers). The alt code 15-2051 reflects the deep technical AI/ML evaluation expertise also required.
The title 'Senior Director' is treated as the level advertised in the title of Senior (the word 'Senior' in the title), but the actual scope — org-wide AI evaluation platform strategy, 50+ person org, cross-functional executive stakeholder engagement — supports Principal-level seniority.
A Ph.D. or Master's in Computer Science, Machine Learning, or a related field is listed as 'a plus,' not a hard requirement, so the degree requirement is set to None.
The compensation range ($279,100–$488,400/year) is stated as a base pay guideline for a specific location and may vary by work location.
Work location is not explicitly pinned to a single city; ServiceNow's headquarters is Santa Clara, CA, and the compensation note references geographic location. The CBSA reflects that headquarters location. The role may support flexible/remote work personas per the posting's Work Personas section, but full remote is not confirmed.
AI evaluation methodologies, benchmarking frameworks, and automated scoring pipelines are treated as hard gates based on the 'To be successful in this role you have' requirements section. PyTorch/TensorFlow is listed as 'hands-on experience' in the same section and is therefore a hard gate, though the 'or TensorFlow' alternative is captured.
Agentic evaluation and safety/alignment evaluation appear as innovation areas the role will 'drive' or 'stay ahead of' — framed as forward-looking scope rather than explicit candidate gates, so marked preferred.
Ignored 2 non-technology phrase(s) as skills (responsibilities/concepts, not named tools): AI evaluation methodologies, safety and alignment evaluation.
Read the full posting
The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.