This role sits within APEX (AI Foundations), the team building the platform-level infrastructure behind ServiceNow's enterprise AI strategy. You'll work alongside engineering, data, and applied research partners to make evaluation a durable, compounding advantage. Agentic AI is only as trustworthy as the evaluation discipline behind it. ServiceNow is building the evaluation methodology, tooling, and closed-loop production system that will define how good is good enough for AI specialists and multi-agent frameworks across the enterprise — and we're looking for the product leader to help own that discipline. This is a rare chance to build a category-defining evaluation platform from the ground up, at a company where the outcome shapes how every business unit ships AI, not just one team's roadmap. Evaluation is the foundational differentiator for enterprise agentic AI. As frontier models commoditize, defensible advantage shifts to the orchestration and quality layer — how reliably an AI specialist performs in a customer's production environment. This role sits at exactly that inflection point: You set direction in undefined space — there is no existing playbook to inherit, and the discipline you build becomes the standard others build on. Your impact is horizontal and enterprise-wide by design, driven through influence, tooling, and a Center of Excellence rather than headcount. You'll be the definitive point of reference for AI specialist quality across the portfolio — deep technical and methodological ownership without the dilution of people management. Success is visible and concrete: your framework adopted across business units, a closed-loop system demonstrably improving specialists release-over-release, and tooling used self-serve by teams you never directly staffed.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed