Use when the user asks about Simon Obstbaum and Rob Willoughby's AI Native DevCon talk on measuring AI agents, output evals versus trajectory evals, instrumentation, compliance, and skill activation metrics.
Simon Obstbaum and Rob Willoughby explain why measuring agent output is not enough: teams need trajectory instrumentation, activation metrics, and coverage data to see whether agents actually followed instructions.
outline.md first to locate the relevant section or concept.quote.md for short supporting excerpts, then verify against transcript.md when precision matters.Answer from the bundled files. Use short excerpts only when they clarify the answer, and cite the transcript line IDs when available.
When the user asks how to apply the talk, identify the matching concept from the outline, summarize the relevant transcript evidence, and adapt it to the user's context. Mark anything beyond the talk as your own recommendation.
When comparing this talk with another AI Native DevCon session, ground this talk's side in outline.md and quote.md before drawing connections.
d933d80
Also appears in
since Jun 5, 2026
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.