Measuring whether the AI actually helped users accomplish their goals.
18
3%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
Passed
No known issues
Fix and improve this skill with Tessl
tessl review fix ./gemini-extension/evaluation/skills/task-success-metrics/SKILL.mdOutput quality doesn't guarantee task success. The AI might produce a beautiful response that doesn't actually help the user do what they came to do. Task success metrics measure the end-to-end outcome.
For each user task, define:
These can diverge:
f41b650
Also appears in
since May 8, 2026
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.