Optimize your skills and tiles: review SKILL.md quality, generate eval scenarios, run evals, compare across models, diagnose gaps, and re-run until scores improve.
91
90%
Does it follow best practices?
Impact
92%
1.13xAverage score across 25 eval scenarios
Passed
No findings from the security scan
Tessl evals compare success rates of agents with and without our optimized context