Analyze eval results, diagnose low-scoring criteria, fix tile content, and re-run evals — the full improvement loop automated
94
Quality
89%
Does it follow best practices?
Impact
98%
1.30xAverage score across 7 eval scenarios
Passed
No known issues
{
"name": "tessl-labs/eval-improve",
"version": "0.5.0",
"summary": "Analyze eval results, diagnose low-scoring criteria, fix tile content, and re-run evals — the full improvement loop automated",
"private": false,
"skills": {
"eval-improve": {
"path": "skills/eval-improve/SKILL.md"
}
}
}