Content
62%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body delivers an exceptionally clear, well-validated TDD workflow with concrete examples, templates, and checklists. Its weaknesses are significant verbosity with repeated content across sections, and a progressive-disclosure structure that inlines material it says to split while pointing to five bundle files that are not actually present.
Suggestions
Consolidate duplicated content: state the Iron Law and RED-GREEN-REFACTOR cycle once, merge the description guidance in "SKILL.md Structure" with the "Skill Discovery Optimization" section, and cut the body toward its own <500-word target by moving extended examples out.
Actually provide (or remove references to) the five cited bundle files — anthropic-best-practices.md, testing-skills-with-subagents.md, persuasion-principles.md, graphviz-conventions.dot, and render-graphs.js — since the core pressure-scenario methodology currently dead-ends at a missing file.
Inline a minimal fallback for writing pressure scenarios (the one operation the workflow depends on but delegates entirely to a missing file), so the skill remains actionable without the bundle.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body runs ~3,800 words — nearly 8x its own stated "<500 words" target for skills — with structural redundancy: the Iron Law appears in full twice ("The Iron Law" section and "The Bottom Line"), the RED-GREEN-REFACTOR cycle is explained three times (mapping table, dedicated section, checklist), and description-authoring rules are duplicated between "SKILL.md Structure" and "Skill Discovery Optimization". This is "noticeably verbose; several unnecessary explanations or padded sections" rather than the occasional tightening opportunity of anchor 3. | 2 / 5 |
Actionability | Guidance is highly concrete: good/bad YAML description examples, a full SKILL.md structure template, exact commands ("wc -w skills/path/SKILL.md", "./render-graphs.js ../some-skill --combine"), a five-step micro-testing protocol, and per-item checklists. It falls short of fully executable because the core operation — how to actually write and run pressure scenarios — is deferred to testing-skills-with-subagents.md, a file that does not exist in the bundle. | 4 / 5 |
Workflow Clarity | The RED → GREEN → REFACTOR sequence is explicit with validation checkpoints and feedback loops: "Run scenarios WITHOUT skill - document baseline behavior verbatim", "Run same scenarios WITH skill. Agent should now comply", "Re-test until bulletproof", plus a mandatory per-item creation checklist. This matches the top anchor: clear sequence, explicit validation steps, and checklists for a complex process. | 5 / 5 |
Progressive Disclosure | References are one level deep and clearly signaled ("See anthropic-best-practices.md", "See [testing-skills-with-subagents.md](...)"), but the body inlines large sections its own rules say belong in separate files (extended SDO examples, testing methodology, code-example guidance) at ~3,800 words. More importantly, every referenced file — anthropic-best-practices.md, graphviz-conventions.dot, render-graphs.js, persuasion-principles.md, testing-skills-with-subagents.md — is missing from the bundle, so the navigation structure it advertises does not actually exist. | 3 / 5 |
Total | 14 / 20 Passed |