Content
93%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A tight, high-quality process skill: lean prescriptive rules with no filler, concrete per-phase steps and required outputs, and appropriately split detail across two existing one-level-deep reference files. The single gap is the absence of an explicit error-recovery loop (what to do when GREEN or REFACTOR validation fails) that would complete the workflow's feedback cycle.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is lean and purely prescriptive — "Execute one complete TDD cycle per requirement. Never test private internals." — with zero space spent explaining what TDD, mocking, or KISS/DRY are. Every section adds rules Claude does not already know (e.g., "Shameless Green is acceptable for a single new test"), matching the 5 anchor; the only near-redundancy (two overlapping UI-query bullets) is too minor to drop it to the 4 anchor. | 5 / 5 |
Actionability | For an instruction-only skill the guidance is fully executable: each phase has an Input, numbered steps, and explicit present-items, including a copy-paste-ready literal ("The behavioral test `<test name>` still passes.") and concrete selector commands ("Prefer `getByRole` with accessible name"), with examples delegated to references/tests.md. This matches the 5 anchor's specificity and coverage; the 4 anchor would require missing key execution details, which are absent. | 5 / 5 |
Workflow Clarity | The RED→GREEN→REFACTOR sequence is explicit with inputs, outputs, and validation checkpoints ("Confirmation that the test now passes, or the exact blocker if it cannot be run", the stop-gate "Stop after RED and wait for confirmation", and a Constraints checklist). It falls short of the 5 anchor because there is no error-recovery loop — e.g., what to do when the refactor breaks the test — only blocker reporting, which fits the 4 anchor (clear sequence, most checkpoints, minor gaps). | 4 / 5 |
Progressive Disclosure | Clear overview sections with two well-signaled, contextual, one-level-deep references — "Read `references/mocking.md` when dependency boundaries or mocks are involved" and "Read `references/tests.md` for examples and red flags" — both verified to exist as real, substantive files (65 and 58 lines), with no nesting. This matches the 5 anchor exactly; the 4 anchor would require buried references or inline content that belongs in a separate file, which is not the case. | 5 / 5 |
Total | 19 / 20 Passed |