Content
71%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-structured, opinionated TDD skill with a clear vertical-slice workflow, strong checklists, and commendable brevity. Its biggest defect is that all five of its reference links point to files that do not exist in the bundle, leaving the promised examples and guidelines unreachable, and the workflow lacks recovery guidance for cycles that get stuck in RED.
Suggestions
Create the five referenced files (tests.md, mocking.md, deep-modules.md, interface-design.md, refactoring.md) or remove/inline the dangling links — currently every external reference in the body is broken.
Add one short worked example of a behavior-focused test versus an implementation-coupled test, since the Philosophy section currently gestures at the distinction without showing it.
Add error-recovery guidance for the incremental loop, e.g., what to do when a test cannot reach GREEN after a few attempts (delete and rethink the test, or revisit the interface design).
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense and directive with no padding — the anti-pattern section ('DO NOT write all tests first...') delivers non-obvious judgment efficiently. The Philosophy section spends three paragraphs explaining good vs. bad tests, some of which (mocking, private methods) Claude already knows, so it is not quite anchor 5's 'every token earns its place' but comfortably above anchor 3. | 4 / 5 |
Actionability | As an instruction-only skill the guidance is concrete and executable: 'Write ONE test that confirms ONE thing', explicit rules ('Only enough code to pass current test', 'Never refactor while RED'), a per-cycle checklist, and a literal question to ask the user. It stays at anchor 4 rather than 5 because there is no worked example of a behavior-focused test, and the files that would have carried examples (tests.md, mocking.md) do not exist in the bundle. | 4 / 5 |
Workflow Clarity | The sequence (Planning with user approval → Tracer Bullet → Incremental Loop → Refactor) is clearly staged, the RED→GREEN cycle is itself a validate→fix feedback loop, and checklists plus 'Run tests after each refactor step' provide checkpoints. It falls short of anchor 5 because there is no error-recovery guidance for a cycle that cannot reach GREEN (e.g., what to do when stuck). | 4 / 5 |
Progressive Disclosure | The body links five files (tests.md, mocking.md, deep-modules.md, interface-design.md, refactoring.md), but no bundle files exist at all — every reference is dangling, so navigation is broken despite the links being inline and clearly signaled. Scoring against the actual bundle structure, this lands at anchor 3 ('structure present but references do not enable navigation'), not anchor 4, whose references resolve. | 3 / 5 |
Total | 15 / 20 Passed |