Content
63%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A clearly structured, opinionated TDD workflow with strong sequencing, checklists, and concrete rules of engagement. Its weaknesses are a padded Philosophy section that re-explains what Claude already knows, missing error-recovery branches, and five dead file references that promise content the bundle does not deliver.
Suggestions
Trim the Philosophy section to 3-4 directive bullets (e.g. 'Test behavior via public interfaces only; if a refactor breaks a test, the test was testing implementation') and cut the discursive prose in the Anti-Pattern section.
Add the referenced files (tests.md, mocking.md, deep-modules.md, interface-design.md, refactoring.md) to the bundle — or remove the dead links and inline the one or two that carry essential guidance.
Add brief error-recovery branches to the loop: what to do when a new test passes immediately (delete or question the test), when the plan proves wrong mid-cycle, and when growing duplication signals it is time to refactor mid-loop.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The Philosophy section spends three paragraphs explaining concepts Claude already knows (good vs. bad tests, mocking internal collaborators, tests breaking on refactor), and the Anti-Pattern section is similarly discursive ('You outrun your headlights'). This is 'mostly efficient but includes some unnecessary explanation' rather than anchor 4 — the Philosophy section alone could be trimmed to a few directive bullets without losing guidance value. | 3 / 5 |
Actionability | As an instruction-only skill the guidance is concrete and directive: 'Write ONE test that confirms ONE thing', 'Only enough code to pass current test', 'Never refactor while RED', a specific question to ask the user ('What should the public interface look like?'), and a worked behavior-naming example ('user can checkout with valid cart'). Not score 5 because no example test or RED/GREEN cycle in a real framework is shown and the referenced example files (tests.md) do not exist, leaving minor gaps. | 4 / 5 |
Workflow Clarity | The four-step workflow (Planning → Tracer Bullet → Incremental Loop → Refactor) is clearly sequenced with three checklists and explicit cycle validation ('test fails' → 'test passes', 'Run tests after each refactor step'). Not score 5 because there are no error-recovery branches — e.g. what to do when a new test unexpectedly passes, when implementation reveals the plan was wrong, or when to abandon a cycle — so feedback loops exist but only on the happy path. | 4 / 5 |
Progressive Disclosure | The body is well-sectioned and its inline references (tests.md, mocking.md, deep-modules.md, interface-design.md, refactoring.md) are clearly signaled and one level deep, which is better organized than anchor 2. However, none of the five referenced files exist in the bundle (no references/, scripts/, or assets/ directories), so navigation leads nowhere and the promised detail content is simply missing — a real organizational defect that keeps it below anchor 4. | 3 / 5 |
Total | 14 / 20 Passed |