Content
71%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A tight, opinionated body that gives concrete rules, a user-confirmation checkpoint, and crisp anti-pattern guidance without padding. Its one real defect is structural: it points readers to tests.md and mocking.md, which are absent from the skill bundle, so the promised detail is unreachable.
Suggestions
Add the referenced files (tests.md with good/bad test examples, mocking.md with mocking guidelines) to the skill bundle, or remove the links and inline the one or two examples that matter.
Add an explicit 'verify the test actually fails (red) before implementing' step to the Rules of the loop to close the validation gap in the loop sequence.
Trim concept re-explanations Claude already knows (the definition of a seam, the mechanics of tautological assertions) to tighten the body further.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is lean and directive with no filler, no library tours, and no basic-concept padding; sections like 'Rules of the loop' are one line each. A few sentences re-explain concepts Claude already knows (e.g., defining 'seam', explaining what a tautological test is at length with examples like `expect(add(a, b)).toBe(a + b)`), which keeps it just below the 'every token earns its place' anchor. | 4 / 5 |
Actionability | For an instruction-only skill the guidance is concrete: an exact question to ask ('What's the public interface, and which seams should we test?'), explicit rules ('Write the failing test first, then only enough code to pass it', 'One seam, one test, one minimal implementation per cycle'), and a named tool invocation ('call the Skill tool with "codebase-design"'). It misses anchor 5 because the concrete good-test and mocking examples are delegated to files that don't exist in the bundle, leaving minor gaps in executable specificity. | 4 / 5 |
Workflow Clarity | The loop is clearly sequenced — agree seams with the user, red (failing test) before green (minimal implementation), one slice at a time — and includes an explicit checkpoint ('write down the seams under test and confirm them with the user. No test is written at an unconfirmed seam'). It stops short of anchor 5: there is no explicit verify-the-test-fails step or error-recovery loop, and refactoring is delegated out without saying when review happens. | 4 / 5 |
Progressive Disclosure | Sections are well organized and the references are clearly signaled ('See [tests.md](tests.md)... [mocking.md](mocking.md)'), but the bundle contains no such files — no references/, scripts/, or assets/ directories exist — so both links are dangling and navigation breaks. Scored against the actual (empty) bundle structure, this is 'structure present but references not resolvable', between the broken/minimal-structure anchors and the good-structure anchor 4. | 3 / 5 |
Total | 15 / 20 Passed |