Content
71%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-organized, opinionated instruction skill with concrete loop rules, a useful seam-confirmation checkpoint, and genuinely instructive anti-pattern examples — no filler. Its one real defect is that two of its references (tests.md, mocking.md) point to files absent from the bundle, which undercuts both the promised examples and the mocking guidance.
Suggestions
Ship the referenced files: create tests.md and mocking.md in the bundle, or remove the links and inline the essential content — dangling references currently leave the good-test section's examples and all mocking rules missing.
Inline one short example of a specification-style test next to the '什么是好 test' section so the section is self-sufficient even if references are not loaded.
State the red->green cycle as an explicit ordered sequence including 'run the test and confirm it fails' to make the loop's built-in validation checkpoint explicit.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is lean and adds judgment the model can't infer (pre-approved seams, tautological-test detection with a concrete `expect(add(a, b)).toBe(a + b)` example, vertical-slices rule), with no padding explaining what TDD is at textbook length. A few sentences could be trimmed — e.g., the explanation of what the `codebase-design` skill is ('是供查阅的 reference,而不是要运行的 session') — so it sits at anchor 4 ('efficient; minor instances of over-explanation') rather than anchor 5's every-token-earns-its-place. | 4 / 5 |
Actionability | Concrete, executable instruction guidance throughout: '先写 failing test,再只写足够让它通过的代码', the pre-test seam-confirmation step with the exact question to ask ('What's the public interface, and which seams should we test?'), and recognizable anti-pattern signatures. It does not reach 5 because key detail is deferred to files that are not in the bundle — mocking rules ('mocking 规则见 mocking.md') and test examples ('示例见 tests.md') point at files that do not exist. | 4 / 5 |
Workflow Clarity | The loop is clearly sequenced with rules that function as checkpoints: red before green, one seam/one test per cycle, refactoring explicitly routed to the review stage, plus an upfront checkpoint (confirm seams with the user before writing any test). This is an instruction skill with no destructive or batch operations, so no validation cap applies; it falls just short of anchor 5 because the red->green cycle is stated as rules rather than an explicit ordered cycle with a verify-the-test-fails step. | 4 / 5 |
Progressive Disclosure | Section structure is clear and the two references ('[tests.md](tests.md)', '[mocking.md](mocking.md)') are one level deep and clearly signaled, but they are dangling — the bundle contains no tests.md or mocking.md, so the promised examples and mocking rules are unreachable. Per the guideline to score against the actual bundle structure, broken references leave this at anchor 3 (structure present but organization undermined) rather than 4, and it cannot take the under-50-lines simple-skill exception because the skill does declare a need for external references. | 3 / 5 |
Total | 15 / 20 Passed |