Content
88%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-crafted workflow skill: tight imperative rules, executable specifics (globs, naming examples, mocking decision rules), and a fully sequenced RED-GREEN-REFACTOR procedure with validation and error-recovery checkpoints. The only deductions are repeated code-quality routing explanations and references to `rules/*.md` files that are not present in the bundle for verification.
Suggestions
Consolidate the three explanations of Skill('code-quality') routing (Step 2 REFACTOR, GREEN phase note, and Code Quality section) into one authoritative statement to cut redundancy.
Ship the referenced `rules/red.md`, `rules/green.md`, `rules/refactor.md`, and `rules/test-after.md` files in the bundle so the clearly signaled one-level references resolve.
The GREEN-phase readability primitives duplicate content implied by the code-quality skill; a one-line pointer would suffice and reduce token cost.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense, imperative rule-setting with almost no tutorial padding ("Write exactly ONE failing test. Run it. Confirm it fails with the expected error. Do NOT write implementation code."). Not 5 because the `Skill('code-quality')` routing rationale is explained three times (Step 2 REFACTOR, GREEN phase, Code Quality section) and could be consolidated; not 3 because nothing explains concepts Claude doesn't need explained. | 4 / 5 |
Actionability | Guidance is copy-paste executable: exact glob patterns ("**/*.test.*", "**/*.spec.*"), runner discovery sources (package.json, Makefile, pyproject.toml...), a concrete naming example ("describe('createOrder')" → "it('should reject order when inventory is zero')"), factory pattern (`buildUser(overrides?)`), and a 3-mock design rule. Per the rubric's instruction-skill note, absence of code isn't penalized when guidance is this actionable — and it covers the common cases. | 5 / 5 |
Workflow Clarity | Steps 0–4 are clearly sequenced with explicit validation checkpoints at every phase: RED ("Confirm it fails with the expected error"), GREEN ("Run the full relevant test suite to check for regressions"), completion ("stop and fix it before proceeding. Never accumulate broken tests"), plus a "When Things Go Wrong" section that gives error-recovery feedback loops and a 2-failed-attempts restart rule. This matches the strongest anchor's validate→fix→retry pattern. | 5 / 5 |
Progressive Disclosure | The overview is well structured and phase details are split into clearly signaled one-level references ("See `rules/red.md`", "See `rules/green.md`", "See `rules/refactor.md`", "See `rules/test-after.md`"). Not 5 because no bundle directories (references/, scripts/, assets/, or rules/) are present, so the referenced rule files are dangling and navigation cannot be confirmed against the actual bundle; not 3 because the references that exist in the text are clearly signaled and content placement is appropriate. | 4 / 5 |
Total | 18 / 20 Passed |