Content
88%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A high-quality decision guide: actionable, correctly sequenced, with explicit per-state verification and hard-won pitfalls stated as rules rather than prose. The only weaknesses are mild redundancy around the off-state cost argument and no use of bundle files to split reference-grade detail out of SKILL.md.
Suggestions
Deduplicate the off-state cost discussion: directive items 3–4 and the 'Is the off-state run worth its cost?' section make the same argument, and 'test/e2e/app-dir/use-offline/' is cited three times — one canonical treatment would cut tokens without losing information.
Move the Pitfalls section (or the detailed test-axis pattern walkthrough) into a references/ file, keeping a one-line pointer and the highest-frequency pitfalls in SKILL.md, so the decision guide loads leaner while details stay one level deep.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense and assumes Claude's competence — it never explains what Jest, skip pragmas, or test fixtures are; every section carries non-obvious repo knowledge ('lazy conditions read the resolved config, never `process.env`', 'a pragma on a skipped test errors as ambiguous'). It is not a 5 because of repetition: the off-state cost calculus appears in directive item 3 and again nearly verbatim in 'Is the off-state run worth its cost?', and the working example 'test/e2e/app-dir/use-offline/' is cited three times. It is not a 3 because there is no padding and no explanation of concepts Claude already knows. | 4 / 5 |
Actionability | Guidance is fully executable: a concrete next.config.js axis-keying example, the pragma form '// @gate concurrentRouterQueue', copy-paste verification commands with env vars ('NEXT_SKIP_ISOLATE=1 pnpm test-start-webpack test/e2e/app-dir/<suite>/<suite>.test.ts', '__NEXT_TEST_AXIS=A …', 'pnpm test-unit test/unit/gate/'), and named working examples to copy from. Specific examples cover the common cases (behavior flag, new API, impossible-to-run), matching the anchor for copy-paste-ready commands. | 5 / 5 |
Workflow Clarity | The decision workflow is clearly sequenced by question type (items 1–5 under 'Choosing the directive'), and the verification section provides explicit validation checkpoints with expected outcomes per state ('expect normal passes, no warnings', 'expect `⚠ gated test failed as expected (@gate …)`', 'expect `○ skipped` at collection, no fixture boot') plus a unit-test command. The Pitfalls section supplies error-recovery guidance (hard errors, cascade failures, stalling bodies). This matches the anchor for clear sequence with explicit validation steps and feedback for recovery. | 5 / 5 |
Progressive Disclosure | A single clearly-signaled, one-level-deep reference leads the document ('Full reference: test/lib/gate/README.md') and sections are well-organized with headers and a scannable anti-pattern table. It is not a 5 because the skill keeps all ~180 lines inline with no bundle files — the Pitfalls section and the detailed test-axis pattern are reference-grade material that could be split out; it is not a 3 because the one reference that does exist is prominently signaled and nothing is buried. | 4 / 5 |
Total | 18 / 20 Passed |