Content
77%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is an exceptionally actionable, well-sequenced review workflow: exact commands, templates, validation checkpoints, and fail-open handling of every degenerate input case. Its weaknesses are token efficiency (repeated rationale passages that could be stated once) and the total absence of progressive disclosure — everything, including large output-format templates, is inlined in a single very long file.
Suggestions
Move the stable output templates (RTM file format, reflexion-log entry, session-state block, traceability-index format) into a references/ file and link to them from the phases, cutting the main body substantially.
State the "skipped check is indistinguishable from a passed check" rule once and reference it from Phases 5-7 instead of re-explaining it three times; the same applies to the extended blockquote justifications in Phase 3 and the Error Recovery preamble.
Trim the meta-explanations of why rules exist (e.g. why 🟡 caps the verdict, why parentheticals break status matching) to a single sentence each — the rule itself is actionable without the narrative.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense with project-specific, non-obvious semantics (denominator interpretation, fail-open rules, verdict gating), so it is not noticeably padded overall — but there is real trim material: the "a skipped check is indistinguishable from a check that passed" rationale is repeated in Phase 5, Phase 6, and Phase 7, and several long blockquotes (e.g. Phase 3's 🟡 justification, the Error Recovery preamble) explain reasoning the executing model could take on trust. Not 4 because the repetition and meta-justification are unnecessary tokens; not 2 because most content is load-bearing project knowledge rather than concepts Claude already knows. | 3 / 5 |
Actionability | Guidance is copy-paste ready throughout: exact Grep patterns with glob and -A counts, exact Bash invocations (review-receipts.sh check/hash, adr-dep-graph.sh), complete output templates for the matrix, RTM file, conflict entries, and session-state block, plus scripted AskUserQuestion option lists. Specific examples cover the common cases of every mode. | 5 / 5 |
Workflow Clarity | Nine clearly sequenced phases with explicit validation checkpoints everywhere: a freshness/receipt check before any scan, denominator counts with fail-open interpretation tables (including the 0-match malformed-ADR case), per-ADR escalation rules, artifact-on-disk verification before treating a phase as done, and an error-recovery protocol that requires a partial report. Feedback loops (re-scan, re-ask, re-validate) are explicit at every write. | 5 / 5 |
Progressive Disclosure | No bundle files exist — the entire ~860-line skill lives in one inline SKILL.md. Internal structure is strong (phased headers, per-mode branching, tables), so it is not a 2, but content that clearly belongs in separate reference files (the RTM output format, reflexion-log entry format, session-state block, collaborative protocol) is inlined, and the only external pointers are to project docs rather than skill-managed references — so it does not reach 4. | 3 / 5 |
Total | 16 / 20 Passed |