Content
77%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is a well-sequenced, highly actionable two-stage pipeline with explicit gates, status labels, decision tables, and executable bash throughout — its workflow clarity is exemplary. Its weaknesses are structural: everything lives in one long file with no reference/script split, and a couple of sections (Bottom Line, PR posting) add tokens or depend on undefined variables like ${COMBINED_REPORT}.
Suggestions
Move the self-contained stub-detection and PR-posting bash blocks into scripts/ files (e.g. scripts/stub-detect.sh, scripts/post-review.sh) and reference them from SKILL.md, reducing the body to an overview plus orchestration guidance.
Define or populate ${COMBINED_REPORT} in the PR-posting snippet, and replace 'Wait for external reviews to complete' with a concrete wait/poll command so the guidance is fully copy-paste executable.
Cut 'The Bottom Line' and fold any non-redundant content into the intro to trim tokens without losing information.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense with executable bash and tables and does not explain concepts Claude already knows, but has minor padding: 'The Bottom Line' restates the pipeline in three redundant lines, and the multi-LLM section includes justification prose ('A Claude-only review pipeline misses what external models catch — Codex excels at...'). This fits 'Efficient; minor instances of over-explanation that could be trimmed' rather than 5, where every token earns its place. | 4 / 5 |
Actionability | Concrete, mostly copy-paste-ready bash is given for loading the intent contract, stub detection, dispatching external providers, and posting to a PR, plus explicit status labels and report templates. Minor gaps keep it at 4 rather than 5: ${COMBINED_REPORT} is referenced but never defined, and the external-provider step says 'Wait for external reviews to complete' without a wait/poll command. | 4 / 5 |
Workflow Clarity | The two-stage sequence is explicit and gated: Stage 1 validates success criteria and boundaries with PASS/FAIL/PARTIAL and RESPECTED/VIOLATED statuses, a decision table routes outcomes (proceed / ask user fix-or-override), and the Error Handling table plus validation-gate frontmatter close the loop. This matches 'Clear sequence with explicit validation steps; feedback loops for error recovery' — the gate is an explicit checkpoint and failure paths route to fix-or-override. | 5 / 5 |
Progressive Disclosure | The single SKILL.md is ~320 lines with no bundle files (references/, scripts/, assets/ are absent) and no pointers to separate material; sizable self-contained scripts (the ~55-line stub-detection block, the PR-posting block) are inlined where scripts/ files would fit. This matches 'Some structure but could be better organized; content that should be separate is inline' — the section structure itself is good, but nothing is split out, keeping it below anchor 4. | 3 / 5 |
Total | 16 / 20 Passed |