Content
88%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-built procedural skill body: executable, verified code examples; a fully specified PASS/REVIEW/FAIL decision flow with an explicitly bounded repair loop; and no filler content. The only improvements are consolidating the thrice-repeated model-only honesty rule and mentioning the bundled patterns.js dependency that validate.js silently loads.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense operational prose with zero padding about concepts Claude already knows — every section states contract rules, statuses, or commands. However, the honesty rule about not claiming execution ('Never claim it ran unless the current host executed it', 'Do not claim the deterministic validator checked semantic representation details that it does not implement', 'label the result as model_only') is repeated across three sections and could be consolidated. This matches anchor 4 — efficient with minor trimming possible — not the every-token-earns-its-place anchor 5. | 4 / 5 |
Actionability | Both code paths are copy-paste ready and verified against the actual bundle: the CLI invocation 'node scripts/validate.js --residual-policy warn before.md after.md' matches the script's real argument parser, and 'const { validate } = require("./scripts/validate.js")' with 'validate(original, rewritten, { residualPolicy: "warn" })' matches the real export signature. Concrete statuses (PASS/REVIEW/FAIL), envelope fields (execution_evidence.verifier, verification_summary.status, pass.index/pass.max), and named routing targets cover the common cases, meeting the score-5 anchor. | 5 / 5 |
Workflow Clarity | The sequence is clear — accept inputs (require both versions, else return to router), run the deterministic validator or fall back to model-only, apply the semantic and representation lenses, then route PASS/REVIEW/FAIL outcomes — with an explicit bounded feedback loop: 'At most one repair re-entry is allowed... After repair, verify once more. If that check still fails, stop and report.' Validation is the core of the skill and every branch has explicit checkpoints, matching the score-5 anchor with feedback loops for error recovery. | 5 / 5 |
Progressive Disclosure | Structure is good: the body stays a concise overview, the referenced scripts/validate.js (534 lines) and the cross-skill router references are one level deep and clearly signaled with concrete commands. The gap is that the bundled scripts/patterns.js (3443 lines, a runtime dependency auto-required by validate.js for residual scoring) is never mentioned anywhere in SKILL.md, leaving a large bundle file without navigation. That minor organization gap fits anchor 4 rather than the fully well-signaled anchor 5. | 4 / 5 |
Total | 18 / 20 Passed |