Content
73%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is a strong, disciplined behavioral gate: a crisp step sequence with hard validation checkpoints, concrete evidence requirements per claim type, and a correct/incorrect example pair. Its weaknesses are redundancy (four sections restate the same rule) and a dangling cross-reference section pointing at files outside the bundle.
Suggestions
Merge the Rationalization Table and the Red Flags table — both map the same excuses/thoughts to "run fresh verification" — and fold "When to Apply" into the Gate section to cut roughly a third of the body.
Drop or prune the Multi-Provider Context section (or move its bash checks into the Evidence table); it is niche padding for most invocations.
Remove or verify the 'Integration with Other Skills' cross-references — none of the six named files exist in this bundle, so the section is unresolvable for the model.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is terse and never explains concepts Claude already knows, but the message "run fresh verification before claiming success" is restated across four overlapping sections — the Rationalization Table ("I ran the tests earlier this session… Earlier is not fresh"), the Red Flags table ("Should work now → Run the verification"), "When to Apply", and The Gate — plus a niche Multi-Provider section that could be tightened. This is "mostly efficient but could be tightened" rather than minor trimming. | 3 / 5 |
Actionability | Concrete, executable guidance dominates: the five-step gate (IDENTIFY/RUN/READ/VERIFY/ONLY THEN), copy-paste bash checks ("ls -la ~/.claude-octopus/results/*-synthesis-*.md | tail -1", "wc -l …"), an evidence table mapping each claim to its required output, and a worked red-green sequence. It misses anchor 5 only because the core directive "execute the full command" stays generic (no per-context command guidance beyond the multi-provider case). | 4 / 5 |
Workflow Clarity | The Gate gives a clear, numbered five-step sequence with an explicit validation checkpoint ("VERIFY — Does output actually confirm the claim?"), a hard skip rule ("Skip any step = the claim is unverified"), a failure-path example (red → green → revert-to-red → green proves the test isn't a false positive), and per-phase checkpoints for the orchestrate.sh workflow. This matches the anchor for clear sequence with explicit validation and feedback loops. | 5 / 5 |
Progressive Disclosure | The body is well-organized with clear section headers, appropriately sized for a single-purpose gate skill, and no bundle files exist to reference — the whole content reasonably lives inline. It misses anchor 5 because the "Integration with Other Skills" section lists six sibling files (flow-develop.md, skill-tdd.md, etc.) that are not part of this bundle, leaving dangling references the model cannot navigate. | 4 / 5 |
Total | 16 / 20 Passed |