Content
88%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body delivers high-value, opinionated design guidance with excellent sequencing, validation checkpoints, and a concrete worked example — everything in it is actionable and little of it is filler. Its weaknesses are minor: some rhetorical flourish that could be trimmed for token efficiency, and a monolithic single-file layout where a references/ bundle (e.g., the worked example or the failure-mode detail) would lighten the main file.
Suggestions
Trim rhetorical flourishes (e.g., "the system tends toward entropy, so maintain it", the aphoristic one-line close) to reclaim tokens without losing guidance.
Move the worked example and/or the failure-mode detail into a references/ file (e.g., references/worked-example.md) linked from the checklist section, keeping SKILL.md as a lean overview.
Tighten Step 4 (damping) into a concrete decision table — which damping to apply for servo vs. regulator, and default retry-cap values — instead of a single prose paragraph.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense with non-obvious judgment content (the 4-condition gate, "Build may not edit the acceptance conditions to pass", the failure-mode table) rather than things Claude already knows, matching 'efficient'. It falls short of the lean anchor-5 because of rhetorical padding such as "the system tends toward entropy, so maintain it" and the aphoristic one-line close, which could be trimmed without losing guidance. | 4 / 5 |
Actionability | For an instruction-only skill the guidance is fully actionable: a concrete 4-condition gate ("① the task repeats weekly or more ② verification can be automated…"), decision tables for loop type and skeleton, explicit rules ("Three failed retries → escalate to a human"), a five-row checklist with literal review questions, and a worked example showing the naive vs. fixed goal. This covers common cases concretely; score 4 would require missing key details, and none are missing. | 5 / 5 |
Workflow Clarity | Action 1 is a clearly sequenced 5-step process with explicit checkpoints (the 4-condition veto gate, the Step-1 self-check "can they run one command and tell whether it's done?", the staged landing plan), and Action 2 is a per-row review checklist with explicit failure conditions and antibodies. Validation and feedback loops (retry caps, escalation, three-stage rollout) are present, matching the anchor-5 pattern of sequence + explicit validation + checklists. | 5 / 5 |
Progressive Disclosure | Structure is good: clear section headers, a two-level 'use / don't use' scope statement, and well-signaled one-level pointers to the mechanism layer ("see `autonomous-loops` / `continuous-agent-loop`"). It sits below anchor 5 because the single ~137-line file inlines content that could be split into references (the worked example, the failure-mode/lineage notes), and there are no reference files at all to spread detail into. | 4 / 5 |
Total | 18 / 20 Passed |