Content
80%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is a lean, well-structured router that practices progressive disclosure exactly as intended: an overview plus trigger-conditioned, one-level-deep references with no token waste. Its weaknesses are inherited from its minimalism — no inline quick-start command and no visible validation checkpoint for the destructive merge/rebase/push workflows it routes to, which caps workflow clarity.
Suggestions
Add one validation checkpoint to the body for destructive workflows, e.g. "Before merge/rebase, confirm the stack is in sync (see commands.md preconditions) and re-verify after push" — this lifts workflow clarity past the rubric's batch-operation cap.
Inline a single quick-start command (e.g. the `gh stack create` or `gh stack submit` invocation) so the most common case is executable without opening a reference.
State what success looks like after a push or submit (e.g. expected PR links or updated stack status) so failures are detectable at the body level.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The ~15-line body contains only routing: the single context sentence ("`gh stack` is a GitHub CLI extension for stacked branches and pull requests") establishes non-obvious tool identity, and every other line directs a file read. No explanation of concepts Claude already knows, no padding. | 5 / 5 |
Actionability | Concrete, conditional routing ("Read [references/usage.md]... for setup, the non-interactive command table, core workflow loops, and exit codes"; "Open the reference whose trigger matches the task") tells Claude exactly which file to open when, but no command or example appears inline — the executable content is entirely delegated to the references, leaving it just short of the 5 anchor's "specific examples cover the common cases". | 4 / 5 |
Workflow Clarity | The body sequences reads clearly (usage.md first, then trigger-matched reference), but the skill manages push, rebase, and merge across stacked PRs — batch/destructive operations — and the body itself shows no validation or verification checkpoint. Per the rubric cap, a destructive/batch skill without validation cannot score above 3. | 3 / 5 |
Progressive Disclosure | A clear overview with four well-signaled, one-level-deep references (all verified to exist: usage.md, stack-design.md, commands.md, troubleshooting.md), each paired with an explicit trigger condition ("before creating a stack", "on unexpected failures", "on conflicts, divergence, restructuring"). Easy navigation, no nesting, nothing inlined that belongs in a separate file. | 5 / 5 |
Total | 17 / 20 Passed |