Content
92%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A dense, executable procedure: every step carries copy-paste commands, explicit validation checkpoints, and fallbacks for each failure mode, with field-level detail properly delegated to a real, well-signaled reference file. The only criticism is stylized prose and some repetition that could be tightened without losing meaning.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body teaches nothing Claude already knows — every sentence is a repo-specific rule or trap (e.g. ".closingIssuesReferences[0].number drops the repository", "`null` means unobserved, not 'all of it'"). However, stylized prose ("It is gitignored scratch, like every other skill's report; the card is the half that the repository keeps") and some repetition (the no-transcripts case is argued in full in "Before Step 1" and again in Step 1) could be trimmed. Not 5: minor over-elaboration remains; not 3: there is no padding or basic-concept explanation. | 4 / 5 |
Actionability | Copy-paste-ready commands throughout: the three-line environment probe, `bun run --cwd tools/pr-metrics card -- --resolve-issue` (and the `--format markdown` variant), the slug-transformed `git add ".claude/metrics/$(echo "$BRANCH" | tr -c 'a-zA-Z0-9._-' '-').json"`, the append-not-rebuild body edit, the `grep -c '^<details><summary>Session metrics'` guard, and a full RETRO.md template. The judgment steps are backed by concrete signal-to-fix mapping tables. Not 4: the common cases are covered by executable commands with specific flags, including the fallbacks. | 5 / 5 |
Workflow Clarity | Steps 1–7 are clearly sequenced with explicit checkpoints: the pre-run probe ("Three answers, once"), the check-before-append guard with its own failure note ("match the summary line, not the tag: a bare `grep -c '<details>'` then reports a block that is not there"), and per-failure handling ("When a step here cannot run, record which and carry on"). The card is deliberately written first so an early-ending run still leaves the record — an explicit feedback/robustness design. Not 4: validation checkpoints and error-recovery loops are explicit, not implicit. | 5 / 5 |
Progressive Disclosure | Field-level detail is correctly offloaded to `references/card-signals.md` (a real bundle file), clearly signaled in Step 2 ("`references/card-signals.md` maps each of those to the question it answers… Read it before Step 3") — one level deep, no nesting. Repo-level references (`wiki/conventions/session-metrics.md`, `tools/pr-metrics/README.md`) are kept cleanly separate from the run procedure, and the skill names its boundary with `analyse-sessions`. Not 4: the split is appropriate and navigation is easy throughout. | 5 / 5 |
Total | 19 / 20 Passed |