Content
82%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A highly actionable, well-structured skill body: executable gh and proof-runner commands, explicit output templates, validation checkpoints, and a batch-close guardrail. Remaining weaknesses are mild redundancy across the evidence sections and an implicit (rather than explicit) end-to-end workflow ordering, plus proof-mode mechanics that could be split into a reference file.
Suggestions
Consolidate the overlapping evidence requirements in "Review Evidence Bar", "Enforce Bug-Fix Evidence", and the best-fix loop into one section referenced from the others.
State the end-to-end review pipeline (live state -> evidence -> best-fix analysis -> proof -> final comment) explicitly at the top so section order does not have to be inferred.
Move the detailed proof:ui runtime mechanics (flags, baseline/candidate URLs, artifact layout) into a reference file, keeping SKILL.md as the overview.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | Dense imperative bullets with no explanation of concepts Claude already knows, but with some redundancy: evidence requirements appear in three sections ("Review Evidence Bar", "Enforce Bug-Fix Evidence", "Best-Fix Review Loop") and the origin/main comparison appears in both "Read Beyond The Diff" and step 6 of the loop. Not a 5 because clear trims are available; not a 3 because it is mostly efficient. | 4 / 5 |
Actionability | Fully executable, copy-paste-ready commands cover the common cases: `gh pr view <number> --repo openclaw/clawhub --json title,body,author,...`, `gh api users/<login> --jq ...`, `bun run proof:ui -- --runner local --mode feature --scenario ...`, and `gh pr comment ... --attach ...`, plus concrete output templates like `LOC: +x/-y (N files)` and `Best-fix verdict:`. | 5 / 5 |
Workflow Clarity | A numbered 7-step best-fix loop, explicit validation checkpoints (verify live state before commenting, inspect every final image/video, verify the final comment), a feedback loop for uninspected paths, and a batch-operation confirmation guardrail for closing more than five issues/PRs. Not a 5 because the end-to-end sequence (live state -> review -> proof -> comment) is implied by section order rather than explicitly framed as a pipeline. | 4 / 5 |
Progressive Disclosure | No bundle files exist, and the single external reference ([proof-video](../proof-video/SKILL.md)) is clearly signaled and one level deep. Sections are well-organized, but the ~200-line body inlines UI-proof runtime mechanics that could live in a separate reference file, keeping it below a 5. | 4 / 5 |
Total | 17 / 20 Passed |