Content
46%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is highly actionable with a clear phase structure, but it is bloated by quadruple-stated safety/evidence rules and ~250 lines of inlined prompt templates, and it ignores its own scripts/gh_fetch.py while re-implementing its pagination inline. The batch workflow also lacks any report-validation or failure-recovery checkpoints, capping workflow clarity.
Suggestions
Replace the ~50 lines of inline pagination bash in Phase 1 with a single call to the already-bundled scripts/gh_fetch.py (which implements the same exhaustive pagination) and reference it explicitly.
Move the six subagent prompt templates into a single references/subagent-prompts.md file (or collapse them to a shared template plus short per-type deltas) and reference it once, eliminating the repeated zero-action and permalink boilerplate.
Add validation to Phase 4: verify each expected {REPORT_DIR}/{issue|pr}-{number}.md exists and contains its required sections before task_update, and define a retry path for failed subagents.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The 587-line body repeats the same rules many times: the zero-action policy appears in the <role> block, the 'Zero-Action Policy' section, the 'ABSOLUTE RULES for Subagents', the 'Common Preamble' ('NEVER run gh issue comment, gh issue merge...'), per-prompt footers ('NEVER merge. NEVER comment.'), and again in the Anti-Patterns table; the evidence/permalink rule is likewise restated four times. This is noticeably verbose with substantial duplicated padding, though it never degrades into explaining concepts Claude already knows. Not 1 because the material is operational rather than tutorial filler; not 3 because the repetition is pervasive, not occasional. | 2 / 5 |
Actionability | Guidance is largely concrete and executable: copy-paste bash for setup and paginated fetching, exact task_create/task call signatures, and fully specified per-type report templates ('Verdict: [CONFIRMED_BUG | NOT_A_BUG | ALREADY_FIXED | UNCLEAR]'). Not 5 because the 'task(category="quick", run_in_background=true, ...)' calls are harness-specific pseudocode rather than a real tool interface, and the PR prompts reference fields (mergeable, reviewDecision, statusCheckRollup) never fetched in Phase 1. Not 3 because nearly everything else is copy-paste ready. | 4 / 5 |
Workflow Clarity | Phases 0-5 give a clear sequence, but this batch operation (spawning up to 1000 background subagents) has no validation or error-recovery checkpoints: Phase 4 says only 'Poll background_output() per task... Parse report' with no step to verify each report file exists and is well-formed, no handling of failed or timed-out subagents, and no feedback loop. Per the rubric cap, a batch workflow without validation cannot score above 3. Not 2 because the sequence itself is coherent and well-ordered with explicit per-phase outputs. | 3 / 5 |
Progressive Disclosure | The bundle ships scripts/gh_fetch.py (398 lines, 'Fetches ALL issues and/or PRs... Implements proper pagination') yet it is never referenced anywhere in SKILL.md — instead ~50 lines of inline pagination bash duplicate its function in Phase 1. Additionally, roughly 250 lines of near-identical subagent prompt templates are inlined in the body rather than split into a references/ file. Content that clearly belongs in separate files is inlined, matching anchor 2. Not 3 because this is not a marginal organization issue: an existing bundle file is orphaned while its logic is duplicated inline. | 2 / 5 |
Total | 11 / 20 Passed |