Content
77%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
An unusually rigorous, executable workflow with excellent sequencing, validation gates, and safety controls — the operational core is top-tier. Weaknesses are length/redundancy (several concepts defined twice) and progressive disclosure: ~370 lines inlined where templates and schemas belong in bundle files, plus citations to shared-references files that are absent from the bundle.
Suggestions
Ship the referenced bundle files or inline their content: 'shared-references/reviewer-routing.md', 'shared-references/review-tracing.md', 'shared-references/integration-contract.md', and 'save_trace.sh' are cited but absent from the bundle, so review tracing (Policy C) is currently unexecutable.
Move stable schemas and templates out of SKILL.md into references/ (issue-card field definitions, the REVISION_PLAN checklist template, and the reviewer calling convention configs) to cut the body from ~370 lines to an overview.
Deduplicate repeated definitions: 'structural_distinction' (Phase 2 vs. Reviewer-defensive moves), pivotal reviewer allocation (Constants vs. Phase 3), and the VENUE_MODE output structure (Phase 4 vs. Phase 7) are each explained twice — keep one canonical definition.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | Mostly efficient operational detail ("Sentence 1: direct answer / Sentence 2-4: grounded evidence"), but with noticeable duplication: pivotal reviewers are defined in both the Constants ("STRESS_TEST_ROUNDS_BASE") and Phase 3 step 4; 'structural_distinction' is explained in full twice (Phase 2 response_mode list and 'Reviewer-defensive moves'); the VENUE_MODE draft structure is stated in Phase 4 and restated in Phase 7; REVISION_PLAN rules appear in both Phase 4 and the checklist rationale. This fits 'mostly efficient but includes some unnecessary explanation or could be tightened' rather than the minor-trim profile of a 4. | 3 / 5 |
Actionability | Fully executable guidance: exact MCP call templates with configs ("config: {\"model_reasoning_effort\": \"xhigh\"}"), a copy-paste stress-test prompt, exact artifact paths (`rebuttal/ISSUE_BOARD.md`, `rebuttal/PASTE_READY.txt`), a concrete markdown checklist example with issue_id/commitment/status fields, and explicit slash-command invocations ("/experiment-bridge \"rebuttal/REBUTTAL_EXPERIMENT_PLAN.md\""). Matches 'copy-paste ready; specific examples cover the common cases'. Not a 4 because the few abstract steps ('pause and ask') are intentional control flow, not missing detail. | 5 / 5 |
Workflow Clarity | Phases 0-9 are clearly sequenced with resume handling ("If rebuttal/REBUTTAL_STATE.md exists → resume from recorded phase"), an explicit validation phase (Phase 5's eight lints: coverage, provenance, commitment, tone, consistency, limit, thread-local context, adversarial scan), and real feedback loops ("If any hard safety blocker remains → revise before finalizing", re-run lints in follow-up rounds). This is the anchor-5 profile of sequence + explicit validation + error-recovery loops. | 5 / 5 |
Progressive Disclosure | The body has good section structure, but the bundle contains no references/, scripts/, or assets/ directories while the text cites "shared-references/reviewer-routing.md", "shared-references/review-tracing.md", "shared-references/integration-contract.md §2", and "save_trace.sh" — dangling references to files not shipped. Additionally, substantial content that belongs in reference files is inlined (issue-card field schema, REVISION_PLAN template, reviewer calling conventions). Fits 'references present but not clearly signaled; content that should be separate is inline'. Not a 2 because the body itself is well-organized with navigable headers rather than a reference-buried or structureless wall. | 3 / 5 |
Total | 16 / 20 Passed |