Content
77%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
This is a highly actionable, well-sequenced workflow skill with strong validation gates and copy-paste-ready commands throughout, though the ~500-line body is monolithic and repeats several core rules across up to five sections. Splitting domain-specific procedures into reference files and deduplicating the PR-plan rules would fix its main weaknesses.
Suggestions
Move the Security Advisory Hotfixes procedure and the Task-Style PR Body template into separate reference files (e.g. references/security-advisory.md, references/pr-body.md) linked one level deep, and load them only when those sources apply — this would cut the main body substantially.
Deduplicate the PR-plan rule (state it once in Verification or the PR Body section instead of restating it in Core Rules, Intake, Verification, PR Body, and Success Criteria) and merge the `--with <pack>` catalog that appears in both Intake step 9 and the Skill Diet `autogoal` bullet.
Trim the Success Criteria section, which largely restates gates already defined in Intake, Review, and Verification, down to the few outcomes not already checkable elsewhere.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense and repo-specific with essentially no explanation of concepts Claude already knows, but the same rules are restated repeatedly — the PR-plan/`🧭 Task plan` requirement appears in Core Rules, Intake, Verification, Task-Style PR Body, and Success Criteria, and the `--with <pack>` catalog appears in both Intake step 9 and the Skill Diet `autogoal` bullet. It is mostly efficient but could be meaningfully tightened, fitting the anchor exactly; it is not 2 because there is no concept-explaining filler, and not 4 because the duplication across sections is more than minor. | 3 / 5 |
Actionability | The guidance is fully executable with copy-paste-ready commands and exact flags: `gh issue view`, `gh api repos/<owner>/<repo>/security-advisories/<GHSA_ID>`, `node .agents/skills/autogoal/scripts/create-goal-scratchpad.mjs --template <task|docs> --with <pack> --title "..."`, `npm view <package>@<version>`, `pnpm run reinstall`, and `gh pr view --json body`. Concrete examples (e.g. `🐛 Fixes #123`, `🟢 95-100% confidence`, the exact table header) cover the common cases; as an instruction-only skill this fully satisfies the rubric's code-vs-instruction note. | 5 / 5 |
Workflow Clarity | The multi-step process is explicitly sequenced (14-step numbered Intake, per-shape Execution Paths, Verification, Final Handoff) with explicit validation checkpoints and feedback loops: reproduce-before-fix with a four-level escalating repro ladder, 'run `pnpm run reinstall` once and rerun the exact failing command', autoreview loop 'keep going until there are no accepted/actionable findings', and hard-stop verdicts (`not reproduced`, `invalid`) before code. Batch work is covered with per-PR validation rather than aggregate evidence, so the batch cap does not apply. | 5 / 5 |
Progressive Disclosure | Section headers are clear and well-ordered, but the skill is a ~500-line monolith with no bundle files at all: content that clearly belongs in separate one-level-deep references — the Security Advisory Hotfixes procedure, the Task-Style PR Body template, the Skill Diet catalog — is all inlined. This fits 'some structure but content that should be separate is inline'; it is above 2 because navigation via headers is genuinely easy, and below 4 because no reference files exist to split the bulk into. | 3 / 5 |
Total | 16 / 20 Passed |