Content
86%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A tightly engineered skill body: an input gate, a routed five-step workflow, fully executable validator commands, real one-level-deep references, and a required handoff checklist. Its only real gaps are a missing failure-handling loop for validator errors and some duplicated boundary language and inline dated version facts that cost tokens.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body uses routing tables, tight bullet lists, and copy-ready commands with no padding or explanation of concepts Claude already knows, so it sits at anchor 4 ('efficient; minor instances that could be trimmed'). Not 5 because boundary language repeats across sections (e.g. 'Never generate findings, impressions, diagnoses...' in Diagnostic scaffolds restates the Non-Negotiable Boundary list) and dated version facts (CONSORT 2025, E2D(R1) adopted 15 September 2025, E6(R3) adopted 16 June 2026) appear inline rather than consolidated in a versions/deprecated section; not 3 because the repetition is limited and every section carries non-obvious constraints. | 4 / 5 |
Actionability | Every script invocation is fully executable as written, e.g. 'PYTHONDONTWRITEBYTECODE=1 python3 scripts/generate_report_template.py --type case-report --output ./case-report-draft.json' and 'PYTHONDONTWRITEBYTECODE=1 python3 scripts/format_adverse_events.py ./aggregate-ae.csv --metadata ./safety-aggregate.json --output ./aggregate-ae-table.md', with concrete template paths (assets/provenance_manifest_template.json) and explicit per-field population rules ('Replace null only when a verified fact ID supports the field'). This matches anchor 5 — copy-paste ready commands covering the common cases; all 15 assets, 11 references, and 8 scripts referenced in the body exist on disk. | 5 / 5 |
Workflow Clarity | A clear 5-step sequence (source-fact manifest → generate template → populate verified fields only → run deterministic checks → apply qualified review) sits behind a 7-condition Input Gate, with an explicit validation stage (step 4's six validators) and a required Final Handoff checklist — matching anchor 4 ('clear sequence with most checkpoints present'). Not 5: there is no feedback loop telling Claude what to do when a validator fails (fix which field, re-run which check), only the boundary-crossing case ('stop the unsafe portion, offer a blank template') is spelled out; not 3 because checkpoints are explicit and gating ('Proceed only when all conditions are true'), not implicit. | 4 / 5 |
Progressive Disclosure | The body is a genuine overview: a routing table that maps each artifact to its guidance file, short sections that each end by pointing to exactly one clearly-named reference ('Read references/privacy_and_deidentification.md', 'Use assets/case_report_template.json and references/case_report_guidelines.md'), plus a consolidated Assets and References listing where every named file exists in the bundle. References are one level deep (references/README.md is itself a file map, not a chain), matching anchor 5 ('clear overview with well-signaled one-level-deep references; easy navigation'). | 5 / 5 |
Total | 18 / 20 Passed |