Content
48%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body contains genuinely useful, concrete material (usage commands, parameter table, error-handling fallback), but it is buried in large amounts of generic template boilerplate and internal-repo artifacts. Referenced bundle files (scripts/main.py, references/audit-reference.md, requirements.txt) are missing from the bundle, and duplicated command sections plus a non-functional cd path undermine executability.
Suggestions
Cut the generic boilerplate sections (Security Checklist, Risk Assessment, Evaluation Criteria, Lifecycle Status, Output Requirements) or move them to a reference file; they add tokens without skill-specific value.
Consolidate the duplicated command sections (Quick Check, Audit-Ready Commands, Example Usage) into one usage section, and remove the non-functional 'cd "20260318/scientific-skills/..."' internal path.
Ship the referenced bundle files (scripts/main.py, references/audit-reference.md, requirements.txt) or remove the references, and document '--hazard-ratio' and other test-specific flags in the Parameters table.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is noticeably verbose with several padded, generic sections: a Security Checklist with items irrelevant to a local stats script ("API requests use HTTPS only", "API timeout and retry mechanisms"), a Risk Assessment table, Lifecycle Status with time-sensitive dates ("Next Review Date: 2026-03-06"), and Evaluation Criteria boilerplate. The same two commands are also repeated across Quick Check, Audit-Ready Commands, and Example Usage. | 2 / 5 |
Actionability | There are concrete, specific commands ("python scripts/main.py --test ttest --effect 0.5 --alpha 0.05 --power 0.8") and a parameter table, but guidance is incomplete: the Example Usage 'cd "20260318/scientific-skills/..."' path is non-functional, '--hazard-ratio' appears in Usage examples but not in the Parameters table, and the referenced scripts/main.py and requirements.txt are not present in the bundle. | 3 / 5 |
Workflow Clarity | A clear sequence exists with most checkpoints present: a pre-execution parse check ("python -m py_compile scripts/main.py"), stop-early scope validation, and an explicit fallback path in Error Handling ("If scripts/main.py fails, report the failure point..."). It falls short of 5 because the workflow is scattered and duplicated across overlapping sections (Workflow, Example Usage run plan, Implementation Details) with circular 'See ## Usage above' cross-references. | 4 / 5 |
Progressive Disclosure | There is real section structure and a clearly signaled References section linking references/audit-reference.md, but that file does not exist in the bundle, and substantial generic boilerplate (Security Checklist, Risk Assessment, Evaluation Criteria, Lifecycle Status) that belongs in a separate reference is inlined in SKILL.md. | 3 / 5 |
Total | 12 / 20 Passed |