Content
96%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
An exemplary operational skill body: fully executable commands, a clear run-to-verify sequence with explicit validation and error-recovery guidance, and near-zero token waste. The only structural improvement is offloading the query-format reference and troubleshooting rows to one-level-deep reference files now that the body has grown past overview length.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense and assumption-respecting throughout — a prerequisites table, exact install/download commands, a query-JSON spec with a validation note ('extra: forbid — unknown keys reject'), a flag table, an output tree, and a compact troubleshooting table. No section explains concepts Claude already knows (no 'what is a protein' or library tutorials); every line carries non-obvious operational detail such as the DeepSpeed-vs-cuEquivariance kernel swap and the eager-imported boto3 gotcha. | 5 / 5 |
Actionability | Everything is copy-paste executable: 'pip install \'openfold3[cuequivariance]==0.4.1\'', the 'huggingface-cli download' command with a pinned checkpoint path, a complete 'run_openfold predict' invocation, a valid query JSON example, and verification commands ('grep -E \'Successful|Failed\' out/summary.txt', 'find out -name '*_model.cif' | wc -l') with the expected-count formula. The troubleshooting table maps exact error strings to exact fixes ('pip install nvidia-cutlass', 'apt-get install libxrender1 libxext6 libsm6'). | 5 / 5 |
Workflow Clarity | The flow is clearly sequenced — prerequisites, install, weights, run, query format, output interpretation — and closes with an explicit 'Verify' section containing feedback criteria ('Successful Queries: N matching your input count', the queries × seeds × samples count check) plus quality thresholds (avg_plddt, ptm/iptm, has_clash) and a diagnosis/fix table for failure recovery. Not anchor 4 because validation checkpoints are explicit, not merely implied. | 5 / 5 |
Progressive Disclosure | Sections are well-organized and the SKILL.md works standalone (no references/ scripts/ assets/ bundle exists, so nothing is buried or nested). It stops short of anchor 5 because at ~150 lines the query-format reference table and the 8-row troubleshooting table are prime candidates for one-level-deep reference files (e.g. TROUBLESHOOTING.md) that would keep the top-level file a lean overview; the guideline's 5-with-sections-alone exception applies only to skills under 50 lines. | 4 / 5 |
Total | 19 / 20 Passed |