Content
100%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is exemplary: executable code, exact input/output formats, a hard operational gotcha (CHAI_DOWNLOADS_DIR), a VRAM misconception correction, and a diagnostic error table — all lean and free of padding. Nothing needs restructuring or splitting into bundle files.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense with non-obvious information ('a multi-entity FASTA in, mmCIF plus pTM/ipTM/pLDDT out', the ~5 GB mid-run download, the 3B-parameter ESM VRAM gotcha) and never explains concepts Claude already knows. The intro's positioning against Boltz-2/AlphaFold3 is decision-relevant, not padding, so it matches the score-5 anchor (every token earns its place) rather than 4's 'minor instances of over-explanation'. | 5 / 5 |
Actionability | The guidance is fully executable: a complete run_inference call with imports and parameters, the exact FASTA header grammar ('>{entity_type}|name={id}'), the shell equivalent ('chai-lab fold complex.fasta out/ --use-msa-server'), concrete output files, and a ranking threshold ('treat iptm > 0.5 as a soft pass'). This matches the score-5 anchor (copy-paste ready, covers common cases); the truncated sequences with '...' are inherently user-specific inputs, not pseudocode. | 5 / 5 |
Workflow Clarity | For a single-task skill the action is unambiguous: write the multi-entity FASTA, call run_inference, rank by aggregate_score, threshold on iptm, and clear output_dir between calls — with post-hoc confidence filtering as the validation checkpoint. The simple-skill exception applies, so it matches the score-5 anchor rather than 4; the score-3 cap for destructive/batch operations does not apply since inference is neither destructive nor an unvalidated batch loop. | 5 / 5 |
Progressive Disclosure | There are no bundle files (references/, scripts/, assets/ do not exist) and none are needed: the body is a compact, well-sectioned single page ('Running it', two gotcha sections, 'Errors worth recognizing') with no content that belongs in a separate file and no nested references. This matches the score-5 anchor for well-organized single-page skills rather than 4, which requires organization gaps. | 5 / 5 |
Total | 20 / 20 Passed |