Content
71%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A tight, well-structured instruction-only skill whose strengths are token efficiency, clear sequencing, and genuinely non-obvious method guidance (the dispersion-consistency defaults and the use-evidence-not-launch-success checkpoint). Its main gap is actionability: it names tools and parameters but never shows an example invocation or a concrete keep/drop rule for the VASP handoff it requires as output.
Suggestions
Add one example invocation for each tool (argument names and typical values for model, dispersion, relax_lattice, input_dir, output_root) so the guidance is executable rather than only descriptive.
Define a concrete keep/drop rule or shortlist criterion (e.g. an energy-ranking threshold from batch_summary_rel, or 'relax top-N by mace_sp_batch energy') instead of leaving the rule entirely to the reader.
Specify what to check in batch_state_rel/status files to decide rerun vs. accept partial outputs (e.g. which status values mean dispatch failure vs. per-structure failure), closing the workflow's validation loop.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The 47-line body is lean, assumes domain competence, and wastes no tokens explaining known concepts — "Do not compare relax and SP outputs as if they were the same screening stage" is exactly the kind of non-obvious guidance the rubric rewards. It falls short of 5 due to minor redundancy: Quick Start steps 2-3 restate Workflow sections 1-2, and the tool list appears in both frontmatter metadata and a separate 'Suggested tools' section. | 4 / 5 |
Actionability | Concrete guidance exists — named tools ("mace_relax_batch", "mace_sp_batch"), named parameters ("model", "head", "dispersion", "relax_lattice"), explicit constraints ("Keep output_root outside input_dir"), and named return fields ("batch_state_rel", "batch_summary_rel") — but there are no executable invocations: no example tool call, no argument shapes, and no concrete keep/drop rule (the Output Contract only requires returning one without defining what a good rule is). This matches anchor 3 ('some concrete guidance but incomplete; missing key details') better than anchor 4, which expects concrete code or commands with only minor gaps. | 3 / 5 |
Workflow Clarity | The Quick Start gives a coherent 4-step sequence, and Workflow section 3 supplies genuine validation checkpoints for the batch operation ("Use collected evidence, not launch success alone"; "inspect those before deciding to rerun"), so the batch-operations cap does not apply. It is not 5 because the keep/drop decision — the workflow's terminal step — is left implicit, and there is no validate-then-retry loop spelling out when a rerun is warranted versus a partial-output salvage. | 4 / 5 |
Progressive Disclosure | There are no bundle files (no references/, scripts/, or assets/ directories), and the body is under 50 lines with no content that needs offloading, so per the rubric's simple-skill guidance well-organized sections alone warrant a 5. Sections (Overview, Quick Start, Suggested tools, Workflow, Method-critical defaults, Output Contract, References) are clearly headed, one level deep, and easy to navigate; the 'References' section points to a sibling skill (vasp-input-preparation) as a handoff boundary, not a nested file. | 5 / 5 |
Total | 16 / 20 Passed |