Content
88%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A strong, execution-first install guide: fully copy-paste-ready commands, an explicit verification checkpoint with expected output, and a symptom→fix troubleshooting table as a feedback loop. Its only weaknesses are mild — a small amount of demonstrative padding and a monolithic single-file structure where a short reference file could offload detail.
Suggestions
Trim the intro sentence that restates the frontmatter description and drop or compress the two sample natural-language questions; they demonstrate usage without adding executable value.
Move the API-key section and expanded troubleshooting entries into a one-level-deep reference file (e.g. references/troubleshooting.md), keeping SKILL.md to install/verify/update essentials and signaling the reference clearly from the Troubleshooting section.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dominated by executable commands, expected-output blocks, and two tight tables — largely lean. Minor over-explanation remains: the intro line "an MCP server exposing 1000+ scientific tools plus 120+ research skills — all auto-configured" duplicates the frontmatter description, and the sample questions ("What are the top mutated genes in breast cancer?" / "Research the drug metformin.") are demonstration padding. This matches anchor 4 ('efficient; minor instances of over-explanation that could be trimmed') better than 5, since not every token earns its place, and is far above the 'mostly efficient but some unnecessary explanation' level 3. | 4 / 5 |
Actionability | Every step is a concrete, copy-paste-ready command: `uv --version`, `agy plugin install /path/to/ToolUniverse/plugin`, the full `git clone https://github.com/mims-harvard/ToolUniverse.git` sequence, `agy plugin list` with expected JSON output, `uv cache clean tooluniverse`, and `uv python install 3.12`. The common cases (fresh install, verify, update, troubleshoot) are each covered by an executable example — anchor 5; the only placeholder (`/path/to/...`) is inherent to user-specific paths and Option 2 removes even that. | 5 / 5 |
Workflow Clarity | The sequence is clear (Prerequisites → Install → Restart → Verify) with an explicit validation checkpoint — "Expect output showing `tooluniverse` under `imports`" gives the exact success criterion — and the Troubleshooting table provides symptom→fix feedback loops for error recovery (`uvx: command not found`, "MCP server won't start", "tools missing"). This matches anchor 5; no destructive or batch operations exist that would cap the score, and validation is explicit rather than implicit, ruling out 4. | 5 / 5 |
Progressive Disclosure | Sections are well organized and self-contained with no nested or dead references (no references/, scripts/, or assets/ exist), so navigation is easy. However, the body runs ~100 lines — above the sub-50-line simple-skill exception — and detail such as the "What you get" component table, the API-key section, and the expanded troubleshooting entries could live in one-level-deep reference files to keep SKILL.md a leaner overview. That places it at anchor 4 ('good structure; most content appropriately placed; minor organization gaps') rather than 5. | 4 / 5 |
Total | 18 / 20 Passed |