Content
85%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
Excellent actionability and workflow clarity — every step has executable commands, verification checkpoints, and error-recovery loops, and the reference quality rules read as a genuine checklist. The two weaknesses are mild routing-rule redundancy and a monolithic single-file layout that inlines backend details and artifact specs that belong in separate reference files.
Suggestions
Move the backend compatibility details ("Important compatibility behavior", "Output compatibility") and the per-backend sections (Chat, Perplexity, deep research) into a references/backends.md and keep only the routing table in SKILL.md.
Extract the packet artifact descriptions (step 4) into a references/packet-format.md linked from the workflow step, leaving SKILL.md a concise overview.
De-duplicate the routing rules stated in both the "Parallel-first routing" table and the per-backend sections — state each routing rule once and cross-reference it.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is almost entirely operational content — commands, flags, packet file listings, quality rules, failure handling — with no explanations of concepts Claude already knows. It does not reach 5 because routing rules are repeated across the "Parallel-first routing" table, "Important compatibility behavior", and each backend section, and the "Output compatibility" field list plus citation-fetching instructions could be trimmed or moved. It stays above 3 because nearly every section carries non-obvious, task-specific detail. | 4 / 5 |
Actionability | Every workflow step ships a copy-paste-ready command with real flags (--academic, --target-references 60, --context-file, --packet-dir, --force-backend, --extract-limit), a concrete JSON context example, the exact packet artifact filenames, pinned install commands, and named failure remedies. This matches the fully-executable anchor covering the common cases. | 5 / 5 |
Workflow Clarity | The five-step manuscript workflow is clearly sequenced with explicit checkpoints: extraction verification in step 3, "inspect coverage.json; refine the question, date range, terminology, or domains" as a shortfall feedback loop, labeled single-source/conflicting claims held "until reviewed", and a failure-handling section with per-error recovery steps. Batch mode isolates failures per query with errors kept in each result envelope, so the batch-operation cap does not apply. | 5 / 5 |
Progressive Disclosure | Sections are well-organized with clear headers and the script bundle is real (scripts/research_lookup.py and scripts/manuscript_packet.py exist and the documented invocation path resolves to them), but at ~336 lines everything lives inline in SKILL.md with no reference files at all — backend compatibility minutiae, packet artifact specs, and citation instructions are content that could be split into one-level-deep reference docs. It is above the 2 anchor because the content is well-sectioned workflow guidance rather than an unstructured API dump, but below 4 because a document this long should offload detail rather than keep it all on the front page. | 3 / 5 |
Total | 17 / 20 Passed |