Content
82%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A tight, actionable skill body with a clean workflow, explicit edge-case handling, and a verified report-template reference. The main defect is the missing `agents/query-judge.md` file referenced in step 3, which is a dangling reference and weakens both actionability and progressive disclosure.
Suggestions
Add the referenced `agents/query-judge.md` to the bundle (or remove the reference and inline the judging criteria), since step 3 tells Claude to prefer a sub-agent whose file does not exist.
Include a short worked example of one domain-file query with its `must`/`should` expectations and the corresponding verdict, so the parse-then-judge contract is unambiguous.
Add a brief fallback rule for malformed domain files or partial MCP failures during `eval all` (e.g., report the domain as skipped vs. failed) to close the last workflow gap.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is lean and imperative throughout ("Skip empty files or files without `##` query headings", "Verdicts are `pass`, `partial`, or `fail`; never inflate ambiguous results") with no explanation of concepts Claude already knows and every line earning its place. The trigger-phrase list adds terms beyond the frontmatter rather than duplicating it. | 5 / 5 |
Actionability | Concrete, executable rules are given: file mapping (`eval/index/<domain>.md`), heading-parse conventions (`- must ...` / `- should ...` bullets), named search tools (`mcp_fusion_search_framework`, then `mcp_fusion_search`), verdict vocabulary, and a real report template. The gap is the reference to an `agents/query-judge.md` sub-agent that is not present in the bundle, though the inline fallback keeps the flow recoverable. | 4 / 5 |
Workflow Clarity | The four-step sequence (resolve → parse → judge → report) is clearly ordered with edge-case handling (skip empty files, sub-agent fallback, "Stop clearly if MCP is unavailable or rate-limited"), which function as checkpoints. Minor gaps: no guidance for malformed domain files or partial failures midway through an `eval all` run, and aggregation across domains is only implied by the report instruction. | 4 / 5 |
Progressive Disclosure | The ~58-line body is well-sectioned (When To Use, Inputs, Workflow, Safety, Expected Output) and points to `assets/report-template.md`, a real one-level-deep bundle file that exists. However, `agents/query-judge.md` is referenced but absent from the bundle, a dangling navigation path. | 4 / 5 |
Total | 17 / 20 Passed |