Content
85%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A strong, highly operational body: the Q-MODE interview → bootstrap → spec mutation → setup → test pipeline is explicitly sequenced with race-condition warnings, validation checkpoints, and a real troubleshooting table, and progressive disclosure is exemplary with verified one-level-deep references. The only deductions are duplicated gotcha warnings and minor executable gaps in the Evaluate/Deploy tail sections.
Suggestions
Conciseness: state each environment gotcha (`.venv/bin/python` vs bare `python`, `adk web "$SKILL_DIR/scripts"`) once at first use and let the Troubleshooting table carry the repeats, instead of restating them in both Workspace Setup and Test.
Actionability: define `<repo-root>` in the Evaluate section and show where `EVAL.yaml` lives (it is absent from the Project Tree), so `./vs eval retail-product-search --project-id $PROJECT` is runnable as written.
Actionability: add a concrete example `gcloud run deploy` invocation (service name, source, the two required role bindings) in Deploy, while keeping the explicit human-approval gate.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | Largely lean — tables, exact commands, and no explanation of concepts Claude already knows — but the same gotchas are repeated: the "Use `.venv/bin/python`, not bare `python`" warning appears in Workspace Setup and again in Test, the `adk web "$SKILL_DIR/scripts"` warning appears in Test and again in the troubleshooting table, and the `'NoneType'` error is explained in Mode 1 step 2 and again in Troubleshooting. Not 5: these duplications could be trimmed to a single pointer; not 3: there is no padded or known-concept explanation anywhere. | 4 / 5 |
Actionability | The main workflow is copy-paste ready end to end: `bash "$SKILL_DIR/scripts/bootstrap.sh"`, exact Edit/sed substitutions for `gcp_project_id: ""`, `.venv/bin/python "$SKILL_DIR/scripts/setup.py" --config ./design-spec.md`, and a smoke-test one-liner. Minor gaps remain: the Evaluate section says `cd <repo-root>` without defining where that is, references an `EVAL.yaml` that is not in the shown Project Tree, and Deploy offers only "Deploy via `gcloud run deploy` or your org's existing tooling" with no concrete invocation. Not 5: those are concrete minor gaps; not 3: everything in the common path is fully executable with expected errors and fixes. | 4 / 5 |
Workflow Clarity | Sequencing is explicit ("Run these steps SEQUENTIALLY — do not parallelize... running them concurrently is a race", "Run bootstrap first and wait for completion"), validation checkpoints are built in ("On non-zero exit, surface the error and check references/troubleshooting.md", a dedicated Test section, EVAL assertions, and a Completion Checklist), and feedback loops exist via the inline error→fix table. Not 4: batch/destructive operations (setup ingestion, cleanup.py with `--confirm`) are all paired with verification steps, so the anchor-5 pattern of validate→fix→retry is fully present. | 5 / 5 |
Progressive Disclosure | The body is an operational overview with all deep material split into six real, one-level-deep reference files (install-paths.md, dependencies.md, architecture.md, troubleshooting.md, agent-example.md, ingestion-scripts.md — all verified present), each linked at the moment it becomes relevant and summarized in a "Load on demand" section, plus a Project Tree and per-script annotations. Not 4: there are no nested references and no bulk content inlined that belongs in a bundle file. | 5 / 5 |
Total | 18 / 20 Passed |