Content
88%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is a strong orchestrator: precise, executable gating and state logic, explicit validation checkpoints, honest build-status disclosure, and disciplined routing into a large, genuinely organized reference bundle. Its main weaknesses are mild cross-section repetition of a few rules (Generate consent, Graviton default) and a reference chain that nests more than one level deep by design.
Suggestions
Consolidate the Generate/run_mode consent rule into one canonical section (e.g., Execution) and reduce the Philosophy, State Management, and Sidebar mentions to one-line pointers, mirroring the no-duplication policy the body already applies to mapping tables.
Trim the repeated x86_64-vs-Graviton statements to the Philosophy bullet plus a single pointer from Defaults and Sidebar Placement.
In the Phase Structure section, add a one-line map of the reference levels (SKILL.md → phase orchestrator → fragment/shared ref) so the multi-level chain is explicit navigation rather than something the reader reconstructs.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense, skill-specific orchestration with no explanations of concepts Claude already knows, and it assumes competence throughout (e.g., "$AZURE_SUBSCRIPTION = The subscription id passed explicitly on every `az` command. Never rely on the CLI's ambient active-subscription context."). However, a few rules are restated across sections — the Generate/run_mode consent rule appears in Philosophy, Execution, and State Management, and the x86_64-vs-Graviton default appears in Philosophy, Sidebar Placement, and Defaults — which is more than minor trimming would remove. Not 3: the overwhelming majority of the 245 lines is unique, load-bearing orchestration rather than unnecessary explanation; not 5: the cross-section repetition and some defensive justifying prose could be tightened. | 4 / 5 |
Actionability | For an instruction-only orchestration skill, the guidance is fully executable: exact state transitions with conditions ("If `current_phase == "estimate"` AND `phases.estimate == "completed"` AND `phases.workshop` is `"pending"` or `"in_progress"`, **do not recompute Estimate**"), an exact copy-paste sidebar prompt with labeled options and per-option state outcomes, precise file paths to load per decision point, and exact state-key semantics ("`run_mode`" | "decide_and_execute""). Not 4: there are no gaps — every rule is stated with the specific files, keys, and conditions needed to execute it. | 5 / 5 |
Workflow Clarity | The multi-phase workflow is clearly sequenced with explicit validation checkpoints: the interpreter loop is summarized ("reads `.phase-status.json`, determines the current phase, runs each phase's `_preconditions` / fragments / `_assemble` / `_postconditions`, advances on `HANDOFF_OK` via `_advances_to`, and validates state"), cold-start and warm-start entry rules are precise, interrupted-state recovery is handled (the mandatory workshop-resume rule), and a halt-on-error checkpoint is stated ("If unable to complete a step, stop and report the specific issue. Do not fabricate or infer data."). Not 4: validation is explicit at every gate (Clarify mandatory, `run_mode` NOT consent, workshop must complete before Generate), not merely implicit or partial. | 5 / 5 |
Progressive Disclosure | The body is a well-structured overview that routes to a real, well-organized bundle: "begin at `references/phases/discover/discover.md`, this skill's entry phase", "the vendored `references/vendored/dsl/INTERPRETER.md`: ... **Load it first**", and per-dialect files "it loads only for the dialects actually present" — and every referenced path I checked (discover-live.md, shared/graviton.md, shared/extract-terraform.md, vendored/pricing/aws-infra-pricing.json, design-refs/ai.md, vendored/ai/*) exists. Not 5: reference chains run 2+ levels deep (SKILL.md → phase orchestrator → fragments → shared refs), which the top anchor flags, even though navigation signals and the ~800-line budget rules keep it navigable. | 4 / 5 |
Total | 18 / 20 Passed |