Content
75%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-structured, actionable skill file: concrete CLI commands, a compact workflow graph, realistic examples, and a useful mistakes table. The main weaknesses are duplicated scan-modules guidance, a small `$PYTHON`/`$PYTHON_PATH` inconsistency, and the absence of explicit verification steps around hypothesis creation and deletion.
Suggestions
Consolidate the scan-modules guidance: state it once (e.g., in Module Selection) and trim the duplicate rows in Common Mistakes and Pipeline Position to short cross-references.
Fix the `$PYTHON` vs `$PYTHON_PATH` inconsistency in the Progress Tracking command and explain placeholders like `N` (e.g., 'replace N with the number of hypotheses added').
Add an explicit verification checkpoint to the workflow (e.g., run `hypothesis list --json` after `hypothesis add` to confirm creation before updating TOPIC.md) and note that `hypothesis delete` is irreversible and should be confirmed with the user first.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is lean and table-driven with no explanations of concepts Claude already knows, but scan-modules guidance is repeated in 'Module Selection' ('use `scan-modules` to discover or validate them'), twice in 'Common Mistakes' ('Run `scan-modules` to confirm valid class and module names', 'Copy exact names from `ags.py scan-modules list --short` output'), and again in 'Pipeline Position', which could be consolidated. | 4 / 5 |
Actionability | The Quick Reference table gives full executable commands with all flags, the Group JSON Format example is concrete and copy-paste ready, and the module combination table covers common cases. Minor gaps: the Progress Tracking command uses `$PYTHON` while every other command uses `$PYTHON_PATH`, and `--metadata '{"hypotheses_count": N}'` leaves the placeholder `N` unexplained. | 4 / 5 |
Workflow Clarity | The dot digraph gives a clear, ordered sequence (read topic -> clarify -> names known? -> scan -> groups -> write -> sync) with a conditional checkpoint (scan-modules when module names are unknown) and CLI-side validation implied by `--skip-module-validation`. However there is no explicit post-add verification step in the flow, and the destructive `hypothesis delete` command has no confirmation or safety guidance. | 4 / 5 |
Progressive Disclosure | Sections are well-organized and content is appropriately inline for a skill of this size (~112 lines) with no nested references. Minor gaps: the bundle's `scripts/hypothesis.py` is never referenced from the body (commands go through the external `.agentsociety/bin/ags.py`), and setup details defer to an external `CLAUDE.md` rather than a bundled reference. | 4 / 5 |
Total | 16 / 20 Passed |