Content
80%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is an exceptionally lean and actionable single-file catalog: every command is a one-liner and the examples are copy-paste ready. Its weaknesses are the absence of validation/error-recovery checkpoints in the recommended flows and the fact that the entire command and parameter reference is inlined in SKILL.md rather than split into reference files.
Suggestions
Add validation checkpoints and error-recovery guidance to the recommended flow, e.g. 'If `click --on B3` fails, re-run `see --annotate` to refresh element IDs and confirm the element exists' and a permissions pre-flight check before automating.
Move the full command catalog and common-parameter tables into a references/ file (e.g. COMMANDS.md), keeping SKILL.md to the overview, quickstart, and the most common flows with a clearly signaled 'See COMMANDS.md' link.
State how to verify results after acting (e.g. re-capture or query the element state after click/type) so the automation loop has an explicit feedback step.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | Every section is a terse one-line-per-command catalog with copy-paste examples and no explanation of concepts Claude already knows — matching 'Lean and efficient; assumes Claude's competence; every token earns its place'. It is not score 4 because there is no padding to trim: even the parameter sections are compact flag lists rather than prose. | 5 / 5 |
Actionability | The body provides fully executable, copy-paste-ready commands throughout ('peekaboo see --app Safari --window-title "Login" --annotate --path /tmp/see.png', the Quickstart happy path, and eight worked example blocks), fitting 'Fully executable; copy-paste ready code or commands; specific examples cover the common cases'. It is not score 4 because the examples cover the realistic common cases with real flag combinations rather than leaving minor gaps. | 5 / 5 |
Workflow Clarity | Sequences are present and the recommended flow is labeled ('See -> click -> type (most reliable flow)' and the note 'Use `peekaboo see --annotate` to identify targets before clicking'), but validation checkpoints and error-recovery guidance are missing — e.g. what to do when `see` finds no matching element, when permissions are missing (only 'Requires Screen Recording + Accessibility permissions' is stated), or how to verify a click landed. This matches 'Steps listed but validation gaps; sequence present but checkpoints missing or implicit' rather than score 4, which requires most checkpoints present. | 3 / 5 |
Progressive Disclosure | No bundle files (references/, scripts/, assets/) exist, and the body inlines the entire ~60-command catalog plus common-parameter tables in one file with good headers, but no one-level-deep references to offload detail — matching 'Some structure but could be better organized; content that should be separate is inline'. It is above score 2 because the file is well-sectioned and navigable (not a header-less wall of text), and below score 4 because the command catalog and parameter tables are exactly the bulk that belongs in a separate reference file for a CLI this large. | 3 / 5 |
Total | 16 / 20 Passed |