Content
82%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A lean, highly actionable skill body: executable tool-call examples, concrete error codes, and strong verification discipline with feedback loops for a risky GUI-automation domain. Main gaps are mild — the screenshot-verification rule is repeated across sections instead of stated once, and there is no single ordered end-to-end flow.
Suggestions
State the screenshot-as-proof rule once (e.g., in Wayland Notes) and reference it elsewhere instead of restating it three times.
Add one short ordered overview of the full loop (detect features -> snapshot -> target -> act -> screenshot verify) so the sequence doesn't have to be reconstructed from three sections.
Consider moving the GNOME shortcut list and the permission/error-code catalog into a one-level-deep reference file to keep SKILL.md an overview.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense, imperative, and free of concepts Claude already knows — every line carries a rule, caveat, or concrete parameter (e.g., "Pass `window_id` whenever one is known so unrelated applications cannot consume the node budget"). It sits at 4 rather than 5 because the screenshot-as-proof rule is restated several times ("Use screenshots for proof after every state-changing action", "Treat every shortcut as an attempt. Inspect the fresh screenshot before saying it worked", "an accepted AT-SPI call without active/focused state is not activation proof") and could be consolidated. | 4 / 5 |
Actionability | Two complete, copy-paste-ready `computer_use_remote` JSON tool calls (`ax_snapshot`, `ax_action`) with concrete parameters, an enumerated operation list (`press`/`focus`/`set_value`), concrete error codes (`COMPUTER_USE_WINDOW_REQUIRED`, `COMPUTER_USE_TARGET_NOT_FOCUSED`, `COMPUTER_USE_AX_UNAVAILABLE`), specific shortcuts, and an ordered fallback chain. This covers the common cases fully, matching the top anchor; score 4 would require missing key details, and there are none of consequence. | 5 / 5 |
Workflow Clarity | There is an explicit conditional flow ("prefer the generic background loop from `host-computer-use`: `list_windows` -> `get_window_state` -> `element_action`. If those features are absent, use the AT-SPI snapshot/action flow below") plus snapshot-choose-act-verify feedback loops ("If an action reports ambiguity, take a fresh snapshot and narrow the target") and validation checkpoints ("Continue only when the result says `focus_verified=true`"). It stays at 4 rather than 5 because the end-to-end sequence is distributed across sections rather than presented as one ordered checklist, so a reader must assemble the order themselves. | 4 / 5 |
Progressive Disclosure | No references/, scripts/, or assets/ directories exist, and the single ~95-line body is well-sectioned (AT-SPI Targeting, Wayland Notes, Permissions) with no nested or buried references, so there is no navigation failure. It is a 4 rather than 5 because the GNOME shortcut list and the permission/error-code catalogs are inline reference material that could live in a one-level-deep reference file, and the skill exceeds the under-50-lines simple-skill case. | 4 / 5 |
Total | 17 / 20 Passed |