Content
78%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is a highly actionable, token-efficient command reference with sensible golden rules and a well-organized guide index. Its one serious defect is that the entire referenced bundle is missing: all 10 linked guides are absent from the skill directory, leaving dead links and duplicated inline reference material.
Suggestions
Ship the 10 referenced guide files (core-commands.md, test-generation.md, screenshots-and-media.md, tracing-and-debugging.md, request-mocking.md, running-custom-code.md, storage-and-auth.md, session-management.md, device-emulation.md, advanced-workflows.md) — or remove the Guide Index if they will not exist.
Move the bulk of the inline Command Reference into core-commands.md (and storage-and-auth.md for the cookie/storage commands), keeping only the most common commands in SKILL.md to eliminate duplication with the linked guides.
Add explicit verification steps to key workflows, e.g. re-run snapshot after fill/click to confirm the page state changed, or check console/network output before declaring a debugging step complete.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is a dense, one-line-per-command cheat sheet with no explanations of concepts Claude already knows — efficient but not perfectly lean. Minor duplication (snapshot/fill/click appear in both Quick Start and Command Reference, and Golden Rules restate command facts) keeps it at anchor 4 rather than 5. | 4 / 5 |
Actionability | Every entry is an executable, copy-paste-ready command ("playwright-cli open https://playwright.dev", "playwright-cli fill e5 \"search query\"", "playwright-cli screenshot --filename=checkout-step3.png"), and the Quick Start covers the common case end-to-end. Fully concrete with no pseudocode — anchor 5. | 5 / 5 |
Workflow Clarity | Quick Start gives a clear sequence (open → snapshot → interact → screenshot → close) and Golden Rules encode checkpoints ("Always snapshot first — never guess ref numbers", "tracing-start before the failing step, not after"). It stays at 4 rather than 5 because there are no explicit verify-after-action or error-recovery steps, though the operations are interactive rather than destructive batch jobs. | 4 / 5 |
Progressive Disclosure | The Guide Index is well organized and clearly signaled, but the bundle contains none of the 10 linked guide files (core-commands.md, test-generation.md, storage-and-auth.md, etc.) — every reference is dead, so navigation cannot resolve. Additionally, ~130 lines of inline command reference duplicate what the linked core-commands.md should hold. Structure exists and signaling is clear (above 2), but broken references and inlined reference content keep it below 4. | 3 / 5 |
Total | 16 / 20 Passed |