Content
92%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A strong CLI skill body: action-dense, explicitly sequenced workflows with validation checkpoints, and textbook progressive disclosure with two real, well-signaled reference files. The only flaw is minor duplicated guidance around runner billing/defaults that a consolidation pass would remove.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense and assumes competence — no tutorial padding, no explaining what a browser or git is — and the per-command table is the economical way to cover ~100 commands. It loses a point only for minor redundancy: runner billing and termination guidance ('it is billed while it runs... terminate a runner you launched' vs 'Launching a runner starts a billed pod, so reuse one id... and stop a runner when you are done') is stated twice in overlapping words, and the runner-default/--runner/QAWOLF_RUNNER_ID rules echo across sections. Not a 5 because these repeats could be consolidated; not a 3 because there is no concept-level over-explanation anywhere. | 4 / 5 |
Actionability | Every section gives executable direction: exact flags ('--flow-statuses failed', '--env', '--env-id', '--file-paths <path>', '--screenshot <path>'), a copy-paste bash block for resuming an agent session, the upload→PUT→pass-path sequence with concrete commands, and an explicit first-action rule ('run qawolf <command> --help once'). Common cases (auth check, run reporting, browser driving) are each covered with specific commands. | 5 / 5 |
Workflow Clarity | Multi-step processes are explicitly sequenced with checkpoints and feedback loops: the git-backed flow (stage → commit → push → 'poll environment get until lastSyncedCommitHash includes your commit' → read flowId/url from a named command), the agent flow (send with --follow, interpret the exit-0 question state, answer with --session, re-follow), and error-recovery guidance (retry run stop only after run.get returns the run, report deployment under a new providerDeploymentId to re-evaluate). Destructive/batch operations get safety validation (never blind-retry writes, verify with auth whoami, inspect git status before publishing), so the workflow cap does not apply. | 5 / 5 |
Progressive Disclosure | The body is a true overview: cross-cutting rules (auth, output, safety, billing) inline, the deep material — run-result/trace reading and the full runner workflow — split into exactly two one-level-deep references, both present on disk and clearly signaled with bold 'Read references/... before...' directives plus a fetch fallback URL if the path is absent. The generated commands table is inline navigation, appropriate for an overview, and marked as generated. | 5 / 5 |
Total | 19 / 20 Passed |