Content
100%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
An exemplary short, single-purpose skill: lean prose, an executable evidence-manifest specification, explicit validation criteria, and a clear failure heuristic with no padding or dangling references.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is lean and assumes Claude's competence — "Verify the real artifact. Do not infer from proxies, self-reports, or 'it compiles.'" — with no explanation of known concepts; the only example (the manifest JSON) teaches a project-specific format Claude could not already know. Every token earns its place. | 5 / 5 |
Actionability | Guidance is fully concrete: an exact output path ('.factory/evidence/<ticket>/<task>/manifest.json'), a copy-paste-ready manifest example with command, cwd, exitCode, and artifact, a hard acceptance rule ("A criterion without a command and a zero exit code is not proof"), and a specific UI variant ("the Surf screenshot and the read-back of the rendered page, taken against the worktree's own server"). | 5 / 5 |
Workflow Clarity | This is a simple, single-purpose skill and the single action is unambiguous: check the real artifact, then record evidence. It includes an explicit validation standard (zero exit code) and an error-recovery heuristic ("When a check fails, suspect the observation method before suspecting the product"), so the simple-skill exception applies cleanly. | 5 / 5 |
Progressive Disclosure | The skill is 37 lines, needs no external references (none exist and none are referenced), and is well organized as principle → verification checklist → evidence format → commit rule. Per the under-50-line guideline, this earns the top score without section headers or bundle files. | 5 / 5 |
Total | 20 / 20 Passed |