Content
81%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The content is a strong, lifecycle-complete guide with executable commands, explicit verification steps, and error-recovery loops — workflow clarity is excellent. Weaknesses are minor: some repetitive phrasing, the absence of a filled-in example invocation, and a template/failure playbook inlined in SKILL.md that could be split into reference files.
Suggestions
Trim repetition and rhetorical emphasis — e.g., collapse 'hack on code, run build/tests, hack on code, etc.' and the 'YOU are in charge of bookkeeping...' lines into one directive sentence.
Add one filled-in example of a real sub-agent invocation (model, a sample task description, expected reply) so the commands and template can be used verbatim.
Move the task-description template and the 'When Things Go Wrong' playbook into a reference file (e.g., references/task-template.md) and keep SKILL.md as a lean overview with one-level-deep pointers.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is mostly lean — concrete commands, a compact Do/Don't list, and a template — but has minor trims available: 'hack on code, run build/tests, hack on code, etc.' repeats 'hack on code', and lines like 'YOU are in charge of bookkeeping, not them. YOU have the big picture, they don't.' add emphasis over information. This fits 'efficient; minor instances of over-explanation that could be trimmed' rather than the every-token-earns-its-place 5 anchor. | 4 / 5 |
Actionability | It gives executable commands ('cursor-agent --print --model <model-name> create-chat', '--resume <conversation-uuid>'), a concrete default model ('sonnet-4.5'), a model-discovery trick, an explicit timeout setting (600000), and a structured task template. The gap versus the 5 anchor is that no filled-in worked example of a real invocation/task is shown, and the commands still require the user to substitute placeholders without an instance of doing so. | 4 / 5 |
Workflow Clarity | The sections follow the full lifecycle — create chat, model selection, invocation with timeout, setup, giving instructions, post-completion verification, and failure recovery — with explicit validation ('Always verify the sub-agent's work: check the diff, review code, run tests/lint') and feedback loops for error recovery ('Read the output, fix any blocking issues, retry with more context', 'Continue the work yourself or spawn another sub-agent with clarified instructions'). This matches the anchor with explicit validation steps and retry loops, not merely 'most checkpoints present'. | 5 / 5 |
Progressive Disclosure | The body (~95 lines) has no bundle files and is well-sectioned with clear headers covering each lifecycle phase. However, it exceeds the under-50-lines simple-skill threshold, and content such as the full task-description template and the failure-mode playbook could plausibly live in a reference file, leaving SKILL.md a leaner overview. This fits 'good structure; most content appropriately placed; minor organization gaps' rather than the ideally-split 5 anchor. | 4 / 5 |
Total | 17 / 20 Passed |