Content
88%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A tight, high-signal body: it assumes Claude's competence, gives a runnable command with a deliberate deferral of volatile flags to --help, and provides an explicit triage-and-re-run feedback loop. The only soft spot is the occasional high-level hint ('an option lets you grade an already-running agent') where the specific flag would make guidance fully copy-paste ready.
Suggestions
Name the option/flag for grading an already-running agent (e.g. 'lk agent simulate --agent <name> text --scenarios scenarios.yaml') so that path is copy-paste ready rather than a hint.
For the export/list workflow, include one concrete example invocation (e.g. 'lk agent simulate export <run-id>') since '--help names the flags' currently leaves the reader to discover the exact syntax mid-task.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | Lean and efficient throughout: no explanation of concepts Claude already knows, no padding, and every sentence carries skill-specific information ('Text is the default... faster, cheaper and more deterministic'; 'The gap between the two is what the user experiences'). Deferring flags to '--help' avoids restating volatile details. | 5 / 5 |
Actionability | Mostly executable guidance: a copy-paste command ('lk agent simulate text --scenarios scenarios.yaml'), named subcommands (list, export), and a concrete triage decision list. However, 'An option lets you grade an already-running agent by name' names neither the option nor its flag, leaving a hint where a specific invocation belongs. | 4 / 5 |
Workflow Clarity | Clear sequence with explicit validation and feedback loops: triage failures into three categorized causes with fixes, then 'After a fix, run the whole file, not only the scenario you were working on', plus 'exits non-zero when any scenario fails' for CI and 'Move repeat failures down the stack' for escalation. Recovery paths are explicit and checklist-like. | 5 / 5 |
Progressive Disclosure | Well-organized single-file skill with clear section headers and well-signaled pointers to sibling skills (reading-livekit-docs for flags, writing-livekit-scenarios for authoring). It references no detail files of its own, and at ~103 lines it exceeds the under-50-line simple-skill case, so structure is good but there is no one-level-deep reference organization to reward. | 4 / 5 |
Total | 18 / 20 Passed |