Content
85%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A highly actionable, well-sequenced body with executable commands, explicit validation checkpoints, and excellent coverage of both job lanes and the durable-execution ladder. Its main weaknesses are a monolithic single-file structure with no progressive disclosure into reference files (and a project docs dependency instead), plus deliberate but costly repetition of the deadman doctrine across three sections.
Suggestions
Split the durable-execution material into a reference file (e.g. references/durable-execution.md) — the capability ladder, deadman pattern, and stage-checkpoint appendix — and keep SKILL.md to the routing table plus a one-paragraph summary with clear links, reducing the ~480-line body substantially.
Move the brain-tools allowlist enumeration and the full flag inventories (gbrain jobs submit / gbrain agent run) into a short reference file, keeping only the 2-3 most common tools and flags inline.
Consolidate the deadman-doctrine statements: state the rules once (Contract or Durable execution) and have the other two sections reference them by link/anchor instead of restating the same three failure modes.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense and nearly every line carries project-specific facts Claude cannot know (the GBRAIN_ALLOW_SHELL_JOBS gate, MCP trust boundary, coalescing behavior, lock TTL semantics, flag inventories) — there is no padding explaining generic concepts. It misses 5 because the durable-execution doctrine is stated three times (Contract, Durable execution section, Anti-Patterns) — an acknowledged deliberate mirror, but still duplication that could be consolidated. | 4 / 5 |
Actionability | Guidance is copy-paste executable throughout: full CLI invocations ('gbrain jobs submit shell --params "{\"cmd\":\"echo hello\",\"cwd\":\"/abs/path\"}"'), argv forms, monitor/control command blocks, the fanout-manifest invocation, lifecycle ops with parameter examples (replay_job with data_overrides), and a concrete timeout table for common long operations. Common cases (submit, monitor, steer, cancel/replay, long-op routing) are each covered by a specific example. | 5 / 5 |
Workflow Clarity | Multi-step processes are explicitly sequenced — Phases 1-5 for subagent jobs, the three-rung capability ladder ordered by deployment capability, and a numbered deadman pattern — with explicit validation checkpoints throughout: 'run gbrain jobs stats to confirm the worker is registered', the deadman's reported-check with a four-branch decision tree (reported/unreported/still-running/dead), checkpoint freshness checks, integrity-on-restore checks, and disarm-on-completion feedback. This is batch/destructive-adjacent work and validation is present, so no cap applies. | 5 / 5 |
Progressive Disclosure | No bundle files exist (no references/, scripts/, or assets/), so all ~480 lines live in SKILL.md itself. Section headers and a routing table give reasonable navigation, but content that clearly belongs in separate one-level-deep reference files is inlined — the content-addressed stage-checkpoint appendix, the 14-tool brain allowlist enumeration, and the full flag inventory — matching the level-3 anchor ('some structure... content that should be separate is inline') rather than 4, where most such content would be split out. | 3 / 5 |
Total | 17 / 20 Passed |