Content
78%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A dense, highly operational body with real validation checkpoints and no wasted tokens, weakened by navigation that doesn't resolve: the per-job pages and authoring reference the body depends on are absent from the bundle. Fixing the file layout would move this from good to excellent.
Suggestions
Add the missing job pages (jobs/finding.md, jobs/enriching.md, jobs/researching.md, jobs/automating.md) and shared/authoring.md to the bundle, or repoint those references to files that actually exist — they are the primary per-task instructions and currently resolve to nothing.
Reconcile references/api-reference.md (60KB, present but never referenced) with references/sdk-reference.md — either reference it where the body currently points to sdk-reference.md or remove the orphan.
Expand the one-line "contract → compare → exploit → recover → export → price" pipeline into a short numbered sequence naming the command and validation checkpoint at each step, so the end-to-end order doesn't have to be inferred from across sections.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is relentlessly lean and assumes competence — "Ordinary TypeScript, no DSL", "Marginal, never amortized", "Quality gates precede economics" — with zero padding or explanation of concepts Claude already knows. Even the war-story sentences ("Resolving that conflict silently cost one run ~30 minutes") carry operative lessons rather than filler; every token earns its place. | 5 / 5 |
Actionability | Mostly executable: concrete invocations throughout ("deepline plays run <file.play.ts> --input '<json>' --debug", "python3 <skill-root>/scripts/scaffold-search-experiment.py ./deepline/data/<task-slug> --name <task-slug> --input-csv <rows.csv>", the feedback-send template). Falls short of fully copy-paste-ready because key guidance is aphoristic rather than executable — "Write `unit + decision + required facts + scale` before touching tools", "A concept is an information geometry, never a vendor" — and the actual Play-file shape is delegated to the scaffold script and job pages. | 4 / 5 |
Workflow Clarity | Validation checkpoints are genuinely present for this batch-cost workflow: "run `deepline preflight --json` as one standalone command", "run-and-export does the structural check, Play check, completed Play, run-bound export", "{ok: true, runId, output}" as the completion receipt, "sentinel-probe one row before scaling", and receipt labels (CUT CANDIDATE / NEVER REACHED / cached calls) that drive the next action. Held at 4 rather than 5 because the end-to-end sequence "contract → compare → exploit → recover → export → price" is given as a single cryptic line — the reader must assemble the actual step order from across the Quick Start, job-page, and Build-and-run sections. | 4 / 5 |
Progressive Disclosure | The routing pattern itself is textbook progressive disclosure — job pages "consulted on a trigger rather than read up front", two lookup files, SDK surface pushed to references/sdk-reference.md. But scored against the actual bundle, 5 of the 7 referenced paths do not exist: jobs/finding.md, jobs/enriching.md, jobs/researching.md, jobs/automating.md, and shared/authoring.md are all missing from the skill directory, and references/api-reference.md exists but is never referenced. Not 4 — broken navigation to the primary per-job instructions is more than a minor organization gap; not 2, because the structure and signaling are clearly good. | 3 / 5 |
Total | 16 / 20 Passed |