Content
67%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-organized, knowledgeable body with genuine fail-fast validation checkpoints and load-bearing craft detail, held back by a complete absence of executable run commands — the agent must open the scripts to learn the CLI — plus some repetition of the signature mechanic across sections and one unsurfaced bundle file (layout.py).
Suggestions
Add runnable invocations to the Scripts section, e.g. `python scripts/build_assets.py --config config.json --work-dir work/` and `python scripts/compose.py --config config.json --work-dir work/ --out out.mp4`, so the build→compose pipeline is copy-paste executable from SKILL.md.
Make the step ordering explicit: a short numbered sequence (1. bind config from recipe choices → 2. run build_assets.py → 3. run compose.py → 4. `watch` QC the master, with the shorten-copy/reduce-count loop on validation failure) would convert the implied workflow into a checklisted one.
Surface `scripts/layout.py` in the Scripts list (it is the shared geometry/wrapping/timing contract both other scripts import), and consider splitting the dense 'Wrapped notifications and software endings' section into a one-level-deep reference file to slim the overview.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense with craft specifics and explains nothing Claude already knows (no 'what is PIL', no library tutorials) — every section carries load-bearing detail like the geometry contract ("W=1080 H=1920, SIDE=135 → banner width BODY_W=810, BANNER_H=176, PAD=60, row pitch H=214, bottom anchor YB=1200"). The reason it is not a 5: the bottom-up push mechanic and the banner styling ("green Apple Messages icon, warm TRANSLUCENT greige... soft dark box-shadow") are each repeated across the intro, Scripts, and Craft rules sections, and the final "Wrapped notifications and software endings" section is padded prose with vague referents ("It holds at least 1.5 seconds") that could be tightened. These are minor over-explanations, matching level 4 rather than level 5's 'every token earns its place'. | 4 / 5 |
Actionability | There is real concrete guidance — exact script names with their outputs ("scripts/build_assets.py — draws the assets from config: nb-1..N.png... pill.png... endcard.png"), exact font sizes ("title Semibold ~38, body Regular ~36, NOW/handle ~25"), the count rule ("3–5 notifications"), and a pointer to the config shape ("scripts/config.example.json"). But the body contains zero runnable commands: no `python scripts/build_assets.py --config ... --work-dir ...` or compose.py invocations, so the actual CLI usage is only discoverable by opening the scripts. Per the rubric's 'evaluate what is written' guideline this is concrete guidance with a key executable detail missing — the level-3 anchor — not level 4's 'concrete code or commands with minor gaps'. | 3 / 5 |
Workflow Clarity | The pipeline sequence is legible (build_assets.py produces the PNGs from config; compose.py animates them: "Ken-Burns push-in on the plate → each banner springs in... → pill rides above → ✕-clear... → serif end card fades in → optional audio → encode h264 + aac") and validation checkpoints are genuinely present: fail-fast on oversized copy ("fail with a correction request. Shorten copy or reduce the notification count; never silently truncate"), a manifest gate ("Compose requires that manifest and rejects changed inputs until assets are rebuilt"), and a QC step ("Requires: watch (QC the final master)"). It falls short of level 5 because the build→compose ordering is implied rather than an explicit numbered sequence, and there is no explicit 'if validation fails, do X then re-run' loop for the QC step — level 4's 'clear sequence with most checkpoints; minor validation gaps'. | 4 / 5 |
Progressive Disclosure | Scored against the actual bundle: scripts/ contains build_assets.py, compose.py, config.example.json, and layout.py, and the body's referenced paths ("scripts/build_assets.py", "scripts/compose.py", "scripts/config.example.json") all exist. Sections are well-organized and the SKILL.md stays an overview of ~80 lines. Two gaps keep it at level 4 rather than 5: layout.py — which both scripts import and which enforces the body's own geometry/wrapping/timing contract — is never named in the body's Scripts section, and the dense final section (resolution objects, b2b end-card requirements) reads as inline reference material that could live in a separate one-level-deep file. | 4 / 5 |
Total | 15 / 20 Passed |