Content
85%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A tight, highly actionable rulebook with an unambiguous review workflow and a well-designed output contract — the content quality is strong. The one material defect is bundle structure: STANDARDS.md is referenced four times as the source of exact values but does not exist in the bundle, leaving the skill unable to deliver the precise citations it repeatedly promises.
Suggestions
Ship the referenced STANDARDS.md (easing curves, per-element duration budgets, spring configs, gesture tables) in the bundle, or inline the essential values into SKILL.md and remove the dead links.
Deduplicate the escalation-trigger list against the Ten Standards (e.g., ease-in, scale(0), layout properties, symmetric timing) by marking the standard numbers they escalate, trimming tokens without losing the flag-on-sight signal.
State a fallback for when a needed value is not in STANDARDS.md (e.g., 'if no exact value is catalogued, flag it as unspecified rather than approximating') so the missing-catalog case degrades gracefully.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense and rule-driven with no padding or explanations of concepts Claude already knows — concrete values like `scale(0.9–0.97)`, sub-300ms budgets, and `@media (hover: hover) and (pointer: fine)` appear inline. It stays at 4 rather than 5 because the "Aggressive Escalation Triggers" list substantially re-states the Ten Non-Negotiable Standards (ease-in, scale(0), layout properties, hover gating, symmetric timing all appear twice), which could be tightened into one list with pointers. | 4 / 5 |
Actionability | Guidance is copy-paste concrete throughout: the example findings table gives exact replacements (`transition: transform 200ms ease-out`, `transform: scale(0.95); opacity: 0`, `var(--radix-popover-content-transform-origin)`), and standards cite specific thresholds (sub-300ms, scale 0.9–0.97, 30–80ms stagger). It exceeds the 4 anchor because the example rows cover the common cases end-to-end, not just isolated snippets. | 5 / 5 |
Workflow Clarity | This is a single-purpose review skill and the action is unambiguous: apply the ten standards, flag the escalation triggers, propose fixes via the ordered remedial hierarchy, then emit the two-part required output ending in an explicit Block/Approve decision with concrete criteria. The mandatory findings table plus explicit verdict checklist function as validation checkpoints for the review's own output, matching the simple-skill exception for a 5. Not below 5 because nothing in the sequence is ambiguous and the review task carries no destructive/batch risk requiring additional validation. | 5 / 5 |
Progressive Disclosure | SKILL.md itself is well-structured with clear sections and appropriately defers detailed tables (easing curves, duration budgets, spring config) to a one-level-deep, clearly-signaled [STANDARDS.md](STANDARDS.md) — good design. However, the bundle contains no STANDARDS.md (no references/, scripts/, or assets/ directories exist), so all four links to the rule catalog are broken and the instruction "pull the exact one from STANDARDS.md" cannot be followed, defeating navigation to the detailed material. | 3 / 5 |
Total | 17 / 20 Passed |