Content
57%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body delivers genuinely useful, concrete output templates for all three modes, but it pads itself with design-system concept explanations Claude already knows and never specifies the audit methodology or any validation checkpoints. It is also a monolith: with no reference files, every mode's ~40-line template is loaded regardless of which subcommand runs.
Suggestions
Cut the 'Components of a Design System' and 'Principles' sections (or compress to a few lines) — Claude already knows what design tokens, component variants, and UI patterns are; the token budget is better spent on methodology.
Add an explicit audit workflow with steps and a verification checkpoint (e.g., 1. enumerate components, 2. grep for hardcoded hex/px values, 3. cross-check naming against tokens, 4. re-validate findings against source before writing the report), since the 'Score: [X/100]' output currently has no defined basis.
Move each mode's output template into a one-level-deep reference file (e.g., references/audit-template.md, references/document-template.md, references/extend-template.md) so SKILL.md stays a lean overview and only the relevant template is loaded.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The output templates earn their tokens, but ~30 lines are spent explaining concepts Claude already knows — 'Design Tokens: Atomic values that define the visual language: Colors... Typography... Spacing...' plus a generic 'Components of a Design System' taxonomy and advice like 'Consistency over creativity'. This matches 'mostly efficient but includes some unnecessary explanation'; not a 2 because the majority of the body is template content rather than padding, and not a 4 because the concept-explainer section and Tips are clearly trimmable. | 3 / 5 |
Actionability | For an instruction-only skill the guidance is concrete: exact slash-command invocations ('/design-system audit') and full markdown output templates with tables and worked rows (e.g., 'Button | ✅ | ✅ | ⚠️ | 8/10'). It stops short of a 5 because the audit mode never says how to detect issues — no grep patterns for hardcoded hex, no commands to scan component files — leaving the core execution method implicit. | 4 / 5 |
Workflow Clarity | The three modes each have a target output shape, but the actual sequence of an audit (what to scan, in what order, how to score the 100-point total) is unstated, and there are no validation or verification checkpoints — only an implicit 'Start with an audit' tip. This matches 'steps listed but validation gaps; sequence present but checkpoints missing or implicit'; not a 2 because the mode separation and output contracts do provide a coherent rough sequence. | 3 / 5 |
Progressive Disclosure | The body is well-sectioned with headers, but everything lives in one 185-line SKILL.md with no bundle files (references/, scripts/, assets/ are absent) — ~120 lines of per-mode output templates that a given invocation won't use are always loaded into context. The single reference is an external path ('../../CONNECTORS.md') outside the skill bundle. This fits 'some structure but could be better organized'; not a 2 because the sections are clearly navigable, and not a 4 because the per-mode templates are prime candidates for one-level-deep reference files. | 3 / 5 |
Total | 13 / 20 Passed |