Content
73%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A strong instruction-only skill body with an exemplary iteration loop, validation gates, and a concrete output contract (report template, design-qa.md schema, passed/blocked rule). The main costs are token efficiency — repeated rules and an inlined fidelity-surface checklist that duplicates the qa-rubric reference — and a couple of high-level deferral steps where an inline example would help.
Suggestions
Conciseness: state the same-image-comparison rule once (either in the Workflow preamble or step 2) and delete the duplicate; drop the restated exclusions in the opening paragraphs since the frontmatter description already carries them.
Progressive disclosure: collapse the 'Required Fidelity Surfaces' section to a short list of the five surface names plus any non-negotiable rule, and move the per-surface detail into references/qa-rubric.md where it is already largely duplicated.
Actionability: add one concrete inline example of an evidence citation (e.g. a filled-in finding line with screenshot path and viewport) so the capture/cite steps are copy-paste executable rather than only described.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is mostly efficient, dense instructions, but includes notable repetition and padding: the description's exclusions are restated ('Do not use this skill for broad UX critique...'), the same-image-comparison rule appears twice ('Do not pretend separate image views are side-by-side comparison' and again in step 2 as 'Capturing screenshots is not enough...'), and wordy lines like 'It is incredibly important to check fonts carefully for fidelity, including looking up similar typefaces...' could be trimmed. This matches anchor 3 ('mostly efficient but includes some unnecessary explanation or could be tightened') — above anchor 2 because there is no tutorial-style explanation of concepts Claude already knows, below anchor 4 because several duplicated passages earn no new information. | 3 / 5 |
Actionability | Concrete, executable guidance throughout: a copy-paste report template with per-finding fields (Location/Evidence/Impact/Fix), an explicit design-qa.md contents checklist, exact severity definitions P0–P3, and a hard binary gate ('final result must be exactly passed or blocked'). Anchor 5 is not fully met because a few steps stay high-level — e.g. 'use design context and screenshot tools when available' and 'follow the Browser Choice rule' defer specifics without inline examples — so anchor 4 ('mostly executable guidance... minor gaps') is the closest fit. | 4 / 5 |
Workflow Clarity | The six-step workflow is clearly sequenced with an explicit feedback loop (record finding → apply fix → recapture at same viewport/state → recompare), a blocked/pass gating rule, and validation checkpoints ('If either artifact cannot be opened... write design-qa.md with final result: blocked'; 'Do not say a design matches... until the required fidelity surfaces have been checked'). This matches anchor 5: clear sequence, explicit validation, feedback loops, and a final checklist. | 5 / 5 |
Progressive Disclosure | Structure is good: the deep checklists live in the real, one-level-deep references/qa-rubric.md, which is clearly signaled ('Read qa-rubric when the QA pass spans more than a quick visual check'), and cross-skill pointers (audit, critical-overrides, index#browser-choice) are linked inline. It falls short of anchor 5 because the 'Required Fidelity Surfaces' section inlines detailed per-surface checklists (fonts, colors, imagery, including the long fail-QA asset-substitution rule) that substantially duplicate references/qa-rubric.md content and could largely live in the reference file. | 4 / 5 |
Total | 16 / 20 Passed |