CtrlK
BlogDocsLog inGet started
Tessl Logo

critique-theater

Five-dimension design quality review — score the artifact against craft, brand, accessibility, and copy, then fix what falls short before handing it over.

59

Quality

69%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/_official/atoms/critique-theater/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured, concise, instruction-only skill with a clear review-fix-iterate workflow, explicit stop criteria, and strong anti-drift guardrails (no fake transcripts, no critique.json). Its only weaknesses are minor: a little plugin meta-commentary that could be trimmed and no concrete numeric thresholds or example defects for the accessibility/craft checks.

DimensionReasoningScore

Conciseness

The body is lean and assumes Claude's competence — it never explains what design, contrast, or typography are — and every section carries instruction. A few sentences could be trimmed (the OD plugin/scheduler meta-paragraph and the enumerated anti-patterns), keeping it just below anchor 5's 'every token earns its place'.

4 / 5

Actionability

For an instruction-only skill the guidance is concrete: five named dimensions each with concrete checks ('contrast ratios, touch-target sizes, focus rings, labels on interactive controls and dialogs, keyboard reachability, sensible heading order'), 'fix it in the same turn', a 'three passes is a sensible ceiling' stop rule, and explicit output instructions. It falls short of anchor 5 only in lacking measurable acceptance criteria (e.g. specific contrast ratios) or example defects.

4 / 5

Workflow Clarity

The workflow is a clearly sequenced loop with explicit validation and stop conditions: review five dimensions → record and fix each defect in the same turn → 'Keep iterating while defects remain' → 'Stop when the artifact clears all five dimensions, or when further rounds stop producing real improvements', with a stated ceiling. This matches the feedback-loop-with-checkpoints anchor, and the non-destructive context means no batch-validation cap applies.

5 / 5

Progressive Disclosure

The skill is under 50 lines, needs no external references (none exist in the bundle), and is cleanly organized under three well-labeled sections (The five dimensions, What to do with the findings, How to report it). This meets the rubric's simple-skill exception for a full score on well-organized sections alone.

5 / 5

Total

18

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description communicates a clear, specific purpose with an action verb sequence (score, fix) and named dimensions, but it is internally inconsistent ('Five-dimension' vs four listed) and provides no trigger guidance for when to use the skill. Adding a 'Use when...' clause and the missing fifth dimension would lift it substantially.

Suggestions

Add an explicit trigger clause, e.g. 'Use when handing over a designed artifact (UI, HTML, image, or document) and the user expects polished, on-brand output.'

Fix the internal inconsistency: either list all five dimensions (craft, critique/problem-fit, brand, accessibility, copy) or say 'multi-dimension'.

Include natural user phrasings as trigger terms ('design review', 'polish this', 'does this look right', 'check contrast/accessibility') so the skill fires on real requests.

DimensionReasoningScore

Specificity

The description names the domain ('design quality review') and two concrete actions — 'score the artifact against craft, brand, accessibility, and copy, then fix what falls short' — but claims 'Five-dimension' while enumerating only four, so coverage is incomplete. This matches anchor 3 ('1-2 concrete actions, but not comprehensive'); it is above anchor 2 because the actions and dimensions are specific, not generic.

3 / 5

Completeness

The 'what' is clear (score against five dimensions, then fix shortfalls), but the 'when' is entirely missing — there is no 'Use when...' clause or equivalent trigger guidance, which caps completeness at 3 per the judging guidelines. Not score 2 because the 'what' half is concrete and unambiguous.

3 / 5

Trigger Term Quality

Relevant keywords are present ('design', 'quality', 'review', 'craft', 'brand', 'accessibility', 'copy'), but there are no natural trigger phrases a user would say and no synonyms or variations (e.g. 'design review', 'check the design', 'polish the artifact'). Fits anchor 3 ('some relevant keywords but missing common variations or synonyms').

3 / 5

Distinctiveness Conflict Risk

The niche — reviewing a produced artifact for craft, brand-token adherence, accessibility, and copy — is mostly distinct, with distinctive vocabulary. Minor overlap risk remains with generic code-review or design skills that could also fire on 'review' requests; not score 5 because no explicit trigger boundary separates it from those.

4 / 5

Total

13

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
nexu-io/open-design
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.