CtrlK
BlogDocsLog inGet started
Tessl Logo

sketch

Throwaway HTML mockups: 2-3 design variants to compare.

57

Quality

67%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./optional-skills/creative/sketch/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong, highly actionable skill body: the workflow is explicitly sequenced with a real validation-and-fix loop, and every stage comes with concrete templates, code, or tool syntax. Its weaknesses are minor — some trimmable secondary sections, duplicated attribution notes, and the assumption that Hermes browser tools are available.

DimensionReasoningScore

Conciseness

The body is dense with actionable guidance — variant axes, a directory template, a CSS reset, a README template, a comparison-table template, and a tool sequence — and it never explains concepts Claude already knows (no 'what HTML is', no library tutorials). A few sections could be trimmed (the duplicated GSD attribution/archival notes appear twice, and 'frontier mode' plus theming are secondary), which keeps it at 'Efficient; minor instances of over-explanation that could be trimmed' rather than a 5.

4 / 5

Actionability

Guidance is largely executable: exact tool-call syntax ('browser_navigate(url="file:///...")', 'browser_vision(question=...)'), a copy-paste CSS reset, a tokens.css example, concrete file layout ('NNN-stance-name/index.html + README.md'), and platform-specific open commands. It falls short of a 5 only in minor gaps: no complete sample variant HTML is provided, and the Hermes browser tools are assumed present without a fallback when they aren't — matching 'Mostly executable guidance; concrete code or commands with minor gaps'.

4 / 5

Workflow Clarity

The pipeline is explicit ('intake → variants → head-to-head → pick winner (or iterate)') with numbered, well-defined stages, and it includes a genuine validation feedback loop: visually verify each variant with browser_vision, 'Fix and re-navigate until each variant looks right', plus 'Open it in a browser. If it looks broken, fix it before showing the user.' This matches the top anchor — clear sequence, explicit validation steps, and an error-recovery loop; nothing in the workflow is destructive or batch, so no cap applies.

5 / 5

Progressive Disclosure

The skill is a single self-contained file with no bundle directories (no references/, scripts/, or assets/ exist), so all guidance is inline by design; section headers are clear and each section is scannable and independently useful. It fits 'Good structure; most content is appropriately placed; minor organization gaps' — the secondary sections (theming, frontier mode, attribution) could arguably live in reference files, and the ~220-line body exceeds the 'under 50 lines' simple-skill exception that would allow a 5 on organization alone.

4 / 5

Total

17

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise and its 'what' is concrete, but it answers only half the required question — there is no explicit 'when to use' guidance, and the natural trigger phrases that appear in the body (sketch, wireframe, show me variants) are absent from the description itself. It sits just above the midpoint: distinctive enough to avoid most conflicts, but incomplete as a trigger surface.

Suggestions

Append an explicit 'when' clause with natural trigger phrases, e.g. 'Use when the user says sketch/mockup this screen, wants 2-3 takes on a UI, or wants to compare layout directions before committing' — this directly lifts completeness and trigger-term quality.

Include the common synonyms users actually say (sketch, wireframe, prototype, takes on a UI) so the description triggers on natural language, not just the terms 'mockups' and 'variants'.

State the boundary in one clause ('not for production or polished one-off pages') to reduce overlap with general design/build skills and push distinctiveness toward 5.

DimensionReasoningScore

Specificity

The description names its domain ("Throwaway HTML mockups") and one concrete action ("2-3 design variants to compare"), but stops there — it lists a single capability rather than several specific actions, matching the anchor 'Names domain and 1-2 concrete actions, but not comprehensive'. It is not a 4 because it does not enumerate several actions, and not a 2 because 'mockups' and 'variants to compare' are concrete, not generic.

3 / 5

Completeness

The 'what' is clear ("Throwaway HTML mockups: 2-3 design variants to compare") but there is no 'when' — no 'Use when...' clause or equivalent trigger guidance, which per the judging guidelines caps completeness at 3. Not a 4 because 'when' is entirely absent rather than merely imprecise; not a 2 because the 'what' half is concrete and unambiguous.

3 / 5

Trigger Term Quality

It includes relevant keywords ("HTML mockups", "design variants", "compare") but misses the natural phrases a user would actually say — 'sketch this screen', 'wireframe', 'prototype', 'mockup this UI' — which the body itself lists as triggers yet the description omits. This fits 'Some relevant keywords but missing common variations or synonyms'; it is above a 2 (not merely generic) but below a 4 (no natural trigger phrasing present).

3 / 5

Distinctiveness Conflict Risk

'Throwaway' plus '2-3 design variants to compare' carves out a clear niche — disposable comparison mockups, explicitly not production code — with mostly distinct triggers. Minor overlap risk remains with related design/prototype skills (claude-design, excalidraw-style diagramming) since the description alone doesn't state those boundaries. This fits 'Mostly distinct; minor overlap risk with closely related skills'; not a 5 because the exclusion of production/one-off design work is only stated in the body, not the description.

4 / 5

Total

13

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

13

/

16

Passed

Repository
NousResearch/hermes-agent
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.