CtrlK
BlogDocsLog inGet started
Tessl Logo

claude-design

Design one-off HTML artifacts (landing, deck, prototype).

56

Quality

67%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/creative/claude-design/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

77%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured, highly actionable design doctrine with an unusually strong workflow: an explicit sequence, tiered verification, and a genuine diagnose→repair→re-score feedback loop. Its main weakness is that everything lives inline in one large SKILL.md with no reference files, costing token budget on every invocation, plus minor redundancy between overlapping sections.

Suggestions

Move the low-frequency detail blocks into one-level-deep reference files (e.g. references/slop-diagnostic.md, references/deck-rules.md, references/format-standards.md) and keep short summaries plus explicit links in SKILL.md, so the ~650-line body shrinks to an overview.

Deduplicate overlapping guidance: the "When To Use" list restates the top decision table, and the "Portable Opening Prompt Pattern" paragraph re-summarizes the entire skill — both can be trimmed to a pointer or removed.

Make abstract composition guidance concrete with one- or two-line inline examples (e.g. what 'rhythm via scale and interruption' looks like in a section), mirroring how the Surface-First section already pairs rules with concrete instances.

DimensionReasoningScore

Conciseness

The body is directive and assumes competence (no explanation of what CSS or HTML is; every section issues instructions rather than tutorials), with only minor trimming opportunities — the "When To Use" list overlaps the top decision table, and the "Portable Opening Prompt Pattern" re-summarizes the whole skill. This fits 'efficient; minor instances of over-explanation that could be trimmed' rather than the fully lean anchor 5.

4 / 5

Actionability

Concrete thresholds and templates abound — "Default slide size: 1920×1080, 16:9", "Mobile hit targets should be at least 44px", "For print documents, text should be at least 12pt", exact versioning filenames ("Name v2.html"), a copy-paste final-response example, and a ten-tell diagnostic with per-tell repair mapping. Not anchor 5 because some guidance stays abstract (e.g. "Design with rhythm: scale, whitespace, density, alignment...") without examples of what rhythm looks like in practice.

4 / 5

Workflow Clarity

An explicit 8-step workflow (understand brief → gather context → commit to surface → define system → choose format → build → verify → report) includes an explicit Verify step with tiered minimums, and the Slop Diagnostic adds a true feedback loop: score out of 10, repair matched to the diagnosis, then "Re-score after repairing. Do not declare done while compositional tells (3, 8, 10) are still firing". This matches the anchor requiring explicit validation steps and feedback loops.

5 / 5

Progressive Disclosure

The skill has no bundle files at all: all ~650 lines are inline in SKILL.md under well-organized section headers. Long blocks that could live in one-level-deep reference files (the ten-tell Slop Diagnostic, Deck/Prototype rules, Typography/Color doctrine, hosted-tool remap list) are inlined, fitting 'content that should be separate is inline' — anchor 3, not 4, because there are no references to be clearly signaled and the overview-to-detail split the rubric expects is absent.

3 / 5

Total

16

/

20

Passed

Description

57%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise and grammatically clean with a clear 'what', but it omits any 'when to use' trigger guidance and under-sells the skill's scope, and it is only weakly distinguishable from sibling design skills. Adding an explicit use-when clause with richer trigger terms would raise most dimensions at once.

Suggestions

Add an explicit trigger clause, e.g. "Use when the user asks for a landing page, teaser page, slide deck, interactive prototype, mockup, or component exploration as a designed HTML artifact, outside the hosted Claude Design UI."

Include natural synonyms and related formats users actually say — "web page", "slides/presentation", "mockup", "UI", "design system preview" — so trigger matching covers common phrasings.

Add a distinguishing cue versus sibling skills (e.g. "for one-off designed artifacts, not token-spec files or brand-clone styling") to reduce overlap risk with other design skills.

DimensionReasoningScore

Specificity

"Design one-off HTML artifacts (landing, deck, prototype)" names the domain and one concrete action with three example artifact types, but does not list several specific actions or cover the skill's actual breadth (variants, tweaks panels, verification, repo implementations), matching the '1-2 concrete actions, not comprehensive' anchor rather than the 'several specific actions' anchor above.

3 / 5

Completeness

The 'what' is clear (design one-off HTML artifacts of the listed kinds), but there is no 'Use when...' clause or equivalent explicit trigger guidance; the parenthetical type list only weakly implies when, capping completeness at 3 per the rubric guideline.

3 / 5

Trigger Term Quality

"landing", "deck", "prototype", "HTML", and "design" are natural phrases users would say when needing this skill, but common synonyms like "web page", "slides", "presentation", "mockup", "UI", or "design system" are missing, fitting 'good keyword coverage; a few natural terms missing' rather than the comprehensive-synonyms anchor.

4 / 5

Distinctiveness Conflict Risk

"Design one-off HTML artifacts" is somewhat specific, but it provides no cue distinguishing it from other design/web/HTML skills in an ecosystem (e.g. token-spec or brand-clone design skills), so it could still trigger for the wrong design skill — anchor 3, not the 'minor overlap risk with closely related skills' of 4 since it doesn't carve out a clear niche.

3 / 5

Total

13

/

20

Passed

Validation

75%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 12 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (651 lines); consider splitting into references/ and linking

Warning

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

12

/

16

Passed

Repository
NousResearch/hermes-agent
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.