CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-extract

Reverse-engineer design systems, tokens, and components from live products or screenshots

44

Quality

44%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./skills/skill-extract/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

36%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body functions as a design/specification document for the skill rather than an operational guide: it catalogs capabilities, algorithms, and output formats, but provides no executable steps, commands, or code Claude could actually run, and mixes in project-management material (roadmap, contributing phases, PRD references). Structure is decent, but detail is inlined that belongs in reference files, and the one cross-reference points to a nonexistent file.

Suggestions

Replace the descriptive pipeline sections with a sequenced, executable workflow (detect inputs → extract tokens → analyze components → generate outputs) with explicit validation checkpoints tied to the quality gates, plus fix-and-retry guidance for the listed error codes.

Move implementation-status, contributing/roadmap, and research-sources material out of SKILL.md — none of it helps Claude execute the skill — and drop the time-sensitive "(2026)" reference.

Split the output-structure tree, error-code catalog, and architecture/API heuristics into one-level-deep reference files under references/ (with clear links from SKILL.md), and either ship `skills/blocks/codex-host-adapter.md` or remove the dangling reference to it.

DimensionReasoningScore

Conciseness

The body is mostly dense, structured material (tables, lists, pipelines) rather than explanations of known concepts, but it is padded with developer-project content that earns no tokens at runtime — "Implementation Status" checklists, week-by-week "Contributing" phases, "Research Sources", "Modern reverse-engineering practices (2026)" (time-sensitive), and a PRD reference. This fits anchor 3 (mostly efficient with some unnecessary sections) rather than 4, where only minor trimming would be needed.

3 / 5

Actionability

The body describes what the skill does ("AST parsing for TypeScript/JavaScript", "Uses CIEDE2000", K-means clustering, service-boundary heuristics) but never instructs Claude how to execute any of it — the only commands are invocations of the skill itself (`/octo:extract ./my-app --mode design`), and no runnable code or concrete steps appear. This matches anchor 2 (high-level hints, missing the specific steps to execute); it is not a 3 because there is no even partially executable guidance for performing an extraction.

2 / 5

Workflow Clarity

No extraction workflow is sequenced: there is no step order, no validation checkpoints wired into a process, and no fix-and-retry loop — error codes (`ERR-001`..`VAL-004`) and quality gates are cataloged but no handling guidance connects them. Quality gates do exist as a validation concept, so this sits between anchor 1 (steps missing, no validation) and anchor 3 (steps listed with checkpoint gaps) — anchor 2's "rough sequence... validation absent" fits best since the usage examples imply an order but define no operational steps.

2 / 5

Progressive Disclosure

The single SKILL.md is well-sectioned but monolithic: content that clearly belongs in separate reference files (the 40-line output-directory tree, error-code catalog, architecture heuristics, performance tables) is inlined, and no references/, scripts/, or assets/ directories exist. The one referenced file, `skills/blocks/codex-host-adapter.md`, is not present in the bundle. Anchor 3 (some structure, content that should be separate is inline) fits; it is not a 2 because section headers make it navigable.

3 / 5

Total

10

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise and names a concrete, fairly distinctive capability, but it omits any "when to use" trigger guidance and lacks natural trigger-term variations. It reads as a capability label rather than a complete skill description.

Suggestions

Add an explicit trigger clause, e.g., "Use when the user mentions reverse-engineering, extracting, or auditing a design system, design tokens, or component library from a website, screenshot, or codebase."

Enumerate 1-2 more concrete actions (e.g., "extract colors, typography, and spacing tokens; catalog component props and variants") to strengthen specificity.

Include natural synonyms and formats users would say — "design tokens", "style guide", "audit a UI", ".css", "Tailwind config" — to broaden trigger-term coverage.

DimensionReasoningScore

Specificity

The description names the domain ("design systems, tokens, and components") and the concrete action "reverse-engineer" with sources ("live products or screenshots"), but stops short of listing several specific actions like extracting colors, typography, or spacing. This matches anchor 3 (domain plus 1-2 concrete actions), not anchor 4 which requires several specific actions with only minor gaps.

3 / 5

Completeness

The "what" is clear (reverse-engineer design systems, tokens, and components), but there is no "Use when..." clause or any equivalent explicit trigger guidance, capping completeness at 3 per the judging guidelines. It is not a 2 because the "what" half is concrete and specific, not vague.

3 / 5

Trigger Term Quality

Relevant keywords are present ("design systems", "tokens", "components", "screenshots") but common variations users would naturally say are missing — e.g., "extract", "audit", "style guide", "design tokens from a website", or file formats like CSS. This fits anchor 3 (some relevant keywords, missing variations/synonyms) rather than anchor 4's good coverage.

3 / 5

Distinctiveness Conflict Risk

"Reverse-engineer design systems... from live products or screenshots" carves a mostly distinct niche with clear triggers, though there is minor overlap risk with general design/branding or code-analysis skills. Anchor 4 fits better than 5, which requires a fully distinct niche with minimal conflict risk.

4 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.