CtrlK
BlogDocsLog inGet started
Tessl Logo

create-skill

Scaffolds, reviews, upgrades, or diagnoses agent skills against best-practice frontmatter, progressive disclosure, token-aware structure, and the agent-skills.git symlink + inventory wiring. Modes: `scaffold` (default — new skill), `review` (audit existing skill), `upgrade` (split a single-file skill into multi-file), `diagnose` (retrospective failure analysis that emits a confidence-gated unified diff against any skill declaring a diagnostic surface). Triggers on "create a skill", "scaffold a skill", "new SKILL.md", "review this skill", "audit my skill", "upgrade this skill", "split this skill", "diagnose this skill", "why did the skill miss this", "/create-skill".

72

Quality

91%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered meta-skill body: gated seven-phase workflow, concrete commands and output templates, and a genuine on-demand reading table. The main defects are duplicated content (validator invocation, category lists, diagnose detail) and — more seriously — a bundle that omits nearly all of the rules/ and templates/ files the body relies on, breaking the progressive-disclosure index it advertises.

Suggestions

Ship the referenced bundle files: the body points to ~17 rules/*.md and several templates/*.md (frontmatter.md, quality-checklist.md, diagnose-mode.md, templates/SKILL.minimal.md, templates/evals.json, ...) that are absent — either include them or prune the pointers to files that actually exist.

De-duplicate the body: state the validator command and the category list once (Phase 5 / Phase 4) and reference them elsewhere instead of repeating them in the Review Workflow and Definition of Done.

Compress the Diagnose Workflow section to a short summary plus the pointer to rules/diagnose-mode.md, since it currently re-describes invocation flags, the diagnostic-surface contract, fallback behavior, and the self-improvement loop that the referenced rule file already covers.

DimensionReasoningScore

Conciseness

The body is mostly efficient — tables, one-liners, and pointers ("This SKILL.md is a thin index ... Reading them all up-front would burn tokens you do not need yet") — but has trimmable redundancy: the category list ("workflow, quality, delivery, testing, design, analysis, authoring") appears in both Phase 1 and Phase 4; `node ${CLAUDE_SKILL_DIR}/scripts/validate-skill.mjs <dir>` is restated in Phase 5, the Review Workflow, and the Definition of Done; and the ~35-line Diagnose Workflow section re-describes invocation, surface contract, fallback, and self-improvement loops after saying "The full procedure ... live in rules/diagnose-mode.md". This is 'efficient; minor instances of over-explanation that could be trimmed' rather than the score-3 anchor's 'some unnecessary explanation', since nothing explains concepts Claude already knows.

4 / 5

Actionability

Guidance is copy-paste ready throughout: exact commands ("bash scripts/sync-symlinks.sh", "node ${CLAUDE_SKILL_DIR}/scripts/validate-skill.mjs <dir> [--portable]", "Verify both hops with readlink"), a literal output template ("Mode: scaffold / Target: skills/<category>/<proposed-name>/"), a concrete FAIL example with error text, a 10-item batched interview, decision tables for structure and mode detection, and per-file authoring specs. Not the score-4 anchor, which tolerates 'minor gaps' — the common cases (scaffold and review) are fully covered with executable steps.

5 / 5

Workflow Clarity

The scaffold workflow is a seven-phase pipeline with an explicit gate per phase ("do not proceed until it passes"), an explicit validation loop ("Treat any unchecked item as a defect — fix it before declaring the skill done"; "On failure:" with a concrete FAIL transcript), user-confirmation checkpoints (Phase 0 answers confirmed "verbatim", upgrade layout "show it to the user for approval before writing", "--apply honored only at ≥ 90 %"), and Definition of Done checklists. This matches the score-5 anchor: clear sequence, explicit validation steps, feedback loops, and checklists.

5 / 5

Progressive Disclosure

On paper the body is a model thin index — a 'Required Reading by Phase' table, optional references clearly flagged ("load only when the user asks"), and the two files that DO exist (references/skill-archetypes.md, references/good-vs-bad-examples.md) plus scripts/validate-skill.mjs are real and correctly linked. But scored against the actual bundle: none of the ~17 referenced rules/*.md files (frontmatter.md, quality-checklist.md, diagnose-mode.md, ...) nor any templates/*.md files (SKILL.minimal.md, evals.json, ...) exist in the bundle, so the majority of navigation pointers break. That is beyond the score-4 anchor's 'minor organization gaps', while the strong in-body signaling and the valid references keep it above the score-3 anchor's 'references present but not clearly signaled'... it sits at 3-to-4 with the missing-file problem pulling it to 3.

3 / 5

Total

17

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: third-person, front-loaded, enumerates all four modes with their concrete behavior, states both what and when, and closes with ten natural trigger phrases including synonyms and the slash command. No fluff or over-claims; comfortably within the 1024-char budget.

DimensionReasoningScore

Specificity

The description enumerates four concrete capabilities ("Scaffolds, reviews, upgrades, or diagnoses agent skills") and then concretizes each mode: `review` ("audit existing skill"), `upgrade` ("split a single-file skill into multi-file"), `diagnose` ("emits a confidence-gated unified diff against any skill declaring a diagnostic surface"). Coverage of what the skill does is comprehensive with no vague filler, matching the anchor 'Lists multiple specific concrete actions; comprehensive coverage' rather than the score-4 anchor, which requires minor gaps.

5 / 5

Completeness

Both 'what' and 'when' are explicit: what — "Scaffolds, reviews, upgrades, or diagnoses agent skills against best-practice frontmatter, progressive disclosure, token-aware structure, and the agent-skills.git symlink + inventory wiring"; when — an explicit "Triggers on ..." clause listing concrete trigger phrases. This is the exact pattern of the score-5 anchor example; the 'when' is neither missing (score 3) nor merely present-but-generic (score 4).

5 / 5

Trigger Term Quality

"Triggers on 'create a skill', 'scaffold a skill', 'new SKILL.md', 'review this skill', 'audit my skill', 'upgrade this skill', 'split this skill', 'diagnose this skill', 'why did the skill miss this', '/create-skill'" — ten natural phrases a user would actually say, including synonyms (create/scaffold/new, review/audit) and the slash command. This matches the anchor 'Comprehensive coverage of natural terms including synonyms'; the score-4 anchor requires 'a few natural terms missing', which is not the case.

5 / 5

Distinctiveness Conflict Risk

The niche is clear (meta skill-authoring: scaffolding/auditing/upgrading other agent skills) and the triggers are keyed to that niche ("diagnose this skill", "why did the skill miss this", "split this skill"), so collision with unrelated skills is minimal. It is fully third-person with no over-claims. The score-4 anchor ('minor overlap risk with closely related skills') would apply if a trigger like bare 'review' were used, but every trigger is skill-scoped.

5 / 5

Total

20

/

20

Passed

Validation

75%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 12 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

relative_links

Relative link issues: 4 missing

Warning

referenced_paths_exist

Referenced path issues: 1 missing

Warning

Total

12

/

16

Passed

Repository
mthines/agent-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.