CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-improve

Improve a skill via a test-fix-retest loop — static checks, targeted fixes, keep or revert on score change.

61

Quality

77%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/skill-improve/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong, highly actionable procedure: the test-fix-retest loop has explicit validation checkpoints, measured retests, and a careful revert path for the destructive write. The only weaknesses are minor redundancy in the NOT ASSESSED explanations and the absence of a brief overview of the loop up front.

DimensionReasoningScore

Conciseness

The body is lean and procedural with no explanation of concepts Claude already knows; the only padding is the NOT ASSESSED rationale explained twice (Phase 2 and Phase 6) and rationale sentences like '0 FAILs from a file nobody could read is not a clean result' that could be trimmed — anchor 4, not the every-token-earns-its-place lean of anchor 5.

4 / 5

Actionability

Concrete, executable guidance throughout: exact commands ('/skill-test static [name]'), copy-paste display templates, exact stop messages and user questions, and per-check diagnosis mappings. Minor gaps (e.g., category metric identification is shown only by example) keep it at anchor 4 rather than fully copy-paste-ready anchor 5.

4 / 5

Workflow Clarity

A clear 7-phase sequence that is itself an explicit validation feedback loop: baseline test → targeted fix → re-measured retest → keep or revert, with checkpoints for NOT ASSESSED results, a measured-not-stated retest rule, and a safe restore path for the destructive overwrite. Matches anchor 5 exactly.

5 / 5

Progressive Disclosure

No bundle files exist and none are needed; the single-file body is well organized with clear per-phase section headers, and external pointers (catalog.yaml, automation-modes.md, the yaml-helper hook) are one level deep and clearly signaled. Anchor 4 rather than 5 because the structure, while good, is plain sequential sections with no overview/summary orientation at the top.

4 / 5

Total

17

/

20

Passed

Description

66%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific and concrete, naming four distinct actions and a clear niche, but it lacks any 'use when' trigger guidance, which caps completeness and limits its discoverability. Trigger terms are good though not exhaustive.

Suggestions

Append an explicit trigger clause, e.g., 'Use when a skill fails /skill-test checks or you want to raise a skill's static/category test score.'

Mention the category-rubric dimension in the description so the 'what' matches the body's actual coverage (static AND category checks).

Add one or two natural synonyms users would say, such as 'skill quality' or 'make the skill pass its checks'.

DimensionReasoningScore

Specificity

Lists several concrete actions — 'test-fix-retest loop', 'static checks', 'targeted fixes', 'keep or revert on score change' — but omits the category-rubric checking that the body actually performs, so coverage has a minor gap (anchor 4) rather than being comprehensive (anchor 5).

4 / 5

Completeness

Has a clear 'what' (improve a skill via a test-fix-retest loop with static checks, targeted fixes, keep/revert on score change) but no 'Use when...' clause or equivalent explicit trigger guidance, so per the rubric guideline completeness is capped at 3.

3 / 5

Trigger Term Quality

Good natural keyword coverage: 'improve a skill', 'static checks', 'test', 'fix', 'retest', 'score'. A few common user phrasings are missing (e.g., 'skill quality', 'make the skill pass', 'audit'), which fits anchor 4 rather than the comprehensive synonym coverage of anchor 5.

4 / 5

Distinctiveness Conflict Risk

'Improve a skill via a test-fix-retest loop' carves a clear niche distinct from testing/auditing skills, with only minor overlap risk against a companion skill-test skill — anchor 4, not 5 since it never names its triggers explicitly.

4 / 5

Total

15

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
Donchitos/Claude-Code-Game-Studios
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.