CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-tester

Validate, test, and score the quality of skills within the claude-skills ecosystem. Comprehensive meta-skill: structure validation, Python script testing (syntax + imports + runtime + output format), multi-dimensional quality scoring with letter grades and tier classification (BASIC/STANDARD/POWERFUL). Use when authoring a new skill, auditing existing skills for tier promotion, setting up pre-commit hooks for skill quality, or integrating skill QA into CI.

72

Quality

90%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

The risk profile of this skill

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is highly actionable with clear workflow checkpoints and well-organized references to bundle files. Slight conciseness padding and un-linked reference paths are the only minor gaps.

DimensionReasoningScore

Conciseness

Mostly lean — quick-start commands, tool-check bullets, tier table, and troubleshooting earn their tokens; a few minor instances of over-explanation (e.g. 'stdlib-only is the repo policy') could be trimmed.

4 / 5

Actionability

Provides exact, copy-paste-ready commands with full paths and flags, documented JSON outputs, and a concrete CI yaml snippet covering the common cases.

5 / 5

Workflow Clarity

The 'Verification Loop' gives a clear three-step pass sequence with explicit criteria and a feedback loop (apply the top roadmap item and re-run all three — never report a partial pass).

5 / 5

Progressive Disclosure

Body is an overview that points to real references/ (structure spec, tier matrix, scoring rubric) and scripts/ one level deep, but individual reference paths are not explicitly linked, leaving minor organization gaps.

4 / 5

Total

18

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, complete, and distinctive, clearly stating both capabilities and trigger contexts. Minor room to broaden trigger synonyms, but it is otherwise strong.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — 'structure validation', 'Python script testing (syntax + imports + runtime + output format)', 'multi-dimensional quality scoring with letter grades and tier classification' — giving comprehensive coverage of capabilities.

5 / 5

Completeness

Explicitly answers both 'what' (validate/test/score with detail) and 'when' ('Use when authoring a new skill, auditing... pre-commit hooks... CI') with concrete trigger phrases.

5 / 5

Trigger Term Quality

Natural triggers like 'authoring a new skill', 'auditing existing skills for tier promotion', 'pre-commit hooks for skill quality', and 'integrating skill QA into CI' are strong, though a few common synonyms/variations are absent.

4 / 5

Distinctiveness Conflict Risk

Occupies a clear niche (skill QA / tier-promotion meta-skill) with distinct triggers and minimal overlap risk with other skills.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

referenced_paths_exist

Referenced path issues: 1 missing

Warning

Total

15

/

16

Passed

Repository
alirezarezvani/claude-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.