CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-validator

Ensure Claude Code skills meet quality standards through validation operations for structure, content, patterns, and production readiness. Task-based validation with pass/fail criteria, automated checks, and compliance reporting. Use when validating skills before deployment, ensuring standards compliance, certifying production readiness, or quality gating skill releases.

63

Quality

79%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/skill-validator/SKILL.md

The canonical home for this skill is skill-validator in fernandezbaptiste/Skrillz

SKILL.md
Quality
Evals
Security

Quality

Content

66%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body delivers a genuinely clear, well-sequenced validation workflow with strong feedback loops and a concrete report template, but it is noticeably redundant — restating its pass criteria and its comparison to review-multi multiple times — which inflates token cost without adding guidance. Trimming the repetition would raise overall quality substantially.

Suggestions

Remove the "Minimum Standards" section and the closing paragraph, both of which restate the per-Operation pass criteria and the Overview almost verbatim; keep one authoritative statement of the criteria.

Consolidate the three separate "difference from review-multi" discussions (Overview, its own subsection, and "Integration with review-multi") into a single short comparison, and drop Best Practices that restate obvious behavior ("Validate Before Deploy", "Re-Validate After Fixes").

State the assumption or verification step for the external dependency "review-multi/scripts/validate-structure.py" (e.g., check the script exists and where to find it), since the only concrete command in the skill depends on it.

DimensionReasoningScore

Conciseness

The pass criteria are restated three times — once per Operation, again nearly verbatim in the "Minimum Standards" section, and summarized once more in the Quick Reference table — and the "difference from review-multi" point is made in three separate places. Best Practices entries like "Re-Validate After Fixes" and "Validate Before Deploy" state the obvious, making this noticeably padded rather than just loosely trimmed.

2 / 5

Actionability

Concrete guidance is present: a runnable command ("python3 review-multi/scripts/validate-structure.py <skill>"), a full copy-paste validation report template, and specific check instructions like "Count examples (look for ``` code blocks, minimum 3)". It falls short of fully executable because only one command is given and it assumes an external skill's script exists.

4 / 5

Workflow Clarity

Operations 1-4 are clearly sequenced, each has explicit pass criteria, Operation 4 composes the earlier ones, and error recovery is explicit via "Validate → Fix → Re-validate cycle" and the Deployment Decision Tree's "After fixes → Re-validate → Deploy if pass". Checklists and a final DEPLOY/HOLD decision match the top anchor.

5 / 5

Progressive Disclosure

As a single-file skill with no bundle files, the body is well-organized with clear section headers, a Quick Reference table, and a decision tree. It is not exemplary structure because the duplicated Minimum Standards content should have been consolidated, and the external reference to review-multi's script is not verifiable within this bundle.

4 / 5

Total

15

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with explicit what-and-when structure and concrete enumerated validation areas. Trigger coverage is good but could add the more casual phrasages users naturally say when they want a skill checked.

DimensionReasoningScore

Specificity

"validation operations for structure, content, patterns, and production readiness" plus "pass/fail criteria, automated checks, and compliance reporting" lists several specific actions with broad domain coverage, but the actions are all variants of "validate" and "meet quality standards" is mildly generic, keeping it just below the comprehensive 5 anchor.

4 / 5

Completeness

The description clearly answers both "what" (validation operations for structure, content, patterns, and production readiness; pass/fail criteria; automated checks; compliance reporting) and "when" via an explicit "Use when..." clause with four concrete trigger scenarios, matching the top anchor.

5 / 5

Trigger Term Quality

Triggers like "validating skills before deployment", "standards compliance", and "quality gating skill releases" give good keyword coverage, but common natural variations a user would actually say ("check skill quality", "review a skill", "is this skill ready") are missing.

4 / 5

Distinctiveness Conflict Risk

The pass/fail deployment-gating niche is mostly distinct with dedicated triggers, but there is minor overlap risk with scoring/review-type skills (the body itself references review-multi) and the description does not articulate the pass/fail-vs-scoring distinction that separates them.

4 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

Total

15

/

16

Passed

Repository
fernandezbaptiste/Skrillz
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.