CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-quality-reviewer

This skill should be used when the user asks to "analyze skill quality", "evaluate this skill", "review skill quality", "check my skill", or "generate quality report". Evaluates local skills across description quality, content organization, writing style, and structural integrity.

60

Quality

71%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

High

Do not use without reviewing

Fix and improve this skill with Tessl

tessl review fix ./skills/skill-quality-reviewer/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

56%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The workflow itself is well-sequenced and actionable with concrete templates and validation checks, but the body carries its own reference manual: scoring tables, a 14-row grade map, benchmarks, restated red flags, and usage examples inflate it to ~1,700 words and duplicate references/scoring-criteria.md. Moving that material out and adding explicit error-recovery and script-invocation examples would lift the weakest dimensions.

Suggestions

Move the per-dimension scoring breakdown tables (Steps 3–6), the letter-grade mapping, and the "Grade Benchmarks" section into references/scoring-criteria.md, keeping only the weights and a pointer in SKILL.md — they currently duplicate that file.

Trim the "Common Quality Issues" and "Usage Examples" sections, which restate the red flags and workflow already given, and surface the unused references/batch-review-template.md from the batch-portfolio mode section.

Replace the four Usage Examples with one compact example, and add an explicit error-recovery branch to Step 1 (what to do when the path is missing or SKILL.md is absent) plus a sample invocation of scripts/skill-audit.py.

DimensionReasoningScore

Conciseness

The ~1,700-word body inlines several padded sections that duplicate material Claude can derive or that lives in references: per-criterion scoring tables (Steps 3–6), a 14-row letter-grade table, a "Grade Benchmarks" section restating it, a "Common Quality Issues" section repeating the red flags already given in Steps 3–5, and four "Usage Examples" that mostly restate the workflow. This is noticeably verbose rather than just 'some' tightening — matching the 2 anchor, and above it sits the reference-heavy duplication that defines the gap.

2 / 5

Actionability

The 8-step workflow gives concrete, executable guidance: validation checklists per step, a worked bash example, an explicit weighted-score formula, exact output filenames ("quality-report-{skill-name}.md"), and full markdown templates. Not a 5 because the bundled scripts (e.g. scripts/skill-audit.py) are listed but never shown how to invoke, leaving a minor gap.

4 / 5

Workflow Clarity

Steps 1–8 are clearly sequenced with most checkpoints present (Step 1 path/validity validation, Step 2 field checks, per-step Check bullets). Not a 5: there is no explicit error-recovery loop (e.g. what to do when SKILL.md is missing or YAML is invalid) and no post-generation verification of the two report files, so feedback loops are implicit rather than explicit.

4 / 5

Progressive Disclosure

References are real and clearly signaled one level deep ("Reference: references/scoring-criteria.md", plus the Additional Resources section; all listed files exist). However the body inlines substantial content that belongs in those files — the scoring breakdown tables and grade-mapping duplicate references/scoring-criteria.md, and the output templates could be a reference file — and references/batch-review-template.md is never surfaced. That 'content that should be separate is inline' pattern matches the 3 anchor better than the 4 anchor's 'minor organization gaps'.

3 / 5

Total

13

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person voice, explicit 'should be used when' trigger clause, five concrete user phrases, and a clear statement of the four evaluation dimensions. The only improvement space is mentioning the outputs it produces (scores, grades, improvement plans) and a few more trigger synonyms.

DimensionReasoningScore

Specificity

"Evaluates local skills across description quality, content organization, writing style, and structural integrity" names the domain and enumerates the four concrete evaluation dimensions. It falls short of the 5 anchor because capabilities the skill actually performs (weighted scoring, letter grades, improvement plans) are absent, leaving minor gaps in coverage.

4 / 5

Completeness

It explicitly answers both parts: what ("Evaluates local skills across description quality, content organization, writing style, and structural integrity") and when ("This skill should be used when the user asks to...") with concrete trigger phrases — a direct match to the top anchor. The 'Use when...' guidance is explicit, so the completeness cap does not apply.

5 / 5

Trigger Term Quality

Five quoted natural phrases — "analyze skill quality", "evaluate this skill", "review skill quality", "check my skill", "generate quality report" — are phrasings a user would plausibly say. Not a 5: "skill quality" is repeated across phrases and common variations like "audit my skills" or "grade this skill" are missing.

4 / 5

Distinctiveness Conflict Risk

The meta-niche of evaluating skills themselves is distinct, and every trigger phrase contains "skill", keeping triggers in-niche with minimal conflict risk against code-review or general review skills — matching the clear-niche anchor.

5 / 5

Total

18

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
Galaxy-Dawn/claude-scholar
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.