CtrlK
BlogDocsLog inGet started
Tessl Logo

release-notes

Generates polished GitHub release notes for a ToolHive release by analyzing every merged PR, cross-referencing linked issues, dispatching expert agents to assess breaking changes, and producing a formatted release body. Use when the user provides a GitHub release URL, tag name, or says "release notes".

69

Quality

87%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

77%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered workflow document with concrete commands, a sound triage priority order, and strong validation gates around a publish operation. Its main defects are the missing TEMPLATE.md bundle file that Phase 4 depends on, and minor redundancy across the Important Notes and Usage Examples sections.

Suggestions

Add the missing TEMPLATE.md to the skill bundle (referenced in Phase 4) — or, if the template is short, inline its section structure so the composition step is self-contained.

Trim redundancy: the Usage Examples section duplicates the Arguments examples, and 'Omit empty sections' / 'Read every PR body' in Important Notes restate guidance from Phase 4 and Phase 1 Step 4.

Consider moving the Phase 2 flagging heuristics and the Phase 3 agent-dispatch table into a references/ file to keep SKILL.md as a leaner overview, and verify every referenced path exists.

DimensionReasoningScore

Conciseness

The body is dense with earned content (exact gh commands, jq filters, priority-ordered classification signals, an agent dispatch table) with almost no explanation of concepts Claude already knows. It falls short of the score-5 'every token earns its place' anchor because of minor redundancy: the Usage Examples section repeats the Arguments block, and 'Omit empty sections' plus 'Read every PR body' in Important Notes restate guidance already given in Phase 4 and Step 4. It is well above score 3, which expects unnecessary explanation.

4 / 5

Actionability

Mostly executable throughout: copy-paste-ready gh/git commands with exact --jq filters, concrete classification criteria in a table, and a specific publish command. It does not reach score 5 because Phase 4 — the composition step that produces the actual deliverable — depends entirely on 'Read the template at [TEMPLATE.md](TEMPLATE.md)', a file that does not exist in the bundle, leaving the final assembly step without concrete guidance.

4 / 5

Workflow Clarity

Five clearly sequenced phases with explicit validation checkpoints: present the draft, surface uncertain breaking-change assessments, wait for user approval with explicit options, always save a reviewable file before publishing, and trust expert-agent verdicts over initial classification (a feedback loop). This matches the score-5 anchor (explicit validation steps, feedback loops, structured process for a complex workflow) and exceeds score 4's 'minor validation gaps'.

5 / 5

Progressive Disclosure

Sections are well organized and the single reference is clearly signaled and one level deep, but the only referenced file, [TEMPLATE.md](TEMPLATE.md), is absent from the bundle (no references/, scripts/, or assets/ directories exist), so the disclosure path dead-ends and the formatting detail it should carry is neither inline nor available. This fits the score-3 anchor ('could be better organized; references present but...') rather than score 4's 'references mostly clear', since a broken reference is a structural gap, not a minor organization one.

3 / 5

Total

16

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person, specific, and comprehensive, with an explicit 'Use when...' clause carrying concrete trigger phrases. The only gap is modest synonym coverage (e.g., 'changelog').

DimensionReasoningScore

Specificity

Phrases like "analyzing every merged PR, cross-referencing linked issues, dispatching expert agents to assess breaking changes, and producing a formatted release body" enumerate multiple concrete, specific actions covering the entire pipeline with no generic filler. This matches the anchor 'lists multiple specific concrete actions; comprehensive coverage' and clearly exceeds the score-4 anchor, which allows minor coverage gaps.

5 / 5

Completeness

The 'what' is explicit ("Generates polished GitHub release notes... producing a formatted release body") and the 'when' is explicit with concrete trigger phrases ("Use when the user provides a GitHub release URL, tag name, or says 'release notes'"), matching the anchor for clearly answering both what AND when. It is not score 4 because the when-clause is already fully explicit rather than needing more specificity.

5 / 5

Trigger Term Quality

"Use when the user provides a GitHub release URL, tag name, or says 'release notes'" gives good natural keyword coverage including the canonical phrase 'release notes'. It falls short of the score-5 anchor because natural synonyms like 'changelog', 'what's changed', or 'release announcement' are absent.

4 / 5

Distinctiveness Conflict Risk

"GitHub release notes for a ToolHive release" pins a clear niche with distinct triggers (release URL, tag name, 'release notes'), giving minimal conflict risk with other skills. It sits above the score-4 anchor's 'minor overlap risk with closely related skills' because the domain and triggers are narrowly scoped rather than broad.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 1 missing

Warning

Total

15

/

16

Passed

Repository
stacklok/toolhive
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.