CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-creator

Create, improve, and evaluate agent skills (SKILL.md plus reference files). Use this skill whenever the user wants to build, scaffold, or design a new skill, improve or fix an existing skill that isn't working well, score or benchmark a skill's quality or run evals on it, or turn a repeated manual workflow into a skill ("I keep doing X manually", "can you remember how to do X", "turn this into a skill").

76

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered meta-skill: clean mode routing, decision tables instead of prose, concrete numeric constraints, explicit validation gates, and a textbook progressive-disclosure split verified against real, substantive reference files. The only weakness is redundancy — the Step 5 checklist restates Step 3's rules and the rubric scorecard is demonstrated twice — which keeps conciseness at anchor 4 rather than 5.

Suggestions

Collapse the Step 5 Quick Checklist to reference the Step 3 Key Rules (or vice versa) — the 17 checklist items largely restate the 7 key rules and could be replaced by a short delta-list of items not already covered.

Merge the two rubric-scorecard demonstrations: keep the full scorecard format only in Step 7c and have Step 6b link to it instead of repeating a second illustrative score table.

The Step 7 'Reference Skills' table and the pattern examples in Step 2 overlap in purpose; consider moving the reference-skill annotations to `references/skill-examples.md` to shave lines from the body.

DimensionReasoningScore

Conciseness

The body is dense and table-driven with almost no explanation of concepts Claude already knows, but there is real redundancy that could be trimmed: the Step 5 Quick Checklist largely restates the Step 3 Key Rules (17 checklist items vs 7 rules), and the rubric-scorecard presentation appears twice in near-identical form (Step 6b's illustrative table and Step 7c's full scorecard). This fits anchor 4 ("efficient; minor instances of over-explanation that could be trimmed") better than anchor 5's every-token-earns-its-place.

4 / 5

Actionability

As an instruction-only skill its guidance is fully actionable: a mode-routing table with jump targets, a requirements-gathering table, pattern-selection table, concrete detection probes ("command -v tool", "tool auth status", "echo $API_KEY", "curl -s endpoint"), explicit numeric constraints (name max 64 chars, description max 1024 chars, reference sizes 50-150/150-400/400-900 lines), and per-mode deliverables in Step 8. All detail delegated to the six references is real and substantive, satisfying the rubric's note that instruction-only skills are not penalized for lacking code.

5 / 5

Workflow Clarity

Steps 1-8 are clearly sequenced with a routing table for the three modes and an explicit ask-if-ambiguous rule. Validation is explicit: Step 5 is a quality gate ("If any item fails, fix it before delivering"), Step 6d requires user approval before editing files, and Step 7 defines the scoring procedure — matching anchor 5 ("explicit validation steps; feedback loops... checklists for complex processes").

5 / 5

Progressive Disclosure

The ~277-line body stays an overview: workflow, routing tables, and key rules inline, with all deep material in references/, and every one of the six referenced files exists and is substantive (173-424 lines). Navigation is well signaled — backtick path pointers at each step plus a terminal "## Reference Files" section with one-line descriptions — and references are one level deep (cross-pointers between references are sibling navigation, not nested chains). This matches anchor 5 ("clear overview with well-signaled one-level-deep references; content appropriately split; easy navigation").

5 / 5

Total

19

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary routing description: it names the artifact and the three lifecycle actions in the first sentence, then enumerates intent categories with natural user phrasings, including directly quoted user language for the workflow-to-skill case. Third person, no fluff, no over-claims. Only minor overlap risk with adjacent prompt/agent-building requests keeps it from a perfect mark.

DimensionReasoningScore

Specificity

The description states three concrete lifecycle actions — "Create, improve, and evaluate agent skills (SKILL.md plus reference files)" — and elaborates each with specific verbs (build, scaffold, design, fix, score, benchmark, run evals). This matches the top anchor ("multiple specific concrete actions; comprehensive coverage") and is more comprehensive than the anchor-4 example, which admits gaps in coverage.

5 / 5

Completeness

The "what" is explicit in the first sentence and the "when" is explicit via "Use this skill whenever the user wants to..." with concrete trigger phrases for each intent category — exactly the anchor-5 pattern ("clearly and explicitly answers both what AND when with concrete trigger phrases"). The 'Use when' clause is present, so the completeness cap of 3 does not apply, and the voice is third person ("Create, improve, and evaluate"), so no specificity penalty applies.

5 / 5

Trigger Term Quality

It covers the full intent vocabulary users would naturally say: "build, scaffold, or design a new skill", "improve or fix an existing skill that isn't working well", "score or benchmark a skill's quality or run evals on it", plus direct quotes of user phrasings ("I keep doing X manually", "can you remember how to do X", "turn this into a skill"). Quoted user language plus synonyms across all three intent categories matches the comprehensive-coverage anchor; it is well above anchor 4 ("a few natural terms missing").

5 / 5

Distinctiveness Conflict Risk

The domain vocabulary ("agent skills", "SKILL.md", "run evals on it") forms a clear niche with distinct triggers, but requests about neighboring work (e.g., general prompt engineering or agent/assistant building) could plausibly route here — fitting anchor 4 ("mostly distinct; minor overlap risk with closely related skills") rather than anchor 5's "minimal conflict risk" file-type-specific niche.

4 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
himself65/finance-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.