CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-creator

Create or update GoClaw agent skills with eval-driven iteration. Use for new skills, skill scripts, references, benchmark optimization, description optimization, eval testing, extending agent capabilities.

64

Quality

76%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Critical

Do not install without reviewing

Fix and improve this skill with Tessl

tessl review fix ./skills/skill-creator/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, token-efficient meta-skill with concrete script commands and a well-sequenced eval-driven workflow. Its main weaknesses are bundle hygiene: several referenced paths (agents/, evals/, eval-viewer/) are missing from the bundle, about 11 of 26 reference files are orphaned with no navigation from SKILL.md, and validation is not explicitly wired into the creation sequence.

Suggestions

Fix the dangling references: either add the missing 'agents/grader.md', 'agents/comparator.md', 'agents/analyzer.md', 'evals/evals.json', and 'eval-viewer/generate_review.py' files or remove/redirect those pointers (generate_review.py also lives only in the scripts table, where its path is likewise wrong).

Add a short 'Further references' or index section linking the ~11 orphaned files (plugin-marketplace-*, mcp-skills-integration, troubleshooting-guide, writing-effective-instructions, testing-and-iteration, yaml-frontmatter-reference) so they are discoverable from SKILL.md.

Insert an explicit validation checkpoint into the Creation Workflow (e.g., 'Run scripts/quick_validate.py after writing SKILL.md; only publish when validation passes') and a brief error-recovery note for failed evals.

Include at least one concrete sample invocation (e.g., 'scripts/run_eval.py <args>') in the Eval & Testing section instead of deferring all execution detail to references.

DimensionReasoningScore

Conciseness

The body is lean and imperative throughout — tables ('Quick Reference', 'Scripts'), terse bullets ('Sacrifice grammar for brevity', 'No duplication: Info lives in SKILL.md OR references, never both') — and every explanation is GoClaw-specific knowledge Claude would not already have (e.g., the publish_skill tool behavior, composite score weights). It does not fall to 4 because there is no section of over-explanation that could be trimmed without losing actionable content.

5 / 5

Actionability

Most guidance is executable: 'scripts/init_skill.py <name> --path <dir>', 'publish_skill(path: "~/.goclaw/skills-store/<name>")', and the scripts table give copy-paste-ready commands. It is not a 5 because several workflow steps are high-level hints ('Run eval suite, grade outputs, compare with/without skill', 'Draft assertions while runs execute') that defer wholesale to references without even a sample command (e.g., no shown invocation of run_eval.py or the grader agent).

4 / 5

Workflow Clarity

The 10-step numbered Creation Workflow and the numbered Eval & Testing process give a clear sequence with built-in feedback loops ('Test & Evaluate' → 'Optimize Description' → 'Iterate — Generalize from feedback'), and validation tooling exists (quick_validate.py, validation-checklist.md, 'After creating and validating a skill'). Not a 5 because validation is not wired into the sequence itself — quick_validate.py appears only in the scripts table, and the workflow does not state a 'validate before publish' checkpoint or an error-recovery path (e.g., what to do when evals fail).

4 / 5

Progressive Disclosure

Structure and signaling of the references that ARE cited is good (one level deep, inline pointers like 'Full anatomy: references/skill-anatomy-and-requirements.md'), but scoring against the actual bundle reveals real problems: referenced paths 'agents/grader.md', 'agents/comparator.md', 'agents/analyzer.md', 'evals/evals.json', and 'eval-viewer/generate_review.py' do not exist in the bundle, and roughly 11 of the 26 reference files (plugin-marketplace-*, mcp-skills-integration, troubleshooting-guide, writing-effective-instructions, testing-and-iteration, yaml-frontmatter-reference) are never linked from SKILL.md, making them undiscoverable. Not a 4 because broken references and orphaned files are more than 'minor organization gaps'.

3 / 5

Total

16

/

20

Passed

Description

75%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A solid, appropriately 'pushy' description that names a distinct niche (GoClaw skill authoring), covers several concrete capabilities, and includes an explicit 'Use for...' trigger list. It sits just below top marks because the trigger list is terse rather than scenario-rich and a few natural trigger phrases (e.g., 'write a SKILL.md', 'publish/package a skill') are missing.

Suggestions

Expand the 'Use for...' clause into scenario-style triggers ('Use when the user asks to create a new skill, write or improve a SKILL.md, package or publish a skill, or optimize skill descriptions/benchmarks').

Add one or two natural synonyms or file-extension triggers such as 'SKILL.md', 'skills-store', or 'eval suite' to broaden keyword coverage.

DimensionReasoningScore

Specificity

The description lists several concrete actions and objects: 'Create or update GoClaw agent skills', 'skill scripts, references, benchmark optimization, description optimization, eval testing'. It is not a full 5 because 'eval-driven iteration' and 'extending agent capabilities' are somewhat abstract, leaving minor gaps in coverage.

4 / 5

Completeness

Both parts are present: 'what' ('Create or update GoClaw agent skills with eval-driven iteration') and an explicit trigger clause ('Use for new skills, ... extending agent capabilities'). It falls short of the 5 anchor because the 'when' is a terse comma list rather than the fuller explicit trigger scenarios ('Use when the user asks to...', 'when the user mentions...') seen in the top anchor.

4 / 5

Trigger Term Quality

'new skills, skill scripts, references, benchmark optimization, description optimization, eval testing, extending agent capabilities' are natural phrases a user would say when wanting to build or improve a skill. Not 5 because common variations like 'create a skill', 'write a SKILL.md', 'package/publish a skill' are absent; not 3 because coverage is genuinely good, not just 'some relevant keywords'.

4 / 5

Distinctiveness Conflict Risk

'GoClaw agent skills' carves out a clear niche (skill authoring/meta-tooling) with distinct triggers like 'benchmark optimization' and 'description optimization'. Minor overlap risk remains: 'extending agent capabilities' is broad enough to collide with general agent-configuration or MCP-type skills, so it is not a clean 5.

4 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
nextlevelbuilder/goclaw
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.