CtrlK
BlogDocsLog inGet started
Tessl Logo

create-tooluniverse-skill

Create high-quality ToolUniverse skills following test-driven, implementation-agnostic methodology.

53

Quality

58%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/create-tooluniverse-skill/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-sequenced, lean, actionable 7-phase workflow with explicit validation and error-recovery hooks, appropriately keeping implementation detail in bundle files. Its main defect is navigation: most of the ten referenced filenames do not match the actual bundle files, so an agent following the pointers would have to guess or search for the real files.

Suggestions

Align reference names with the actual bundle (e.g. references/tool_testing_workflow.md, references/skill_standards_checklist.md, references/devtu_optimize_integration.md, references/implementation_agnostic_format.md) and include the directory path for each.

Resolve or remove references to files that don't exist in the bundle (PARAMETER_VERIFICATION.md, CODE_TEMPLATES.md, PACKAGING_TEMPLATE.md) or add those files; point SKILL_TEMPLATE.md and QUICKSTART_TEMPLATE.md at assets/skill_template/ with correct names.

Add an explicit failure-handling step in Phase 6 (e.g. 'If validation fails, fix and re-run the test suite before packaging') to close the workflow's last implicit feedback loop.

DimensionReasoningScore

Conciseness

The body is dense and assumes competence: phase tables, one-line pillar statements, and file paths with no concept over-explanation. Not a 5 because the 10-pillar list and the 'Quality Indicators' section partially duplicate content delegated to OPTIMIZE_INTEGRATION.md and VALIDATION_CHECKLIST.md and could be trimmed.

4 / 5

Actionability

Concrete guidance throughout: exact data path ('/src/tooluniverse/data/*.json'), specific deliverables per phase ('python_implementation.py', 'test_skill.py'), named templates, and named integration skills. Absence of inline code is justified by the implementation-agnostic principle, so it is not penalized; not a 5 because there are no copy-paste commands for steps like running the test suite or invoking the templates.

4 / 5

Workflow Clarity

The 7-phase workflow is explicitly sequenced with durations, a mandatory test-first checkpoint in Phase 2, a dedicated validation phase, and an error-recovery loop via devtu-fix-tool. Not a 5 because the recovery path when Phase 6 validation fails is implicit rather than spelled out in the sequence.

4 / 5

Progressive Disclosure

Structure is good — a dedicated reference table, one level deep, with an overview body — but scored against the actual bundle, most referenced filenames do not resolve: only test_tools_template.py matches (in scripts/), while TESTING_GUIDE.md, VALIDATION_CHECKLIST.md, OPTIMIZE_INTEGRATION.md, and IMPLEMENTATION_AGNOSTIC.md only loosely correspond to differently-named files in references/, and PARAMETER_VERIFICATION.md, CODE_TEMPLATES.md, PACKAGING_TEMPLATE.md, SKILL_TEMPLATE.md, and QUICKSTART_TEMPLATE.md do not match any bundle file by name. Broken reference names block navigation, which is more than the 'minor organization gaps' of a 4.

3 / 5

Total

15

/

20

Passed

Description

46%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description states a clear, niche-specific 'what' but completely omits any 'when to use' guidance, and its methodology jargon ('test-driven, implementation-agnostic') provides no natural user trigger terms. It is distinguishable from unrelated skills but reads more like an internal process label than an invocable capability description.

Suggestions

Add an explicit trigger clause, e.g. "Use when creating or restructuring a ToolUniverse skill for a new scientific domain."

Replace or supplement the methodology jargon with natural phrases a user would actually say, such as "build a ToolUniverse skill", "document ToolUniverse tools", or "set up a new tooluniverse-[domain] skill".

Briefly enumerate 1-2 more concrete actions (e.g. "test tools, write implementation-agnostic SKILL.md and QUICK_START.md") to raise specificity from one action to several.

DimensionReasoningScore

Specificity

Names the domain ("ToolUniverse skills") and a single concrete action ("Create"), but "test-driven, implementation-agnostic methodology" describes process attributes rather than additional concrete actions, matching the '1-2 concrete actions, not comprehensive' anchor.

3 / 5

Completeness

The 'what' is clear (create high-quality ToolUniverse skills), but there is no 'Use when...' clause or any trigger guidance, capping completeness at 3 per the judging guidelines. Not a 4 because 'when' is entirely absent rather than just under-specified.

3 / 5

Trigger Term Quality

Beyond "create" and "ToolUniverse skills", the description offers only methodology jargon ("test-driven, implementation-agnostic methodology") that users would not naturally say, so natural trigger phrases are largely missing. Not a 1 because "ToolUniverse" and "skills" are domain terms a user could plausibly utter.

2 / 5

Distinctiveness Conflict Risk

"ToolUniverse skills" carves out a clear niche with minimal conflict against generic skills, but it overlaps with closely related sibling skills it itself references (devtu-create-tool, devtu-optimize-skills). Not a 5 because that sibling overlap is real; not a 3 because the ToolUniverse qualifier keeps it well above generic skill-creation descriptions.

4 / 5

Total

12

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
mims-harvard/ToolUniverse
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.