CtrlK
BlogDocsLog inGet started
Tessl Logo

run-unit-tests

Use when asked to run unit tests.

53

Quality

67%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/run-unit-tests/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a lean, appropriately scoped instruction: one command, one path note, no filler. Its weaknesses are the ambiguous `{root}/sdk/python` placement note (with a typo) that blurs where the command should be run, and the implication that other packages may have tests while only `flet` is covered. Fixing the path context would make it fully copy-paste ready.

Suggestions

Replace the ambiguous `{root}` note with an explicit instruction, e.g. "Run from the repo root after `cd sdk/python`, or use the absolute path `sdk/python/packages/flet/tests`".

Fix the typo 'commend' → 'command' in the path note.

Clarify scope: state whether other Flet packages have test suites and give their commands, or say `flet` is the only package with unit tests.

DimensionReasoningScore

Conciseness

The body is ~10 lines: a one-line scope statement ("run unit tests of the Python part of Flet framework"), the exact command, and a path note — no concept teaching, no padding. Every token earns its place, matching the lean/efficient anchor; it does not include the unnecessary explanation that anchor 4 allows.

5 / 5

Actionability

A concrete, executable command (`uv run --group test pytest packages/flet/tests`) is provided, but two minor gaps keep it below fully copy-paste ready: the `{root}` placeholder in "The directory in that commend is relative to `{root}/sdk/python`" is undefined (the command fails unless run from the right directory), and 'commend' is a typo. This is anchor 4's minor gaps, not anchor 5's fully executable guidance.

4 / 5

Workflow Clarity

As a simple single-task skill it would qualify for 5 if the single action were unambiguous, but the relative-path note leaves ambiguity about where the command must be run from (undefined `{root}`, no cd instruction, typo 'commend'). The action is clear with minor ambiguity — anchor 4; running tests is non-destructive so no validation cap applies.

4 / 5

Progressive Disclosure

The skill is under 50 lines, needs no external references (none exist in the bundle: references/, scripts/, assets/ are all absent), and is cleanly organized with a heading and code block. Per the simple-skill guideline, well-organized short content with no need for external files scores 5.

5 / 5

Total

18

/

20

Passed

Description

32%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is a bare trigger clause: it answers only 'when' and leaves 'what', scope, and concrete actions unstated. Without any Flet/Python/pytest specificity it risks triggering for unrelated unit-test requests. A one-sentence 'what' plus scope-anchored triggers would move it into good-example territory.

Suggestions

Add a 'what' clause stating the concrete action, e.g. "Runs the Python (flet) package unit tests with uv + pytest in the Flet repo. Use when asked to run unit tests for the Flet framework."

Anchor the trigger to the skill's niche (Flet/Python) so it does not fire on unit-test requests for other projects or languages.

Include natural trigger variations users would say, such as 'run the tests', 'pytest', or 'test suite' for the Flet Python package.

DimensionReasoningScore

Specificity

"Use when asked to run unit tests" names the domain but offers no concrete actions or scope — it never says what the skill actually does (run the Flet Python package tests via uv/pytest). It sits above the entirely-vague anchor because the domain is named, but below anchor 3 because no distinct concrete actions are listed.

2 / 5

Completeness

Only the 'when' half is present ("Use when asked to run unit tests") with no 'what' — the description never states what the skill does (e.g., running the Flet Python unit tests with a specific command). This is precisely anchor 2's "only 'when' is present without 'what'"; it is not anchor 3 because no clear 'what' exists even weakly.

2 / 5

Trigger Term Quality

"run unit tests" is a phrase users would naturally say, giving some relevant keywords, but common variations are missing: "pytest", "test suite", "tests are failing", "run the tests". This matches anchor 3 (some relevant keywords, missing variations) rather than 4's good coverage.

3 / 5

Distinctiveness Conflict Risk

"run unit tests" is very broad — with no mention of Flet, Python, or uv/pytest, this description would trigger for any unit-testing request in any project, creating high overlap with generic test-running skills. That matches anchor 2 (very broad, high overlap risk), not the somewhat-specific anchor 3.

2 / 5

Total

9

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
flet-dev/flet
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.