CtrlK
BlogDocsLog inGet started
Tessl Logo

run-tests

Use this skill when the user asks to "run tests", "test this", "check if tests pass", "cargo test", "run clippy", "lint this", "check formatting", "cargo fmt", "CI checks", "verify changes", "does this pass tests", "run the full check", "pre-commit check", or wants to verify that code changes are correct. Use this even when the user says something like "make sure this works" or "check for issues" in the context of code changes.

69

Quality

87%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

This is a model skill body: fully executable commands, a change-type-driven decision procedure, an explicit pre-PR validation gate, and error-class-to-fix feedback loops — all delivered with near-zero token waste. Organization is clean and the scope boundary with the add-command skill is explicitly drawn.

DimensionReasoningScore

Conciseness

Lean and efficient throughout: the Quick Reference table, the exact one-line CI sequence, and terse per-failure diagnostics carry maximum information per token. Nothing explains concepts Claude already knows — even the `--locked` justification ("ensures the lockfile is respected - this matches .github/workflows/test.yml") is earned context, not padding. Anchor 5; nothing to trim, so not 4.

5 / 5

Actionability

Every instruction is a copy-paste-ready command: the exact CI chain (`cargo fmt --check && cargo clippy --locked -- -D warnings && cargo test --locked`), a change-type-to-command mapping ("API types (src/commands/<cmd>/api.rs) - run that command's tests first"), precise test filters (`cargo test --test dataprime input`), and concrete failure fixes. Fully executable with common cases covered — anchor 5, not 4.

5 / 5

Workflow Clarity

Clear decision procedure ("What to Run After a Change" maps each change type to its verification level), an explicit checkpoint ("Run it before committing or creating a PR"), and genuine feedback loops in "Interpreting Failures" (failure class → likely cause → fix, e.g., deserialization failures → check serde attributes). Anchor 5 — validation steps and error-recovery loops are both explicit, so not 4.

5 / 5

Progressive Disclosure

The body is ~55 lines with no bundle files and no content that belongs in separate files; sections are well organized (table, CI check, decision guide, test targeting, failure interpretation) and it cleanly signals its boundary ("for writing new tests, see the add-command skill"). Per the rubric's simple-skill exception, well-organized sections with no external-reference needs warrant 5; not 4 since nothing is misplaced.

5 / 5

Total

20

/

20

Passed

Description

73%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A trigger-term powerhouse: the "when" guidance is among the most explicit possible, with excellent natural-phrase and tool-command coverage. The main weakness is that the description never explicitly states what the skill does — the capability is only implied through the quoted user requests — which caps completeness and leaves slight overlap risk on generic verification phrases.

Suggestions

Open with an explicit capability statement before the trigger list, e.g., "Runs tests, lint, and formatting checks for the cx CLI with cargo" — this would raise completeness from 3 to 5.

Tie the description to the cx CLI specifically ("for the cx CLI using cargo") so it distinguishes itself from generic test/verify skills in other repositories.

Consider trimming the broadest catch-alls ("check for issues", "make sure this works") or scoping them explicitly to test execution, since they risk triggering code-review or verify skills instead.

DimensionReasoningScore

Specificity

Concrete actions are named throughout the trigger list ("run tests", "run clippy", "check formatting", "cargo fmt", "CI checks", "verify changes") — several specific actions with minor gaps. It falls short of 5 because the skill's own capabilities are never stated directly; the actions appear only as quoted user requests, not as a capability statement.

4 / 5

Completeness

The "when" is maximally explicit, but the "what" is only weakly implied through the quoted user phrases — the description never states what the skill does (e.g., "Runs tests, lint, and format checks for the cx CLI"). Per the guideline to score only what is explicitly stated, this mirrors anchor 3 (one half clear, the other weakly implied); it is not 4 because no explicit capability statement exists, and not 2 because the trigger list does convey the action domain.

3 / 5

Trigger Term Quality

Comprehensive natural-term coverage including synonyms and exact tool commands: "run tests", "test this", "check if tests pass", "cargo test", "run clippy", "lint this", "check formatting", "cargo fmt", "CI checks", "does this pass tests", "pre-commit check", plus the softer phrasings "make sure this works" and "check for issues". No natural term for this domain is missing, matching the anchor 5 example's synonym-plus-tool-command coverage.

5 / 5

Distinctiveness Conflict Risk

The cargo/clippy/fmt/CI terms carve a distinct niche, but broad catch-all phrases like "verify changes", "make sure this works", and "check for issues" overlap with code-review and general verification skills. Anchor 4 (mostly distinct, minor overlap risk with closely related skills) is the best fit — not 5 due to those generic triggers, not 3 since the majority of trigger terms are unambiguous.

4 / 5

Total

16

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

13

/

16

Passed

Repository
coralogix/cx-cli
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.