CtrlK
BlogDocsLog inGet started
Tessl Logo

verify-changes

Verify code changes by running the project's typecheck, build, lint, and targeted tests, then fix and re-run until clean. Use after editing any source file.

76

Quality

95%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary instruction-only skill body: dense with per-ecosystem executable commands, a numbered procedure with an explicit validate→fix→retry loop, and guardrails (never run the full suite, don't delete failing tests). It is concise, actionable, and appropriately scoped for a self-contained skill with no bundle files.

DimensionReasoningScore

Conciseness

The 29-line body is lean and assumes Claude's competence — it gives 'npx tsc --noEmit', 'mvn -q -pl <module> compile', etc. without explaining what typecheckers or build tools are. Every line is instruction; there is no padding to trim, so it fits the top anchor rather than the 'minor instances of over-explanation' of score 4.

5 / 5

Actionability

Fully executable commands are given per ecosystem: 'npx tsc --noEmit', 'npx eslint <changed files>', 'ruff check <path>', 'mypy <path>', 'pytest <specific test file>', 'mvn -q -pl <module> compile ... -Dtest=<ClassName>', './gradlew :<module>:compileJava', plus 'NO_COLOR=1'. Placeholders like <module> are parameters, not pseudocode, and the commands cover the common cases, matching the copy-paste-ready top anchor; score 4 would require minor gaps in command coverage.

5 / 5

Workflow Clarity

The 5-step procedure is clearly sequenced with an explicit validation checkpoint ('Read the output as ground truth. If it reports errors, they are real — do not claim success') and an explicit feedback loop ('Fix the root cause, then re-run the SAME command. Repeat until it passes'), reinforced by the rule 'Only mark a todo completed once its verification command passes'. This matches the top anchor including error-recovery loops; no destructive/batch cap applies since checks are read-only.

5 / 5

Progressive Disclosure

The skill is under 50 lines with a single task and no need for external references; the bundle directories (references/, scripts/, assets/) do not exist, and the body references none. Per the rubric's simple-skill note, well-organized sections (intro, Procedure, Rules) score 5 — the structure is clear and nothing that belongs in a separate file is inlined.

5 / 5

Total

20

/

20

Passed

Description

88%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person voice, concrete multi-action capability statement, and an explicit 'Use after...' trigger clause. The only risks are breadth of the trigger (any source-file edit) and a few missing natural synonyms.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — 'running the project's typecheck, build, lint, and targeted tests, then fix and re-run until clean' — covering the full capability surface of the skill with no vague filler. It is not score 4 because coverage is comprehensive for this domain, not just 'several specific actions with minor gaps'.

5 / 5

Completeness

It explicitly answers both questions: what ('verify code changes by running typecheck, build, lint, and targeted tests, then fix and re-run until clean') and when ('Use after editing any source file'). Both are concrete and explicit, matching the top anchor exactly; score 4 would require the 'when' clause to be less specific than it is.

5 / 5

Trigger Term Quality

Natural terms users would actually say are present: 'verify code changes', 'typecheck', 'build', 'lint', 'tests', 'editing any source file'. It is not 5 because a few natural synonyms are missing (e.g. 'run the checks', 'CI', 'run the tests' phrasing), though no jargon-only or generic terms appear, keeping it above 3.

4 / 5

Distinctiveness Conflict Risk

The niche (post-edit verification loop) is distinct from review or test-writing skills, but 'Use after editing any source file' fires on nearly every edit and could overlap with closely related verify/test/debug skills. It is more distinct than score 3 ('could still overlap with similar skills') but not the minimal-conflict clarity of 5.

4 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
agentscope-ai/agentscope-java
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.