CtrlK
BlogDocsLog inGet started
Tessl Logo

verification-loop

A comprehensive verification system for Claude Code sessions.

47

Quality

50%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/verification-loop/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a strong, actionable verification playbook: concrete commands in every phase, explicit stop conditions and a re-run feedback loop, and a clear report format. Its only weaknesses are mild padding in the Continuous Mode and hooks sections and a slightly long all-inline structure. No changes are strictly required.

DimensionReasoningScore

Conciseness

The body is dominated by lean, copy-paste commands with terse instructions ('If build fails, STOP and fix before continuing'), with only minor trimmable content such as the 'Set a mental checkpoint' block, the closing hooks-comparison section, and the restated one-line summary under the title. This matches 'Efficient; minor instances of over-explanation that could be trimmed' rather than level 5, which requires every token to earn its place.

4 / 5

Actionability

Every phase ships a concrete executable command ('npx nx run-many --target=build --all --parallel', 'npx tsc --noEmit', grep/git-diff commands) plus a concrete report template. It stops short of level 5 because a few parts are non-executable guidance — the 'Set a mental checkpoint' instructions, 'Run: /verify', and the X-placeholder reporting lines.

4 / 5

Workflow Clarity

Six phases are clearly sequenced with explicit stop conditions ('If build fails, STOP and fix before continuing', 'Fix critical ones before continuing'), a mandatory pre-push gate, and a genuine feedback loop ('After fixing failures, re-run the full suite to confirm before pushing'). This matches 'Clear sequence with explicit validation steps; feedback loops for error recovery; checklists' — the report format even acts as a checklist.

5 / 5

Progressive Disclosure

The single-file body is well-sectioned with clear headers (When to Use, phases, Output Format, Pre-Push Gate) and no buried or nested references; no bundle files exist to verify. It scores 4 rather than 5 because at ~130 lines with some inline content (Continuous Mode, hooks integration) that could be split out, there are minor organization gaps rather than ideal placement throughout.

4 / 5

Total

17

/

20

Passed

Description

20%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is a vague category label: it names a domain (verification) but states no concrete capabilities and gives no 'when to use' guidance, so it would rarely be selected by trigger matching and overlaps heavily with similar skills. The body of the skill is far more informative than the description suggests. It needs a rewrite listing the concrete verification phases and explicit trigger conditions.

Suggestions

Replace the abstract label with the concrete actions the skill performs, e.g. 'Runs build, type-check, lint, test, security, and diff-review verification phases across an Nx monorepo and produces a pass/fail readiness report.'

Add an explicit trigger clause, e.g. 'Use when the user asks to verify changes, check quality gates, or confirm code is ready before a PR or git push.'

Include natural trigger terms users would actually say — 'verify', 'quality gates', 'pre-push check', 'is this ready for a PR' — to distinguish it from the built-in /verify skill and other testing skills.

DimensionReasoningScore

Specificity

The description reads only "A comprehensive verification system for Claude Code sessions" — a category label with no concrete action verbs at all (nothing like 'runs builds', 'checks types', or 'runs tests'). This matches the anchor 'Entirely vague; no concrete actions; pure abstract language' rather than level 2, which still expects at least a minimal/generic action such as 'Processes PDF files'.

1 / 5

Completeness

There is a vague 'what' ('comprehensive verification system') and no 'when' — no 'Use when...' clause or equivalent trigger guidance anywhere, which alone caps completeness at 3. It matches anchor 2 ('Has a vague what and no when') rather than anchor 1 only because a domain is at least named; the 'what' is not a description of what the skill actually does.

2 / 5

Trigger Term Quality

The only keyword is 'verification' plus the jargon-y 'Claude Code sessions'; natural phrases a user would actually say ('verify before pushing', 'run the tests', 'quality gates') are absent. This sits at 'One or two generic keywords; missing the natural phrases users say' rather than level 1, since 'verification' is a genuinely natural term, just unaccompanied.

2 / 5

Distinctiveness Conflict Risk

'A comprehensive verification system for Claude Code sessions' is very broad and would collide with the built-in /verify skill, any testing/CI skill, and general QA guidance — high overlap risk with many similar skills, matching anchor 2 ('Very broad; high overlap risk'). It is not level 1 because it is not so generic that it conflicts with virtually any skill regardless of domain.

2 / 5

Total

7

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
devrev/meerkat
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.