CtrlK
BlogDocsLog inGet started
Tessl Logo

skvm-general

Drive the skvm CLI on behalf of a user to profile models, AOT-compile skills, run skill-assisted tasks, run benchmarks, and manage compiled proposals. Trigger when the user asks to "profile", "aot-compile", "bench", "run a single ad-hoc task with a skill", or asks about skvm proposals. Do NOT trigger for `jit-optimize` or when the user wants to optimize/improve a skill — use the sibling `skvm-jit` skill instead.

76

Quality

96%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

90%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, accurate CLI-driving skill: fully executable commands, prerequisite self-checks, and strong safety rules for costly and destructive operations. The main gaps are the absence of post-run validation/error-recovery loops and no use of progressive disclosure to move reference material out of the main body.

Suggestions

Add a short post-run verification loop to Steps 5–6, e.g. after bench: 'check the session summary in .skvm/log/bench/<sessionId>/ and surface failures before reporting success', and for detached runs: 'if run-status.json shows failed, show the tailed run.log errors and stop rather than accepting'.

Move the --conditions catalog, proposal-id format details, and the environment-variable table into a single reference file (e.g. references/cli-reference.md) linked from the relevant steps, keeping SKILL.md as a lean overview.

Include an explicit error-recovery path for prerequisite failures beyond exiting (e.g. after the key check, suggest where the user finds their OPENROUTER_API_KEY) so the workflow has a feedback loop rather than a dead stop.

DimensionReasoningScore

Conciseness

Every section carries skvm-specific facts Claude cannot know — real flag sets, pass semantics, condition strings, proposal id format, env vars — with zero padding and no explanation of general concepts. It matches 'lean and efficient; every token earns its place' rather than level 4, which reserves room for trimmable over-explanation.

5 / 5

Actionability

All seven steps give copy-paste-ready executable bash commands with the real flag set (e.g. 'skvm bench --model=<id> --conditions=original,aot-compiled', 'skvm proposals accept <id> --round=2') plus inline comments stating each example's purpose. Fully executable and covering the common cases, matching the level-5 anchor.

5 / 5

Workflow Clarity

Steps 1–7 are clearly sequenced with pre-flight checks and explicit confirmation gates for expensive/destructive operations ('Never run bench or profile across many models without explicit user confirmation', 'Never run proposals accept unless the user explicitly asked to deploy'). It falls short of level 5 because validation is purely pre-run: there are no post-run verification or error-recovery loops (e.g., checking a bench session's results or handling a failed detached run beyond describing run-status.json).

4 / 5

Progressive Disclosure

The single SKILL.md is well structured with clear per-command sections and a rules section, and no external references are buried or nested. It sits at level 4 rather than 5 because at ~145 lines some self-contained blocks (the env-var table, the --conditions catalog, proposal-id format details) could be split into one-level-deep reference files, and the simple-skill (<50 line) exception does not apply.

4 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: concrete action list, natural quoted trigger terms matching the CLI's command vocabulary, an explicit 'Trigger when' clause, and a negative-trigger boundary that disambiguates a sibling skill. Voice is imperative/third-person as required.

DimensionReasoningScore

Specificity

The description lists five concrete, distinct actions — 'profile models, AOT-compile skills, run skill-assisted tasks, run benchmarks, and manage compiled proposals' — giving comprehensive coverage of the CLI's surface. It exceeds the level-4 anchor ('lists several specific actions; minor gaps') because no major capability of the skill is left unmentioned.

5 / 5

Completeness

Both questions are explicitly answered: what ('Drive the skvm CLI... to profile models, AOT-compile skills...') and when ('Trigger when the user asks to...') with concrete trigger phrases. This matches the level-5 example structure and exceeds level 4, whose 'when' is less explicit.

5 / 5

Trigger Term Quality

It quotes the exact terms a user driving this CLI would naturally say — 'profile', 'aot-compile', 'bench', 'run a single ad-hoc task with a skill', 'skvm proposals' — and these map directly to the command vocabulary. It fits the level-5 anchor's comprehensiveness; there is no meaningful synonym or extension the domain requires that is missing.

5 / 5

Distinctiveness Conflict Risk

The niche is unambiguous (driving the skvm CLI) and the negative boundary — 'Do NOT trigger for `jit-optimize` or when the user wants to optimize/improve a skill — use the sibling `skvm-jit` skill instead' — actively routes away the most likely confusion. Minimal conflict risk, matching the level-5 anchor.

5 / 5

Total

20

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
XiaoLuoLYG/GOD
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.