CtrlK
BlogDocsLog inGet started
Tessl Logo

review-fix-loop

This skill should be used when the user asks to "address review feedback", "fix PR comments", "close findings", "respond to reviewer notes", or "reduce review churn". Enforces a finding-to-evidence closure loop for every review round.

71

Quality

89%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exceptionally lean, well-structured process skill: a sequenced loop with genuine validation checkpoints, an error-recovery feedback loop, and a handoff checklist. The only gap is concreteness of examples — severity labels are undefined and there is no worked closure-table row or sample verification command.

Suggestions

Add one example closure-table row (e.g. '| Null deref in parse() | P1 | src/parse.py | Guard clause + test | `pytest tests/test_parse.py` | Pass |') so the table format and expected evidence are unambiguous.

Define the severity tiers in one line each (what makes a finding P1 vs P2 vs P3) or state that severity is taken from the review source when already labeled.

Clarify what counts as acceptable 'proof' per finding — e.g. command output, diff excerpt — since the handoff requires 'pass/fail evidence per finding' without specifying its form.

DimensionReasoningScore

Conciseness

The body is ~43 lines of pure directive content — a six-step loop, a table spec, terse decision rules, and a checklist — with zero padding and no explanation of concepts Claude already knows (nothing about what PRs or code review are). Every section earns its place, matching the 'lean and efficient; assumes Claude's competence' anchor; there is nothing to trim, ruling out 4.

5 / 5

Actionability

Concrete artifacts are given: the closure-table column schema is copy-paste-ready, the fail-twice rule ('If the same command fails twice, stop rerunning') is a crisp executable decision rule, and the verification ladder orders concrete check categories (unit/module, compile/lint/type, repo gates). Not 5 because there are no worked examples — no filled table row, no sample verification command — and severity labels 'P1, P2, P3' are used without definition; not 3 because the guidance is structured and directive rather than pseudocode or high-level hints.

4 / 5

Workflow Clarity

The six-step Execution Loop is a clear sequence with validation woven in ('Run targeted verification', 'Update closure table with outcomes', 'Run broader gate checks'), the Re-Run Control section is an explicit error-recovery feedback loop (fail twice → isolate repro → patch smallest cause → re-run narrow before broad), and the Handoff Requirements section is a closing checklist. This matches the anchor requiring explicit validation steps, feedback loops, and checklists; validation is present, so no cap applies.

5 / 5

Progressive Disclosure

The skill is under 50 lines with no external references needed, and the content is cleanly organized into six labeled sections (Goal, Execution Loop, Closure Table, Re-Run Control, Verification Ladder, Handoff Requirements). Per the rubric's simple-skill guidance, well-organized sections alone warrant 5 here; nothing that belongs in a separate file is inlined.

5 / 5

Total

19

/

20

Passed

Description

81%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with excellent, natural trigger phrases that clearly delimit a distinct niche. Its main weakness is the 'what': one jargon-heavy clause ('finding-to-evidence closure loop') instead of a concrete enumeration of what the skill actually does.

Suggestions

Replace or expand the 'what' clause with 2-3 concrete actions, e.g. 'Parses review findings into a closure table, implements fixes in the smallest coherent batch, and runs a narrow-to-broad verification ladder with pass/fail evidence per finding.'

Spell out what 'finding-to-evidence closure loop' means in plain terms, since the jargon assumes the reader already knows the skill's method.

DimensionReasoningScore

Specificity

The 'what' is a single abstract capability — 'Enforces a finding-to-evidence closure loop for every review round' — with no listing of concrete actions (no mention of parsing findings, building a closure table, or running verification). It names the domain and one jargon-heavy action, which matches the 'names domain and 1-2 concrete actions, but not comprehensive' anchor; it is not 4 because it does not list several specific actions, and not 2 because it goes beyond a bare domain label.

3 / 5

Completeness

Both parts are explicit: 'This skill should be used when the user asks to...' is a clear trigger clause, and 'Enforces a finding-to-evidence closure loop' states what it does. Not 5 because the 'what' is expressed in one abstract, jargon-laden sentence rather than concrete capabilities a user can evaluate; not 3 because the 'when' is fully explicit with concrete trigger phrases rather than weakly implied.

4 / 5

Trigger Term Quality

Five natural, quotable phrases users would actually say — 'address review feedback', 'fix PR comments', 'close findings', 'respond to reviewer notes', 'reduce review churn' — cover the request space comprehensively with synonyms ranging from formal to informal. No file extensions apply to this process skill, so nothing natural is missing; this matches the comprehensive-coverage anchor and is clearly above 'a few natural terms missing'.

5 / 5

Distinctiveness Conflict Risk

The niche — remediating/closing review findings — is clearly distinct from skills that perform reviews, and the triggers ('fix PR comments', 'respond to reviewer notes') would not naturally fire for a review-generation or general-coding skill. This matches 'clear niche with distinct triggers; minimal conflict risk'; the neighboring anchor 4 would require identifiable overlap with a closely related skill, which is not the case.

5 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
spacedriveapp/spacebot
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.