CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-debug

Debug a reproducible symptom with a bounded feedback loop and original-scenario verification

56

Quality

62%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/skill-debug/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a lean, disciplined debugging procedure with strong validation gates and concrete escalation rules; it stays well within token budget. Its main weaknesses are the reliance on external plugin files and commands whose contents and behavior are not restated, and a workflow that is narrated rather than explicitly sequenced.

Suggestions

State what the referenced block files and the freeze guard actually enforce in one line each (or vendor them into references/), so the skill remains actionable when the plugin paths are unavailable.

Number the core loop steps explicitly (reproduce → minimize → hypothesize → fix → verify both scenarios → clean up) so the sequence and its checkpoints are unambiguous without reconstruction from prose.

DimensionReasoningScore

Conciseness

The body is dense, procedural, and free of explanations of concepts Claude already knows; rules like the HARD-GATE, 3-Strike Rule, and WTF thresholds are all load-bearing. Minor trimmable padding remains in the host-adapter note and license attribution.

4 / 5

Actionability

Mostly executable guidance: a runnable freeze-guard bash snippet, numeric thresholds ('+15% per revert and +20% for touching unrelated files', STOP above 20%), and a copy-ready report format line. It falls short of fully executable because several mechanisms (the contents of skills/blocks/debug-feedback-loop.md, /octo:unfreeze, the freeze guard's enforcement) are referenced but never specified.

4 / 5

Workflow Clarity

A clear sequence is present (reproduce, retain failure signature, minimize, test one named hypothesis, verify minimal repro and original scenario, remove instrumentation, keep a regression test) with an explicit HARD-GATE validation before production changes and feedback loops (3-strike rotation, WTF stop threshold). It is not a 5 because the sequence must be assembled from prose rather than being explicitly ordered end-to-end.

4 / 5

Progressive Disclosure

References are one level deep and clearly signaled with exact paths ('Read skills/blocks/debug-feedback-loop.md from the installed plugin root'), keeping SKILL.md lean. However, no bundle files exist alongside this SKILL.md and all referenced paths point to an assumed external plugin, so navigation depends on unverifiable paths.

4 / 5

Total

16

/

20

Passed

Description

50%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description communicates a clear methodological identity (bounded loop, original-scenario verification) but reads more like a subtitle than a trigger-rich description. It lacks any 'when to use this' guidance and the natural vocabulary users employ when asking for debugging help.

Suggestions

Append an explicit trigger clause such as 'Use when the user reports a reproducible bug, failing test, error, or unexpected behavior and asks why it happens or how to fix it.'

Enumerate the concrete actions the skill performs (reproduce the symptom, minimize the scenario, test named root-cause hypotheses, verify the fix in the original scenario, preserve a regression test) so the 'what' is comprehensive rather than a single qualified action.

Add natural trigger synonyms users would say — 'bug', 'error', 'crash', 'failing test', 'root cause', 'regression' — to reduce overlap with adjacent skills and improve routing.

DimensionReasoningScore

Specificity

The description names the debugging domain and qualifies it with 'bounded feedback loop' and 'original-scenario verification', but this is one action described with method qualifiers rather than a list of several distinct concrete actions.

3 / 5

Completeness

There is a clear 'what' ('Debug a reproducible symptom with a bounded feedback loop and original-scenario verification') but no 'Use when...' clause or equivalent explicit trigger guidance, which caps completeness at 3.

3 / 5

Trigger Term Quality

'Debug' and 'symptom' are natural user terms, but the description is missing common variations users actually say such as 'bug', 'error', 'crash', 'failing test', or 'root cause'.

3 / 5

Distinctiveness Conflict Risk

'Debug a reproducible symptom' with methodology qualifiers is somewhat specific, but the thin trigger language leaves real overlap risk with generic debugging or verification skills.

3 / 5

Total

12

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.