CtrlK
BlogDocsLog inGet started
Tessl Logo

shell-error-debug-workflow

Systematic workflow for diagnosing and resolving unknown errors from run_shell commands

54

Quality

68%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

High

Do not use without reviewing

Fix and improve this skill with Tessl

tessl review fix ./benchmarks/gdpval/skills/shell-error-debug-workflow/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

67%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is highly actionable with concrete commands and a useful troubleshooting table, organized into a clear stepwise workflow. Its main weakness is redundancy from restating the diagnostic sequence multiple times, which lightly hurts conciseness and organization.

Suggestions

Collapse the 'Complete Diagnostic Sequence' and the Python example into a single worked example to remove the triplicated diagnostic steps.

Add explicit validation/checkpoint language between steps (e.g. 'If stderr is empty, proceed to Step 2') to tighten the feedback loop.

Trim restating sentences like 'This ensures error messages are captured...' since the code already shows the intent.

DimensionReasoningScore

Conciseness

The body is mostly efficient with tight code blocks, but it repeats the same diagnostic sequence three times (workflow steps, 'Complete Diagnostic Sequence', and the Python example) and includes a few restating sentences that could be trimmed.

3 / 5

Actionability

It provides concrete, copy-paste-ready run_shell commands for each step plus a symptom/cause/solution table; minor gaps come from templated placeholders like 'some-command' and 'relevant_var'.

4 / 5

Workflow Clarity

Steps 1-6 are clearly sequenced and the causes/solutions table supplies an implicit feedback loop, though explicit validation checkpoints between steps are only implied rather than stated.

4 / 5

Progressive Disclosure

The skill is self-contained with no bundle files and is organized into clear, well-labeled sections; the only organization gap is the duplicated diagnostic content that could be consolidated.

4 / 5

Total

15

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is clear and reasonably specific about its niche, but it omits an explicit 'when to use' trigger clause and lacks synonym coverage of natural user phrasings. It sits solidly at the midpoint across most dimensions.

Suggestions

Add an explicit trigger clause, e.g. 'Use when run_shell returns an unknown error or a command fails with no clear output'.

Broaden trigger terms to include synonyms users actually say, such as 'shell failure', 'command failed', or 'ambiguous error'.

Optionally surface one or two concrete diagnostic techniques (e.g. 'capture stderr, verify paths') to lift specificity above 3.

DimensionReasoningScore

Specificity

The description names the domain ('unknown errors from run_shell commands') and two concrete actions ('diagnosing and resolving'), but stops short of comprehensive coverage of the workflow's actual techniques.

3 / 5

Completeness

It gives a clear 'what' (a systematic diagnostic workflow) but lacks an explicit 'Use when...' trigger clause; per the rubric, a missing explicit trigger caps completeness at 3.

3 / 5

Trigger Term Quality

It includes the natural phrase 'unknown errors' and the tool name 'run_shell', but misses common variations users might say such as 'shell failure', 'command failed', or 'ambiguous error'.

3 / 5

Distinctiveness Conflict Risk

The narrow focus on 'unknown errors from run_shell commands' carves a distinct niche with low conflict risk, though it could be sharpened to fully separate from general debugging skills.

4 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
HKUDS/OpenSpace
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.