CtrlK
BlogDocsLog inGet started
Tessl Logo

gstack-openclaw-investigate

Use when asked to debug, fix a bug, investigate an error, or do root cause analysis, and when users report errors, stack traces, unexpected behavior, or say something stopped working.

61

Quality

71%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./openclaw/skills/gstack-openclaw-investigate/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered debugging workflow: clearly phased, rich in feedback loops and verification checkpoints, and written in lean imperative prose with concrete templates and commands. The only real weaknesses are mild duplication between the phases and the closing Important Rules section, and a pattern catalog that could be split into a reference file.

DimensionReasoningScore

Conciseness

The body is lean and imperative — "Trace the code path from the symptom back to potential causes", "Minimal diff: Fewest files touched, fewest lines changed" — with no explanations of concepts Claude already knows. It is not 5 because of repeated material: 'Recurring bugs in the same files are an architectural smell' appears in both Phase 1 and Phase 2, the '>5 files' flag appears in Phase 4 and again in Important Rules, and the 3-strike rule is stated twice.

4 / 5

Actionability

For an instruction-only skill, guidance is fully concrete and executable: a copy-paste command (`git log --oneline -20 -- <affected-files>`), exact output formats ("Root cause hypothesis: ..."), a verbatim user-facing 3-strike message, defined statuses (DONE | DONE_WITH_CONCERNS | BLOCKED), and a structured DEBUG REPORT template. It is not 4 because no key execution detail is missing.

5 / 5

Workflow Clarity

Five clearly sequenced phases with explicit validation checkpoints and feedback loops: hypothesis must be confirmed with evidence before any fix, a wrong hypothesis loops back to Phase 1, the 3-strike rule stops runaway guessing, the regression test must fail without the fix and pass with it, and Phase 5 mandates fresh reproduction. This matches the top anchor (explicit validation steps, feedback loops, error-recovery paths).

5 / 5

Progressive Disclosure

No bundle files exist (references/, scripts/, assets/ are absent), and the ~130-line body is cleanly sectioned into distinct phases, so nothing is misfiled or buried. It is not 5 because the skill exceeds 50 lines and the Phase 2 pattern catalog and the DEBUG REPORT format are self-contained chunks that could live in one-level-deep reference files, keeping SKILL.md a tighter overview.

4 / 5

Total

18

/

20

Passed

Description

55%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

This is a trigger-only description: it excels at when to fire (natural, comprehensive user phrasing) but never states what the skill does, capping completeness and specificity low. Adding a one-clause capability summary (e.g., a systematic root-cause-investigation workflow with hypothesis testing, minimal-diff fixes, and regression tests) would move it toward the good examples.

Suggestions

Add a leading 'what' clause before the trigger clause, e.g.: 'Systematic debugging workflow: investigate root cause before fixing, confirm hypotheses with evidence, apply minimal fixes with regression tests. Use when asked to debug...' — this addresses both specificity (currently 2) and completeness (currently 2).

Name the skill's concrete outputs (root cause hypothesis, debug report, regression test) so the description distinguishes what this skill produces beyond generic debugging help.

Narrow the broadest triggers ('debug', 'fix a bug') with context that signals the systematic-investigation approach, reducing overlap with general code-review or verification skills.

DimensionReasoningScore

Specificity

The description names the debugging/root-cause domain via phrases like "debug, fix a bug, investigate an error, or do root cause analysis", but describes no concrete capabilities of the skill itself — what it actually does (systematic investigation phases, hypothesis testing, regression-test requirements, debug reports) is entirely absent. It matches 'names the domain but actions are minimal or generic' rather than level 3, which requires 1-2 concrete actions to be stated.

2 / 5

Completeness

Only the 'when' is present — the entire description is a 'Use when...' trigger clause with no statement of what the skill does. This matches anchor 2 ('only when is present without what') exactly. It is not 1 because the 'when' is concrete and multi-faceted, and not 3 because anchor 3 requires a clear 'what', which is missing entirely.

2 / 5

Trigger Term Quality

Trigger coverage is comprehensive and natural: "debug", "fix a bug", "investigate an error", "root cause analysis", "stack traces", "unexpected behavior", and "something stopped working" — these are exactly the phrases users say, including synonyms and both user-report and direct-request framings. It is not level 4 because no natural terms are missing.

5 / 5

Distinctiveness Conflict Risk

Phrases like "root cause analysis" and "stack traces" carve out a distinct debugging niche with low conflict risk against most skills. It is not 5 because broad triggers like "debug" and "fix a bug" could overlap with general code-review, testing, or verification skills.

4 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
garrytan/gstack
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.