CtrlK
BlogDocsLog inGet started
Tessl Logo

oma-debug

Diagnose a reproducible failure, fix its cause, and verify the regression. Use for crashes, incorrect behavior, and failing tests.

60

Quality

69%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/oma-debug/SKILL.md

The canonical home for this skill is oma-debug in first-fluke/oh-my-agent

SKILL.md
Quality
Evals
Security

Quality

Content

62%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-sequenced debugging protocol with strong workflow clarity, explicit verification checkpoints, and recovery paths, but roughly half the body is abstract agent-schema metadata (SSL primitives, resource scopes, control-flow features) that spends tokens without adding executable guidance, and its many references point to files absent from the bundle. Tightening the meta sections into concrete commands and consolidating the duplicated dependency/reference listings would lift both conciseness and actionability.

Suggestions

Cut or compress the schema-descriptive sections (Actions/SSL-primitive table, Resource scope, Control-flow features) into the concrete workflow they summarize, keeping only directives Claude can act on.

Replace abstract rows like "Infer root cause | INFER | Diagnostic reasoning" with executable guidance, e.g. example reproduction commands, a sample rg search against a stack trace, and how to run and scope the regression test.

Consolidate the Dependencies and References sections into one clearly navigable list, and ensure every referenced file (resources/*.md, ../_shared/core/*) actually ships in the bundle or drop the dead paths.

DimensionReasoningScore

Conciseness

The body is mostly lean, but meta-descriptive sections add tokens without instructing: the Actions table's "SSL primitive" column (`CALL_TOOL`, `READ`, `INFER`), the "Resource scope" table (CODEBASE/LOCAL_FS/PROCESS/MEMORY), and "Control-flow features" describe the agent's shape rather than telling Claude what to do, and the Scheduling section's "Intent signature"/"When to use" duplicate the frontmatter description.

3 / 5

Actionability

Concrete guidance exists ("rg \"<error-message-or-symbol>\"", "Document in `.agents/results/bugs/`", "run the smallest reproduction command first"), but the bulk of "Logical Operations" is an abstract taxonomy — "Infer root cause | INFER | Diagnostic reasoning" — with no executable commands or examples, and key details like what a reproduction or verification command looks like are left to the referenced files that are not present in the bundle.

3 / 5

Workflow Clarity

The Entry → Scenes (PREPARE/ACQUIRE/REASON/ACT/VERIFY/FINALIZE) → Exit arc is clearly sequenced with an explicit validation stage ("VERIFY: Re-run failing and related checks"), a feedback loop ("If the first fix fails verification, return to root-cause analysis"), and failure/recovery paths for missing environments and infeasible regression tests. This matches the top anchor: clear sequence, explicit checkpoints, error-recovery loops.

5 / 5

Progressive Disclosure

The References section is one level deep and signals purpose per file ("Checklist (pre-submit self-verification): `resources/checklist.md`"), but no bundle files exist in `references/`, `scripts/`, or `assets/` and every referenced path (`resources/*.md`, `../_shared/core/*`, `../oma-observability/SKILL.md`) resolves to nothing in the bundle, and the Dependencies section duplicates the reference list loosely, leaving navigation unreliable.

3 / 5

Total

14

/

20

Passed

Description

76%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A concise, third-person description that clearly states both what the skill does (diagnose, fix, verify a regression) and when to use it with enumerated triggers. Its main gap is trigger coverage: for a debugging skill it lacks the most natural user terms like "bug", "error", and "debug".

Suggestions

Add the natural trigger terms users actually say — "bug", "error", "debug" — e.g. 'Use for bugs, crashes, errors, incorrect behavior, and failing tests'.

Optionally surface the regression-test deliverable in the what-clause ("...and add a regression test") to round out coverage.

DimensionReasoningScore

Specificity

"Diagnose a reproducible failure, fix its cause, and verify the regression" names the domain and three concrete actions — diagnose, fix, verify — which fits the 'several specific actions; minor gaps in coverage' anchor; it stops short of the 5 anchor's comprehensiveness (no mention of regression-test authorship or root-cause documentation).

4 / 5

Completeness

It explicitly answers both questions: what ("Diagnose a reproducible failure, fix its cause, and verify the regression") and when ("Use for crashes, incorrect behavior, and failing tests") with concrete trigger phrases, matching the top anchor; it is not the 4 anchor because the 'when' clause is specific and enumerated rather than generic.

5 / 5

Trigger Term Quality

"Use for crashes, incorrect behavior, and failing tests" provides three relevant trigger phrases, but for a debugging skill it omits the most natural terms users say — "bug", "error", "debug", "exception", "broken" — matching 'some relevant keywords but missing common variations or synonyms' rather than the good-coverage anchor.

3 / 5

Distinctiveness Conflict Risk

Crash/failing-test triggers carve out a mostly distinct debugging niche with only minor overlap risk against test-writing or code-review skills, but "incorrect behavior" is broad enough to invite overlap, so it fits 'mostly distinct; minor overlap risk' rather than the minimal-conflict top anchor.

4 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
first-fluke/oh-my-agent
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.