CtrlK
BlogDocsLog inGet started
Tessl Logo

qa

General-purpose QA verdict for any artifact type

54

Quality

60%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude-plugin/skills/qa/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

85%

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured, actionable QA workflow with concrete tool calls, explicit validation/feedback checkpoints, a complete output template, and a worked example. Its only real weakness is moderate verbosity in a few explanatory and meta sections that could be trimmed without losing clarity.

DimensionReasoningScore

Conciseness

The body is mostly efficient and avoids explaining basic concepts Claude already knows, but sections like 'How It Works', the Step 3.5 acting-verification prose, the iterative-loop meta, and the RFC breadcrumb footer add length that could be tightened. It is not score 1 (no padded beginner explanations like 'a PDF is a file format') and not score 3 (some passages restate or over-explain rather than earning every token).

2 / 3

Actionability

It gives concrete commands (`ooo qa`, `tool discovery query: "+ouroboros qa"`), a fully specified MCP tool call with named arguments, a complete verdict output template, and a worked end-to-end example. It is not score 2 (the guidance is specific and copy-paste ready, not pseudocode) and clearly meets the score-3 'executable commands + specific examples' anchor for an instruction skill.

3 / 3

Workflow Clarity

The process is clearly sequenced (Step 0 mode selection, QA Steps 1-5 with sub-step 3.5), includes explicit validation (acting-verification: run, observe real effects, probe adversarial classes, capture evidence) and feedback loops (iterative QA loop until pass/fail, re-run after revise). It is not score 2 (checkpoints are explicit, not implicit) and matches the score-3 anchor with validation steps and error-recovery loops.

3 / 3

Progressive Disclosure

The body is organized into clear navigable sections (Usage, How It Works, Instructions, Fallback, Example) and its single external reference (`<project-root>/src/ouroboros/agents/qa-judge.md`) is one level deep and clearly signaled; no nested 2+-level reference chains exist. It is not score 2 (organization is clean and the reference is well signaled, with no inline content that obviously belongs in a separate file) and not below 3.

3 / 3

Total

11

/

12

Passed

Description

35%

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description names the skill's purpose (a QA verdict) and scope but omits natural trigger terms and an explicit 'when to use' clause, making it weak on discoverability and trigger quality. It is concise and in correct third person, but too generic to clearly differentiate from sibling evaluation skills.

Suggestions

Add an explicit trigger clause, e.g. 'Use when the user asks for a quality check, qa check, or a quick verdict on code, docs, API responses, or test output.'

Include natural trigger keywords directly in the description ('quality check', 'qa check', 'ooo qa') rather than only in the body.

Tighten distinctiveness by contrasting scope, e.g. 'Fast single-pass QA verdict for any artifact — use instead of ooo evaluate when you need a quick check rather than 3-stage verification.'

DimensionReasoningScore

Specificity

The phrase 'QA verdict' names a concrete output action and 'any artifact type' scopes the domain, but only one action is named rather than a comprehensive list of specific concrete actions, matching the score-2 anchor 'Processes PDF files and extracts content'. It is not score 1 (it is not abstract like 'Helps with documents') nor score 3 (no multiple specific actions such as 'parse, score, render verdict, suggest fixes').

2 / 3

Completeness

It states what the skill does ('QA verdict for any artifact type') but has no 'Use when...' clause or explicit trigger guidance, so per the judging guidelines completeness is capped at 2. It is not score 1 (the 'what' is present) and not score 3 (the 'when' is entirely missing rather than explicit).

2 / 3

Trigger Term Quality

The description 'General-purpose QA verdict for any artifact type' contains no natural user-facing trigger keywords; the natural terms ('ooo qa', 'qa check', 'quality check') appear only in the body, not the description. It is not score 2 because not even 'some relevant keywords' are present in the description itself.

1 / 3

Distinctiveness Conflict Risk

'QA verdict' gives it a recognizable niche, but 'General-purpose ... any artifact type' is broad and could overlap with other evaluation/QA skills. It is not score 1 (it is more specific than 'Helps with code and documents') and not score 3 (the broad scope raises conflict risk rather than carving a distinct trigger niche).

2 / 3

Total

7

/

12

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
Q00/ouroboros
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.