CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-analyze-code-quality

Agent skill for analyze-code-quality - invoke with $agent-analyze-code-quality

55

1.49x
Quality

32%

Does it follow best practices?

Impact

94%

1.49x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/agent-analyze-code-quality/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

36%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The instructional core of this skill is a compact, sensible role prompt with a good output template, but it is buried under a ~125-line inlined YAML agent-config blob that wastes most of the body's tokens, and it never defines the analysis workflow or how to produce the scores its template demands. Removing the config blob and adding a short, sequenced analysis procedure would transform the file. The strongest immediate fix is deleting or externalizing the YAML block.

Suggestions

Delete the ~125-line embedded YAML config block (triggers, capabilities, constraints, hooks, optimization settings) from the body — it consumes the majority of the skill's tokens without instructing Claude on anything.

Add a short sequenced workflow (e.g., 1. glob allowed file types within allowed paths, 2. review each file against the criteria, 3. assign severity and score, 4. emit the report in the given template) so the output template is actually derivable.

Specify how to compute the report's quantitative fields ('Overall Quality Score: X/10', 'Technical Debt Estimate: X hours') so the guidance is executable rather than aspirational.

DimensionReasoningScore

Conciseness

Roughly 125 of the ~180 body lines are an embedded YAML agent-config blob (triggers, capabilities, constraints, hooks, memory limits, emoji usage) that instructs Claude on nothing — heavy padding that crowds the context window. The trailing markdown section is leaner, but it also re-lists code smells (long methods, duplicate code, god objects) that Claude already knows, fitting anchor 2 (noticeably verbose, several unnecessary/padded sections).

2 / 5

Actionability

There is some concrete guidance — code-smell thresholds (">50 lines", ">500 lines"), a five-criteria checklist, and a fully specified markdown output-report template — but the analysis process itself is never specified: how to select files, how to compute the 'Overall Quality Score: X/10', or how the 'Technical Debt Estimate: X hours' is derived. This matches anchor 3 (some concrete guidance but incomplete, missing key details).

3 / 5

Workflow Clarity

No sequenced steps exist anywhere in the body — only lists of responsibilities and criteria; the flow from analysis to report is merely implied by document order. The skill also declares batch operation (max_file_operations: 100, batch_size: 20) with no validation or verification checkpoint, which caps workflow_clarity at 3; the absence of any real sequence pulls it to anchor 2 ('rough sequence present but many gaps; validation absent').

2 / 5

Progressive Disclosure

No bundle files exist (references/, scripts/, assets/ are absent) and the body references no external files, so there is no nesting problem. The markdown section is reasonably organized with clear headers (Key responsibilities, Analysis criteria, Code smell detection, Review output format), but the 125-line YAML config blob is inlined content that clearly belongs in a separate config file, fitting anchor 3 (some structure but content that should be separate is inline).

3 / 5

Total

10

/

20

Passed

Description

28%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is boilerplate metadata that identifies the skill's domain but communicates no capabilities, no usage triggers, and no 'when to use' guidance. It would rarely be selected correctly because it contains none of the natural phrases a user would say when needing code quality analysis. A concrete action list plus an explicit 'Use when...' clause is the highest-leverage fix.

Suggestions

Replace the boilerplate with concrete third-person actions, e.g., 'Identifies code smells, evaluates complexity and maintainability, and generates a code quality report with refactoring recommendations.'

Add an explicit trigger clause: 'Use when the user asks for a code review, code quality analysis, refactoring suggestions, or technical debt assessment.'

Include the natural trigger terms users actually say (code review, refactor, code smell, technical debt, best practices) instead of the '$agent-analyze-code-quality' invocation jargon.

DimensionReasoningScore

Specificity

The description only names the domain ("Agent skill for analyze-code-quality") without listing any concrete actions — it does not say what the skill actually does (e.g., identify code smells, assess technical debt, generate reports). This matches anchor 2 ("Names the domain but actions are minimal or generic"); it is above anchor 1 because the domain itself is specifically identified, and below anchor 3 because zero concrete actions are stated.

2 / 5

Completeness

The 'what' is only weakly implied by the domain name ("Agent skill for analyze-code-quality") and the 'when' is entirely absent — there is no 'Use when...' clause or equivalent trigger guidance, which alone caps completeness at 3. This fits anchor 2 (vague 'what' and no 'when').

2 / 5

Trigger Term Quality

"analyze-code-quality" is one generic keyword phrase, but the only other term is the jargon invocation hint "invoke with $agent-analyze-code-quality", which no user would naturally say. Natural trigger phrases users would actually use ("code review", "refactor", "technical debt", "code smell", "best practices") are all missing, matching anchor 2.

2 / 5

Distinctiveness Conflict Risk

The named niche (code quality analysis) is somewhat specific, but the description establishes no distinguishing trigger phrases, so it would compete with generic code-review and refactoring skills. This fits anchor 3 ("Somewhat specific but could still overlap with similar skills"); it is above anchor 2 because the domain is not fully generic, and below anchor 4 because nothing in the description sets it apart from adjacent skills.

3 / 5

Total

9

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
ruvnet/ruflo
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.