CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-tester

Agent skill for tester - invoke with $agent-tester

50

1.14x
Quality

26%

Does it follow best practices?

Impact

88%

1.14x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/agent-tester/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

28%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body reads as a generic testing tutorial padded with knowledge Claude already has, plus a QA-agent config block. Code examples are structurally useful but corrupted with escaped characters and pseudo-MCP syntax, and there is no stepwise workflow with validation/feedback loops telling the agent how to actually run and iterate on tests.

Suggestions

Delete the textbook material (test pyramid diagram, F.I.R.S.T. list, definitions of unit/integration/E2E, generic Jest scaffolding, and the best-practices list) and keep only project-specific conventions such as the coverage thresholds and the memory-coordination keys.

Fix the corrupted examples so they are executable: restore '/users', '/register', '</script>' escapes, and replace 'mcp__claude-flow__memory_usage { ... }' pseudo-syntax with actual tool call format; also fix the post hook's '2>$dev$null' shell.

Replace the capability list with an explicit operating workflow with validation checkpoints, e.g.: run suite -> parse failure report from memory/store -> fix failing tests -> re-run until green -> store final results via memory coordination.

DimensionReasoningScore

Conciseness

The ~290-line body extensively teaches concepts Claude already knows: the test pyramid diagram, what unit/integration/E2E tests are, basic Jest scaffolding, F.I.R.S.T. characteristics ('Fast', 'Isolated', 'Repeatable'), and generic best practices like 'Descriptive Names' and 'Arrange-Act-Assert'. This is 'severely verbose; extensively explains concepts Claude already knows; heavily padded' — anchor 1 — since nearly every section re-states textbook testing knowledge.

1 / 5

Actionability

There is real, concrete guidance (Jest test structures, coverage thresholds, MCP coordination snippets), but much of it is not executable as written: URLs corrupted to '$users' and '$register', broken shell like 'npm test -- --reporter=json 2>$dev$null', pseudo-syntax 'mcp__claude-flow__memory_usage { ... }' that is not a real tool invocation, and undefined helpers (validate(), processItems(), sanitizeInput(), generateItems()). This matches anchor 3 ('some concrete guidance but incomplete; pseudocode instead of executable code') rather than 4, whose code is copy-paste runnable.

3 / 5

Workflow Clarity

There is no operating sequence for the agent at all — 'Core Responsibilities' is an unordered capability list, not steps, and there are no validation checkpoints (run tests, interpret failures, fix, re-run) despite testing being an inherently feedback-loop-driven task. This fits anchor 2 ('rough sequence present but many gaps; validation absent'), and not anchor 1 only because the best-practices list and post-hook sketch a rough sense of what the agent does.

2 / 5

Progressive Disclosure

The body has reasonable section structure (Testing Strategy, Test Quality Metrics, Performance, Security, MCP Integration) but is a monolithic 290-line document with no bundle files and no references — all example corpora and MCP details are inlined where they could be split. This matches anchor 3 ('some structure but could be better organized; content that should be separate is inline') rather than 2, since section headers do exist and navigation within the body is possible.

3 / 5

Total

9

/

20

Passed

Description

25%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The frontmatter description is a stub: it names a role but no capabilities, no use-when trigger clause, and no natural search terms. The only substantive capability statement ('Comprehensive testing and quality assurance specialist') sits in the body's second YAML block where it does not function as the skill description.

Suggestions

Rewrite the frontmatter description to list concrete capabilities, e.g. 'Writes and runs unit, integration, e2e, performance, and security test suites; analyzes coverage and reports results.'

Add an explicit trigger clause: 'Use when the user asks to write or run tests, improve coverage, fix failing tests, or set up a test framework.'

Move the capability summary currently buried in the body's second YAML block (e.g., 'Comprehensive testing and quality assurance specialist' plus the capabilities list) into the frontmatter description itself.

DimensionReasoningScore

Specificity

The description 'Agent skill for tester - invoke with $agent-tester' names the domain (tester) but lists zero concrete capability actions — nothing about writing unit/integration/E2E tests, coverage, or QA validation. It matches 'Names the domain but actions are minimal or generic' rather than anchor 1, since a domain is at least named, and not anchor 3, which requires 1-2 explicit concrete actions.

2 / 5

Completeness

The 'what' is vague ('Agent skill for tester') and there is no 'when' clause at all — no 'Use when...' trigger guidance, which caps completeness at 3, and the vague 'what' pushes it to anchor 2 ('Has a vague what and no when'). The richer capability description exists only in the body's second YAML block, not in the frontmatter description field.

2 / 5

Trigger Term Quality

The only keyword is 'tester' plus a self-referential invocation token; users would naturally say 'write tests', 'testing', 'QA', 'unit tests', or 'test coverage'. This is one or two generic keywords missing the natural phrases users actually say — anchor 2 — and not anchor 3, which requires some relevant domain keywords beyond a bare role label.

2 / 5

Distinctiveness Conflict Risk

'Agent skill for tester' is very broad and would overlap with any testing, QA, or validation-related skill. It matches 'Very broad; high overlap risk with many similar skills' and lacks the distinct concrete triggers (e.g., 'run the test suite', 'improve coverage') that would earn anchor 4 or 5.

2 / 5

Total

8

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
ruvnet/ruflo
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.