CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-integration-testing

Use when the user requests integration testing, feature validation, or test plan execution

55

Quality

61%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

High

Do not use without reviewing

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/agent-integration-testing/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is a well-structured, largely executable instruction skill: a concrete five-step process, an explicit anti-pattern section, and a realistic example test spec. Its main weaknesses are minor duplication between the Overview, Core Process, and Quick Reference, and the absence of a post-failure re-validation loop.

DimensionReasoningScore

Conciseness

The body is largely lean and assumes competence (no concept explanations, no library tutorials), but the Overview paragraph and the Quick Reference table partially restate the Core Process steps. These are minor trimmable instances matching anchor 4 rather than the every-token-earns-its-place ideal of 5.

4 / 5

Actionability

Concrete guidance throughout: exact file path './tests/<name>.md', a mandatory Prerequisites section, and a copy-paste-ready example spec with a POST request and a sqlite3 verification query. Minor gaps — subagent spawning is described loosely ('using the Task tool or @mention subagent system') with no invocation example, and the results-collection format is unspecified — keep it at anchor 4.

4 / 5

Workflow Clarity

A clear 5-step sequence with expectations validated by subagents and Pass/Fail results collected; the Red Flags section adds error-recovery guidance. It falls short of anchor 5 because there is no explicit feedback loop after failures — the fix step is optional and user-gated, with no re-run/re-validate checkpoint.

4 / 5

Progressive Disclosure

No bundle files exist, and none are needed: the ~64-line body is cleanly sectioned (Overview, Core Process, Quick Reference, Red Flags, Example) with the single short example appropriately inlined. Nothing belongs in a separate file and there are no buried or nested references.

5 / 5

Total

17

/

20

Passed

Description

43%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description has a clean, explicit trigger clause with reasonably natural keywords, but it is trigger-only: it never states what the skill does. Adding a concrete action statement before the 'Use when' clause would raise both completeness and specificity substantially.

Suggestions

Add a 'what' clause before the trigger, e.g. "Creates and executes agent-verifiable integration test specs via subagents. Use when...", to fix the completeness gap (currently trigger-only).

List 2-3 concrete actions in the opening (write test specs to ./tests/<name>.md, spawn subagents per test, collect Pass/Fail results) to raise specificity from domain-naming to action-describing.

Include common trigger synonyms such as "run integration tests", "end-to-end testing", or "validate this feature" to improve keyword coverage and reduce missed invocations.

DimensionReasoningScore

Specificity

The description names the domain ("integration testing, feature validation, or test plan execution") but lists no concrete actions — what the skill actually does is entirely unstated. It matches 'names the domain but actions are minimal or generic' rather than the vague-anchor 1, and lacks the 1-2 concrete actions needed for 3.

2 / 5

Completeness

Only the 'when' half is present ("Use when the user requests..."); the 'what' — creating and executing verifiable test specs via subagents — never appears. This matches anchor 2's example ('Use when working with documents') exactly; anchor 3 requires a clear 'what', which is absent.

2 / 5

Trigger Term Quality

"integration testing", "feature validation", and "test plan execution" are natural phrases a user would say, giving good keyword coverage. Common variations like "run integration tests", "e2e testing", or "validate this feature" are missing, so it falls short of the comprehensive synonym coverage of anchor 5.

4 / 5

Distinctiveness Conflict Risk

"Integration testing" and "feature validation" carve out a somewhat specific niche, but the phrases would also plausibly trigger general test-execution or verification skills. It is not 'mostly distinct' enough for anchor 4 given the broad overlap with any testing/QA-style skill.

3 / 5

Total

11

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
av/harbor
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.