CtrlK
BlogDocsLog inGet started
Tessl Logo

designing-tests

Designs and implements testing strategies for any codebase. Use when adding tests, improving coverage, setting up testing infrastructure, debugging test failures, or when asked about unit tests, integration tests, or E2E testing.

59

Quality

74%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/designing-tests/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

56%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is highly actionable with concrete templates and commands, and the workflow has a real validation loop, but it spends significant tokens re-explaining testing fundamentals Claude already knows and keeps everything inline in one long file. Trimming known-concept sections and splitting per-language templates into reference files would improve it most.

Suggestions

Cut sections that re-teach concepts Claude already knows — the Testing Pyramid diagram, the When to Mock ✅/❌ list, and the What to Cover philosophy — and keep only the skill's own decisions (chosen frameworks, thresholds) to reduce conciseness padding.

Move the per-language framework tables and test structure templates into a references/ file (e.g., references/templates.md) with one-level-deep, clearly signaled links, keeping SKILL.md as a workflow overview.

Add matching pytest/testify example templates for the Python and Go stacks the skill recommends, since only JavaScript templates are currently provided.

DimensionReasoningScore

Conciseness

Multiple sections re-teach concepts Claude already knows: the ASCII "Testing Pyramid" diagram, the ✅/❌ "When to Mock" list, "What to Cover" coverage philosophy, standard Arrange/Act/Assert structure, and MSW/faker usage. Run commands also appear twice (validation checklist and bash block), and the two checklists overlap. This is several unnecessary/padded sections (anchor 2) — more than the "some" tightening of anchor 3, though the compact framework tables partially earn their place.

2 / 5

Actionability

Provides executable, copy-paste-ready templates for unit, integration, and E2E tests, concrete coverage-threshold config, test factories, and specific commands like `npm test -- --coverage`. Minor gaps keep it at anchor 4: only JavaScript templates are shown despite recommending pytest/testify/Playwright for Python and Go, so the recommended stacks lack matching examples.

4 / 5

Workflow Clarity

A clear 6-step sequenced checklist (identify → select type → write → run → check coverage → fix) plus a validation loop with an explicit feedback step ("If any tests fail, fix them before proceeding"). The steps themselves are generic (no detail on how to identify what to test or select a type), fitting anchor 4 (clear sequence, most checkpoints) rather than 5.

4 / 5

Progressive Disclosure

No bundle files exist and the entire ~255-line body is inline with good section headers. Content that would work better as separate references (framework selection tables, per-language test templates, mocking examples) is inlined rather than split out. Anchor 3 ("content that should be separate is inline") fits; structure is too well-organized for anchor 2, and the skill exceeds the simple <50-line exception for a 5.

3 / 5

Total

13

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with an explicit and well-phrased "Use when" trigger clause covering the main testing scenarios. The capability statement is slightly abstract compared to the best examples, and a few natural trigger variations are missing, but it clearly communicates both what the skill does and when to use it.

DimensionReasoningScore

Specificity

"Designs and implements testing strategies for any codebase" names the domain and 1-2 actions, but the capability statement is somewhat abstract ("testing strategies") rather than listing several concrete actions. Matches anchor 3; not anchor 4 because the what-clause lacks a list of specific actions like "writes unit/integration/E2E tests, configures coverage".

3 / 5

Completeness

Explicitly answers both what ("Designs and implements testing strategies for any codebase") and when ("Use when adding tests, improving coverage, setting up testing infrastructure, debugging test failures, or when asked about unit tests...") with concrete trigger phrases. This is a direct anchor 5 match; the when-clause is explicit and specific, so anchor 4's caveat does not apply.

5 / 5

Trigger Term Quality

Includes natural phrases users would say: "adding tests", "improving coverage", "debugging test failures", "unit tests, integration tests, or E2E testing". Missing common variations such as "write tests", "test suite", or "regression tests", so it fits anchor 4 (good coverage, a few natural terms missing) rather than 5.

4 / 5

Distinctiveness Conflict Risk

Testing is a distinct niche with clear, testing-specific triggers (coverage, unit/integration/E2E), giving minimal conflict risk. "For any codebase" is slightly broad and could overlap with code-review or CI-related skills, so anchor 4 rather than 5.

4 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
CloudAI-X/claude-workflow-v2
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.