CtrlK
BlogDocsLog inGet started
Tessl Logo

tdd

Test-driven development with red-green-refactor loop. Use when user wants to build features or fix bugs using TDD, mentions "red-green-refactor", wants integration tests, or asks for test-first development.

82

1.18x
Quality

75%

Does it follow best practices?

Impact

96%

1.18x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/tdd/SKILL.md

The canonical home for this skill is tdd in mattpocock/skills

SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, opinionated TDD skill with a clear vertical-slice workflow, strong checklists, and commendable brevity. Its biggest defect is that all five of its reference links point to files that do not exist in the bundle, leaving the promised examples and guidelines unreachable, and the workflow lacks recovery guidance for cycles that get stuck in RED.

Suggestions

Create the five referenced files (tests.md, mocking.md, deep-modules.md, interface-design.md, refactoring.md) or remove/inline the dangling links — currently every external reference in the body is broken.

Add one short worked example of a behavior-focused test versus an implementation-coupled test, since the Philosophy section currently gestures at the distinction without showing it.

Add error-recovery guidance for the incremental loop, e.g., what to do when a test cannot reach GREEN after a few attempts (delete and rethink the test, or revisit the interface design).

DimensionReasoningScore

Conciseness

The body is dense and directive with no padding — the anti-pattern section ('DO NOT write all tests first...') delivers non-obvious judgment efficiently. The Philosophy section spends three paragraphs explaining good vs. bad tests, some of which (mocking, private methods) Claude already knows, so it is not quite anchor 5's 'every token earns its place' but comfortably above anchor 3.

4 / 5

Actionability

As an instruction-only skill the guidance is concrete and executable: 'Write ONE test that confirms ONE thing', explicit rules ('Only enough code to pass current test', 'Never refactor while RED'), a per-cycle checklist, and a literal question to ask the user. It stays at anchor 4 rather than 5 because there is no worked example of a behavior-focused test, and the files that would have carried examples (tests.md, mocking.md) do not exist in the bundle.

4 / 5

Workflow Clarity

The sequence (Planning with user approval → Tracer Bullet → Incremental Loop → Refactor) is clearly staged, the RED→GREEN cycle is itself a validate→fix feedback loop, and checklists plus 'Run tests after each refactor step' provide checkpoints. It falls short of anchor 5 because there is no error-recovery guidance for a cycle that cannot reach GREEN (e.g., what to do when stuck).

4 / 5

Progressive Disclosure

The body links five files (tests.md, mocking.md, deep-modules.md, interface-design.md, refactoring.md), but no bundle files exist at all — every reference is dangling, so navigation is broken despite the links being inline and clearly signaled. Scoring against the actual bundle structure, this lands at anchor 3 ('structure present but references do not enable navigation'), not anchor 4, whose references resolve.

3 / 5

Total

15

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with an explicit and multi-trigger 'Use when...' clause and good natural keyword coverage. Its main limitation is that the 'what' half relies on naming the methodology rather than enumerating the concrete actions the skill performs, and the 'integration tests' trigger slightly widens conflict risk toward general testing skills.

DimensionReasoningScore

Specificity

The description names the domain and one concrete practice ("Test-driven development with red-green-refactor loop") but lists no further specific actions like writing failing tests first or refactoring after green. This matches anchor 3 ('names domain and 1-2 concrete actions, but not comprehensive') and falls short of anchor 4, which expects several specific actions.

3 / 5

Completeness

It explicitly answers both questions: the 'what' ("Test-driven development with red-green-refactor loop") and an explicit 'Use when...' clause with multiple concrete trigger phrases. This matches anchor 5; the 'when' clause is fully explicit, so anchor 4's caveat ('when could be more explicit') does not apply.

5 / 5

Trigger Term Quality

Natural trigger phrases are well covered: "build features or fix bugs using TDD", "mentions 'red-green-refactor'", "wants integration tests", "asks for test-first development". A few common variants are missing (e.g., 'write the tests first', 'unit tests'), so it fits anchor 4 rather than anchor 5's comprehensive synonym coverage, and is clearly above anchor 3's partial coverage.

4 / 5

Distinctiveness Conflict Risk

Triggers like 'red-green-refactor' and 'test-first development' carve out a clear TDD niche, but 'wants integration tests' overlaps with general testing skills. This is anchor 4 ('mostly distinct; minor overlap risk with closely related skills') rather than anchor 5's minimal-conflict profile.

4 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 5 missing

Warning

Total

15

/

16

Passed

Repository
openstatusHQ/data-table-filters
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.