CtrlK
BlogDocsLog inGet started
Tessl Logo

dev-story

Implement a story: ADR guidelines, right programmer agent, code plus test. Then /story-done (/story-readiness before, /code-review after, at standard/full).

56

Quality

71%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/dev-story/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

77%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exceptionally actionable, well-gated orchestration skill: every phase carries concrete commands, exact prompts, and explicit validation with feedback loops, and it consistently pushes deep procedures out to one-level-deep .claude/docs references. Its weaknesses are rhetorical padding that could be tightened throughout, and a monolithic single-file layout that inlines routing tables, engine rosters, and per-engine standards that clearly belong in reference files.

DimensionReasoningScore

Conciseness

The body is almost entirely project-specific policy with no generic-concept teaching, but it carries noticeable rationale padding — extended justifications ("Telling a user to confirm the passing of tests that do not exist is worse than saying nothing"), the "omitted vs forgotten" principle stated at length in multiple places, and a stray double-period typo in Phase 3. Anchor 3 ("mostly efficient but includes some unnecessary explanation or could be tightened") fits: it is not anchor 2 because nothing explains concepts Claude already knows, and not anchor 4 because the rhetorical elaboration exceeds minor trimmable instances.

3 / 5

Actionability

Fully executable throughout: exact Grep invocations with pattern, path, output_mode and -A flags; exact per-engine verification commands (godot --headless -s parse-check, Unity -batchmode smoke, UnrealBuildTool/Build.sh per platform); verbatim AskUserQuestion prompts with enumerated options; exact file paths, exit-code semantics, and a copy-paste summary/checkpoint template. This matches anchor 5 — copy-paste-ready commands covering the common cases.

5 / 5

Workflow Clarity

Phases 1–7 are clearly sequenced with explicit validation checkpoints everywhere: file-existence gates with STOP/WARN semantics per tier, ADR status and version-mismatch handling, dependency status checks, exit-code-based parse verification, run-and-observe with retained screenshots, INCOMPLETE detection for agents that stop early, and a dedicated error recovery protocol requiring partial reports. This matches anchor 5 — explicit validation steps, feedback loops, and error recovery.

5 / 5

Progressive Disclosure

No bundle files exist (references/, scripts/, assets/ are all absent), so everything lives in one ~660-line SKILL.md. External references (.claude/docs/run-and-observe.md, error-recovery-protocol.md, automation-modes.md, code-root-resolution.md, test-standards.md) are one level deep and clearly signaled, but substantial material that belongs in separate files — agent routing tables, engine specialist rosters, per-platform build commands, per-engine test-naming standards — is inlined. Anchor 3 ("content that should be separate is inline") fits: not anchor 4 because the inline detail is more than a minor organization gap, not anchor 2 because section structure and reference signaling are good rather than minimal.

3 / 5

Total

16

/

20

Passed

Description

55%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description communicates a concrete, domain-specific capability with terse, action-oriented phrasing, but it lacks any explicit trigger guidance ("Use when...") and offers minimal keyword variety. Its distinctiveness from the other story-lifecycle skills rests on one verb, leaving moderate mis-trigger risk within that family.

Suggestions

Add an explicit 'Use when...' trigger clause, e.g. "Use when the user asks to implement, start, or continue work on a story" — this would lift completeness from 3 to 4–5 and satisfy the rubric's cap.

Include natural keyword variations users would actually say ("user story", "start the story", "continue story", "implement TR-XXX-NNN") to broaden trigger-term coverage.

Sharpen distinctiveness from sibling skills by stating the lifecycle position explicitly, e.g. "Implements a story after /create-stories writes it and before /story-done closes it" rather than relying on the verb alone.

DimensionReasoningScore

Specificity

The description lists several concrete actions — "Implement a story", "ADR guidelines", "right programmer agent", "code plus test" — plus the downstream command routing (/story-done, /story-readiness, /code-review). It falls short of anchor 5 because it omits significant capabilities (context loading, phase-gated verification, session-state updates), but exceeds anchor 3 since more than 1–2 specific actions are named.

4 / 5

Completeness

The "what" is clear (implement a story following ADR guidelines, routing to the right programmer agent, writing code plus test), but there is no "Use when..." clause or equivalent explicit trigger guidance — the sequencing of sibling commands only weakly implies when to invoke it. Per the rubric guideline, a missing 'Use when...' clause caps completeness at 3.

3 / 5

Trigger Term Quality

"Implement a story" is a natural phrase a user would say, but coverage is thin — no variations or synonyms such as "user story", "start work on", "continue the story", or "sprint task". Anchor 4 requires good keyword coverage with only a few natural terms missing; here most natural variants are absent, fitting anchor 3.

3 / 5

Distinctiveness Conflict Risk

The noun "story" is shared across several sibling skills (/create-stories, /story-readiness, /story-done), and the description relies on the single verb "Implement" to differentiate itself rather than explicit positioning. Anchor 3 ("could still overlap with similar skills") fits; it is not anchor 4 because the overlap with the other story-lifecycle skills is more than minor.

3 / 5

Total

13

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (677 lines); consider splitting into references/ and linking

Warning

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

13

/

16

Passed

Repository
Donchitos/Claude-Code-Game-Studios
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.