CtrlK
BlogDocsLog inGet started
Tessl Logo

estimate

Estimate task effort from complexity, dependencies, velocity, risk. Structured estimate with confidence levels.

56

Quality

70%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/estimate/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-sequenced, lean, instruction-only workflow with a concrete output template, an explicit no-sprint-history fallback, and useful estimation guidelines. Main gaps are minor: a few assessment steps lack method, and there is no self-check on the generated estimate (e.g. breakdown total vs. recommended budget).

Suggestions

Add a final self-check before declaring COMPLETE: verify the suggested-breakdown total equals the recommended budget and that confidence is Low whenever the approach is undecided.

Give one line of method for the vague assessment steps — e.g. what counts as an integration point, or how to gauge coupling when scanning affected files.

Move the Guidelines section before the Phase 5 next-steps list so the rounding/range rules are seen before the template is filled, and remove overlap with the template's implicit constraints.

DimensionReasoningScore

Conciseness

The body is lean: factor checklists, a copy-paste output template, and concrete guidelines, with no explanation of concepts Claude already knows. It is not a 5 because there is minor redundancy — the Guidelines section restates constraints partly visible in the template ("Always give a range" vs. the template's optimistic/expected/pessimistic rows), and a few checklist items could be tightened. It is well above anchor 3 since essentially all content is non-obvious, skill-specific instruction.

4 / 5

Actionability

Guidance is concrete and executable for an instruction-only skill: named sources to read (CLAUDE.md, `design/gdd/`, `production/sprints/`), a full output template with tables, and specific rules like "Round to half-day increments" and "The recommended budget should be the expected estimate". It falls short of anchor 5 because a few steps are direction without method — e.g. "Assess complexity (size, dependency count, cyclomatic complexity)" and "Identify integration points" give no indication of how to measure or what counts — leaving minor gaps.

4 / 5

Workflow Clarity

Five clearly sequenced phases with an early clarification checkpoint ("If the description is too vague... ask for clarification before proceeding"), an explicit fallback when sprint history is missing, and conditional next steps in Phase 5. This is a read-only skill, so the destructive-operation validation cap does not apply. It is not anchor 5 because the workflow is largely linear — there is no checkpoint validating the estimate itself (e.g. verifying the breakdown total matches the recommended budget) before declaring COMPLETE.

4 / 5

Progressive Disclosure

The skill has no bundle files and the single-file structure is appropriate: well-organized phase sections, all content at one level, no nested or dangling references. The paths it mentions (`design/gdd/`, `production/sprints/`) are project data sources, not skill bundle files, and no `references/`, `scripts/`, or `assets/` directories exist to verify. It is not anchor 5 because at ~130 lines the large output template dominates the body and the Guidelines section sits at the tail after Phase 5; minor reorganization (leading with the guidelines, or trimming template boilerplate rows) would improve navigation.

4 / 5

Total

16

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise and third-person, with a clear statement of what the skill does and its key inputs, but it entirely lacks a "when to use" trigger clause and misses natural synonyms users would say. Adding trigger guidance and common estimation phrasings would raise it substantially.

Suggestions

Add an explicit trigger clause, e.g. "Use when the user asks how long a task will take, wants an effort/time estimate, or needs sizing for planning."

Include natural synonyms and phrasings users would actually say: "how long will this take", "time estimate", "sizing", "story points".

Spell out what the structured estimate contains (optimistic/expected/pessimistic range, breakdown, risk factors) so the deliverable is concrete rather than just "structured estimate".

DimensionReasoningScore

Specificity

The description names one concrete action ("Estimate task effort") plus its inputs ("complexity, dependencies, velocity, risk") and output shape ("Structured estimate with confidence levels"), but stops short of listing several specific actions. It sits between anchor 3 ("names domain and 1-2 concrete actions, but not comprehensive") and anchor 4 ("lists several specific actions"), closer to 3 because the actual deliverables (range, breakdown, risk table) are only implied by "structured estimate".

3 / 5

Completeness

The "what" is clear — estimate task effort and produce a structured estimate with confidence levels — but there is no "Use when..." clause or equivalent trigger guidance anywhere in the description. Per the judging guidelines, a missing explicit trigger clause caps completeness at 3; the description fits anchor 3 ("has a clear 'what' but 'when' is missing or only weakly implied").

3 / 5

Trigger Term Quality

Relevant keywords like "estimate", "effort", "complexity", "risk", and "confidence" are present, but common natural phrasings users would say are missing — "how long will this take", "time estimate", "sizing", "story points". This matches anchor 3 ("some relevant keywords but missing common variations or synonyms") rather than 4, since several obvious synonyms are absent.

3 / 5

Distinctiveness Conflict Risk

"Estimate task effort from complexity, dependencies, velocity, risk" carves out a fairly distinct niche (effort estimation) with differentiated inputs like velocity and confidence levels. It is mostly distinct with only minor overlap risk against closely related planning skills, matching anchor 4; it is not anchor 5 because the phrase "complexity, dependencies, risk" is generic enough that it could weakly overlap with code-review or risk-analysis skills.

4 / 5

Total

13

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
Donchitos/Claude-Code-Game-Studios
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.