CtrlK
BlogDocsLog inGet started
Tessl Logo

project-sizing-guide

Software project effort estimation assistant. Outputs three-point estimates (optimistic/most-likely/pessimistic values with confidence intervals), T-shirt sizes, or Function Point Analysis (FPA) counts. Triggered when users ask 'how long will this feature take,' need to assess project workload, perform PERT estimation, T-shirt sizing, FPA, sprint planning, or quote-based effort breakdowns.

70

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-structured skill: every method comes with executable tooling that matches the real bundled script, concrete numeric thresholds, and a report template. Its weaknesses are moderate redundancy of textbook material Claude already knows (basic PERT statistics, FPA definitions) and the absence of an explicit validate-and-revise loop in the estimation workflow.

Suggestions

Trim the explanations of concepts Claude already knows — the Variance/σ² definitions, the 'Where: O/M/P' definitions, and the 68/95/99.7 confidence-interval table — keeping only the decision rule of which interval to use for which audience.

Add an explicit feedback loop to the PERT steps, e.g., 'If spread ratio (P/O) > 4, return to the user to clarify requirements before finalizing the estimate,' so the rules of thumb act as checkpoints rather than commentary.

Move the FPA DET/RET complexity matrices and the T-shirt conversion tables into a references/ file (e.g., references/fpa-tables.md), keeping SKILL.md as an overview that links to them one level deep.

DimensionReasoningScore

Conciseness

Much of the body is genuinely useful reference data (FPA weight matrices, T-shirt size tables, adjustment-factor percentages), but it also explains concepts Claude already knows: 'Variance V = σ²', a confidence-interval table restating the empirical rule (68.3% = ±1σ, 95% = ±2σ), and definitions like 'O (Optimistic): Shortest duration assuming everything goes smoothly'. Mostly efficient with some unnecessary explanation that could be tightened, matching anchor 3 rather than the noticeably-padded anchor 2 or the minor-trim anchor 4.

3 / 5

Actionability

The three script invocations for PERT, tshirt, and fpa are copy-paste ready with complete JSON payloads, and I verified the bundled scripts/estimate_calculator.py implements exactly these flags and value tables. Together with the method-selection table, concrete thresholds (e.g., 'O should not be less than 30% of M'), and a fill-in output template, the common cases are fully covered — the top anchor.

5 / 5

Workflow Clarity

The Quick Start gives a clear 4-step sequence, each method has numbered steps, and the O/M/P rules of thumb and adjustment-factor checklist function as verification checkpoints. However, there is no explicit feedback loop (e.g., 'if spread ratio > 4, clarify requirements and re-estimate'), so it fits anchor 4 (most checkpoints present, minor gaps) rather than anchor 5's explicit validate-then-retry pattern.

4 / 5

Progressive Disclosure

The single bundle file (scripts/estimate_calculator.py) is clearly signaled with executable usage in two places, and the body is well-sectioned with a navigation-friendly method-selection table up front. Scored against the actual bundle (one script, no references/), the remaining gap is that ~290 lines of FPA DET/RET lookup matrices and T-shirt conversion tables are inlined where a references/ file would keep the overview lean — good structure with minor organization gaps, i.e., anchor 4 not 5.

4 / 5

Total

16

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states concrete outputs for all three estimation methodologies, includes a verbatim natural user trigger phrase plus explicit 'Triggered when' guidance, and occupies a clear niche. The only weakness is that broadly-applicable trigger terms like 'sprint planning' create minor conflict risk with adjacent planning skills.

DimensionReasoningScore

Specificity

Enumerates multiple concrete capabilities covering its full method space: 'Outputs three-point estimates (optimistic/most-likely/pessimistic values with confidence intervals), T-shirt sizes, or Function Point Analysis (FPA) counts.' All three supported methods are named with their concrete outputs, matching the comprehensive-coverage anchor rather than the minor-gaps anchor at 4.

5 / 5

Completeness

Explicitly answers both questions: what it does ('Outputs three-point estimates... confidence intervals... FPA counts') and when to use it via a concrete 'Triggered when users ask...' clause with quoted trigger phrases. This matches the anchor requiring clear and explicit what AND when with concrete trigger phrases.

5 / 5

Trigger Term Quality

Contains natural phrases a user would actually say — "how long will this feature take", "assess project workload" — plus methodology keywords 'PERT estimation, T-shirt sizing, FPA, sprint planning, or quote-based effort breakdowns'. Coverage includes a verbatim user utterance, synonyms, and all methodology names, so it sits at the top anchor rather than merely 'good coverage with a few missing'.

5 / 5

Distinctiveness Conflict Risk

The effort-estimation niche with method-specific triggers (PERT, FPA, T-shirt sizing) is clearly distinct, but 'sprint planning' and 'assess project workload' are broad phrases that could overlap with general agile-planning or project-management skills — minor overlap risk with closely related skills, fitting the anchor 4 rather than 5.

4 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
zebbern/claude-code-guide
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.