CtrlK
BlogDocsLog inGet started
Tessl Logo

progressive-estimation

Estimate AI-assisted and hybrid human+agent development work with research-backed PERT statistics and calibration feedback loops

49

Quality

53%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/AI-Agents-Safe-Coding-Skills-claude/skills/progressive-estimation/SKILL.md

The canonical home for this skill is progressive-estimation in administrakt0r/AI-Agents-Safe-Coding-Skills

SKILL.md
Quality
Evals
Security

Quality

Content

53%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is well-structured and concise, but its actionability is weak: it names a PERT/multiplier methodology without ever giving the formulas or multiplier values that are the skill's core value. The batch workflow also lacks explicit validation checkpoints.

Suggestions

Provide the actual PERT expected-value formula (e.g. E = (O + 4M + P)/6) and standard deviation, plus a concrete table of the 'research-backed multipliers' for human-only, hybrid, and agent-first modes.

Show a worked example end-to-end (inputs -> PERT calculation -> P50/P75/P90 bands -> formatted output) instead of only quoting sample prompts.

Add a validation/verification checkpoint in the batch workflow, e.g. confirm each issue has size/complexity/risk classified before computing, and flag outliers whose P90 exceeds a threshold.

DimensionReasoningScore

Conciseness

The body is efficiently organized into focused sections with only minor fluff ('produces statistical estimates rather than gut feelings'), matching anchor 4; not 5 because a few descriptive sentences could be trimmed, not 3 because it largely assumes Claude's competence without explaining basics.

4 / 5

Actionability

The body describes the process (mode detection, PERT calculation, 'research-backed multipliers', confidence bands) but never provides the actual PERT formula, the multipliers, or the confidence-band math, so it gives high-level hints while missing the specific steps to execute, matching anchor 2; not 3 because no executable methodology or pseudocode is supplied for the core computation.

2 / 5

Workflow Clarity

'How It Works' lists 7 clearly sequenced steps, but the skill performs batch operations ('handles 5 or 500 issues') with no validation/verification checkpoint inside the workflow, so per the rubric's batch-operation cap it cannot exceed 3; not 4 because explicit checkpoints are absent.

3 / 5

Progressive Disclosure

Sections are well organized and external links (source repo, installation, research references) are clearly signaled, but no bundle files exist and the core research/formulas are only linked externally rather than split into one-level-deep local references, matching anchor 4; not 5 because there are no well-signaled local reference files to navigate to.

4 / 5

Total

13

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description clearly states what the skill does and occupies a distinct niche, but lacks any explicit 'Use when...' trigger guidance and leans on jargon (PERT, calibration) over the natural phrases users actually say. Specificity is adequate but not comprehensive.

Suggestions

Add an explicit trigger clause, e.g. 'Use when estimating AI-assisted or hybrid development tasks, sprint planning with agents, or forecasting release dates with confidence intervals.'

Include natural user phrases such as 'how long will this take', 'story points', 'sprint planning', and 'time estimate' alongside the technical PERT/calibration terms.

List a couple more concrete actions (e.g. 'size a backlog', 'forecast release dates', 'calibrate past estimates') to lift specificity from one main verb to several.

DimensionReasoningScore

Specificity

Names the domain ('AI-assisted and hybrid human+agent development work') and a concrete action ('Estimate') with method modifiers (PERT statistics, calibration feedback loops), but those are features rather than several distinct actions, matching anchor 3; not 4 because no list of multiple specific actions, not 2 because the domain and a real action are present.

3 / 5

Completeness

The 'what' is clear (estimate AI/hybrid dev work with PERT and calibration), but there is no 'Use when...' clause or equivalent trigger guidance, so per the rubric completeness is capped at 3; not 2 because the 'what' is clear rather than vague.

3 / 5

Trigger Term Quality

'Estimate', 'AI-assisted', and 'hybrid' are natural terms, but common variations users say ('sprint planning', 'story points', 'how long will this take') are missing and 'PERT statistics'/'calibration feedback loops' are technical jargon, matching anchor 3; not 4 due to the missing common synonyms.

3 / 5

Distinctiveness Conflict Risk

It carves a clear niche (AI-assisted PERT estimation with calibration) with only minor overlap risk against general estimation skills, matching anchor 4; not 5 because the absent 'when' triggers leave slightly more conflict risk than the anchor-5 example.

4 / 5

Total

13

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
administrakt0r/AI-Agents-Safe-Coding-Skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.