CtrlK
BlogDocsLog inGet started
Tessl Logo

effort

Estimates the implementation effort required to address the given issue.

58

Quality

66%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./tools/caretaker-agent/cloudrun/triage-worker/.gemini/skills/effort/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured, actionable classification rubric that gives Claude a concrete output schema and specific, project-grounded examples for each effort level. Its main weakness is the absence of any explicit sequencing or decision tie-breakers beyond the final Note.

Suggestions

Add a brief decision tie-breaker: when signals span two levels, default to the higher and justify in effort_reasoning.

State the single-step process explicitly (1. Read issue + code exploration; 2. Match to a level; 3. Return JSON) to push workflow clarity to 5.

Clarify what 'code exploration output' refers to (the upstream skill's file listing) so the input contract is unambiguous.

DimensionReasoningScore

Conciseness

The body is efficient and assumes Claude's competence, with bullet examples earning their place as project-specific classification criteria rather than padding; only minor trimming opportunities exist, matching the 'efficient; minor instances of over-explanation' anchor.

4 / 5

Actionability

It provides a concrete JSON output format and specific per-level criteria with real, recognizable project examples (Zod, node-pty, ENXIO, A2A), giving mostly executable guidance with only minor gaps, fitting the 4 anchor.

4 / 5

Workflow Clarity

The single task is unambiguous (analyze issue + code exploration, emit JSON) and the operation is non-destructive so no validation cap applies, but the process is described as a single sentence rather than an explicit sequence, keeping it just below 5.

4 / 5

Progressive Disclosure

The skill is under 50 lines, needs no external references, and is organized into clear sections (JSON Output Format, Effort Levels, Note), meeting the simple-skill exception that allows a 5 with well-organized sections.

5 / 5

Total

17

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is clear and reasonably specific about what the skill does, but it omits any 'when to use' trigger guidance, which limits its completeness and trigger discoverability. Adding a 'Use when...' clause with natural user phrases would lift it toward the top anchors.

Suggestions

Add an explicit 'Use when...' clause with natural trigger phrases (e.g. 'Use when sizing an issue before implementation, estimating effort, or triaging backlog work').

Include common synonyms users actually say ('estimate effort', 'size this task', 'how big is this issue') to improve trigger term coverage.

Optionally name the output (a SMALL/MEDIUM/LARGE estimate) to sharpen the 'what' beyond the generic 'estimates the implementation effort'.

DimensionReasoningScore

Specificity

The phrase 'Estimates the implementation effort required to address the given issue' names the domain (effort estimation for issues) and one concrete action (estimating effort), but offers no further actions, matching the 'names domain and 1-2 concrete actions, not comprehensive' anchor.

3 / 5

Completeness

There is a clear 'what' (estimates implementation effort) but no 'Use when...' clause or equivalent trigger guidance, so per the rubric completeness is capped at 3 ('clear what, when missing or weakly implied').

3 / 5

Trigger Term Quality

Terms like 'implementation effort' and 'issue' are relevant, but common natural variations a user might say (e.g. 'estimate effort', 'size this task', 'how big is this') are missing, fitting the 'some relevant keywords but missing common variations' anchor.

3 / 5

Distinctiveness Conflict Risk

Effort estimation for a given issue is a fairly distinct niche with only minor overlap risk against planning or code-review skills, matching the 'mostly distinct; minor overlap risk' anchor rather than the broader 3.

4 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
google-gemini/gemini-cli
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.