CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-cost-projections

Project remaining workflow cost from per-phase averages — warns on budget ceiling overruns

53

Quality

60%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/skill-cost-projections/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

67%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured, largely executable six-step procedure with good error-handling coverage and concrete display formats. Its weaknesses are redundant repeated display examples, an undefined `total_steps` variable that leaves the core projection not fully copy-paste ready, and slightly excessive length for a single-procedure skill.

Suggestions

Define or derive `total_steps` explicitly (e.g. show how to read it from the workflow config or the Step 3 table) so the Step 3 code is actually runnable.

Remove the verbatim repetition of the HUD display lines — show each format once and consolidate the 'Complete Display Examples' section into a compact matrix of state → output.

Drop the unused `completed_costs` jq call and guard `add` against an empty metrics file (e.g. `[.[].cost] | add // 0`) to keep the scripts copy-paste safe.

DimensionReasoningScore

Conciseness

Steps are compact and formula-driven, but the same HUD display lines ("💰 Spent: $2.40 | Est. remaining: $3.60 | Total: ~$6.00") are repeated verbatim in Step 4, Step 5, Step 6, and again in the 'Complete Display Examples' section, and formulas are restated in prose examples. Anchor 3: mostly efficient but with redundant sections that could be trimmed.

3 / 5

Actionability

Concrete, near-executable bash using jq and bc, with real format rules, env vars (OCTO_BUDGET_CEILING, OCTO_COST_THRESHOLD), and per-workflow step-count tables. Not 5: `completed_costs` is computed but never used, `total_steps` is referenced in Step 3 without ever being defined or derived, and `[.[].cost] | add` yields null on an empty metrics file — minor but real gaps in copy-paste readiness.

4 / 5

Workflow Clarity

Steps 1-6 are clearly sequenced with an explicit insufficient-data checkpoint (<2 steps → actual spend only) and a dedicated Error Handling section covering missing metrics, corrupted entries, and zero remaining steps. Not 5: there is no validate-and-retry loop for corrupted data (only 'skip malformed entries'), and data-source fallback priority is stated but not enforced with checks.

4 / 5

Progressive Disclosure

A self-contained skill with no bundle files and no nested references; all content sits one level deep under clear section headers, and the file listing confirms no references/ or scripts/ to navigate. Not 5: at ~210 lines it exceeds the simple-skill threshold, and the consolidated display examples plus integration sections could be tightened or split.

4 / 5

Total

15

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise, specific, and grammatically clean, giving a clear picture of what the skill does. Its main weakness is that all trigger guidance lives in a separate frontmatter `trigger` field rather than the description, leaving the description itself without a "when to use" clause and thin on natural user phrasings.

Suggestions

Append a 'Use when...' clause to the description itself, e.g. 'Use when the user asks to project remaining workflow cost, forecast spend, or check whether they are over budget.'

Fold one or two natural user phrasings into the description (e.g. 'how much will this cost', 'am I over budget') instead of confining them to the trigger field.

Optionally mention the HUD/statusline display output in the description to further distinguish it from generic cost-lookup skills.

DimensionReasoningScore

Specificity

"Project remaining workflow cost from per-phase averages" and "warns on budget ceiling overruns" name the domain plus 1-2 concrete actions, matching anchor 3. Not 4: it does not list several specific actions (no mention of the HUD display, metrics sourcing, or profile suggestions).

3 / 5

Completeness

The description clearly answers "what" (projects remaining cost from per-phase averages, warns on overruns) but contains no "Use when..." clause or equivalent trigger guidance within the description string, which caps completeness at 3 per the judging guidelines. Not 2: the "what" is concrete, not vague.

3 / 5

Trigger Term Quality

The description carries "cost", "budget ceiling", and "overruns" — some relevant keywords — but the natural phrases users would say ("how much will this cost", "am I over budget", "spending forecast") appear only in the separate frontmatter `trigger` field, not in the description itself. Anchor 3: relevant keywords but missing common variations.

3 / 5

Distinctiveness Conflict Risk

"Per-phase averages" and "budget ceiling overruns" carve a clear niche (in-flight workflow cost projection) that few skills share, with minor overlap risk against generic API-pricing or cost-lookup skills. Not 5: the description alone doesn't state the workflow-session scope that would fully exclude general cost questions.

4 / 5

Total

13

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.