CtrlK
BlogDocsLog inGet started
Tessl Logo

octopus-quick

Quick execution for ad-hoc tasks without full workflow overhead — use for small, self-contained requests

48

Quality

51%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/octopus-quick/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

50%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content provides a clear, actionable 5-step execution recipe with concrete commands and good escalation guidance, but it is roughly 3-4x longer than needed. Padding (benefits, comparison, best-practices, duplicate examples) dilutes the core workflow, and there is no verification step between making a change and committing it.

Suggestions

Cut the "Benefits", "Comparison", "Directory Structure", and "Summary" sections — they restate information evident from the workflow itself and could cut the file by more than half.

Add an explicit verification checkpoint between making the change and committing (e.g. "re-read the diff and run any affected tests before committing") so the workflow has a feedback loop.

Move the troubleshooting, comparison table, and complete worked example into a references/ file (or fold the essential parts into the step instructions), and fix the $TASK_DESCRIPTION placeholder in the summary heredoc.

DimensionReasoningScore

Conciseness

The body is noticeably verbose for a simple 5-step process: sections like "Benefits of Quick Mode" ("No multi-AI orchestration overhead", "Only uses Claude"), the Comparison table, "Directory Structure", "Best Practices", and a "Summary" restating the introduction add little Claude doesn't already know. It avoids generic concept explanations (so above 1) but contains several padded sections, matching the 2-anchor.

2 / 5

Actionability

The guidance is mostly executable: a concrete git commit message template, specific state-manager.sh invocations, and a copy-paste summary-generation heredoc with datestamped filenames. Minor gaps keep it below 5: $TASK_DESCRIPTION in the heredoc is an undefined variable and the metrics value "1" is a placeholder requiring judgment.

4 / 5

Workflow Clarity

The 5-step sequence (understand → change → commit → record state → generate summary) is clearly listed and even includes escalation indicators for scope growth. However, there is no validation checkpoint after making the change (no "confirm the diff / run affected tests" step), matching 'steps listed but validation gaps; checkpoints missing or implicit'.

3 / 5

Progressive Disclosure

The body has section structure but is a ~360-line monolith: the comparison table, troubleshooting, directory layout, and complete worked example are inlined content that belongs in separate reference files, and the host note cites skills/blocks/codex-host-adapter.md, a path not present in this bundle. This matches 'some structure but could be better organized; content that should be separate is inline'.

3 / 5

Total

12

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise and correctly answers both what the skill does and when to use it, but it stays at a generic level. It names no concrete task actions and few natural trigger phrases, limiting its discoverability and distinctiveness.

Suggestions

List 2-3 concrete action types in the description, e.g. "Execute small one-file fixes, config changes, and tiny refactorings directly with atomic commits".

Add natural trigger phrasings users would say, such as "quick fix", "simple change", or "just update X", to improve keyword coverage.

Sharpen the when-clause with concrete task categories (bug fix, typo, config tweak, dependency bump) so it is distinguishable from general task-execution skills.

DimensionReasoningScore

Specificity

"Quick execution for ad-hoc tasks without full workflow overhead" names the domain but the only action is the generic verb "execution" — no concrete actions (fix, edit, update, refactor) are listed, matching the anchor 'names the domain but actions are minimal or generic' rather than the 3-anchor which requires 1-2 concrete actions.

2 / 5

Completeness

Both parts are explicitly present: the what ("Quick execution for ad-hoc tasks without full workflow overhead") and the when ("use for small, self-contained requests"), matching the 4-anchor 'has both what and when; when could be more explicit or specific'. It is not a 5 because the when lacks concrete trigger phrases naming the task types it covers.

4 / 5

Trigger Term Quality

Terms like "quick", "ad-hoc", and "small, self-contained requests" are phrases users might naturally say, but common variations users would actually utter ("quick fix", "simple change", "one-file change", "typo") are missing, matching 'some relevant keywords but missing common variations or synonyms'.

3 / 5

Distinctiveness Conflict Risk

The lightweight-vs-full-workflow niche is stated, but the trigger condition "small, self-contained requests" is relative — nearly any simple request could match it, so it could still overlap with general task-execution skills, matching 'somewhat specific but could still overlap with similar skills'.

3 / 5

Total

12

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.