CtrlK
BlogDocsLog inGet started
Tessl Logo

octopus-research

Thorough research across multiple sources — use for complex topics needing broad synthesis

52

Quality

58%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/skill-deep-research/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

62%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body's strongest asset is its workflow: five gated steps with an executable validation script and explicit per-step error handling. Its weaknesses are significant redundancy in the enforcement language (the same prohibition restated many times, plus a duplicated banner line) and a dangling reference to a skill-security-framing.md that does not exist in any bundle.

Suggestions

State the execution contract once (e.g. keep the HARD-GATE) and remove the repeated prohibition lists in Step 3 and the Error Handling section — one clear statement plus the per-step error table conveys the same information at a fraction of the tokens.

Fix the banner's duplicated provider line ('🟡 Antigravity CLI' and '🧭 Antigravity CLI' both use ${agy_status}) and have orchestrate.sh print the synthesis-file path so Step 4 does not rely on a find-by-timestamp heuristic.

Resolve the dangling reference: either ship skill-security-framing.md in references/ and link it as a one-level-deep reference, or inline the concrete URL validation rules and drop the pointer to the missing file.

DimensionReasoningScore

Conciseness

The enforcement rhetoric is stated five or more times — the HARD-GATE block ('You MUST call orchestrate.sh... If you produce research findings without a Bash call... you have violated this contract'), Step 3's 'CRITICAL: You are PROHIBITED' list, and the Error Handling section's near-duplicate 'Never fall back to direct research' — plus the banner duplicates the Antigravity provider line ('🟡 Antigravity CLI' / '🧭 Antigravity CLI'). That is several padded, redundant sections that do not assume Claude's competence; it is not 3 because the repetition is pervasive rather than occasional.

2 / 5

Actionability

The body provides concrete, near-executable guidance: an exact provider-detection bash snippet, a complete AskUserQuestion call with labeled options, and a copy-paste Step 4 validation script (find ~/.claude-octopus/results -name "probe-synthesis-*.md" -mmin -10 with an exit-1 failure path). It is not 5 because Step 2/3 require interpolating template variables (${depth_choice}, ${codex_status}) and the synthesis-file path is inferred rather than returned by orchestrate.sh, leaving minor assembly gaps.

4 / 5

Workflow Clarity

The five steps are explicitly sequenced with blocking gates ('DO NOT PROCEED TO STEP 2 until all questions are answered', 'DO NOT PROCEED TO STEP 3 until banner displayed'), an executable validation gate (Step 4 fails with exit 1 and error messaging when the synthesis file is missing), and per-step error recovery including log locations and an explicit never-fallback rule. This matches anchor 5 — clear sequence with explicit validation and error-recovery handling.

5 / 5

Progressive Disclosure

No bundle files exist (references/, scripts/, assets/ are all absent), yet the body's Security section references **skill-security-framing.md**, which is a dangling pointer to a nonexistent file. Internal section structure is clear, but content that belongs in separate files (the interactive question template, the banner, security framing) is inlined into a ~230-line monolith, and the one reference given is not loadable — 'some structure but could be better organized' matches anchor 3.

3 / 5

Total

14

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description correctly follows the what-plus-when pattern in third person, but it stays at a high level of abstraction. It conveys the domain without naming a single concrete capability or the natural phrases users would say when they need it, and its broad research framing creates overlap risk with built-in research skills.

Suggestions

Name the concrete actions the skill performs, e.g. 'Orchestrates multi-provider research probes via orchestrate.sh and synthesizes results into a cited report'.

Surface the natural trigger phrases already written in the frontmatter trigger block ('research this topic', 'investigate how X works', 'explore different approaches', 'what are the options for Z') into the description itself.

Sharpen the 'when' clause beyond 'complex topics needing broad synthesis' toward specific signals (e.g. 'when the user asks to investigate, compare approaches, or synthesize findings across many sources').

DimensionReasoningScore

Specificity

The description names the domain ("Thorough research across multiple sources") but lists no concrete actions — it never says what the skill actually does (orchestrate multi-provider probes via orchestrate.sh, synthesize findings into a report). It matches anchor 2 ('names the domain but actions are minimal or generic'); it is not 3 because no specific action is articulated.

2 / 5

Completeness

Both 'what' ("Thorough research across multiple sources") and 'when' ("use for complex topics needing broad synthesis") are explicitly present, satisfying an explicit trigger clause so it scores above 3. It is not 5 because the 'what' is high-level rather than concrete, and the 'when' clause is generic ('complex topics') rather than specific trigger phrases.

4 / 5

Trigger Term Quality

"research" and "synthesis" are relevant keywords, but the description omits natural user phrases like "investigate how X works", "explore", "what are the options", or "deep dive" (which the frontmatter trigger block actually contains but the description does not). Some relevant keywords with missing variations matches anchor 3; it is not 4 because most natural trigger phrasing is absent from the description itself.

3 / 5

Distinctiveness Conflict Risk

"Thorough research across multiple sources" for "complex topics needing broad synthesis" overlaps significantly with generic web-search and deep-research capabilities, and the skill's own aliases ('research', 'deep-research') collide with existing skills. Somewhat specific but with real overlap risk matches anchor 3; it is not 4 because the description's trigger surface is broad enough to compete with several common research skills.

3 / 5

Total

12

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.