CtrlK
BlogDocsLog inGet started
Tessl Logo

octopus-architecture

Review system architecture, simplify boundaries, or compare interface designs from repository evidence

52

Quality

58%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/octopus-architecture/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

61%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A concise, well-structured instruction-only skill with a concrete output contract and clear drafting rules, but its core method lives in external files not present in the bundle, and the workflow lacks explicit validation checkpoints. It reads as a dispatch stub more than a self-contained executable guide.

Suggestions

Inline the essential steps of the architecture-simplification method (how to pin a revision, which callers/tests to inspect, what counts as exact source evidence) instead of delegating entirely to skills/blocks/architecture-simplification.md.

Add explicit validation checkpoints and recovery guidance — e.g., what to do when caller evidence is incomplete or when the two drafts converge — rather than only the final keep/simplify/investigate gate.

Verify or bundle the referenced files (skills/blocks/*, THIRD_PARTY_NOTICES.md, codex-host-adapter.md); none exist in this skill's directory, leaving the reference chain unverifiable.

DimensionReasoningScore

Conciseness

The ~50-line body is lean and assumes competence; no space is spent explaining known concepts. A few abstract sentences ("do not invoke the current command recursively or add provider calls from a seat", "Transport diversity does not prove model family diversity") could be tightened, keeping it below the lean-and-efficient anchor.

4 / 5

Actionability

There is a concrete output contract ("Return `Evidence`, `Caller contract`, `Proposed interface`, `Migration`, `Test impact`, and `Decision`" with a decision enum) and draft criteria, but the core method is delegated to an external file and guidance like "inspect implementation, callers, tests, recent churn, and failure behavior" is high-level without specifics of how to pin a revision or structure the evidence. Not 2 because the completion contract and Draft A/B criteria are more than minimal hints.

3 / 5

Workflow Clarity

A rough sequence exists (select method → apply architecture-simplification block → draft alternatives → return labeled outputs) with a partial gate ("`simplify` requires exact source evidence and a concrete caller example"), but explicit validation checkpoints and error-recovery loops are absent — e.g., no step for what to do when evidence is inconclusive beyond the investigate decision. Fits 'steps listed but validation gaps'.

3 / 5

Progressive Disclosure

Sections (Method, Completion, LSP Integration) are well organized for a sub-50-line skill, and references ("skills/blocks/engineering-method-selection.md", "skills/blocks/architecture-simplification.md", "THIRD_PARTY_NOTICES.md") are one level deep and clearly signaled. Not 5 because none of the referenced files exist in this bundle — they point into an external "installed plugin" — so navigation and content placement cannot be verified.

4 / 5

Total

14

/

20

Passed

Description

55%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A specific, third-person description with concrete named actions, but it omits any "when to use this" trigger clause, which both caps completeness and weakens natural-term coverage. Adding an explicit usage-trigger sentence with common synonyms would move it into the good-example tier.

Suggestions

Append a trigger clause such as: "Use when asked to review or simplify system/module boundaries, compare API or interface designs, or decide whether to keep vs. refactor an abstraction."

Add natural user synonyms — refactor, API design, service boundaries, dependency structure, keep vs. simplify — to improve trigger term coverage and distinctiveness.

Distinguish the skill from generic code review by naming its unique decision output (e.g., a keep/simplify/investigate decision backed by caller evidence).

DimensionReasoningScore

Specificity

Lists several concrete actions — "Review system architecture, simplify boundaries, or compare interface designs" — with a stated evidence basis ("from repository evidence"). Not 5 because coverage has gaps (no refactoring, dependency, or migration terms); not 3 because it names more than 1-2 specific actions.

4 / 5

Completeness

The "what" is clear (review/simplify/compare architecture from repo evidence) but there is no "Use when..." clause or equivalent trigger guidance, which caps completeness at 3 per the judging guidelines. Not 2 because the "what" is specific, not vague.

3 / 5

Trigger Term Quality

"system architecture" and "interface designs" are natural terms users would say, but common synonyms and variations are missing ("refactor", "API design", "code review", "service boundaries", "monolith"). Fits the 'some relevant keywords but missing common variations' anchor rather than the good-coverage anchor above.

3 / 5

Distinctiveness Conflict Risk

"Review system architecture" is somewhat specific but overlaps with code-review and general design-review skills; the description does not carve out a clearly distinct niche trigger. Not 4 because the overlap with related review skills is more than minor.

3 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.