CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-council

Run a configurable multi-LLM council with personas, budget caps, synthesis, veto gates, and optional implementation handoff.

62

Quality

73%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/skill-council/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a strong, execution-focused contract: an exact runner invocation, concrete flag mappings, explicit quorum and budget checkpoints, and gated implementation with veto handling. Its only real weaknesses are mild redundancy between the two interactive-choice sections and a block of sandbox-internals detail that would sit better in a reference file.

DimensionReasoningScore

Conciseness

The body is dense and assumes Claude's competence — no space is spent explaining what LLMs, councils, or sandboxes are, and sections like Preflight and Quorum are tight checklists. Minor trimming opportunities keep it below 5: the 'Interactive Choice Handling' section restates rules already given in 'Phase 0A' ('2-4 mutually exclusive choices... recommended first... do not end with a free-form question'), and the sandbox-allowlist mechanics paragraph runs long.

4 / 5

Actionability

Guidance is fully executable: a copy-paste runner command with the ${CLAUDE_PLUGIN_ROOT} fallback ('"${CLAUDE_PLUGIN_ROOT:-$HOME/.claude-octopus/plugin}/scripts/orchestrate.sh" council <user arguments>'), a literal AskUserQuestion block, and exact flag mappings ('Advice, Decision, Implementation plan, or Review -> --goal advice|decision|plan|review'). Concrete defaults and stop conditions cover the common cases, matching the 'fully executable, copy-paste ready' anchor.

5 / 5

Workflow Clarity

The sequence is explicit — Phase 0A clarification, Phase 0 preflight, quorum check, five-step Council Procedure, Gates A-C — with real validation checkpoints and feedback loops: budget re-checks before critique/revision/synthesis/implementation, dry-run stop after preflight, chair retry on failure, quorum-loss stop, and veto gates before implementation. This matches the top anchor (clear sequence, explicit validation, error-recovery loops, checklists).

5 / 5

Progressive Disclosure

Sections are well organized and no broken or nested references exist (the skill has no references/, scripts/, or assets/ directories, and the only external path — the runner script — is a real plugin path, not a bundle file). It falls short of 5 because the ~15-line sandbox-allowlist mechanics block and the full default clarification example are detail that could live in a one-level-deep reference file rather than the main SKILL.md.

4 / 5

Total

18

/

20

Passed

Description

58%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description gives a concrete, well-scoped picture of what the skill does, with a specific feature list that distinguishes it from generic multi-agent skills. Its main weakness is the complete absence of any 'when to use this' trigger guidance, which both caps completeness and leaves trigger-term coverage thin on natural synonyms.

Suggestions

Append an explicit trigger clause, e.g. 'Use when the user asks for a council, panel, or committee of multiple LLMs to debate, review, or decide on a question.'

Add natural synonyms users would actually say ('panel', 'committee', 'multi-model debate', 'second opinions from multiple models') to broaden trigger-term coverage.

Optionally name one or two of the concrete council actions (cross-critique, chair synthesis, ratify/veto) to round out the capability list toward comprehensive coverage.

DimensionReasoningScore

Specificity

The description lists several concrete capabilities — 'personas, budget caps, synthesis, veto gates, and optional implementation handoff' — anchored by the core action 'Run a configurable multi-LLM council'. It stops short of level 5 because it names features rather than the fuller set of concrete actions the skill performs (cross-critique, revision, chair synthesis, ratification), leaving minor coverage gaps.

4 / 5

Completeness

The 'what' is clear ('Run a configurable multi-LLM council with personas, budget caps, synthesis, veto gates...'), but there is no 'Use when...' clause or equivalent explicit trigger guidance anywhere in the description — the rubric explicitly caps completeness at 3 in that case. It is not level 2 because the 'what' half is concrete and specific rather than vague.

3 / 5

Trigger Term Quality

'multi-LLM council' and 'council' are natural terms a user might say, but common synonyms and variations users actually use — 'panel', 'committee', 'jury of models', 'multi-model debate/review', 'red-team' — are absent, and phrases like 'veto gates' and 'budget caps' lean toward internal jargon. Good keywords exist but coverage is incomplete, fitting the 'some relevant keywords but missing common variations' anchor better than level 4.

3 / 5

Distinctiveness Conflict Risk

'multi-LLM council' with personas, veto gates, and budget caps carves a fairly distinct niche with low risk of firing for unrelated skills; only minor overlap risk remains with generic multi-agent/debate/committee skills. Not level 5 because 'council' and 'personas' are terms adjacent to those broader multi-agent workflows.

4 / 5

Total

14

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.