CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-code-review

Expert multi-AI code review with inline PR comments — use for thorough quality and security analysis

55

Quality

63%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/skill-code-review/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

63%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body delivers mostly executable, well-sequenced guidance with real error recovery in the PR-posting flow, but it is padded in places (the compliance section, inlined stub-detection script, sample output) and its progressive disclosure is weak: external references are unverifiable and inlined content duplicates them. Mid-to-good quality with clear trimming and restructuring opportunities.

Suggestions

Move the ~90-line stub-detection shell script into a bundled references/stub-detection.md file and keep only a summary plus the blocking/non-blocking rules in SKILL.md.

Compress the "MANDATORY COMPLIANCE" section to a short list of the pipeline requirements instead of five restatements of the same prohibition.

Enumerate the full review pipeline's phases (quick mode only says to skip them) and define $REVIEW_SYNTHESIS and $COMMIT_RANGE before they are used, with an existence check for the ~/.claude-octopus plugin scripts.

DimensionReasoningScore

Conciseness

The "MANDATORY COMPLIANCE" section restates one prohibition five ways ("Doing a direct single-model... Substituting two background Sonnet agents... Rationalizing 'a focused audit..."), a ~90-line stub-detection shell script is fully inlined, and a large sample output block is reproduced verbatim. Not 2 because much of the body is lean, concrete instruction Claude does not already know; not 4 because several sections clearly need trimming.

3 / 5

Actionability

Commands are largely copy-paste executable: orchestrate.sh invocations, complete bash blocks for stub detection and PR detection, and a concrete AskUserQuestion script with options. Not 5 because $REVIEW_SYNTHESIS and $COMMIT_RANGE are used but never defined, and everything depends on an external ~/.claude-octopus plugin that is not validated for existence; not 3 because most guidance runs as written.

4 / 5

Workflow Clarity

Stub detection and PR posting are numbered step sequences with explicit failure handling ("if ! ... safe-gh-comment.sh ... check for the review comment before retrying") and a clear auto-post vs. ask-first decision rule. Not 5 because the full pipeline's phases are never enumerated — quick mode is defined only as "skip the full review pipeline" — so the primary path's sequence is incomplete.

4 / 5

Progressive Disclosure

Sections are well headed, but the skill bundle contains no references/ directory, and the body points to external paths (".claude/references/stub-detection.md", "agents/personas/code-reviewer.md") that are unverifiable from the bundle, while ~90 lines of stub-detection script that belong in that reference are inlined instead. Not 2 because structure and section headers are clearly present; not 4 because content that should be separate is inline and the referenced paths are neither bundled nor clearly signaled.

3 / 5

Total

14

/

20

Passed

Description

62%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A clear, concise description in third person that states what the skill does and includes an explicit use-for clause, but the trigger guidance is purpose-based rather than phrase-based, and it undersells several documented capabilities. It sits solidly above mediocre but below the exemplar descriptions.

Suggestions

Replace the purpose-based trigger with concrete phrases, e.g. "Use when the user asks to review a PR, check a diff, audit staged changes, or find security issues in code."

Mention the distinctive capabilities the body documents (stub detection, TDD/autonomy assessment, posting findings as a PR comment) so the description's coverage matches the skill.

Include natural synonyms users say — "pull request" (not just PR), "review my changes", "security audit" — to improve trigger matching.

DimensionReasoningScore

Specificity

"Expert multi-AI code review with inline PR comments" names the domain and two concrete actions (multi-provider review, inline PR comments), but "thorough quality and security analysis" is generic and the description omits documented capabilities like stub detection, TDD assessment, and PR posting. Not 2 because it lists concrete actions beyond naming the domain; not 4 because the coverage gaps are more than minor.

3 / 5

Completeness

The "what" is clear ("multi-AI code review with inline PR comments") and there is an explicit trigger clause ("use for thorough quality and security analysis"). Not 5 because the "when" describes purpose rather than concrete trigger phrases a user would say; not 3 because an explicit "use for" clause is present.

4 / 5

Trigger Term Quality

Relevant keywords are present ("code review", "PR", "quality", "security"), but common natural variations are missing: "review this PR", "pull request", "check my changes", "diff". Not 4 because several phrases a user would naturally say are absent; not 2 because the terms included are ones users actually use.

3 / 5

Distinctiveness Conflict Risk

"Multi-AI" synthesis and "inline PR comments" carve a fairly distinct niche versus generic review skills, though "code review" broadly overlaps many quality/security skills. Not 5 because overlap risk with ordinary PR-review skills remains; not 3 because the multi-LLM and inline-comment angles are distinguishing.

4 / 5

Total

14

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.