CtrlK
BlogDocsLog inGet started
Tessl Logo

skill-review-response

Use when a reviewer, CI bot, or another AI leaves feedback to address

53

Quality

60%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/skill-review-response/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a strong, prescriptive instruction document: a sequenced response workflow with explicit validation checkpoints, a decision-table checklist, concrete forbidden behaviors, and a worked pushback example, plus genuine feedback-loop recovery guidance. Its main weaknesses are a dangling cross-reference to 'skill-verification-gate' (never linked or defined) and a somewhat speculative multi-provider characterization section. As an instruction-only skill with no bundle files, structure and actionability are high.

DimensionReasoningScore

Conciseness

The body is almost entirely prescriptive — the 6-step response pattern, forbidden phrases, and decision table add information Claude does not already have, with no explanation of basic review concepts. Minor trimmable padding exists, e.g., the multi-provider characterizations ('tends toward enterprise patterns, may over-engineer') which are speculative generalizations. Fits anchor 4 ('Efficient; minor instances of over-explanation that could be trimmed'); not 5 because the Multi-Provider Review Context section partially pads, not 3 because there is no genuinely unnecessary explanation of known concepts.

4 / 5

Actionability

Concrete, executable guidance throughout: a numbered 6-step pattern (READ/RESTATE/VERIFY/EVALUATE/RESPOND/IMPLEMENT), an exact list of forbidden phrases, an if-yes/if-no decision table, and a copy-paste-ready pushback response example with line numbers. The minor gap is 'Run verification (skill-verification-gate)' and 'see skill-verification-gate', referenced twice but never linked or explained, leaving the reader without a way to execute that step. Fits anchor 4 ('Mostly executable guidance; concrete code or commands with minor gaps'); not 5 because of that dangling reference, not 3 because guidance is concrete rather than pseudocode.

4 / 5

Workflow Clarity

The response pattern is a clear sequenced workflow with an explicit validation checkpoint (step 3 VERIFY, step 4 EVALUATE), the Evaluation Checklist provides a decision table for complex cases, and the Handling Feedback Loops section gives explicit error-recovery guidance ('Re-read the original feedback — did you address the root cause', 'Each round should have FEWER issues', stop and re-read when the same issue recurs). This matches anchor 5 ('Clear sequence with explicit validation steps; feedback loops for error recovery; checklists for complex processes') exactly; it is not 4 because validation checkpoints and feedback loops are both fully present.

5 / 5

Progressive Disclosure

The body is well-organized with clear section headers, a table, and no monolithic wall of text; no references/, scripts/, or assets/ bundle files exist, and no external files are referenced. Fits anchor 4 ('Good structure; most content is appropriately placed; minor organization gaps') — the gap being the twice-mentioned 'skill-verification-gate' with no path or link, and the Multi-Provider section (~103 lines total, over the under-50-line simple-skill exception) which is context that could be signposted more clearly. Not 5 because content exceeds the simple-skill size guidance and contains unresolved cross-references; not 3 because structure is consistently clear and nothing that belongs in a separate file is buried inline.

4 / 5

Total

17

/

20

Passed

Description

40%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description has a clear, concrete trigger clause but completely omits what the skill does, making it read as a pure 'when' statement. Trigger terms are relevant but lack common variations like 'code review' or 'PR comments'. Adding a 'what' clause (e.g., 'Evaluates review feedback against the codebase, responds with technical acknowledgment or reasoned pushback') would substantially raise completeness and specificity.

Suggestions

Add a 'what' clause describing the skill's concrete actions (e.g., 'Evaluates review feedback against the actual codebase, then responds with technical acknowledgment or evidence-based pushback') to fix the missing half of completeness.

Include natural trigger variations users would actually say: 'code review feedback', 'PR review comments', 'changes requested on a PR', 'CI failure comments'.

State the distinction from review-producing skills explicitly in the description (e.g., 'for responding to reviews, not performing them') to reduce conflict risk with code-review skills.

DimensionReasoningScore

Specificity

The description names the domain — "a reviewer, CI bot, or another AI leaves feedback to address" — but names no concrete actions the skill performs (verify, evaluate, push back, respond). This matches anchor 2 ('Names the domain but actions are minimal or generic'); it is not 3 because no 1-2 concrete actions are listed, and not 1 because the domain is clearly identified rather than pure abstraction.

2 / 5

Completeness

The 'when' is explicit ("Use when a reviewer, CI bot, or another AI leaves feedback to address") but the 'what' — what the skill actually does when invoked — is entirely missing. Per the guideline 'only when is present without what' fits anchor 2; it is not 3 because that anchor requires a clear 'what', and not 1 because the 'when' half is concrete and specific.

3 / 5

Trigger Term Quality

Natural terms present include "reviewer", "CI bot", and "feedback", but common variations users would actually say — "code review", "PR comments", "review feedback", "changes requested" — are absent. Fits anchor 3 ('Some relevant keywords but missing common variations or synonyms'); not 4 because keyword coverage has clear gaps, not 2 because more than one or two generic keywords are present.

3 / 5

Distinctiveness Conflict Risk

Scoping to feedback the agent *receives* ("reviewer, CI bot, or another AI leaves feedback") distinguishes it from skills that *perform* reviews; the overlap with code-review/PR-review skills is minor. Fits anchor 4 ('Mostly distinct; minor overlap risk with closely related skills'); not 5 because the trigger could still fire for other feedback-handling situations (e.g., generic user feedback, CI logs), not 3 because the receiving/responding role is explicitly staked out.

4 / 5

Total

12

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
nyldn/claude-octopus
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.