CtrlK
BlogDocsLog inGet started
Tessl Logo

receiving-code-review

Evaluate code review feedback on its technical merits, verify each suggestion against the codebase, then implement the fixes that hold up and push back with reasoning on the ones that don't. This is the internal review-fix stage of the delivery-flow workflow, run when delivery-flow receives feedback from a reviewer. It is not a standalone entry point, so do not activate it directly for a one-off "address this review comment" or "respond to PR feedback" request; use delivery-flow for those instead.

73

Quality

90%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

The canonical home for this skill is receiving-code-review in obra/superpowers

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced instruction skill with explicit validation checkpoints, error-recovery loops, and concrete response templates. Its weaknesses are repetition of the no-gratitude/no-performative rules across multiple sections and a ~205-line single-file body whose examples could be split into a one-level-deep reference.

Suggestions

Consolidate the anti-performative rules: 'Forbidden Responses', 'Acknowledging Correct Feedback', and 'Real Examples' all restate the same 'You're absolutely right!' / 'Thanks for [anything]' prohibitions — state them once and cross-reference to cut repetition.

Move the 'Real Examples' section (and possibly the ✅/❌ response template lists) into a single references file (e.g. references/response-templates.md) with a clearly signaled one-level-deep link, keeping SKILL.md as a leaner overview.

Trim near-duplicate guidance such as the 'Unclear Item' example, which appears almost verbatim in both 'Handling Unclear Feedback' and 'Real Examples'.

DimensionReasoningScore

Conciseness

The body is compact and assumes Claude's competence (no concept explanations, dense IF-blocks and tables), but the anti-performative/anti-gratitude rules are repeated across three places — 'Forbidden Responses', 'Acknowledging Correct Feedback' ('You're absolutely right!' / 'Great point!' / 'Thanks for [anything]' / 'ANY gratitude expression'), and 'Real Examples'. This fits 'efficient; minor instances... that could be trimmed' rather than the score-5 'every token earns its place' anchor.

4 / 5

Actionability

Guidance is fully executable for an instruction-only skill: literal copy-paste response templates ('I understand items 1,2,3,6. Need clarification on 4 and 5 before proceeding.'), a concrete grep-the-codebase YAGNI check, an exact command ('gh api repos/{owner}/{repo}/pulls/{pr}/comments/{id}/replies'), and a specific fix-ordering priority list. Concrete examples cover the common cases, matching the score-5 anchor.

5 / 5

Workflow Clarity

The Response Pattern gives a clear six-step sequence (READ → UNDERSTAND → VERIFY → EVALUATE → RESPOND → IMPLEMENT) with explicit validation ('VERIFY: Check against codebase reality', 'Test each fix individually', 'Verify no regressions'), a feedback loop for error recovery ('Gracefully Correcting Your Pushback'), and checklists (the five external-reviewer checks, the Common Mistakes table). This matches the score-5 anchor; the score-4 anchor's 'minor validation gaps' does not apply.

5 / 5

Progressive Disclosure

No bundle files exist, and the body is well-sectioned and navigable (Overview, Source-Specific Handling, Implementation Order, Common Mistakes, Real Examples). However, at ~205 lines with no references, the 'Real Examples' and template-heavy sections could be offloaded to a reference file, so it sits at 'good structure... minor organization gaps' rather than the score-5 clear-overview-with-references pattern.

4 / 5

Total

18

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete and comprehensive on capabilities, explicit on both when to use and when not to use, and deliberately disambiguated from the overlapping delivery-flow entry point. The only minor gap is synonym coverage of natural trigger phrases.

DimensionReasoningScore

Specificity

The description enumerates multiple concrete, comprehensive actions for the domain: 'Evaluate code review feedback on its technical merits, verify each suggestion against the codebase, then implement the fixes that hold up and push back with reasoning on the ones that don't.' This covers the full evaluate/verify/implement/push-back lifecycle, matching the 'multiple specific concrete actions; comprehensive coverage' anchor rather than the score-4 anchor, which requires minor gaps in coverage.

5 / 5

Completeness

It explicitly answers both questions: what ('Evaluate... verify... implement... push back with reasoning') and when ('run when delivery-flow receives feedback from a reviewer'), plus an explicit when-not clause. This matches the score-5 anchor with concrete trigger phrases; the score-4 anchor's 'when could be more explicit' does not apply here.

5 / 5

Trigger Term Quality

Good keyword coverage with natural phrases users would say — 'code review feedback', 'address this review comment', 'respond to PR feedback', 'reviewer' — but a few common variations are missing (e.g., 'review comments', 'PR comments'). It fits 'good keyword coverage; a few natural terms missing' rather than the score-5 anchor's comprehensive synonym coverage.

4 / 5

Distinctiveness Conflict Risk

It claims a clear niche — 'the internal review-fix stage of the delivery-flow workflow' — and actively disambiguates the closest conflict case: 'do not activate it directly for a one-off... request; use delivery-flow for those instead.' Minimal conflict risk, matching the score-5 'clear niche with distinct triggers' anchor.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
tesslio/tessl-eval-demo
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.