CtrlK
BlogDocsLog inGet started
Tessl Logo

review-correctness

Review a code change for off-by-one and boundary errors, null propagation, changed sentinel meanings, tooling and provisioning drift, race conditions, invalid state transitions, React effect cleanup gaps, and broken error propagation. Use when reviewing for logic bugs, behavioral correctness, or edge cases that tests miss.

76

Quality

95%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary instruction-only skill body: dense with non-obvious, category-specific bug patterns, concrete methods and reporting requirements, explicit thresholds that gate what to report, and zero padding or basic-concept explanations. Its structure gives an unambiguous review workflow that fits comfortably within a small single-file skill.

DimensionReasoningScore

Conciseness

The body is lean and assumes Claude's competence — every bullet adds non-obvious, specific guidance such as "pagination that misses the final page when the total is an exact multiple of the page size" and "returning an empty array so the caller reads 'no results' instead of 'query failed'". There is no explanation of concepts Claude already knows and no padding, matching the 'every token earns its place' anchor.

5 / 5

Actionability

For an instruction-only skill, the guidance is fully executable: concrete directives like "Trace boundary math with concrete values at the edges. Follow each changed return value to its callers and each mutation to its cleanup" and "enumerate every useEffect exit path and check that each mutation before return has matching cleanup", plus a Reporting section specifying the exact output (trace, wrong outcome, fix). Specific example patterns cover the common cases in each category.

5 / 5

Workflow Clarity

The skill is single-purpose and under 50 lines, with an unambiguous flow: Scope (what to check) → Method (how to trace) → Threshold (what to report) → Reporting (output format). The Threshold section is an explicit validation checkpoint — "Report bugs you can trace from an input, through the branch it takes, to the line that produces the wrong result, where a normal caller will hit it" — and the Scope section functions as a checklist, satisfying the simple-skill exception.

5 / 5

Progressive Disclosure

The body is roughly 30 lines, well-organized into clearly labeled sections, with no content that warrants separate reference files and no nested references. Per the rubric's guideline for skills under 50 lines with no need for external references, well-organized sections alone merit the top score; no bundle files exist to verify.

5 / 5

Total

20

/

20

Passed

Description

88%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that clearly states what the skill does through a comprehensive enumeration of concrete bug classes and provides an explicit, natural 'Use when' trigger clause. The only weaknesses are minor: a few common user phrasings are absent from the triggers, and there is slight overlap in wording with generic code-review skills.

DimensionReasoningScore

Specificity

The description enumerates eight concrete bug classes it reviews for: "off-by-one and boundary errors, null propagation, changed sentinel meanings, tooling and provisioning drift, race conditions, invalid state transitions, React effect cleanup gaps, and broken error propagation" — comprehensive, specific coverage with no gaps, matching the top anchor rather than the 'minor gaps' level 4 anchor.

5 / 5

Completeness

It explicitly answers both questions: the 'what' is the enumerated list of bug classes reviewed, and the 'when' is a concrete "Use when reviewing for logic bugs, behavioral correctness, or edge cases that tests miss" clause with specific trigger phrases — exactly the structure of the level-5 anchor example, not the vaguer 'when' of level 4.

5 / 5

Trigger Term Quality

The 'Use when' clause includes natural phrases users would say — "reviewing for logic bugs, behavioral correctness, or edge cases that tests miss" — plus category keywords like race conditions and null propagation. A few common phrasings (e.g., 'find bugs', 'bug hunt', 'code review') are missing, so it sits between 'good coverage' (4) and 'comprehensive synonyms' (5).

4 / 5

Distinctiveness Conflict Risk

The niche (correctness/logic-bug review of a code change) is well-defined with distinct triggers around logic bugs and edge cases. However, "Review a code change" overlaps with general code-review or security-review skills, so it is mostly distinct with minor overlap risk rather than minimal conflict.

4 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
perihelionhq/perihelion-platform-context
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.