CtrlK
BlogDocsLog inGet started
Tessl Logo

review-reliability

Review a code change for missing error handling on I/O boundaries, unbounded retries, missing timeouts, swallowed errors, resource leaks on error paths, cascading failures, and CI or deploy guards that do not mirror production. Use when reviewing for reliability, failure modes, partial failures, or graceful degradation.

77

Quality

97%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A lean, expert-level review lens: every section adds non-obvious judgment (defect taxonomy with literal code patterns, Release It! antipattern vocabulary, threshold criteria for reporting, and a concrete finding format) with zero padding. The single-task workflow is clearly sequenced with a Threshold section that validates what gets reported, and the short length makes external references unnecessary.

DimensionReasoningScore

Conciseness

At ~36 lines of dense, non-padded guidance, every sentence adds judgment Claude doesn't have: concrete defect patterns, guard-fidelity comparison instructions, threshold criteria, and reporting requirements. No basic concepts (what a retry or timeout is) are explained, matching the 'every token earns its place' anchor.

5 / 5

Actionability

Concrete, executable guidance throughout: literal code patterns to search for ('catch (e) {}', '.catch(() => {})', 'no finally, defer, using, or context manager'), a specific procedure in Method (trace every external call's failure to callers, read the AGENTS.md/CLAUDE.md chain), and concrete finding requirements ('Point to the line missing the protection'). Per the scoring note, absence of code in an instruction-only skill is not penalized when guidance is this actionable.

5 / 5

Workflow Clarity

For a single-purpose review skill, the sequence is unambiguous: Scope defines what to look for, Method defines how to trace and prioritize, Threshold provides explicit include/exclude validation criteria for findings ('Report a gap when... Do not report internal pure functions...'), and Reporting defines output format. The Threshold section acts as the validation checkpoint for findings, and the operation is non-destructive so no stronger feedback loop is needed.

5 / 5

Progressive Disclosure

The skill is under 50 lines, needs no external references (none exist in the bundle), and is organized into five well-labeled sections (intro, Scope, Method, Threshold, Reporting) that each stay at overview altitude. This matches the guideline for short simple skills scoring 5 with well-organized sections.

5 / 5

Total

20

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An excellent description: it enumerates seven concrete, specific defect categories the skill reviews and pairs them with an explicit, natural 'Use when...' trigger clause in third person. The only improvement would be adding common synonyms like 'resilience' or 'robustness' to the trigger list.

Suggestions

Add natural synonyms such as 'resilience', 'robustness', or 'stability' to the 'Use when...' clause to broaden trigger coverage.

DimensionReasoningScore

Specificity

The description lists multiple specific concrete actions — 'missing error handling on I/O boundaries, unbounded retries, missing timeouts, swallowed errors, resource leaks on error paths, cascading failures, and CI or deploy guards that do not mirror production' — giving comprehensive coverage of what the skill reviews. Not 4 because the enumerated defect categories leave no minor gaps in what/coverage.

5 / 5

Completeness

It explicitly answers both: what ('Review a code change for missing error handling on I/O boundaries, unbounded retries, missing timeouts...') and when ('Use when reviewing for reliability, failure modes, partial failures, or graceful degradation') with concrete trigger phrases. This matches the anchor-5 example structure exactly.

5 / 5

Trigger Term Quality

Triggers include natural phrases users would say: 'reviewing for reliability, failure modes, partial failures, or graceful degradation', plus 'error handling', 'timeouts', 'retries', 'resource leaks' in the capability list. Not 5 because common synonyms users might use — 'resilience', 'robustness', 'stability', 'availability' — are absent.

4 / 5

Distinctiveness Conflict Risk

It carves out a clear niche (reliability/failure-mode review of code changes) with distinctive triggers ('failure modes, partial failures, graceful degradation') that separate it from general code-review or security-review skills. Not 4 because the overlap risk with other review lenses is minimal, not 'minor overlap with closely related skills' in a material way.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
perihelionhq/perihelion-platform-context
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.