CtrlK
BlogDocsLog inGet started
Tessl Logo

review-agent-native

Review a code change for gaps where a user can do something an agent cannot, where the agent lacks the context to act, or where agent tools encode workflows instead of primitives. Use when reviewing for agent-native architecture, action and context parity, agent tool design, or system prompt context.

72

Quality

91%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary instruction-only review lens: token-dense, stack-specific search guidance, a well-sequenced method, and an explicit threshold/reporting contract. The only minor gaps are the absence of a worked example finding and an explicit verify-before-report step.

Suggestions

Add one short worked example finding under Reporting (e.g. 'Feed refresh button (app/views/feeds/_actions.html.erb:12) has no matching tool — core priority; fix: add a tool and document it in the prompt') to make the output format copy-paste concrete.

Add an explicit verification step to the Method, e.g. 'Re-confirm each reported gap against the tool registry before writing it up, to avoid flagging tools defined elsewhere in the stack.'

DimensionReasoningScore

Conciseness

Lean and efficient across all ~30 lines: no concept explanations Claude already knows, no padding — every section (Scope, Method, Threshold, Reporting) delivers novel, decision-relevant guidance in dense prose. Phrases like 'A new button, form or gesture with no matching tool is an orphan feature' carry definition, red flag, and vocabulary in one line. Matches the score-5 anchor: every token earns its place.

5 / 5

Actionability

Highly concrete guidance: exact search targets per stack ('onClick, onSubmit, form actions, button_to, form_with', 'tool() and the tools parameter of streamText or generateText', '@tool and StructuredTool', 'agents/*.md and skills/*/SKILL.md') and a precise reporting format with location, priority, and fix. Not 5: no worked example of an actual finding (e.g. a sample report line), so a reviewer must interpolate the output format from the bullet description — a minor gap. Not 3: the guidance is fully executable, not pseudocode or high-level hints.

4 / 5

Workflow Clarity

Clear multi-step sequence: (1) check for agent integration at all, (2) locate UI-action and tool definitions, (3) focus on new/modified code, cross-reference actions against tools, (4) second pass by domain noun, then Threshold and Reporting. The Threshold section acts as a decision checklist. Not 5: the review is read-only so no destructive/batch validation cap applies, but there is no explicit verify-findings feedback loop (e.g. re-confirm each gap against the tool registry before reporting) — a minor checkpoint gap. Not 3: checkpoints are present via the Threshold and out-of-scope rules, not merely implicit.

4 / 5

Progressive Disclosure

The skill is under 50 lines with no external references needed, and it is well-organized into four clearly signaled sections (Scope, Method, Threshold, Reporting). Per the rubric's guideline for short self-contained skills, this warrants a 5. There are no references/ or scripts/ bundle files, so all content appropriately lives inline.

5 / 5

Total

18

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person voice, concrete enumeration of the three gap types it detects, and an explicit 'Use when' trigger clause. The only weakness is slightly narrow trigger vocabulary within its niche (no MCP/ecosystem synonyms).

Suggestions

Add one or two ecosystem synonyms users naturally say to the trigger clause, e.g. 'Use when reviewing for agent-native architecture, MCP/tool design, action and context parity...' to broaden trigger matching without adding vagueness.

DimensionReasoningScore

Specificity

The description lists multiple concrete review capabilities: gaps 'where a user can do something an agent cannot, where the agent lacks the context to act, or where agent tools encode workflows instead of primitives'. Three distinct, specific gap types plus the review action itself — comprehensive coverage for a single-purpose skill. Not 4: unlike 'minor gaps in coverage', every capability of the skill is explicitly enumerated.

5 / 5

Completeness

Explicitly answers both parts: 'what' is 'Review a code change for gaps...' with three concrete gap types, and 'when' is an explicit 'Use when reviewing for agent-native architecture, action and context parity, agent tool design, or system prompt context' clause with concrete trigger phrases. This matches the score-5 anchor exactly.

5 / 5

Trigger Term Quality

Good natural-term coverage within its niche: 'agent-native architecture', 'action and context parity', 'agent tool design', 'system prompt context' — phrases a user requesting this review would plausibly say. Not 5: misses common synonyms and ecosystem terms a user might naturally use, e.g. 'MCP', 'tools', 'agent experience', 'agentic'. Not 3: the terms present are specific and well-matched, not merely 'some relevant keywords'.

4 / 5

Distinctiveness Conflict Risk

Clear niche — an agent-native architecture review lens — with distinct trigger terms ('agent-native', 'action and context parity', 'agent tool design') that no generic code-review skill would claim. Not 4: overlap risk with a generic 'review this change' request is addressed by the 'Use when' clause narrowing to agent-native concerns, leaving minimal conflict risk.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
perihelionhq/perihelion-platform-context
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.