CtrlK
BlogDocsLog inGet started
Tessl Logo

autoreview

Structured Codex, Claude, Amp, Pi, or Kimi code review when explicitly requested.

56

Quality

64%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./templates/plate-template/.agents/skills/autoreview/SKILL.md

The canonical home for this skill is autoreview in openclaw/openclaw

SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, executable reference for the autoreview helper: real commands, tables, and status semantics with verification checkpoints throughout. Its weaknesses are the inlining of long engine-auth and Git-config policy that belongs in reference files, and the absence of a compact step-by-step run sequence that would make the main path easier to follow.

Suggestions

Move the engine config-projection/authentication detail (the '--codex-config', launcher, and catalogue paragraphs) into a reference file linked from the Engines section.

Add a short numbered quick-start sequence (select target → run → read exit code/status → verify findings) at the top so the main path is scannable before the edge-case policy.

Trim restated edge cases in the autocrlf and status-output paragraphs to reduce token load without losing behavior.

DimensionReasoningScore

Conciseness

The body is dense operational policy with essentially no explanation of concepts Claude already knows — every section states flags, constraints, or semantics. Minor trimming is possible (e.g., the auth-route projection and catalogue paragraphs restate edge cases at length), so it sits at 'Efficient; minor instances of over-explanation' rather than the uniformly lean score of 5.

4 / 5

Actionability

Concrete, runnable guidance throughout: the quick-start bash snippet, the merge-base example, flag tables for modes/engines/exits, and a complete JSON sidecar example. Minor gaps keep it below 5 — `<ref>` placeholders need filling and some sections are policy statements (e.g., 'Never work around an isolation failure') rather than commands.

4 / 5

Workflow Clarity

The path is unambiguous: pick a Git target from the mode table, run the helper, interpret the exit-code table and status sidecar, and verification checkpoints are present ('Verify findings against the actual code', 'resolve incomplete before claiming completion', 'A failed pass does not produce a partial clean verdict'). It is not a numbered sequence with error-recovery loops, so it fits 'Clear sequence with most checkpoints present; minor validation gaps' rather than 5.

4 / 5

Progressive Disclosure

Sections are well organized and the heavy detail is appropriately delegated to the script and `--help` (scripts/autoreview exists as referenced), but the bundle has no reference files while ~14KB of engine-routing, authentication, and Git-configuration policy is inlined in SKILL.md — content that would fit the anchor 'content that should be separate is inline'. Not a 2 because the inlined material is well-sectioned and navigable, and no references are buried.

3 / 5

Total

15

/

20

Passed

Description

57%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description identifies a clear niche (multi-engine structured code review) but undersells the skill's actual capabilities and provides almost no trigger guidance beyond 'when explicitly requested'. Adding concrete actions and explicit 'Use when...' trigger phrases would raise both completeness and specificity.

Suggestions

Add explicit trigger phrases, e.g. 'Use when the user asks for a code review, mentions reviewing a PR/diff, or names a specific engine (Codex, Claude, Amp, Pi, Kimi).'

List 2-3 concrete capabilities (e.g., 'reviews local changes, branches, or single commits; filters findings by severity; emits validated JSON reports') so the 'what' is more than one action.

Include natural synonyms users say — 'PR review', 'diff review', 'review my changes' — to strengthen trigger term coverage.

DimensionReasoningScore

Specificity

The description names the domain ('code review') and one concrete action shape ('Structured ... review') plus the five engine names ('Codex, Claude, Amp, Pi, or Kimi'), but lists no other concrete capabilities such as diff analysis, severity filtering, or validated reports. This matches the anchor 'Names domain and 1-2 concrete actions, but not comprehensive'; score 4 would require several specific actions listed, which is not met.

3 / 5

Completeness

The 'what' is clear (structured multi-engine code review), but the only 'when' guidance is 'when explicitly requested', which names no concrete trigger phrases and would apply to nearly any skill, making it effectively weak. The explicit-trigger cap of 3 applies: 'Use when...' or equivalent concrete trigger guidance is missing.

3 / 5

Trigger Term Quality

'code review' is the natural phrase users say, and the engine names (Codex, Claude, Amp, Pi, Kimi) act as strong additional triggers for engine-specific requests. Common variations like 'review this PR', 'diff', or 'review my changes' are missing, matching 'Good keyword coverage; a few natural terms missing' rather than the comprehensive synonym coverage of a 5.

4 / 5

Distinctiveness Conflict Risk

Naming five specific engines gives it a niche, but the generic trigger 'code review ... when explicitly requested' would fire on any review request and overlaps with common built-in review skills. This fits 'Somewhat specific but could still overlap with similar skills'; a 4 would require the overlap risk to be minor only.

3 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
udecode/plate
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.