CtrlK
BlogDocsLog inGet started
Tessl Logo

caveman-review

Compressed code review - one line per finding with location, problem and fix. Use for /caveman-review, "review this PR", or "review the diff".

74

Quality

91%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary terse instruction skill: every rule is concrete, the ❌/✅ examples make the target format unambiguous, and scope exceptions and boundaries are explicitly handled. The body is fully self-contained at ~45 well-organized lines with no wasted tokens.

DimensionReasoningScore

Conciseness

The body is lean throughout: rules are one-liners ("Exact line numbers", "Concrete fix, not 'consider refactoring this'"), the Drop list enumerates filler phrases without explaining why they are bad, and no concept Claude already knows is re-explained. Every token earns its place, matching the anchor-5 'lean and efficient' standard; there is no padded section that would pull it to 4.

5 / 5

Actionability

Guidance is fully concrete and copy-paste ready for an instruction-only skill: an exact output template (`L<line>: <problem>. <fix>.`), four defined severity prefixes ("🔴 bug: — broken behavior, will cause incident"), and three ❌/✅ example pairs showing exactly how to transform verbose comments into the format. This matches anchor 5 ('specific examples cover the common cases') rather than 4, which would leave minor gaps in executable guidance.

5 / 5

Workflow Clarity

This is a simple single-purpose skill (under 50 lines, one task), so per the rubric's simple-skill exception workflow clarity can score 5 when the single action is unambiguous — and it is: format, severities, drop/keep lists, exceptions (Auto-Clarity), and boundaries are each explicitly scoped. The Auto-Clarity section even sequences the exception handling (drop terse mode for security findings, 'then resume terse for the rest'), and the skill involves no destructive or batch operations that would require validation checkpoints.

5 / 5

Progressive Disclosure

Per the rubric's guidance, a sub-50-line skill with no need for external references scores 5 with well-organized sections alone, and this body has exactly that: Rules, Examples, Auto-Clarity, and Boundaries are clearly headed, each small enough to belong inline, with no separate reference files needed or buried. There is no inlined content that should live in another file, so it does not fall to the anchor-4 'minor organization gaps' level.

5 / 5

Total

20

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that answers both 'what' and 'when' concisely with concrete, user-natural trigger phrases and a precise output format. The main gaps are missing synonyms for the trigger terms and generic review phrasing that overlaps with a conventional code-review skill.

Suggestions

Add natural trigger synonyms such as 'code review', 'review my changes', or 'review these commits' to broaden trigger term coverage.

Sharpen distinctiveness by signaling the compression/style angle in the 'Use for' clause (e.g., "Use when the user wants terse, compressed review comments") so it does not compete with a standard review skill for generic 'review this PR' triggers.

DimensionReasoningScore

Specificity

"Compressed code review - one line per finding with location, problem and fix" names the domain and specifies the exact output shape (per-finding line containing location, problem, fix), which is more concrete than the anchor-3 example ('Processes PDF files and extracts content'). It stays at 4 rather than 5 because it describes a single action's format rather than the multiple distinct actions in the anchor-5 example, leaving minor coverage gaps (e.g., severity tiers or multi-file format are not hinted at).

4 / 5

Completeness

It explicitly answers both questions: what ("Compressed code review - one line per finding with location, problem and fix") and when ("Use for /caveman-review, 'review this PR', or 'review the diff'") with concrete trigger phrases, exactly matching the anchor-5 good example pattern. It is not a 4 because the 'when' clause is already explicit and quoted, not merely present-but-vague.

5 / 5

Trigger Term Quality

"Use for /caveman-review, 'review this PR', or 'review the diff'" includes very natural phrases a user would actually say, matching the anchor-4 'good keyword coverage; a few natural terms missing' level. It is not a 5 because common synonyms like 'code review', 'review my changes', or 'look over this PR' are absent, so coverage is good rather than comprehensive.

4 / 5

Distinctiveness Conflict Risk

The named slash-command trigger and the compressed one-liner format carve out a mostly distinct niche, but 'review this PR' and 'review the diff' are generic review triggers that would also fire for any standard code-review skill, matching anchor 4 ('minor overlap risk with closely related skills'). It is not a 5 because those shared triggers leave more than minimal conflict risk with a sibling review skill; not a 3 because the format description and /caveman-review command clearly disambiguate intent.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
JuliusBrussee/caveman
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.