CtrlK
BlogDocsLog inGet started
Tessl Logo

dev-review

AI DevKit · Final code review phase guidance for holistic pre-push review. Use when the user wants code review, final lifecycle review, design alignment checks, integration risk review, or dev-lifecycle phase 9.

66

Quality

83%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

This is a well-built instruction-only skill body: dense, imperative, and free of filler, with concrete commands, an explicit validation loop (verify skill, blocking-findings routing, final checklist), and clean single-file organization appropriate to its size. The only real gap is the underspecified interface to the `task` tracing system, referenced but never shown.

DimensionReasoningScore

Conciseness

The body is a lean, imperative checklist ('Run `npx ai-devkit@latest lint`', 'grep exported names to trace callers', 'Check consistency against 1-2 similar modules') with zero explanation of concepts Claude already knows and no padding. Every line is an instruction, matching 'Lean and efficient; assumes Claude's competence; every token earns its place.'

5 / 5

Actionability

Concrete commands are given (`npx ai-devkit@latest lint --feature <name>`, `git status -sb`, `git diff --stat`) and most steps name exactly what to check. Minor gaps remain: the tracing-event step says to 'emit review phase, progress, blocker/finding... events per `task`' without showing the event format, and no example of a finding's expected shape is given, so it falls just short of fully copy-paste-ready guidance.

4 / 5

Workflow Clarity

The sequence is explicit (Phase Contract 1-6, review steps 1-17, then a Done checklist) with validation checkpoints and error-recovery routing: 'Apply the `verify` skill before claiming readiness', 'if blocking issues remain, return to `dev-implementation` or `dev-testing`', and the final 6-item readiness checklist. This matches the anchor 'Clear sequence with explicit validation steps; feedback loops for error recovery; checklists for complex processes.'

5 / 5

Progressive Disclosure

The body is ~37 lines with no bundle files (references/, scripts/, assets/ are absent), organized into clear sections (Phase Contract, Code Review, Done) that hold all needed content inline. Cross-references to sibling skills (`verify`, `task`) are one level deep and clearly signaled. Per the rubric's simple-skill guideline, this earns a 5.

5 / 5

Total

19

/

20

Passed

Description

66%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description has a clear, explicit trigger clause with good natural keywords and proper third-person phrasing. Its weaknesses are a somewhat abstract statement of what the skill does and a broad 'code review' trigger that risks collision with generic review skills, partially mitigated by the DevKit/phase-9 scoping.

Suggestions

Make the 'what' concrete: replace 'Final code review phase guidance for holistic pre-push review' with the specific actions performed, e.g., 'Runs holistic pre-push review: verifies design alignment, checks contract and boundary integrity, flags breaking changes and rollback risk, and reports severity-ordered findings with file/line references.'

Sharpen the trigger clause so it does not compete with generic review requests, e.g., 'Use when the user wants the final AI DevKit lifecycle review, a design-alignment or integration-risk check before pushing, or dev-lifecycle phase 9.'

Add the natural synonyms users are likely to say ('pre-push review', 'review before PR') to broaden trigger coverage beyond 'code review'.

DimensionReasoningScore

Specificity

Phrases like 'Final code review phase guidance', 'design alignment checks', and 'integration risk review' name the domain and a couple of concrete review activities, but coverage of what the skill actually does is not comprehensive (no mention of findings/severity reporting, breaking-change checks, or rollback safety). This matches the anchor 'Names domain and 1-2 concrete actions, but not comprehensive'; a 4 would require several specific actions with only minor gaps.

3 / 5

Completeness

It has an explicit 'Use when...' clause listing five triggers (the 'when'), and a 'what' in 'Final code review phase guidance for holistic pre-push review'. The 'what' is somewhat abstract ('guidance', 'holistic') rather than concrete actions, so it sits at 'both present; when/what could be more explicit' — below the 5 anchor where both are concrete.

4 / 5

Trigger Term Quality

'Use when the user wants code review, final lifecycle review, design alignment checks, integration risk review' gives good coverage of natural phrases a user would say, anchored by the strong 'code review' trigger. A few natural variations are missing (e.g., 'review my changes', 'pre-PR check'), and 'dev-lifecycle phase 9' is internal jargon, keeping it below the 5 anchor's comprehensive synonym coverage.

4 / 5

Distinctiveness Conflict Risk

'Use when the user wants code review' is a very broad trigger that would overlap with generic code-review skills, though the 'AI DevKit', 'dev-lifecycle phase 9', and 'final lifecycle review' qualifiers carve out a narrower niche. This lands at 'somewhat specific but could still overlap with similar skills' rather than 4's 'mostly distinct'.

3 / 5

Total

14

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
codeaholicguy/ai-devkit
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.