CtrlK
BlogDocsLog inGet started
Tessl Logo

orchard-review

Review a release proposal

44

Quality

55%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./projects/skill-scanner/examples/benign/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

65%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is maximally token-efficient and appropriately self-contained for its size, but it is too underspecified to act on: it never says how to obtain the deployment diff or what the review output should be. Workflow clarity is middling because the sequence exists yet the deliverable and checkpoints are implicit.

Suggestions

Add the concrete step for obtaining the diff, e.g. "Run `git diff <base>..HEAD` (or read the attached patch file) to get the deployment diff" so the guidance is executable.

Specify the output format for the review, e.g. "For each changed replica setting, quote the old and new value with its file and line, then flag any that alter replica counts or resources".

Add a completion checkpoint such as confirming every changed replica setting is quoted before finishing the review, which would also raise workflow clarity.

DimensionReasoningScore

Conciseness

"Read the deployment diff. Quote changed replica settings." is two lean imperative sentences with zero padding, zero explanation of concepts Claude already knows, and nothing to trim. This matches 'Lean and efficient; assumes Claude's competence; every token earns its place'.

5 / 5

Actionability

The body names concrete artifacts (deployment diff, replica settings) but gives no executable specifics: no command to obtain the diff (e.g. `git diff main...`), no file paths, and no statement of the expected output format. This matches 'Minimal concrete guidance; high-level hints but missing the specific steps to execute'. Not 1 because it instructs with named concrete objects rather than only describing; not 3 because there is no partially executable guidance at all — nothing here can be run or followed mechanically.

2 / 5

Workflow Clarity

A rough two-step sequence is present (read the diff, then quote the changed replica settings), but there are no checkpoints and the deliverable is underspecified — where the diff lives and what the final output should look like are left implicit. This matches 'Steps listed but validation gaps; sequence present but checkpoints missing or implicit'. Not 4 because the gaps are not minor (the review's definition of done is unstated); not 2 because the sequence is coherent and both steps are defined at a high level. The operation is read-only, so the destructive/batch validation cap does not apply.

3 / 5

Progressive Disclosure

The skill is under 50 lines, has no bundle files (references/, scripts/, assets/ do not exist), references nothing external, and inlines nothing that belongs in a separate file. Per the rubric's simple-skill guidance, a skill this small with no need for external references scores 5: the entire appropriately-scoped content sits in SKILL.md with nothing buried.

5 / 5

Total

15

/

20

Passed

Description

25%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is a three-word phrase that names a domain but gives no concrete actions, no trigger guidance, and no distinguishable niche. It sits at the second-lowest anchor on every dimension and resembles the rubric's bad example class ("Helps with documents") more than any good example. Voice is acceptable (imperative/third-person), so no voice penalty applies.

Suggestions

State the concrete actions the skill performs, e.g. "Reads the deployment diff and quotes changed replica settings with file/line references" instead of the generic verb "Review".

Add an explicit 'Use when...' clause, e.g. "Use when the user shares a release proposal, rollout plan, or deployment diff for review" — this also lifts the completeness cap.

Include natural trigger terms and synonyms users would actually say (release proposal, rollout, deployment diff, release notes) to improve trigger term quality and distinctiveness.

DimensionReasoningScore

Specificity

"Review a release proposal" names the domain (release proposal) but the single action "Review" is generic — it never says what the review does (read a diff, quote settings, summarize, approve). This matches the anchor 'Names the domain but actions are minimal or generic' ("Processes PDF files"). Not 3 because no concrete action is listed; not 1 because the domain is explicitly named.

2 / 5

Completeness

There is only a vague 'what' ("Review a release proposal" — what the review produces is unstated) and no 'when' clause at all, matching 'Has a vague what and no when'. Not 1 because a domain and action are named; not 3 because the 'what' is not clear — the deliverable of the review is never stated — and the explicit 'Use when...' guidance is entirely missing (which alone caps this dimension at 3).

2 / 5

Trigger Term Quality

The only keywords are "review" and "release proposal" — one bare domain phrase with no synonyms or natural variations a user might say (rollout, deployment, release notes, ship). This sits at 'One or two generic keywords; missing the natural phrases users say' ("Works with files"). Not 3 because the anchor-3 example ("Works with PDF files") carries a specific, recognizable artifact keyword with broader coverage; here coverage is a single thin phrase.

2 / 5

Distinctiveness Conflict Risk

"Review" is an ultra-common verb that would collide with any review-type skill (code review, PR review, security review), and only the phrase "release proposal" narrows the scope. This fits 'Very broad; high overlap risk with many similar skills' rather than anchor 3, because the dominant trigger word is generic and nothing distinguishes this skill from other review skills.

2 / 5

Total

8

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
rohitg00/ai-engineering-from-scratch
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.