CtrlK
BlogDocsLog inGet started
Tessl Logo

panel-review-loop

Iteratively review and improve a Fallow user-facing surface across representative real-world projects until the panel has no blocking concerns.

60

Quality

70%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/panel-review-loop/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an exemplary lean instruction-only skill: a precise setup pointer, a well-sequenced iterative loop with an explicit stop criterion, and stability constraints stated without padding. The only weakness is that a couple of steps (capturing output, invoking panel-review) assume context that is not provided inline.

DimensionReasoningScore

Conciseness

The body is lean and efficient: every line is instruction (a setup pointer with an exact command, a seven-step loop, and two constraint notes), with zero padding and no explanation of concepts Claude already knows. This matches the 'every token earns its place' anchor exactly, and there is nothing to trim that would justify a 4.

5 / 5

Actionability

Concrete, executable guidance dominates: 'run `npm --prefix benchmarks run download-fixtures`', 'Select representative projects from `benchmarks/fixtures/real-world/`', 'Run `panel-review` on the evidence'. Minor gaps keep it below 5 — how to 'Capture actual output' and how `panel-review` is invoked are left unspecified, assuming a sibling skill or prior context.

4 / 5

Workflow Clarity

A clearly numbered 1-7 sequence with an explicit feedback loop (run panel-review on evidence -> implement smallest coherent improvement -> re-run the same corpus and compare) and an explicit stop condition in step 7. It is not a 5 because checkpoints within steps 3-4 (capturing output, invoking panel-review) are implicit, leaving minor validation gaps typical of the 4 anchor.

4 / 5

Progressive Disclosure

This is a short single-purpose skill with no bundle files; its only external pointer is the clearly signaled, one-level-deep link to BENCHMARKS.md for fixture setup, and stability constraints are kept inline where they belong. Per the simple-skill guideline, well-organized short content with no need for external references scores 5.

5 / 5

Total

18

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description communicates a clear, specific purpose in third person with a well-defined convergence condition, but it omits any explicit trigger guidance ('Use when...'), which both caps completeness and weakens natural keyword coverage. Its distinctiveness is carried by project-specific terminology rather than by distinct trigger terms.

Suggestions

Add an explicit trigger clause, e.g. 'Use when improving a Fallow user-facing surface or when the panel reports blocking concerns', to raise completeness from its capped 3.

Include natural synonyms users would say (e.g. 'panel review', 'iterate until no blockers', 'real-world corpus') alongside the domain jargon to improve trigger term quality.

Briefly name the artifacts involved (e.g. benchmark fixtures, captured output, panel verdicts) so the 'what' covers more than the two generic actions 'review' and 'improve'.

DimensionReasoningScore

Specificity

The description names its domain ('a Fallow user-facing surface', 'the panel') and two concrete actions ('Iteratively review and improve ... across representative real-world projects'), matching the anchor for 1-2 concrete actions without comprehensive coverage. It is not a 4 because it does not list several specific actions — what the review covers or what 'improve' entails is left open — and not a 2 because the domain is named with real actions rather than generic processing.

3 / 5

Completeness

The 'what' is clear — iteratively review and improve a surface across representative projects until the panel has no blocking concerns — but there is no 'Use when...' clause or equivalent trigger guidance, which caps completeness at 3 per the judging guidelines. It is not a 2 because the 'what' is specific and well-stated, not vague.

3 / 5

Trigger Term Quality

Relevant keywords exist ('review', 'improve', 'panel', 'real-world projects'), but the natural phrases a user would actually say are largely absent and jargon ('Fallow user-facing surface', 'blocking concerns') dominates. It sits above the 2 anchor ('one or two generic keywords') but below the 4 anchor, which expects good natural-term coverage with only a few missing.

3 / 5

Distinctiveness Conflict Risk

The niche terms ('Fallow', 'the panel', 'blocking concerns') make it mostly distinct from other skills with only minor overlap risk against generic review/iterate skills. It is not a 5 because distinct trigger phrases are absent, so distinctiveness rests on project-specific jargon rather than clearly distinct triggers.

4 / 5

Total

13

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 1 suspicious

Warning

Total

15

/

16

Passed

Repository
fallow-rs/fallow
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.