CtrlK
BlogDocsLog inGet started
Tessl Logo

conformance-loop

Iteratively improve Fallow analysis accuracy by comparing it with competing tools and verified source truth across a stable real-world corpus.

57

Quality

64%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/conformance-loop/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a concise, well-sequenced instruction-only workflow with concrete artifacts and built-in validation gates. Minor abstractness in two steps and a missing explicit regression-failure feedback loop are the only weak spots.

DimensionReasoningScore

Conciseness

Lean, terse steps that assume Claude's competence, with one minor instance of over-explanation in step 5's rationale clause that could be trimmed; not perfectly token-efficient as anchor 5.

4 / 5

Actionability

Names a concrete artifact path, a real command ('Run review'), and a defined classification taxonomy, with minor gaps where tool-run settings and 'one general correction' stay abstract.

4 / 5

Workflow Clarity

Clear 8-step sequence with validation checkpoints (regression fixture, re-run-and-retain-only-net-improvements, review), but lacks an explicit error-recovery feedback loop and checklist, stopping short of anchor 5.

4 / 5

Progressive Disclosure

A short, self-contained skill under 50 lines with no need for external references; the single well-ordered numbered list is well-organized, qualifying for 5 under the simple-skill exception.

5 / 5

Total

17

/

20

Passed

Description

51%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description conveys a specific capability niche but omits explicit trigger guidance, so completeness is capped and trigger-term quality is weak. Specificity and distinctiveness are solid.

Suggestions

Add a 'Use when...' clause naming the natural situations a user would invoke this skill (e.g., 'Use when validating Fallow findings against other tools or adjudicating disagreements').

Include natural user-facing keywords and synonyms alongside the internal jargon so the skill triggers reliably from ordinary requests.

Spell out one or two more concrete actions (e.g., 'record adjudications', 'add regression fixtures') to move specificity toward comprehensive coverage.

DimensionReasoningScore

Specificity

Names several concrete actions ('comparing it with competing tools', 'verified source truth', 'stable real-world corpus', 'iteratively improve'), with minor gaps in coverage; not as comprehensively tangible as anchor 5.

4 / 5

Completeness

States a clear 'what' (improve accuracy via comparison and verification) but provides no 'when'/'Use when' trigger guidance, which caps completeness at 3 per the rubric.

3 / 5

Trigger Term Quality

Relies on internal jargon ('Fallow analysis accuracy', 'corpus') with a couple recognizable terms, but misses the natural phrases a user would say and lacks synonyms or variations.

2 / 5

Distinctiveness Conflict Risk

Carves a fairly specific niche (Fallow conformance/evaluation) with mostly distinct triggers and only minor overlap risk, though triggers are not as concrete as anchor 5.

4 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
fallow-rs/fallow
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.