CtrlK
BlogDocsLog inGet started
Tessl Logo

ac-criteria-validator

Validate acceptance criteria and feature completion. Use when checking if features pass, validating test results, verifying acceptance criteria, or determining feature completion status.

62

Quality

78%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/ac-criteria-validator/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

68%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, mostly lean and actionable body with a real referenced script and concrete examples. The main weakness is the Workflow section, which lists steps without validation checkpoints or feedback loops for failure cases, capping workflow clarity at 3.

Suggestions

Expand the Workflow section with explicit validation checkpoints and failure handling, e.g. '3. Execute: run the test suite — if any test fails, record the failing criterion and mark passes=false; do not retry blindly' and a final verify step before reporting.

Define 'project_dir' in the Quick Start (e.g., 'validator = CriteriaValidator(Path("."))') so the example is fully copy-paste ready, and add one concrete command showing how related tests are discovered for a feature id.

Trim redundancy: remove the duplicated purpose line under the H1 and cut the generic 'Manual Criteria' bullet list, or move the Validation Methods and custom-rules details into a reference file.

DimensionReasoningScore

Conciseness

The body is efficient — no explanations of concepts Claude already knows, and sections like Validation Methods and Validation Rules are terse bullet/config blocks. It falls short of a 5 due to minor redundancy: the line under the H1 duplicates the Purpose section, and the 'Manual Criteria' bullets ('UI/UX requirements, Performance benchmarks...') are generic filler that could be trimmed.

4 / 5

Actionability

The Quick Start gives mostly executable code ('from scripts.criteria_validator import CriteriaValidator... result = await validator.validate_feature("auth-001")') backed by a real script, plus a concrete result JSON and a YAML rules block. Minor gaps keep it from a 5: 'project_dir' is never defined in the example, and key operations like how tests are discovered or how criteria are matched to tests have no concrete commands — matching 'Mostly executable guidance; concrete code or commands with minor gaps'.

4 / 5

Workflow Clarity

The five-step Workflow ('1. Load... 5. Report') is a clear sequence, but each step is a one-word label with no validation checkpoints or failure handling — what to do when tests fail, when a criterion can't be verified, or whether to re-run after fixes. This matches 'Steps listed but validation gaps; sequence present but checkpoints missing or implicit'; the Validation Rules section defines thresholds but the workflow never wires them into the sequence, so it does not reach a 4.

3 / 5

Progressive Disclosure

Good structure: Quick Start up front, well-organized sections, and the API Reference correctly points one level deep to the real 'scripts/criteria_validator.py' rather than inlining the implementation. Not a 5 because detailed material (the Validation Methods breakdown and custom-rule YAML) is inlined in SKILL.md where the rubric favors pointing to reference files, and the skill exceeds the under-50-line simple-skill exception.

4 / 5

Total

15

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person, explicit 'Use when' trigger clause, and natural user phrasings with synonym coverage. Its main limitation is that the enumerated actions are repetitive variations of one capability rather than a set of distinct concrete actions.

DimensionReasoningScore

Specificity

The description names the domain ('Validate acceptance criteria and feature completion') and a couple of concrete actions, but the listed variations ('validating test results', 'verifying acceptance criteria', 'determining feature completion status') are near-synonyms of the same single action rather than several distinct capabilities. This matches the anchor 'Names domain and 1-2 concrete actions, but not comprehensive'; it is not score 4 because there are no genuinely distinct actions like analyzing coverage or generating validation reports.

3 / 5

Completeness

Both what and when are explicitly answered with concrete trigger phrases: 'Validate acceptance criteria and feature completion' (what) followed by 'Use when checking if features pass, validating test results, verifying acceptance criteria, or determining feature completion status' (when). This matches the anchor 'Clearly and explicitly answers both what AND when with concrete trigger phrases'.

5 / 5

Trigger Term Quality

Natural phrases users would say are present: 'checking if features pass', 'validating test results', 'verifying acceptance criteria', 'feature completion status' — good keyword coverage with synonym variation. Not score 5 because common variations a user might say (e.g., 'acceptance tests', 'QA gates', 'definition of done', 'does this feature meet requirements') are missing; not score 3 because the coverage clearly exceeds 'some relevant keywords' with multiple natural phrasings.

4 / 5

Distinctiveness Conflict Risk

'Acceptance criteria' establishes a clear niche with distinct triggers, so conflict risk is low, matching 'Mostly distinct; minor overlap risk'. Not score 5 because 'validating test results' could overlap with a generic test-execution skill, leaving minor overlap risk with closely related skills.

4 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
fernandezbaptiste/Skrillz
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.