CtrlK
BlogDocsLog inGet started
Tessl Logo

gherkin-specification

Elicit or revise software behavior, maintain a recoverable behavior workpiece, and author or review honest Gherkin feature documents. Use for a behavior-specification interview, Gherkin document, executable-specification draft, or review of any of them.

49

Quality

53%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./libs/@hashintel/brunch-agent/evaluations/protocols/gherkin-shape-c-paper-v1/instrument/gherkin-specification/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

38%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body sketches a coherent lifecycle with honest capability-checking language, but it is an instruction shell: all substantive detail is delegated to four reference files that are absent from the bundle, and the inline guidance itself is too abstract to act on without them. Adding the referenced files and a concrete example or template would move both actionability and progressive disclosure up substantially.

Suggestions

Ship the four referenced files (universal-elicitation.md, gherkin-elicitation.md, workpiece-template.md, gherkin-authoring-and-checks.md) in references/ — every procedural instruction currently points at material that does not exist.

Include one short inline example (a worked Given/When/Then snippet or a filled-in workpiece skeleton) so the skill is actionable even before the references are read.

Cut the philosophical meta-commentary (e.g. "Authoring is a thin projection and correction surface, not a separate modelling world") and replace abstract check language with concrete checkpoints, such as how to verify a draft parses or matches a supplied step vocabulary.

DimensionReasoningScore

Conciseness

The body is compact and does not explain Gherkin basics Claude already knows, but it spends tokens on abstract meta-commentary — "Authoring is a thin projection and correction surface, not a separate modelling world", "A target-shaped draft does not replace these distinctions while they remain load-bearing" — that could be tightened, matching "mostly efficient but includes some unnecessary explanation" rather than the lean level-4/5 anchors.

3 / 5

Actionability

Guidance stays high-level: "follow one concrete example through its starting context, one focal event or action, and observable outcome" describes rather than instructs, and there is no sample Gherkin snippet, workpiece excerpt, or concrete question/output example to execute against — matching "minimal concrete guidance; high-level hints but missing the specific steps" rather than level 3, which would require at least some concrete worked detail.

2 / 5

Workflow Clarity

The lifecycle (orient → elicit/revise → maintain workpiece → author Gherkin → check and deliver) and the interactive vs. render-only branches are clearly sequenced, but verification is abstract — "Apply the checks supported by the current capabilities" and "do not claim an unavailable check occurred" name no concrete checkpoint, matching the implicit-checkpoints level-3 anchor; it is above level 2 because the sequence itself is coherent and complete.

3 / 5

Progressive Disclosure

The body signals one-level-deep references cleanly (e.g. "Read `universal-elicitation.md` and `gherkin-elicitation.md`"), but none of the four referenced files — universal-elicitation.md, gherkin-elicitation.md, workpiece-template.md, gherkin-authoring-and-checks.md — exist in the bundle (no references/ directory), so every pointer dangles and the skill's substance is unreachable, matching the minimal-structure anchor rather than level 3 where references are at least present.

2 / 5

Total

10

/

20

Passed

Description

67%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A solid description with explicit what and when clauses in third-person voice and a clear Gherkin niche. Its main weakness is trigger vocabulary that leans on the skill's own coined terms rather than the natural synonyms and file extensions users would actually say.

Suggestions

Add natural trigger synonyms users would say, e.g. "Use when the user mentions Gherkin, feature files, .feature files, BDD scenarios, or Cucumber."

Replace or gloss the coined term "recoverable behavior workpiece" in the description with plain language (e.g. "maintain a running behavior account") so the what-clause is concrete to unfamiliar readers.

DimensionReasoningScore

Specificity

Lists several concrete third-person actions — "Elicit or revise software behavior", "maintain a recoverable behavior workpiece", "author or review honest Gherkin feature documents" — but "recoverable behavior workpiece" is abstract coined jargon, keeping it below the comprehensive level-5 anchor and above the 1-2-action level-3 anchor.

4 / 5

Completeness

Both what (three stated actions) and when ("Use for a behavior-specification interview, Gherkin document, executable-specification draft, or review of any of them") are explicitly present, but the when-clause reuses coined phrases instead of natural user trigger phrasing, so it does not clearly reach the level-5 anchor's "concrete trigger phrases"; it is well above the weakly-implied-when level 3.

4 / 5

Trigger Term Quality

"Gherkin document" is a natural term users would say, but the other triggers ("behavior-specification interview", "executable-specification draft") are the skill's own vocabulary, and common synonyms like "feature file", ".feature", "BDD", "Cucumber", or "scenarios" are missing — matching the "some relevant keywords but missing common variations" anchor rather than the good-coverage level 4.

3 / 5

Distinctiveness Conflict Risk

"Gherkin feature documents" carves a mostly distinct niche, but the broad opener "Elicit or revise software behavior" could overlap with general requirements- or spec-writing skills, matching "mostly distinct; minor overlap risk" rather than the minimal-conflict level 5.

4 / 5

Total

15

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
hashintel/hash
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.