CtrlK
BlogDocsLog inGet started
Tessl Logo

gherkin-specification

Elicit or revise software behavior, maintain a recoverable behavior workpiece, and author or review honest Gherkin feature documents. Use for a behavior-specification interview, Gherkin document, executable-specification draft, or review of any of them.

61

Quality

71%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./libs/@hashintel/brunch-agent/packages/plugin-gherkin/src/skills/gherkin-specification/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

56%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An instruction-only skill with a genuinely clear lifecycle structure and disciplined use of one-level-deep references, undermined by an abstract, jargon-heavy register and a missing `templates/workpiece.md` file that its own resource-discipline rules depend on. The sequencing is strong, but the body reads more like process theory than executable direction in places.

Suggestions

Ship the referenced `templates/workpiece.md` in the bundle (or remove/replace the reference), since the body instructs reading it when creating or materially revising the behavior account and the file is absent.

Rewrite abstract framing sentences ("thin projection and correction surface", "while they remain load-bearing", "signable behavior account") as short plain-language directives so every sentence tells the agent what to do.

Add a small concrete artifact inline — a minimal workpiece skeleton or a 5-line example Gherkin feature — so the authoring and delivery stages have an executable shape without waiting on external resources.

DimensionReasoningScore

Conciseness

The body avoids restating known concepts and defers detail to references, but abstract meta-framing like "Authoring is a thin projection and correction surface, not a separate modelling world" and "A target-shaped draft does not replace these distinctions while they remain load-bearing" is process philosophy that could be tightened into shorter plain directives. This is 'mostly efficient but includes some unnecessary explanation or could be tightened', not 4 because several such sentences add tokens without adding actionable instruction.

3 / 5

Actionability

There is some concrete guidance — "Activate the `elicitation` skill and read `references/gherkin-elicitation.md` before substantive questions", "Do not invent step-definition bindings", "Name unillustrated rules, ambiguous behavior, new or unchecked step phrases" — but key mechanics are deferred to an external `elicitation` skill and "core's" routing rule not in this bundle, and "Read `templates/workpiece.md`" points to a file absent from the bundle. Not a 4: these gaps leave an agent unable to execute parts of the instruction as written; not a 2 because per-stage direction is present and specific in intent.

3 / 5

Workflow Clarity

The lifecycle is clearly sequenced — orient, elicit or revise, maintain the workpiece, author or revise Gherkin, check and deliver — with an explicit runtime-branch selection up front and a "Check and deliver" stage that names what to report. Not a 5 because validation steps are named ("Apply the checks supported by the current capabilities") without any inline checkpoint or error-recovery detail; not a 3 because the sequence and its checkpoints (settle the workpiece before render-only handoff, deliver with gap accounting) are explicit.

4 / 5

Progressive Disclosure

The 52-line body is well-sectioned and points to two real one-level-deep references (`references/gherkin-elicitation.md`, `references/gherkin-authoring-and-checks.md`, both present), but it also instructs reading `templates/workpiece.md`, which does not exist in the bundle — the skill's own "Resource discipline" section makes that a broken navigation path. Not a 4: a dangling advertised resource is more than a minor organization gap; not a 2: most structure and references are appropriately placed and clearly signaled.

3 / 5

Total

13

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person, concise, with an explicit 'Use for...' clause covering both capability and triggers. Its only weakness is missing common synonyms (Cucumber, BDD, .feature files) that users of this domain would naturally say.

DimensionReasoningScore

Specificity

The description lists several concrete actions — "Elicit or revise software behavior, maintain a recoverable behavior workpiece, and author or review honest Gherkin feature documents" — which maps to the 'several specific actions; minor gaps in coverage' anchor. It is not a 5 because supporting activities like checking/executing specs are only implicit, and not a 3 because more than 1-2 distinct actions are named.

4 / 5

Completeness

It explicitly answers both questions: what ("Elicit or revise software behavior, maintain a recoverable behavior workpiece, and author or review honest Gherkin feature documents") and when ("Use for a behavior-specification interview, Gherkin document, executable-specification draft, or review of any of them") with concrete trigger phrases. Not a 4 because the 'when' clause is explicit and enumerates specific triggering artifacts, satisfying the top anchor.

5 / 5

Trigger Term Quality

Natural terms a user would say appear — "Gherkin", "Gherkin document", "executable-specification draft" — giving good coverage, but common synonyms like "Cucumber", "BDD", "feature file", or ".feature" are missing. This fits the 'good keyword coverage; a few natural terms missing' anchor rather than 5 (no synonym/extension coverage) or 3 (more than a few natural terms present here).

4 / 5

Distinctiveness Conflict Risk

"Gherkin feature documents" and "behavior-specification interview" carve out a clear niche with distinct triggers unlikely to collide with other skills. Not a 4: the triggers are domain-specific terms rather than broad categories with minor overlap.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
hashintel/hash
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.