CtrlK
BlogDocsLog inGet started
Tessl Logo

example-argument-substitution

Test harness for Claude Code skill argument substitution — demonstrates capture-block pre-declaration, XML tag referencing, unintentional variable corruption in code blocks, and correct placement of shell examples in reference files. Use when verifying substitution behavior before applying a pattern to other skills, testing how arguments flow from skill invocations, or understanding the pre-declaration and reference file pattern with greet/farewell/inspect actions.

74

Quality

92%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered test-harness skill: executable invocation commands, a hypothesis-driven 7-step testing loop with validation, concrete expected-output annotations for every section, and an appropriately placed one-level-deep reference file. The only weakness is minor redundancy (restated reference-file rules, a long variable enumeration) that could be trimmed for token efficiency.

DimensionReasoningScore

Conciseness

The body is largely efficient — no padding explaining what substitution is at length, and the repeated "Expected with 10 args / Expected with 0 args" annotations are functional test fixtures rather than filler. Not 5 because there are minor instances that could be trimmed: the fact that reference files are not substituted is stated twice in the body ("Reference files are NOT subject to substitution" and again in the Correct Pattern section), and the inspect output block enumerates 13 captured variables when a representative subset plus a pointer would do.

4 / 5

Actionability

Fully executable guidance throughout: exact invocation commands ("/example-argument-substitution CANARY_A CANARY_B ..."), concrete code examples with literal expected corruption, a mermaid routing diagram, and copy-paste-ready action output templates. The common cases (greet, farewell, inspect, empty/unknown action) are all covered with specific outputs.

5 / 5

Workflow Clarity

Clear sequenced workflows with explicit validation: the 4-step usage procedure ends in "Compare... Check whether output matches", and the 7-step add-a-test procedure is a full feedback loop (hypothesis → run → observe → record finding → only then apply), enforced by "Do not document any pattern as safe without completing all 7 steps." This matches the score-5 anchor with error-recovery loops; no destructive or batch operation is involved, so no cap applies.

5 / 5

Progressive Disclosure

Scored against the actual bundle: one reference file (references/argument-substitution-reference.md) exists, is one level deep, contains no nested references, and is clearly signaled at the end of the body with its scope stated ("All substitution variables, pitfall table, and verified escape evidence"). The demonstration content that remains inline must be there by design — substitution only applies to the SKILL.md body, which is the very behavior under test — so the split is appropriate rather than monolithic.

5 / 5

Total

19

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states a precise niche, enumerates four concrete capabilities, and gives an explicit 'Use when...' clause covering verification, testing, and learning use cases. Trigger coverage is good but could add a few more natural synonyms like "$ARGUMENTS" or "positional arguments".

DimensionReasoningScore

Specificity

The description lists four concrete capabilities — "demonstrates capture-block pre-declaration, XML tag referencing, unintentional variable corruption in code blocks, and correct placement of shell examples in reference files" — giving comprehensive, non-generic coverage of what the skill does. Not score 4 because there is no noticeable gap in coverage for this domain; not below 5 because every action named is concrete rather than abstract.

5 / 5

Completeness

Both what and when are explicit: what — "Test harness for Claude Code skill argument substitution — demonstrates..."; when — "Use when verifying substitution behavior before applying a pattern to other skills, testing how arguments flow from skill invocations, or understanding the pre-declaration and reference file pattern". This matches the score-5 anchor exactly with concrete trigger phrases; score 4 would require a less explicit 'when', which is not the case.

5 / 5

Trigger Term Quality

Good keyword coverage with natural phrases like "argument substitution", "reference files", "skill invocations", and the greet/farewell/inspect action names. Not 5 because some natural variations users might say are missing (e.g., "$ARGUMENTS", "positional arguments", "skill arguments"); not 3 because coverage goes well beyond a couple of generic keywords.

4 / 5

Distinctiveness Conflict Risk

"Test harness for Claude Code skill argument substitution" is a clear niche with distinct triggers (substitution verification, pre-declaration pattern, reference-file placement) unlikely to fire for unrelated skills. Not 4 because no meaningful overlap with other skill domains exists; the scope is tightly bounded to skill-authoring internals.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
Jamie-BitFlight/claude_skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.