CtrlK
BlogDocsLog inGet started
Tessl Logo

spike

Throwaway experiments to validate an idea before build.

51

Quality

58%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./skills/software-development/spike/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

73%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a strong, self-contained workflow: an explicit method loop, concrete tool sequences, decision guidance (risk-ordering, comparison spikes, delegation), and a validated verdict format. Its weaknesses are moderate: a few padded meta-sections and placeholder code blocks that keep it from fully lean and fully copy-paste ready.

DimensionReasoningScore

Conciseness

The body is mostly efficient — instructive tables, templates, and tool sequences with no concept explanations — but the 'If the user has the full GSD system installed' and 'Attribution' sections are non-essential padding and several passages could be tightened. Fits anchor 3 ('mostly efficient but includes some unnecessary explanation or could be tightened'); not 4 because those meta-sections consume real tokens without advancing the task.

3 / 5

Actionability

Concrete tool sequences ('terminal("mkdir -p spikes/001-websocket-streaming")'), a full directory layout, and complete verdict and head-to-head templates make the guidance mostly executable. Not 5 because write_file examples carry '...' placeholder bodies, making them sketches rather than copy-paste ready; not 3 because the sequences and templates are concrete and directly followable.

4 / 5

Workflow Clarity

The decompose → research → build → verdict loop is explicit with an alignment checkpoint ('Build all in this order, or adjust?'), an iteration feedback loop ('Observe output, iterate'), and evidence-based validation ('Never declare "it works" after one happy-path run', verdict must be backed by evidence). This matches anchor 5 — clear sequence, explicit validation, feedback loops for error recovery.

5 / 5

Progressive Disclosure

No bundle files exist and none are referenced; sections are clearly headed and well-organized for a self-contained skill of this size. Fits anchor 4 ('good structure; most content appropriately placed; minor organization gaps'); not 5 because at ~185 monolithic lines, content such as the verdict/head-to-head templates and the GSD cross-reference could be split into one-level-deep reference files.

4 / 5

Total

16

/

20

Passed

Description

42%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise and names a distinct niche (disposable pre-build experiments) but is fatally thin for discovery: it has no 'Use when...' trigger clause and no natural user phrasing like 'spike' or 'prototype', so it would rarely be selected at the right moment. What it says is accurate and specific; what it omits is the entire triggering surface.

Suggestions

Add an explicit trigger clause, e.g. 'Use when the user says "spike this out", "quick prototype", "is this even possible?", or wants to compare approaches before committing to a build.'

Include natural synonyms and variations users actually say — 'spike', 'prototype', 'proof of concept', 'feasibility check', 'try this out' — in the description text itself (the body already contains these; surface them in the frontmatter).

Name 1-2 concrete artifacts the skill produces (e.g. 'builds throwaway spike directories with a VALIDATED/PARTIAL/INVALIDATED verdict per question') to lift specificity above a single abstract action.

DimensionReasoningScore

Specificity

Names the domain ('Throwaway experiments') and one concrete action ('validate an idea before build'), but offers no enumeration of concrete outputs or capabilities. It is not 4 because only a single action is given, and not 2 because the domain and action are specifically named rather than generic.

3 / 5

Completeness

The 'what' is clear (throwaway experiments to validate an idea), but there is no 'Use when...' clause or any explicit trigger guidance, which caps completeness at 3 per the judging guidelines. It is not 2 because the 'what' is specific rather than vague.

3 / 5

Trigger Term Quality

The only keyword is the generic 'experiments'; natural user phrases like 'spike this out', 'quick prototype', 'is this even possible', or 'proof of concept' are entirely absent from the description. Anchor 2 fits ('one or two generic keywords; missing the natural phrases users say') since no synonyms or variations appear at all.

2 / 5

Distinctiveness Conflict Risk

'Throwaway experiments... before build' carves out a spike/prototype niche, but without trigger phrases the description could overlap with general research, planning, or prototyping skills. Fits anchor 3 ('somewhat specific but could still overlap with similar skills'); not 4 because the boundary against plain research/exploration is not explicit.

3 / 5

Total

11

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

13

/

16

Passed

Repository
NousResearch/hermes-agent
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.