CtrlK
BlogDocsLog inGet started
Tessl Logo

run-documentation-examples

Extract Python examples from markdown docs and run them (including LLM examples). Use when validating documentation, after doc changes, or to verify all doc examples execute correctly.

83

1.51x
Quality

75%

Does it follow best practices?

Impact

100%

1.51x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.cursor/skills/run-documentation-examples/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A concise, largely actionable skill body with good section structure and a genuine failure-recovery loop. Its main weakness is that its script references don't match the actual bundle layout — one referenced file is missing and the other uses an incorrect path — which undercuts both executability and navigation.

Suggestions

Fix the Manual Two-Step Workflow to reference files that exist in the bundle: `scripts/extract_documentation_examples.py` is not present; either add it or replace the section with the actual invocation (e.g., `.venv/bin/python scripts/run_documentation_examples.py`).

Make script paths consistent with the bundle: replace `.cursor/skills/run-documentation-examples/scripts/run_documentation_examples.py` with the bundle-relative `scripts/run_documentation_examples.py`.

Add a quick prerequisite check before the recommended workflow step (e.g., verify Ollama is serving) so the batch run has an explicit validation checkpoint before it starts.

DimensionReasoningScore

Conciseness

The body is lean and assumes competence: short Environment, Workflow, Prerequisites, and Failure Handling sections with commands rather than prose. Minor trims are possible ("The pytest approach gives standard test output" and the parenthetical about resume duplicate failure-handling info), matching 'Efficient; minor instances of over-explanation that could be trimmed' rather than anchor 5's every-token-earns-its-place.

4 / 5

Actionability

Nearly everything is copy-paste ready (pytest command, CLI, ollama pull, pip install extras, retry procedure), matching 'Mostly executable guidance; concrete code or commands with minor gaps'. It falls short of 5 because the Manual Two-Step Workflow references `scripts/extract_documentation_examples.py`, which does not exist in the bundle, and the skill-script path uses the `.cursor/skills/...` prefix rather than the actual bundle path.

4 / 5

Workflow Clarity

The recommended path comes first with alternatives ordered by preference, prerequisites are itemized per backend, and the Failure Handling section provides an explicit feedback loop (fail index, retry from point, delete file to restart) — the checkpoint anchor-5 demands. It stays at 4 because the workflow presents parallel options rather than one validated sequence, and the recommended step doesn't state how to confirm prerequisites (e.g., Ollama running) before executing.

4 / 5

Progressive Disclosure

Sections are well organized and the skill is short, but scored against the actual bundle: the only bundle file is `scripts/run_documentation_examples.py`, while the body cites it via a mismatched `.cursor/skills/run-documentation-examples/scripts/...` path and additionally references a nonexistent `scripts/extract_documentation_examples.py`. This fits 'references present but not clearly signaled / could be better organized' (anchor 3); anchor 4's 'references mostly clear' is not met when two of the referenced script paths don't resolve to real bundle files.

3 / 5

Total

15

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that concisely states both what the skill does and explicit 'Use when' triggers in third-person voice, with natural trigger terms. The only shortfall is that the action list is compact rather than comprehensive, so it does not fully enumerate what the skill covers.

DimensionReasoningScore

Specificity

"Extract Python examples from markdown docs and run them (including LLM examples)" names the domain and two concrete actions (extract, run), but does not enumerate further capabilities (e.g., reporting failures, retry support). This matches the anchor 'Names domain and 1-2 concrete actions, but not comprehensive' rather than 4, which requires several specific actions; it is clearly above anchor 2, which only names a domain.

3 / 5

Completeness

The description explicitly states what ("Extract Python examples from markdown docs and run them (including LLM examples)") and when ("Use when validating documentation, after doc changes, or to verify all doc examples execute correctly") with concrete trigger phrases — an exact match for anchor 5. Anchor 4 is ruled out because the 'when' clause is explicit and specific, not merely present.

5 / 5

Trigger Term Quality

Natural phrases like "markdown docs", "validating documentation", "doc changes", and "doc examples" mirror what a user would say, and "Python examples"/"LLM examples" are on point. It falls just short of anchor 5 because common variations such as "code examples" or file extensions are absent, and it is well above anchor 3's 'missing common variations'.

4 / 5

Distinctiveness Conflict Risk

"Run documentation examples" carves a clear niche (doc-example validation/CI) with distinct triggers, matching 'Mostly distinct; minor overlap risk with closely related skills' — a generic 'run tests' or 'write documentation' skill could still collide on phrases like 'validating documentation'. It does not reach anchor 5's minimal conflict risk, and is far more distinct than anchor 3's generic overlap.

4 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

referenced_paths_exist

Referenced path issues: 1 missing

Warning

Total

15

/

16

Passed

Repository
sandialabs/talkpipe
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.