CtrlK
BlogDocsLog inGet started
Tessl Logo

bounded-experiment

Plan and validate bounded, fail-closed experiment loops (ExperimentSpec v2 with explicit cycle limits, guards and a sandbox receipt) via scripts/validate_spec.py and scripts/autoresearch.py. Use when the user asks to run an autoresearch loop, an iterative experiment with a budget, or to validate an experiment spec before executing it. [EXPLICIT]

56

Quality

68%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/claude-native-toolkit/skills/bounded-experiment/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

53%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is concise and well-structured as a lean stage contract that delegates depth to references, but it is light on actionable, executable guidance and only abstractly handles validation for a destructive loop. A couple of referenced bundle paths do not exist.

Suggestions

Add at least one concrete, copy-pasteable command in the body (e.g. the validate_spec.py preflight invocation) so the SKILL.md is actionable without opening references.

Make validation checkpoints explicit in the Procedure (a validate-then-proceed step with the script call), since this is a destructive/batch loop.

Reconcile the packet folder list and the playbook's examples/experiment-spec.json reference against the actual bundle, removing or creating the missing paths.

DimensionReasoningScore

Conciseness

The body is lean with short, unpadded sections and assumes Claude's competence; minor trimming possible from the evidence-tag noise and a placeholder TL;DR ("[source identity removed]").

4 / 5

Actionability

The body offers only high-level hints ("Apply its decision tables; pick the strategy explicitly") with no concrete code or commands, deferring all executable detail to references/ and scripts/.

2 / 5

Workflow Clarity

A 3-step Procedure with a Quality Criteria checklist gives a sequence, but for a destructive/batch experiment loop the validation checkpoints are only abstract, so workflow clarity is capped at 3.

3 / 5

Progressive Disclosure

A clear resource-map table points to real one-level-deep references (contracts.md, full-playbook.md), though the packet block and playbook also name folders/files (knowledge/, examples/experiment-spec.json) that are not present in the bundle.

4 / 5

Total

13

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, well-scoped description that concretely states both capability and explicit trigger conditions with named scripts and schema. Minor room to broaden action verbs and add a couple of natural synonyms.

DimensionReasoningScore

Specificity

Names the domain and concrete actions ("Plan and validate bounded, fail-closed experiment loops") plus specific tooling (scripts/validate_spec.py, scripts/autoresearch.py, ExperimentSpec v2), but the verb set is limited to plan/validate rather than fully comprehensive coverage.

4 / 5

Completeness

Clearly answers both what (plan and validate bounded fail-closed experiment loops) and when, with an explicit "Use when the user asks to..." clause listing concrete triggers.

5 / 5

Trigger Term Quality

Includes natural trigger phrases users would say ("run an autoresearch loop", "iterative experiment with a budget", "validate an experiment spec"), with good coverage but a few common synonyms missing.

4 / 5

Distinctiveness Conflict Risk

The niche and triggers are distinct (autoresearch loop, ExperimentSpec v2), but the metadata family lists related agent-loop skills, leaving minor overlap risk with closely related capabilities.

4 / 5

Total

17

/

20

Passed

Validation

68%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 11 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

referenced_paths_exist

Referenced path issues: 1 missing

Warning

Total

11

/

16

Passed

Repository
JaviMontano/claude-plugins
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.