CtrlK
BlogDocsLog inGet started
Tessl Logo

hypothesis-formulation

Structured scientific hypothesis generation from observations. Use when formulating testable hypotheses, competing explanations, or experimental predictions.

80

Quality

100%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-3

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a lean, well-organized set of actionable, sequenced best-practice steps for hypothesis formulation with no padding and no unnecessary external references. It appropriately assumes Claude's competence while providing concrete structural guidance.

DimensionReasoningScore

Conciseness

Lean and efficient with no padding or explanation of concepts Claude already knows; every line is actionable guidance, matching the score-3 anchor and not the score-2 which carries unnecessary commentary.

3 / 3

Actionability

Concrete, specific guidance throughout ("If...then...because" structure, "design experiments to reject H0", "power analysis", "pre-register hypotheses"); per scoring_notes, absence of code in an instruction-only skill is not penalized when the guidance is actionable, so it clears the score-3 bar rather than settling at score-2.

3 / 3

Workflow Clarity

Five clearly sequenced sub-processes with numbered steps and forward-looking checkpoints ("confirm vs. refute", "Plan for both positive and null results"); no destructive or batch operations requiring stricter feedback loops, so it meets the score-3 anchor.

3 / 3

Progressive Disclosure

Under 50 lines, single-purpose, and organized into well-labeled sections with no need for external references; scoring_notes allow a score-3 here for such simple skills, rather than the score-2 that implies poorly-split inline content.

3 / 3

Total

12

/

12

Passed

Description

100%Weight 40%Scale 1-3

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise, concrete, and fully answers both what the skill does and when to use it via an explicit Use-when clause with natural trigger terms. It occupies a distinct scientific niche with low conflict risk.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — "hypothesis generation from observations", "formulating testable hypotheses", "competing explanations", "experimental predictions" — matching the score-3 anchor.

3 / 3

Completeness

Answers both what (structured scientific hypothesis generation) and when via the explicit "Use when formulating..." clause, matching the score-3 example and not the score-2 which lacks an explicit when.

3 / 3

Trigger Term Quality

Natural terms a researcher would actually say ("testable hypotheses", "competing explanations", "experimental predictions") give good coverage, matching the score-3 anchor rather than the narrower score-2.

3 / 3

Distinctiveness Conflict Risk

A clearly defined scientific-hypothesis niche with distinct triggers unlikely to conflict with general skills; not the generic score-1 or the somewhat-overlapping score-2.

3 / 3

Total

12

/

12

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
aiming-lab/AutoResearchClaw
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.