CtrlK
BlogDocsLog inGet started
Tessl Logo

aim-and-hypothesis-designer

Designs primary aims, secondary aims, and testable hypotheses from broad biomedical research ideas. Use this skill when a user needs to convert a loose study idea into a tighter protocol-framing structure with clear aim hierarchy, hypothesis discipline, and separation between hypothesis-driven and exploratory components. Always keep aims answerable, non-overlapping, and aligned to the intended evidence type and study scope.

67

Quality

84%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

70%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-orchestrated, instruction-only skill: explicit ordered workflow with self-validation checkpoints, a mandatory output structure, and clean one-level-deep reference integration. Its main weakness is token efficiency — four sections restate the same rules that the 8-step execution and Hard Rules already establish — and the absence of a worked example showing what a good output looks like.

Suggestions

Consolidate the overlapping directive sections: 'Core Function', 'What This Skill Should Not Do', and 'Quality Standard' largely restate the 8-step execution and 'Hard Rules' — merge them into one section (e.g., keep Hard Rules plus a short should/should-not pair) to cut substantial duplication.

Add one compact worked example in the output guidance — a vague idea transformed into one primary aim, a secondary aim, and a testable vs exploratory classification — so the expected output style is demonstrated rather than implied.

State each core rule once: 'separate confirmatory from exploratory components' is repeated across at least five sections; keep it in the relevant execution step and the hard-rules list, and drop the rest.

DimensionReasoningScore

Conciseness

The body restates the same directives across four separate sections — 'Core Function' (9 items), 'Hard Rules' (15 items), 'What This Skill Should Not Do', and 'Quality Standard' (10 items) — which substantially overlap the 8-step execution and each other; 'separate confirmatory from exploratory' alone appears in at least five sections. This is noticeably verbose with several padded sections, matching the score-2 anchor rather than score 3 ('some unnecessary explanation' — here the redundancy is systemic, not occasional). It is not a 1 because it never explains concepts Claude already knows and each individual section is internally tight.

2 / 5

Actionability

Concrete, executable guidance for an instruction-only skill: 8 ordered execution steps each specifying exactly what to identify, a mandatory output structure with defined content for Sections A–J, input-validation examples, and a scripted out-of-scope redirect. Per the rubric's code_vs_instruction note, the absence of code is not penalized. It falls short of 5 because there is no worked example (e.g., a before/after transformation of a vague idea into a disciplined aim), leaving the expected output style implicit.

4 / 5

Workflow Clarity

The 8 steps are explicitly ordered ('always run in order'), each mapped to specific reference modules, and include validation checkpoints with feedback loops: Step 7 detects failure modes and directs a conservative rewrite, and Step 8 is a self-critical review checklist with five explicit checks before finalizing. Output completeness is further enforced by the mandatory section structure. This matches the top anchor — clear sequence, explicit validation, feedback loop, and checklist.

5 / 5

Progressive Disclosure

All seven reference files cited in the body exist in references/, are compact (11–40 lines), and contain no nested references — exactly one level deep. Each citation is well-signaled with its purpose and mapped to specific output sections (e.g., 'references/aim-hierarchy-framework.md → define primary, secondary, and optional supporting aims in Sections B–D'), and SKILL.md functions as a clear overview. The body is long, but its content is workflow orchestration rather than detail that belongs in reference files.

5 / 5

Total

16

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete third-person action list, an explicit 'Use this skill when...' trigger clause, and a well-delineated biomedical niche. The only notable gap is that a few natural user phrasings ('specific aims', grant-style aim language) appear in the body but not in the description itself.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — 'Designs primary aims, secondary aims, and testable hypotheses', 'convert a loose study idea into a tighter protocol-framing structure', 'separation between hypothesis-driven and exploratory components', 'keep aims answerable, non-overlapping, and aligned to the intended evidence type and study scope' — giving comprehensive coverage of the skill's behavior in third-person voice. It is not a 4 because the action set covers the domain end-to-end rather than leaving minor gaps.

5 / 5

Completeness

It explicitly answers both questions: what ('Designs primary aims, secondary aims, and testable hypotheses from broad biomedical research ideas') and when ('Use this skill when a user needs to convert a loose study idea into a tighter protocol-framing structure...'), with concrete trigger phrasing. The 'when' clause names the exact situation, matching the top anchor.

5 / 5

Trigger Term Quality

Good natural keyword coverage: 'primary aims', 'secondary aims', 'testable hypotheses', 'biomedical research ideas', 'protocol-framing structure', 'hypothesis-driven and exploratory'. A few natural terms users would plausibly say are absent — 'specific aims' (used in the body but not the description), 'grant aims', 'study concept'. It is not 3 because the terms present are the natural phrases a researcher would use, not jargon; it is not 5 because synonym coverage is incomplete.

4 / 5

Distinctiveness Conflict Risk

It occupies a clear niche — biomedical protocol framing and aim/hypothesis design — with distinct triggers (aims, hypotheses, protocol framing, exploratory separation) that are unlikely to fire for unrelated skills. Overlap with a generic 'study design' skill is possible in principle but the biomedical framing and aim-specific vocabulary keep conflict risk minimal.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
aipoch/medical-research-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.