CtrlK
BlogDocsLog inGet started
Tessl Logo

designing-experiments

Design experiments and quasi-experiments before analysis. Use when choosing study design, treatment/control structure, outcomes, assumptions, validation plans after scientific experiment failure, or which of DiD, ITS, synthetic control, or regression discontinuity fits the research question. For fitting models or estimating effects on existing data, use performing-causal-analysis instead.

70

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

72%

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-organized, lean decision skill that assumes Claude's knowledge and needs no bundle files. Its weakness is actionability and workflow clarity: some decision branches are incomplete and validation/verification checkpoints are absent.

Suggestions

Resolve the incomplete decision-tree branches (e.g. 'Multiple Treated Units' without a control group, and the implicit 'No control + single unit' path) so every leaf recommends a specific design.

Add an explicit validation checkpoint and a concrete decision-rule example to the Failed Experiment Recovery section (e.g. a worked rule for continue/revise/stop).

Include one short worked example or minimal specification template for a chosen design (e.g. a DiD parallel-trends check list) to make the guidance copy-paste ready.

DimensionReasoningScore

Conciseness

The body is lean and assumes Claude's competence; it never explains what DiD or ITS are beyond the design surface, and every section earns its place.

3 / 3

Actionability

Provides a concrete decision tree and named method rules, but several branches are incomplete (e.g. 'Multiple Treated Units' with no control group, or the ITS recommendation) and there are no executable artifacts or worked examples.

2 / 3

Workflow Clarity

The decision framework is clearly sequenced, but there are no explicit validation or verification checkpoints; the recovery section lists steps but lacks decision-rule examples or feedback loops.

2 / 3

Progressive Disclosure

Under 50 lines with well-organized sections (Decision Framework, Method Quick Reference, Failed Experiment Recovery) and no need for external references, matching the simple-skill allowance for a top score.

3 / 3

Total

10

/

12

Passed

Description

100%

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, third-person, and free of fluff, with explicit trigger guidance and a clear sibling-skill boundary. It answers both what the skill does and when to use it.

DimensionReasoningScore

Specificity

Lists multiple concrete actions ('Design experiments and quasi-experiments', 'choosing study design, treatment/control structure, outcomes, assumptions, validation plans') rather than vague language, matching the 'lists multiple specific concrete actions' anchor.

3 / 3

Completeness

Explicit 'Use when...' clause answers when, the opening phrase answers what, and a steering clause ('For fitting models... use performing-causal-analysis instead') adds boundary guidance.

3 / 3

Trigger Term Quality

Covers natural terms users would say ('design experiments', 'study design', 'treatment/control') plus method names ('DiD, ITS, synthetic control, regression discontinuity') that a researcher would invoke.

3 / 3

Distinctiveness Conflict Risk

Clear niche (pre-analysis study design) with an explicit redirect of overlapping cases to a sibling skill, making wrong-skill triggering unlikely.

3 / 3

Total

12

/

12

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
foryourhealth111-pixel/Vibe-Skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.