CtrlK
BlogDocsLog inGet started
Tessl Logo

adaptive-trial-simulator

Design and simulate adaptive clinical trials with interim analyses, decision rules, and operating-characteristic summaries; use when planning adaptive designs or comparing stopping, enrichment, or sample-size re-estimation strategies.

62

Quality

78%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./scientific-skills/Protocol Design/adaptive-trial-simulator/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

56%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong executable core — runnable commands, a complete parameter table, design-type/spending-function references, and a concrete output example — buried in heavy audit boilerplate. Roughly 100+ lines of meta-sections, four overlapping error-handling sections, and a References list naming nonexistent files hurt token efficiency and navigation.

Suggestions

Cut the meta-boilerplate sections (Risk Assessment, Security Checklist, Lifecycle Status, Evaluation Criteria, Technical Difficulty) and merge the four overlapping scope/error sections (Error Handling, Failure Handling, Input Validation, When Not to Use) into a single section — this alone removes roughly 40% of the body's tokens.

Delete the second 'References' section that lists files not present in references/ (keep only the working audit-reference.md link), and move the Parameters, Design Types, and Spending Functions tables into a bundled reference file so SKILL.md stays a lean overview.

Bundle an actual requirements.txt, or replace "pip install -r requirements.txt" with "pip install numpy scipy matplotlib", so every documented command runs against the real bundle.

DimensionReasoningScore

Conciseness

The ~305-line body is padded with meta-boilerplate sections that teach Claude nothing ("Risk Assessment", "Security Checklist", "Lifecycle Status" with a next-review date, "Technical Difficulty: **HIGH**", "Evaluation Criteria"), four overlapping error/scope sections ("Error Handling", "Failure Handling", "Input Validation", "When Not to Use"), and commands repeated up to three times ("python -m py_compile scripts/main.py" appears in both Quick Check and Audit-Ready Commands). That is several unnecessary padded sections (score 2), though not 1 since it never explains concepts Claude already knows and the core usage content is dense.

2 / 5

Actionability

Concrete, copy-paste-ready commands ("python scripts/main.py --design adaptive_reestimate --n-simulations 25 --optimize") plus a complete parameter table with defaults, design-type and spending-function tables, and a realistic JSON output example. Not 5 because of minor gaps: "pip install -r requirements.txt" targets a requirements.txt that is not present in the bundle.

4 / 5

Workflow Clarity

The Workflow section gives a sequenced 5-step path with an explicit feedback/fallback loop ("If execution fails or inputs are incomplete, switch to the fallback path and state exactly what blocked full completion"), backed by explicit checkpoints (Quick Check py_compile, "python scripts/main.py --help", the Quick Validation checklist). Not 5 because the checkpoints are scattered across redundant sections and the steps are governance-level rather than a concrete simulate-then-interpret-then-optimize procedure.

4 / 5

Progressive Disclosure

The one real reference is well-signaled ("[references/audit-reference.md](references/audit-reference.md) - Audit-ready assumptions, supported design modes, and fallback boundaries") and scripts/main.py exists as bundled, but a second "References" section lists phantom entries ("Adaptive design statistical theory", "Regulatory guidance documents", "Alpha spending function literature") that match no file in references/, and audit/lifecycle content that belongs in a separate reference is inlined. Some structure, but organization gaps and misleading navigation keep this at 3 rather than 4.

3 / 5

Total

13

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description in third-person voice: concrete capabilities, comprehensive domain coverage, and an explicit use-when clause with specific trigger scenarios. The only weakness is trigger-term synonym coverage (missing terms like 'group sequential', 'futility', or 'alpha spending').

DimensionReasoningScore

Specificity

"Design and simulate adaptive clinical trials with interim analyses, decision rules, and operating-characteristic summaries" lists multiple specific concrete actions (design, simulate, interim analyses, decision rules, OC summaries) with comprehensive domain coverage, matching the score-5 anchor. Not 4 because coverage extends beyond 'several actions with minor gaps' — the when-clause adds stopping, enrichment, and re-estimation strategies.

5 / 5

Completeness

Both questions are explicitly answered: the 'what' ("Design and simulate adaptive clinical trials with interim analyses, decision rules, and operating-characteristic summaries") and an explicit trigger clause ("use when planning adaptive designs or comparing stopping, enrichment, or sample-size re-estimation strategies"). Score 5 rather than 4 because the when-clause is explicit and concretely phrased, not vague or implied.

5 / 5

Trigger Term Quality

Natural user phrases like "adaptive clinical trials", "interim analyses", "stopping", "enrichment", and "sample-size re-estimation" are present, but common variations a user might say — "group sequential", "futility", "alpha spending", "power" — are absent. This is good keyword coverage with a few natural terms missing (score 4), not the comprehensive synonym coverage of score 5.

4 / 5

Distinctiveness Conflict Risk

A clear niche — simulation of adaptive clinical trial designs — with distinct domain triggers ("interim analyses", "sample-size re-estimation") that would not plausibly fire a generic statistics, document, or coding skill. Minimal conflict risk, matching the score-5 anchor.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
aipoch/medical-research-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.