CtrlK
BlogDocsLog inGet started
Tessl Logo

survival-analysis-km

Kaplan-Meier survival analysis tool for clinical and biological research. Generates publication-ready survival curves with statistical tests.

48

Quality

60%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./scientific-skills/Data Analysis/survival-analysis-km/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

38%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The skill has a genuinely useful core (arguments, input/output formats, statistical methods, example results) buried in extensive generic template boilerplate and broken structural references. It reads as an auto-filled template rather than curated guidance, with missing bundle files and duplicated/malformed parameter documentation.

Suggestions

Delete the generic template sections (Risk Assessment, Security Checklist, Evaluation Criteria, Lifecycle Status, Response Template, Output Contract, Inputs to Collect, Validation and Safety Rules) — none contain survival-analysis-specific information Claude cannot infer.

Fix the broken references: remove the bogus 'cd "20260318/..."' path, correct the 'See ## Features/Usage/Workflow above' pointers (those sections appear later), and either ship the referenced scripts/main.py, references/, and requirements.txt or stop citing them.

Merge the malformed duplicate 'Parameters' table (empty descriptions for --event/--group, a broken '--figsize | str | 10' default, and --risk-table inconsistent between Required and Optional) into the single accurate 'Arguments' table, and make the example invocations uncommented, runnable commands.

DimensionReasoningScore

Conciseness

Roughly half of the ~320-line body is generic template boilerplate ('Risk Assessment', 'Security Checklist', 'Evaluation Criteria', 'Lifecycle Status', 'Response Template', 'Output Contract', 'Validation and Safety Rules') that adds nothing specific to survival analysis, plus a duplicated and malformed 'Parameters' table alongside the earlier 'Arguments' table. It is not 1 because the core usage content (Arguments, Input Format, Output Files, Statistical Methods) is itself tight and specific rather than tutorial-style padding.

2 / 5

Actionability

Concrete guidance exists (full argument table, example CSV input format, named output files, 'python -m py_compile scripts/main.py'), but all example invocations are commented out ('# Example invocation: python scripts/main.py ...') rather than executable, and the referenced artifacts (scripts/main.py, references/, requirements.txt) do not exist in the bundle, with a bogus hardcoded 'cd "20260318/..."' path. This incompleteness goes beyond the 'minor gaps' of a 4.

3 / 5

Workflow Clarity

A sequence is present (confirm inputs -> py_compile check -> run scripts/main.py -> review outputs) with an explicit validation checkpoint and error-handling/fallback guidance, but the actual 'Workflow' and 'Example run plan' sections are generic boilerplate ('Validate that the request matches the documented scope'), duplicated across multiple sections, and never concretely tied to the survival-analysis steps. Not 4 because the checkpoints are implicit and scattered rather than forming one clear task-specific sequence.

3 / 5

Progressive Disclosure

The body is a monolithic document with content that belongs in separate files inlined (security checklists, evaluation criteria, lifecycle status), and its file references are broken: 'scripts/main.py', 'references/', and 'requirements.txt' are cited but absent from the bundle, while pointers like 'See `## Features` above' and 'See `## Usage` above' reference sections that actually appear later in the document. It is above 1 only because the many section headers do provide some navigation.

2 / 5

Total

10

/

20

Passed

Description

65%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description communicates a clear, distinctive what-statement with strong natural trigger terms, but it lacks any when-to-use guidance and undersells the skill's actual capabilities. Adding an explicit trigger clause and one more concrete capability would raise it substantially.

Suggestions

Append an explicit trigger clause, e.g. 'Use when the user mentions survival analysis, Kaplan-Meier curves, time-to-event data, or censored clinical trial outcomes.'

Mention one or two more concrete capabilities from the body (e.g., Cox proportional hazards regression with hazard ratios, log-rank tests) so the what-statement is comprehensive.

DimensionReasoningScore

Specificity

The description names the domain ('Kaplan-Meier survival analysis tool for clinical and biological research') and two concrete actions ('Generates publication-ready survival curves with statistical tests'). This matches the 1-2-concrete-actions anchor; it is not 4 because several capabilities documented in the body (Cox regression, hazard ratios, log-rank tests, risk tables) are omitted, leaving coverage incomplete.

3 / 5

Completeness

The 'what' is clear (Kaplan-Meier analysis, publication-ready curves, statistical tests), but there is no 'Use when...' clause or equivalent explicit trigger guidance, which caps completeness at 3 per the judging guidelines. Not 2 because the 'what' is concrete and domain-specific rather than vague.

3 / 5

Trigger Term Quality

Natural phrases users would say are present: 'Kaplan-Meier', 'survival analysis', 'survival curves', 'clinical', 'biological research'. It falls short of 5 because common variations such as 'time-to-event', 'censoring', 'log-rank', 'Cox', and 'hazard ratio' are missing.

4 / 5

Distinctiveness Conflict Risk

'Kaplan-Meier survival analysis' is a clear niche with distinct trigger terms, so the risk of firing for the wrong skill is minimal. The generic phrase 'statistical tests' introduces only negligible overlap, not the 'minor overlap risk with closely related skills' that would justify 4.

5 / 5

Total

15

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

referenced_paths_exist

Referenced path issues: 12 missing

Warning

Total

14

/

16

Passed

Repository
aipoch/medical-research-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.