CtrlK
BlogDocsLog inGet started
Tessl Logo

systematic-review-screener

Automated abstract screening tool for systematic literature reviews with PRISMA workflow support.

48

Quality

60%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./scientific-skills/Evidence Insight/systematic-review-screener/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

50%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The skill's operational core (commands, criteria config, outputs, PRISMA data) is genuinely executable and well-externalized, but it is buried under dozens of lines of duplicated, generic governance boilerplate, and several examples contain errors or point at files missing from the bundle.

Suggestions

Cut the template-filler sections (Risk Assessment, Security Checklist, Evaluation Criteria, Lifecycle Status, and the overlapping Error Handling/Failure Handling/When Not to Use/Input Validation blocks) and consolidate what remains into one short error-handling section.

Fix the concrete defects: the CSV sample whose data row duplicates the header, the shell commands mislabeled as ```python blocks, and the duplicated "--help" line in Audit-Ready Commands.

Remove or replace the missing references (references/prisma_2020_checklist.pdf, requirements.txt) — either ship them or drop the citations — and replace the abstract 5-step Workflow with the actual operational sequence (install deps, run command, inspect output files).

DimensionReasoningScore

Conciseness

The body is ~347 lines, and roughly 40% is generic template boilerplate that teaches Claude nothing new ("Risk Assessment", "Security Checklist", "Evaluation Criteria", "Lifecycle Status", "Output Requirements", "Required Inputs", "User Checkpoints", "Quick Validation"), with heavy duplication across "Error Handling", "Failure Handling", "When Not to Use", and "Input Validation". This matches 'Noticeably verbose; several unnecessary explanations or padded sections'.

2 / 5

Actionability

Commands are concrete and copy-paste ready ("python scripts/main.py --input references.csv --criteria criteria.yaml"), the criteria YAML is shown inline with a real template at references/criteria_template.yaml, and the CLI options and output files are tabulated. Falls short of 5 because of real defects: the CSV example is broken (the header line is repeated verbatim as a data row), shell commands are fenced as ```python, and two referenced prerequisites are missing (references/prisma_2020_checklist.pdf and requirements.txt).

4 / 5

Workflow Clarity

A validation entry point exists ("python -m py_compile scripts/main.py" under Quick Check), but the five-step "Workflow" section is abstract policy ('Confirm the user objective... Validate that the request matches...') rather than the operational sequence (install deps → run script → check outputs), and output verification is only the generic "Quick Validation" checklist. This is a batch operation with implicit rather than explicit checkpoints, matching 'Steps listed but validation gaps; sequence present but checkpoints missing or implicit'.

3 / 5

Progressive Disclosure

Bulk detail is appropriately external (605-line scripts/main.py, references/criteria_template.yaml, references/sample_references.csv), but the body still inlines large generic sections that belong in separate files or nowhere, and one of the three listed references (references/prisma_2020_checklist.pdf) does not exist in the bundle. This fits 'Some structure but could be better organized; references present but not clearly signaled'.

3 / 5

Total

12

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is domain-specific and accurate but terse to a fault: it states one action and a standard, with no use-when trigger clause and no synonyms or file-format cues that would help a user (or Claude) select this skill at the right moment.

Suggestions

Add an explicit trigger clause, e.g. 'Use when screening abstracts for a systematic review or PRISMA flow reporting.'

List one or two more concrete capabilities in the description (e.g., 'screens PubMed MEDLINE, EndNote XML, and CSV inputs against inclusion/exclusion criteria with confidence scores') to lift specificity.

Include natural synonyms users say — 'systematic review', 'literature screening', 'inclusion/exclusion criteria' — to improve trigger-term coverage.

DimensionReasoningScore

Specificity

The description names the domain clearly ("Automated abstract screening tool for systematic literature reviews") and implies one concrete action (screening abstracts) plus a standard it supports ("with PRISMA workflow support"). This matches the anchor 'Names domain and 1-2 concrete actions, but not comprehensive' — it does not list several specific actions like input formats, confidence scoring, or conflict detection.

3 / 5

Completeness

It has a clear 'what' (automated abstract screening with PRISMA workflow support) but no 'when' — there is no 'Use when...' clause or equivalent explicit trigger guidance anywhere in the description. Per the judging guidelines, a missing 'Use when...' clause caps completeness at 3.

3 / 5

Trigger Term Quality

Relevant natural phrases are present ("abstract screening", "systematic literature reviews", "PRISMA"), which a user doing an SLR would plausibly say. However, common variations and synonyms are missing: "systematic review", "SLR", "literature screening", "inclusion/exclusion criteria". Fits the anchor 'Some relevant keywords but missing common variations or synonyms'.

3 / 5

Distinctiveness Conflict Risk

The PRISMA/systematic-review niche is specific and unlikely to trigger for unrelated skills; it sits at 'Clear niche' territory. It falls short of 5 because without explicit trigger phrases there is residual overlap risk with general literature-search or citation-management skills, matching 'Mostly distinct; minor overlap risk with closely related skills'.

4 / 5

Total

13

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

referenced_paths_exist

Referenced path issues: 1 missing

Warning

Total

14

/

16

Passed

Repository
aipoch/medical-research-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.