CtrlK
BlogDocsLog inGet started
Tessl Logo

literature-review

Find, verify, and synthesize scientific literature — from "what's the seminal paper for X" through full multi-source reviews. Covers grounding claims in real retrieved sources, avoiding fabricated citations, handling retractions, and calibrating confidence to evidence strength.

53

Quality

67%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./resources/skills/literature-review/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

56%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong, expert-voiced instruction body with unusually concrete formatting and verification rules and a logical workflow with real checkpoints. Its weaknesses are the missing kernel.py file that every executable instruction depends on, and cross-section redundancy in the prose-style guidance that inflates token cost without adding instruction.

Suggestions

Ship kernel.py in the skill bundle (or inline the critical helper signatures) — the body's only executable code (`exec(open("<this skill's directory>/kernel.py").read())`) fails as bundled, since no kernel.py exists in the skill directory.

State the kernel.py helper list once (in Setup) and drop the restatement in the "Put the answer" section, and merge the overlapping prose-style rules in "Making the prose carry its weight" and "Put the answer in the answer" to cut redundant tokens.

Add one short example invocation per key helper (e.g. `search_openalex("transformer scaling laws")`, `verify_dois(["10.xxxx/xxxx"])`) so the actionability is copy-paste ready rather than name-only.

DimensionReasoningScore

Conciseness

The prose is dense and opinionated with little explanation of concepts Claude already knows, but it is noticeably redundant: the kernel.py helper list is repeated in the "Put the answer" section, and the process-narration / open-on-substance rules span two overlapping sections ("Making the prose carry its weight" and "Put the answer in the answer"), so it fits anchor 3 rather than anchor 4's minor trimming.

3 / 5

Actionability

Guidance is concrete (exact exec line for kernel.py, named env vars, exact citation markdown format with %28/%29 encoding rules, style_pass single-pass usage), but the core executable dependency — kernel.py — is referenced throughout yet absent from the bundle, leaving the executable path incomplete per anchor 3's "missing key details" rather than anchor 4's mostly-executable bar.

3 / 5

Workflow Clarity

The sections sequence a coherent workflow (setup → read the request → retrieval sweep → citation-graph expansion → retraction checks → synthesis → style pass → save) with validation checkpoints such as verify_dois, notice inspection for surprising findings, and style_pass before saving with an explicit fix-in-one-pass rule. Minor gaps — no numbered overview and only implied handling of a missing OPENALEX_API_KEY — keep it at anchor 4 rather than 5.

4 / 5

Progressive Disclosure

Section headers are clear and references are one level deep, but the single referenced bundle file (kernel.py) does not exist in the bundle, and dense inline style rules (citation link formatting, DOI URL-encoding, heading-length rules) read like content that belongs in a separate reference file — fitting anchor 3's structure-present-but-imperfect placement rather than anchor 4.

3 / 5

Total

13

/

20

Passed

Description

66%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A specific, well-scoped description with concrete capability language and good natural trigger terms. Its main gap is the missing explicit "Use when..." trigger clause, which leaves the activation condition only implied by the example question, and a few common synonyms (papers, studies, research) that users would naturally say are absent.

Suggestions

Add an explicit trigger clause, e.g. "Use when the user asks for papers, studies, evidence, citations, or a literature review on a topic, or asks to verify citations or DOIs."

Include common natural synonyms — "papers", "studies", "research", "evidence", "survey of the literature" — alongside "scientific literature".

Consider naming the concrete retrieval sources (Crossref, OpenAlex) to sharpen specificity and distinguish the skill from general web-research skills.

DimensionReasoningScore

Specificity

The description lists several concrete actions — "Find, verify, and synthesize scientific literature", "grounding claims in real retrieved sources", "avoiding fabricated citations", "handling retractions", "calibrating confidence to evidence strength" — but stops short of anchor 5's comprehensive coverage because it names no concrete outputs, tools, or formats.

4 / 5

Completeness

The "what" is clear (find, verify, synthesize scientific literature with grounding/retraction/confidence handling), but there is no explicit "Use when..." clause or equivalent — the quoted "what's the seminal paper for X" only weakly implies when to use it, capping this at 3 per the judging guideline.

3 / 5

Trigger Term Quality

Natural keywords users would say are present ("scientific literature", "seminal paper", "citations", "reviews"), matching anchor 4's good-but-incomplete coverage; common synonyms like "papers", "studies", or "research" are missing, so it does not reach anchor 5, while its coverage clearly exceeds anchor 3's partial set.

4 / 5

Distinctiveness Conflict Risk

The scientific-literature/citation-verification niche is distinct with domain-specific triggers, but there is minor overlap risk with general research or web-synthesis skills, fitting anchor 4 rather than anchor 5's minimal-conflict profile.

4 / 5

Total

15

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

Total

14

/

16

Passed

Repository
aipoch/open-science
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.