CtrlK
BlogDocsLog inGet started
Tessl Logo

research-lit

Search and analyze research papers, find related work, summarize key ideas. Use when user says "find papers", "related work", "literature review", "what does this paper say", or needs to understand academic papers.

60

Quality

73%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./skills/research-lit/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

62%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The workflow is exceptionally well-engineered — explicit sequencing, mandatory verification gates, graceful-degradation fallbacks, and checklists — but the body is roughly three times longer than it needs to be: duplicated resolver boilerplate, twice-stated D2 policy, and design-space justification that belongs in a reference document. Referenced shared-references files are not part of the bundle, so progressive disclosure is only partially realized.

Suggestions

Factor the ~15-line fetcher/ARIS_REPO resolver into one shared snippet (or a helper script in scripts/) and reference it from each source block instead of repeating it six times; this alone would cut hundreds of tokens.

Move the D2 policy rationale, the fan-out/acceptance-gate justification in Step 2, and the per-source de-duplication rules into a references/ file (e.g. sources-and-dedup.md) linked once, and state each rule once instead of "both lines must stay in sync" dual restatements.

Replace placeholder-heavy blocks with one concrete worked example (a real query and arXiv ID), and name the most common Zotero/Obsidian MCP tool patterns in Step 0a/0b instead of "e.g., search".

DimensionReasoningScore

Conciseness

The ~750-line body repeats the identical ARIS_REPO/`$ARXIV_FETCHER` resolver boilerplate six times, restates the D2 policy twice ("the finalization block below restates this rule canonically — both lines must stay in sync"), and spends long passages on internal design rationale ("Policy D2", "cross-model-family rule", Tier 1/2/3 fan-out justifications) the executor does not need — noticeably verbose with several padded sections. Not 1, because it never teaches concepts Claude already knows (no "what is a PDF" material) and the verbosity is duplication and policy meta-commentary rather than pure abstraction.

2 / 5

Actionability

Mostly executable: complete, wrapped bash blocks with `set -e` guards, concrete WebSearch/de-dup rules, an MCP call with the full Gemini prompt, and a fallback python heredoc. Gaps keep it from 5: `QUERY`, `ARXIV_ID`, and the candidate-papers JSON are placeholders rather than copy-paste-ready, and the Zotero/Obsidian steps say only "try calling a Zotero MCP tool (e.g., search)" without concrete tool names.

4 / 5

Workflow Clarity

Steps are explicitly sequenced (0a → 0b → 0c → 1 → 1.5 → 2 → 6) with validation checkpoints throughout: the mandatory Step 1.5 anti-hallucination verify gate with fallback, the D2 empty-aggregate error-and-stop gate, per-source warn-and-continue feedback loops, an explicit wiki-ingest checklist, and retry guidance for network errors — matching the anchor "clear sequence with explicit validation steps; feedback loops for error recovery; checklists".

5 / 5

Progressive Disclosure

Sections and headers exist and there are clearly-signaled links to `shared-references/integration-contract.md`, `citation-discipline.md`, `fan-out-pattern.md`, `output-composition.md`, and `wiki-helper-resolution.md`, but none of those files exist in the skill bundle (no references/, scripts/, or assets/ directories), and the resolver boilerplate, D2 policy prose, and per-source de-duplication rules — content that clearly belongs in a separate reference file — are inlined and repeated. This fits "some structure but could be better organized; content that should be separate is inline"; not 4 because the referenced paths are unverifiable/absent and the bulk stays monolithic.

3 / 5

Total

14

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete actions, explicit third-person voice, and a well-populated 'Use when...' clause with natural trigger phrases. The only gaps are a few missing synonyms (survey, arXiv, .pdf) and slight overlap risk with general paper-reading/summarization skills.

Suggestions

Add one or two high-frequency variants to the trigger list, e.g. "survey the literature", "papers on <topic>", or "arXiv", to lift trigger-term coverage.

Clarify the paper-reading boundary (e.g. "use when researching a body of academic literature, not for reading a single PDF the user already has") to reduce overlap with PDF-reading skills.

DimensionReasoningScore

Specificity

"Search and analyze research papers, find related work, summarize key ideas" lists three to four concrete actions in the research-paper domain, matching the anchor "Lists several specific actions; minor gaps in coverage" — it omits capabilities the body actually has (verification, multi-source dedup, wiki ingest). Not 5 because coverage is incomplete; not 3 because it goes beyond one or two actions.

4 / 5

Completeness

It answers both parts explicitly: "what" ("Search and analyze research papers, find related work, summarize key ideas") and "when" ("Use when user says... or needs to understand academic papers") with concrete quoted trigger phrases — a direct match to the anchor-5 good example. Not 4, because the "when" clause is already explicit and specific rather than improvable.

5 / 5

Trigger Term Quality

Quotes natural phrases users would say — "find papers", "related work", "literature review", "what does this paper say" — giving good keyword coverage, but common variants like "survey the literature", "papers on X", "arXiv", or ".pdf" are missing, so it fits "Good keyword coverage; a few natural terms missing" rather than comprehensive.

4 / 5

Distinctiveness Conflict Risk

"literature review" and "find papers" establish a clear research niche, but "what does this paper say" and "summarize key ideas" could also fire for a generic PDF-reading or summarization skill — minor overlap risk with closely related skills. Not 5 because the overlap is real; not 3 because the core triggers are distinctly academic.

4 / 5

Total

17

/

20

Passed

Validation

75%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 12 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (757 lines); consider splitting into references/ and linking

Warning

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

relative_links

Relative link issues: 7 suspicious

Warning

Total

12

/

16

Passed

Repository
wanshuiyin/Auto-claude-code-research-in-sleep
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.