CtrlK
BlogDocsLog inGet started
Tessl Logo

arxiv

Search, download, and summarize academic papers from arXiv. Use when user says "search arxiv", "download paper", "fetch arxiv", "arxiv search", "get paper pdf", or wants to find and save papers from arXiv to the local paper library.

66

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

77%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, highly actionable skill body with a clearly sequenced workflow and genuine validation and rate-limit recovery for batch downloads. Its main weaknesses are the monolithic structure — the inline Python and bash resolution logic should live in script files, and the external integration-contract reference is unverifiable — plus small gaps in executability (the Step 3 stub and literal placeholders in the download fallback).

Suggestions

Move the inline Python search/download fallbacks (~50 lines) into a scripts/ file (e.g., scripts/arxiv_fetch.py) and reference it, reducing SKILL.md to the resolution logic and workflow steps — this also eliminates the duplicated fetcher-resolution chain in Constants and Step 2.

Make Step 3's fallback fully executable by replacing the '# print full details ...' stub with complete code, and parameterize the download fallback correctly (e.g., pass PAPER_DIR/ARXIV_ID as arguments or use shell-expanded values) so the heredoc runs without manual editing.

Ensure ../shared-references/integration-contract.md exists as a real, reachable file (or inline the few relevant rules from its §2) — three links point to a file that is absent from the bundle, leaving the canonical resolution chain and Policy D1 under-specified for standalone use.

DimensionReasoningScore

Conciseness

The body is an efficient operational procedure with no conceptual padding (it never explains what arXiv or Atom feeds are), but the $ARXIV_FETCHER resolution chain is described twice — once in Constants and again verbatim in Step 2 — and the integration-contract link is repeated three times. These duplications are trimmable, fitting 'efficient; minor instances that could be trimmed' rather than the every-token-earns-its-place anchor of 5.

4 / 5

Actionability

Mostly executable: the Step 2 inline Python fallback is complete and copy-paste ready, and the download fallback includes a User-Agent header, size reporting, and an already-exists guard. It falls short of fully executable because Step 3's fallback is a stub ('# print full details ...') and the download fallback embeds literal PAPER_DIR/ARXIV_ID placeholders inside single-quoted Python that would need manual substitution to run.

4 / 5

Workflow Clarity

Seven clearly sequenced steps (parse → search → fetch → download → summarize → wiki → report) with explicit validation and recovery for the risky batch operation: verify each PDF is > 10 KB and warn-and-delete if smaller, 1-second delay between downloads, retry once after 5 seconds on HTTP 429, and never overwrite an existing PDF. This matches the top anchor's explicit validation steps plus feedback loops for error recovery.

5 / 5

Progressive Disclosure

The body is well-sectioned, but this is a ~210-line monolithic SKILL.md with no bundle files at all (no references/, scripts/, or assets/), and roughly 35 lines of inline Python fetcher code plus the bash resolution chain clearly belong in separate script files. The three links to ../shared-references/integration-contract.md point outside the bundle to a file that does not exist in this workspace, so the reference is signaled but unresolvable. This fits 'some structure but could be better organized; content that should be separate is inline'.

3 / 5

Total

16

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that pairs a concrete three-action capability statement with an explicit, trigger-phrase-driven 'Use when' clause. It is specific, mostly distinct, and clearly answers both what the skill does and when to invoke it, with only minor gaps in action coverage and a couple of generic trigger phrases that could overlap sibling literature skills.

DimensionReasoningScore

Specificity

The description lists three specific concrete actions — "Search, download, and summarize academic papers from arXiv" — which matches the anchor for several specific actions with minor gaps. It is not a 5 because coverage has gaps: no mention of looking up papers by arXiv ID or downloading to a local library location beyond the one trigger phrase.

4 / 5

Completeness

It clearly and explicitly answers both 'what' (search, download, and summarize academic papers from arXiv) and 'when' with a concrete "Use when user says" clause enumerating exact trigger phrases. This matches the top anchor exactly; the 'when' is explicit and trigger-driven, not merely implied.

5 / 5

Trigger Term Quality

Good keyword coverage with natural phrases users would say: "search arxiv", "download paper", "fetch arxiv", "arxiv search", "get paper pdf", plus "find and save papers from arXiv to the local paper library". A few natural variations are missing (e.g., "find papers on arxiv", "get paper by ID", "arxiv papers"), keeping it below the comprehensive-synonym anchor of 5.

4 / 5

Distinctiveness Conflict Risk

The arXiv niche is clear and most triggers are arXiv-specific ("search arxiv", "arxiv search", "fetch arxiv"), but generic phrases like "download paper" and "get paper pdf" carry minor overlap risk with closely related skills (the body itself references a sibling multi-source /research-lit skill). This fits 'mostly distinct; minor overlap risk with closely related skills' rather than the minimal-conflict anchor of 5.

4 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 3 suspicious

Warning

Total

15

/

16

Passed

Repository
wanshuiyin/Auto-claude-code-research-in-sleep
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.