CtrlK
BlogDocsLog inGet started
Tessl Logo

arxiv

Use this skill whenever the user wants to find, read, cite, track, download, or analyze academic papers on arXiv. That includes: searching papers by topic, author, category, or arXiv ID; fetching abstracts or full metadata; generating BibTeX citations; downloading PDFs; listing the latest submissions in a field (e.g. cs.AI daily digest); checking a paper's citation impact; finding who cites a paper, what it references, or related-paper recommendations. Trigger on mentions of 'arXiv', an arXiv ID (e.g. 2601.02780 or hep-th/0601001), an arxiv.org URL, 'paper search', 'literature review', 'find papers about X', 'cite this paper', or 'what's new in cs.LG'.

76

Quality

95%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

90%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-crafted body: fully executable commands dominate, the content is dense with non-obvious operational details (rate limits, ID formats, versioning, withdrawn-paper flags), and workflows are clearly sequenced. The only deductions are minor: no explicit error-recovery checkpoints in the workflows and an inline raw-API section that could be split into a reference file.

DimensionReasoningScore

Conciseness

The body is lean throughout — a goal/command table, three compact workflows, and short raw-API snippets — with zero padding and no explanation of concepts Claude already knows (it never explains what arXiv or BibTeX is). Not 4: there is no over-explanation to trim; every line adds non-obvious information (flags, rate limits, versioning caveats).

5 / 5

Actionability

All guidance is copy-paste executable: 'python scripts/arxiv.py search "GRPO reinforcement learning" --max 10 --sort date', a full command table with common flags, and complete curl examples like 'curl -s "https://api.semanticscholar.org/graph/v1/paper/arXiv:2601.02780?fields=title,citationCount,…"'. Not 4: the common cases (search, get, cite, download, raw queries) are all covered by ready-to-run commands.

5 / 5

Workflow Clarity

Three multi-step workflows are clearly sequenced (literature review, deep-dive, stay current) with concrete per-step commands, and read-only operations mean the destructive/batch validation cap does not apply. Not 5: the workflows are plain step sequences with no explicit checkpoints or error-recovery feedback (e.g., what to do on HTTP 429 or empty results mid-workflow); not 3: sequencing is explicit and complete with only minor validation gaps.

4 / 5

Progressive Disclosure

Structure is good: quick-start table, sectioned workflows, and a real bundled script (scripts/arxiv.py, verified present) referenced at one level deep with clear signaling. Not 5: the ~40-line 'Raw API Reference' section is advanced inline content that would fit better in a references/ file, a minor organization gap matching the anchor-4 'most content is appropriately placed'.

4 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: it comprehensively and concretely states what the skill does, provides two explicit trigger clauses with natural user phrasings and concrete ID/URL examples, and is cleanly distinct from other skills. Voice stays in third person about the user, and despite its length every clause carries a concrete capability or trigger.

DimensionReasoningScore

Specificity

The description enumerates multiple concrete actions — 'searching papers by topic, author, category, or arXiv ID; fetching abstracts or full metadata; generating BibTeX citations; downloading PDFs; listing the latest submissions… checking a paper's citation impact; finding who cites a paper, what it references, or related-paper recommendations' — comprehensively covering every capability of the skill. Not 4: there are no gaps in coverage relative to the skill's actual commands.

5 / 5

Completeness

Both questions are explicitly answered: 'what' via the full enumeration of capabilities and 'when' via two explicit trigger clauses — 'Use this skill whenever the user wants to find, read, cite, track, download, or analyze academic papers on arXiv' and 'Trigger on mentions of…'. Not 4: the 'when' is fully explicit with concrete trigger phrases, matching the anchor-5 example pattern.

5 / 5

Trigger Term Quality

Trigger terms are comprehensive and natural: "'arXiv', an arXiv ID (e.g. 2601.02780 or hep-th/0601001), an arxiv.org URL, 'paper search', 'literature review', 'find papers about X', 'cite this paper', or 'what's new in cs.LG'" covers synonyms, both ID formats, URLs, and natural user phrasings. Not 4: no commonly used natural term is missing.

5 / 5

Distinctiveness Conflict Risk

The skill occupies a clear niche (arXiv/Semantic Scholar academic-paper research) with distinct triggers ('arXiv', arXiv IDs, arxiv.org URLs) unlikely to fire for other skills. Not 4: overlap risk with generic research or web-search skills is minimal given the arXiv-specific trigger vocabulary.

5 / 5

Total

20

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
XiaomiMiMo/MiMo-Code
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.