CtrlK
BlogDocsLog inGet started
Tessl Logo

daily-papers-fetch

论文抓取(3 步流水线的第 1 步)。抓取 arXiv + HuggingFace 最新论文,打分筛选,富化信息, 输出 daily_papers_enriched.json 到共享临时目录,供后续 skill 使用。 触发词:"论文抓取"、"跑一下论文抓取" 支持多天模式:"过去3天论文推荐"、"过去一周论文推荐"、"过去一周的论文"、"抓 3 天的论文"、"最近5天"

69

Quality

87%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-built operational skill: concrete commands for each phase, an explicit multi-day argument contract, multi-layered output validation (exit code, stale-file provenance, JSON validity, empty-array handling, stderr diagnostics), and honest failure semantics for the ranking API. Main weaknesses are mild redundancy in the config/days explanations, recovery guidance that ends at diagnosis without a retry step, and unverifiable out-of-bundle script paths with some inlined implementation detail.

Suggestions

Remove the duplicated days-argument explanation (the paragraph after the Phase 1+2 code block restates the '解析天数' section) and fold the '配置来源' section into Step 0, since both restate what '执行环境' already establishes.

Close the feedback loop on failures: after '检查 stderr 诊断问题', add an explicit instruction to fix the identified issue and re-run the failed phase, re-checking exit code and output validity before proceeding.

Move script-internal details (the enrichment merge-priority rules and the full output-field schema) into a reference file next to the scripts and link to it, keeping SKILL.md as a lean interface overview — this would also make the ../daily-papers script paths verifiable within the bundle.

DimensionReasoningScore

Conciseness

The body is largely operational — command blocks, script-behavior bullets, a days-parsing table, and an output-field schema — and assumes Claude's competence (no explaining what arXiv/JSON/asyncio are). It sits between anchors 3 and 4, above the midpoint: the padding is limited to a few redundant spots — the DAYS_ARG rule explained in "解析天数" and re-explained after the Phase 1+2 code block ("根据前面解析的 DAYS_ARG,如果用户指定了天数就加 --days N,否则不加"), and the "配置来源" section restating what "执行环境" and Step 0 already established ("后续统一以共享配置和上面的变量为准"). Not score 5 because those repeats could be trimmed; not score 3 because the great majority of lines carry non-inferable operational detail.

4 / 5

Actionability

Concrete, mostly executable commands are given for every phase — `python3 ../_shared/user_config.py`, both the default and `--days N` invocations of `fetch_and_score.py` with explicit output paths, and `enrich_papers.py` with a clear two-path-argument contract ("使用两个文件路径参数(输入 + 输出)") — plus an explicit days-parsing mapping ("过去一周...→ --days 7"). It is not score 5 because the commands are not copy-paste ready as written: `{TEMP_DIR}` and `N` are placeholders the agent must substitute, and the `../daily-papers/` script paths resolve only within the parent multi-skill layout; it is well above score 3 since nothing is pseudocode and the common cases (default day and multi-day) are both covered.

4 / 5

Workflow Clarity

The sequence is clear (执行环境 → Step 0 配置 → 解析天数 → Phase 1+2 → Phase 3 → 输出) and validation is unusually explicit for a batch operation: "必须退出码为 0,确认输出来自本次运行;旧文件不能作为成功证据。确认...存在且包含有效 JSON 数组。如果为空数组或文件不存在,检查 stderr 诊断问题", plus a defined failure policy for the ranking API ("缺失或 API 失败时停止并说明,不静默退回"). It is not score 5 because the error-recovery loops stop at diagnosis ("检查 stderr 输出诊断问题") without an explicit fix-and-re-run step, matching the anchor with minor validation gaps rather than the closed feedback-loop anchor.

4 / 5

Progressive Disclosure

The SKILL.md works as an overview that delegates implementation to scripts and clearly signals its external references one level deep — [Agent 运行约定](../_shared/agent-runtime.md), `user_config.py`, `fetch_and_score.py`, `enrich_papers.py` — with well-organized sections. It is not score 5 because the referenced files live outside the skill directory in sibling paths (`../_shared/`, `../daily-papers/`) that are not part of this bundle and cannot be verified, and a moderate amount of script-internal detail (merge-priority rules, the full enrichment output-field schema) is inlined; it is above score 3 because references are clearly signaled and the split between workflow overview and script implementation is appropriate.

4 / 5

Total

16

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states concrete capabilities (arXiv + HuggingFace fetch, scoring, enrichment, named JSON output) and gives an explicit trigger-word list with natural multi-day phrasing variations. The only notable risk is mild overlap with sibling pipeline skills on '论文推荐'-style triggers.

DimensionReasoningScore

Specificity

The description lists four concrete actions with named sources and a named artifact: "抓取 arXiv + HuggingFace 最新论文,打分筛选,富化信息,输出 daily_papers_enriched.json 到共享临时目录,供后续 skill 使用", and positions the skill as "3 步流水线的第 1 步". This matches the comprehensive-coverage anchor; it is not score 4 because no meaningful capability of the fetch stage is left unnamed.

5 / 5

Completeness

It explicitly answers both questions: what ("抓取...打分筛选,富化信息,输出 daily_papers_enriched.json" with the pipeline-step framing) and when (a dedicated "触发词" line with concrete trigger phrases). Voice is third-person verb phrases, so no voice penalty applies; this clearly matches the anchor requiring both what and when with concrete triggers, not the anchor where 'when' is only weakly implied.

5 / 5

Trigger Term Quality

Explicit natural trigger phrases are listed: "论文抓取"、"跑一下论文抓取" plus multi-day variants "过去3天论文推荐"、"过去一周论文推荐"、"过去一周的论文"、"抓 3 天的论文"、"最近5天" — covering the bare command, colloquial phrasing, and synonym/number variations a user would actually say. This is comprehensive natural-term coverage including synonyms, matching the top anchor rather than the 'a few natural terms missing' anchor.

5 / 5

Distinctiveness Conflict Risk

The skill has a clear niche (fetch/score/enrich as step 1 of a 3-step pipeline) with distinct triggers like "论文抓取"、"跑一下论文抓取". It is not score 5 because recommendation-flavored triggers such as "过去3天论文推荐"、"过去一周论文推荐" plausibly overlap with the downstream recommendation/review skills of the same pipeline; it is above score 3 because the fetch-specific triggers and the explicit "第 1 步" framing keep it mostly distinct.

4 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 2 suspicious

Warning

Total

15

/

16

Passed

Repository
huangkiki/dailypaper-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.