CtrlK
BlogDocsLog inGet started
Tessl Logo

citation-verification

This skill provides reference guidance for citation verification in academic writing. Use when the user asks about "citation verification best practices", "how to verify references", "preventing fake citations", or needs guidance on citation accuracy. This skill supports ml-paper-writing by providing detailed verification principles and common error patterns.

51

Quality

56%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/citation-verification/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

42%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is well-sectioned and contains genuinely useful concrete guidance (authority order, matching tolerances, failure markers), but it is significantly undermined by an unresolved internal contradiction — the new canonical-source principle (DOI/CrossRef/arXiv first, Google Scholar demoted) coexists with old Google-Scholar-centric instructions in the example, Best Practices, and Summary — plus heavy cross-section duplication and a complete failure to link the existing references/ and scripts/ bundle files.

Suggestions

Resolve the Google Scholar contradiction: update the worked example, Best Practices, and Summary to fetch BibTeX from CrossRef/arXiv/publisher metadata per the Core Principle, or explicitly present the Google Scholar path as fallback-only the way Verification Principle 2 does.

Link the bundle from SKILL.md (e.g., "**Verification rules**: See references/verification-rules.md", "**API usage**: See references/api-usage.md", "**Common errors**: See references/common-errors.md") instead of inlining that detail, and deduplicate the twice-written failure-handling and Best Practices sections into one.

Delete the closing Summary section (or compress it to the Core Principle plus the [CITATION NEEDED] convention) — it restates the entire skill a third time.

DimensionReasoningScore

Conciseness

The ~200-line body is noticeably verbose with clear duplication: "Handling Verification Failures" appears as a full ## section and again nearly verbatim as a ### under Best Practices ("Don't guess", "[CITATION NEEDED]", "Notify the user"); the Best Practices items ("Never generate citations from memory", "Use WebSearch to find", "Verify promptly") restate the Verification Principles section; and the closing Summary re-explains the entire skill a third time. This matches anchor 2 ("Noticeably verbose; several unnecessary explanations or padded sections") rather than anchor 3, since multiple whole sections could be deleted with no information loss.

2 / 5

Actionability

There is real concrete guidance — the preferred authority order (DOI → arXiv → CrossRef → Semantic Scholar → Zotero → Google Scholar), matching tolerances ("Year (±1 year difference allowed)"), a concrete failure marker ("[CITATION NEEDED]"), and example WebSearch queries ("Attention is All You Need Vaswani 2017"). But the guidance is undermined by direct contradictions: the Core Principle states "Google Scholar... is not the canonical verification authority", yet the example instructs "Click 'Cite' on Google Scholar → Select BibTeX format → Copy BibTeX entry" and Best Practices demand "Confirm on Google Scholar" / "Copy BibTeX from Google Scholar". As written, an agent cannot execute a single consistent procedure, which keeps this at anchor 3 ("Some concrete guidance but incomplete; missing key details") rather than anchor 4.

3 / 5

Workflow Clarity

A clear ASCII workflow with a validation gate exists ("Add to bibliography" only after "Confirm paper details"), and failure handling forms a feedback loop (check spelling → try different queries → alternative sources → mark pending → notify user), so this is above anchors 1-2. It falls short of anchor 4 because two competing workflows coexist: the flow diagram says "Find DOI / arXiv ID / publisher page... Verify metadata with CrossRef / arXiv / Semantic Scholar", while the worked example and Summary route through Google Scholar ("Google Scholar to verify existence"). The batch/destructive cap at 3 does not apply (this is neither batch nor destructive), but the incoherent dual sequence prevents the clear single sequence anchor 4 requires.

3 / 5

Progressive Disclosure

Scored against the actual bundle: references/ contains common-errors.md, verification-rules.md, and api-usage.md, and scripts/ contains verify-citations.py, api-clients.py, and format-checker.py — yet the SKILL.md body references none of them (zero links or path mentions), so the bundle is undiscoverable from the skill entry point. The body itself is well-sectioned with headers, so it is not the anchor-2 "minimal structure" / wall of text; but the ~200 lines inline detailed verification rules and error patterns that duplicate the reference files while leaving those files completely unsignaled. This matches anchor 3 ("references present but not clearly signaled; content that should be separate is inline") and rules out anchor 4 ("references mostly clear").

3 / 5

Total

11

/

20

Passed

Description

70%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A solid description with a genuinely useful, explicit "Use when..." trigger clause containing natural quoted phrases, which is its strongest asset. The main weakness is that the capability statement is generic ("provides reference guidance") rather than naming the concrete verification actions the skill performs.

Suggestions

Replace "provides reference guidance" with 1-2 concrete operations, e.g., "Verifies citations against canonical sources (DOI/CrossRef, arXiv, Semantic Scholar), checks title/author/year/venue match, and flags unverifiable citations".

Add common trigger synonyms such as "check citations", "fake references", "hallucinated citations", or "BibTeX verification" to broaden natural keyword coverage.

DimensionReasoningScore

Specificity

The description names the domain ("citation verification in academic writing") and gestures at two deliverables ("detailed verification principles and common error patterns"), but the actual action verb is the generic "provides reference guidance" — no concrete operations like verifying against DOI/CrossRef/arXiv or flagging unverifiable citations. It sits between anchor 2 ("actions are minimal or generic") and anchor 3 ("1-2 concrete actions"): there is slightly more concrete content than the anchor-2 example, but less actionable specificity than the anchor-3 example, so 3 is the best fit.

3 / 5

Completeness

Both parts are present and explicit: the "what" ("provides reference guidance for citation verification... providing detailed verification principles and common error patterns") and a full "Use when the user asks about... or needs guidance on..." clause with concrete quoted triggers — so the anchor-3 cap (missing/weak "when") does not apply. It does not reach anchor 5 because the "what" is generic ("provides reference guidance") rather than stating concretely what the skill does when invoked, matching anchor 4 ("'what' present but could be more specific").

4 / 5

Trigger Term Quality

Quoted trigger phrases are natural and well-chosen: "citation verification best practices", "how to verify references", "preventing fake citations", plus "citation accuracy". A user needing this skill would plausibly say these. It falls short of anchor 5 because common synonyms are missing — e.g., "check citations", "fake references", "hallucinated citations", "BibTeX" — leaving a few natural entry points uncovered, which matches anchor 4 ("Good keyword coverage; a few natural terms missing").

4 / 5

Distinctiveness Conflict Risk

This is a clear niche — citation verification in academic writing — with distinct trigger phrases, so anchor 5 territory. However, the sentence "This skill supports ml-paper-writing" creates minor overlap/conflict risk with the closely related ml-paper-writing skill (a paper-writing request could plausibly trigger either), matching anchor 4 ("Mostly distinct; minor overlap risk with closely related skills").

4 / 5

Total

15

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
Galaxy-Dawn/claude-scholar
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.