CtrlK
BlogDocsLog inGet started
Tessl Logo

source-audit

外部高风险 claim 与研究贡献审计。Use when: 数字、benchmark、因果、趋势、模型能力、 外部论文或会进入 docs/ADR/PPT 的结论。Not for: 低风险常识、只读官方原文且不外推、 已进入 deep-research 的重调研。Output: claim ledger + source / non-triviality / decision-fit 三轴 verdict + provenance。

68

Quality

84%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, well-structured instruction-only audit methodology with concrete templates, explicit evidence-sufficiency gates, and worked pressure tests. The main gaps are the absence of a fully filled worked example and some length that could be offloaded to reference files.

Suggestions

Add one fully worked example: a completed claim ledger row plus the resulting three-axis verdict for a concrete claim (e.g. the MemU 65% case), so the templates are demonstrated end-to-end.

Move the detailed minimal execution format and/or Pressure Tests into a separate reference file referenced one level deep from SKILL.md to tighten the core overview.

Trim the Common Mistakes table to the highest-impact rows or collapse low-frequency entries to reduce length while preserving the failure-mode coverage.

DimensionReasoningScore

Conciseness

Densely packed operational methodology with no padding of concepts Claude already knows, though at ~208 lines some sections (Common Mistakes table, Pressure Tests) could be tightened without losing substance.

4 / 5

Actionability

Provides concrete copy-paste templates (claim ledger table with column specs, minimal execution format field blocks, provenance line format) and explicit verdict criteria, but lacks a fully worked filled example of a completed ledger row or verdict for a real claim.

4 / 5

Workflow Clarity

Clear sequence (Trigger → Claim Ledger → 八问 checklist → three-axis Verdict → Provenance) with explicit validation gates (最低证据面 caps that downgrade verdicts) and Pressure Test '合格输出必须' criteria serving as feedback checkpoints.

5 / 5

Progressive Disclosure

Well-organized into clearly headed sections with no nested references and no bundle files to navigate; over the 50-line simple-skill threshold, and some detailed material (execution format, Pressure Tests) could live in reference files, so it stops short of a 5.

4 / 5

Total

17

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, well-structured description that clearly states what the skill does, when to use it, when not to use it, and what it outputs. Minor room to expand the audit action verbs and trigger synonyms, but it cleanly answers both 'what' and 'when' with low conflict risk.

DimensionReasoningScore

Specificity

Names the domain ('外部高风险 claim 与研究贡献审计') and lists several concrete deliverables ('claim ledger + source / non-triviality / decision-fit 三轴 verdict + provenance'), but the action verb '审计' alone is somewhat generic without enumerating the audit sub-steps.

4 / 5

Completeness

Explicitly answers both what ('外部高风险 claim 与研究贡献审计' + Output contract) and when ('Use when: ...' with concrete trigger phrases), and adds a 'Not for' boundary.

5 / 5

Trigger Term Quality

Good natural trigger coverage ('数字、benchmark、因果、趋势、模型能力、外部论文或会进入 docs/ADR/PPT 的结论') plus a 'Not for' anti-trigger, though a few common synonyms (e.g. statistics, citation) are absent.

4 / 5

Distinctiveness Conflict Risk

Clear niche (external high-risk claim provenance hygiene) with distinct triggers and an explicit 'Not for: 已进入 deep-research 的重调研' boundary that minimizes overlap with adjacent skills.

5 / 5

Total

18

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
zts212653/clowder-ai
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.