CtrlK
BlogDocsLog inGet started
Tessl Logo

paper-spine

Research, write, review and deliver evidence-bound papers in one task, with user choices, real files, editable outputs and same-task revision.

52

Quality

57%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./src/skill/SKILL.md

The canonical home for this skill is paper-spine in WUBING2023/PaperSpine

SKILL.md
Quality
Evals
Security

Quality

Content

60%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-referenced, well-sequenced orchestrator document with genuine validation checkpoints and read-back loops, and all linked playbooks exist. Its main weaknesses are verbosity — dense, repetitive policy prose that buries the operational signal — and a bundle inconsistency (a required update script with no scripts/ directory present).

Suggestions

Trim the repeated prohibitions and policy digressions ("Never..." clauses appear in nearly every section) into a single boundary section or the referenced playbooks; the body could be roughly halved without losing operational content.

Resolve the bundle inconsistency: either ship scripts/paperspine_update.py (and the author_voice_check.py / humanize_check.py helpers) or reword the update and quality-check steps so they do not invoke files absent from the skill.

Inline the exact call sequence for the launch/configure/open-task loop (or a minimal example of one paperspine_open_task payload) since the body currently says "Follow the exact calls in product-v1-workflow.md" for its most safety-critical step.

DimensionReasoningScore

Conciseness

The body is ~280 lines of dense, policy-laden prose with heavy padding: prohibitions are repeated many times ("Never invent...", "Never fake review/delivery completion", "Never repeatedly reinstall"), and abstract policy statements like "The manuscript is the product; the task state is only its support" compete with operational content. It does not explain concepts Claude already knows (so not level 1), but several sections are noticeably padded and could be cut or pushed to the referenced playbooks.

2 / 5

Actionability

Concrete anchors are present: "scripts/paperspine_update.py --preflight --yes", "launch --no-open", tool names ("paperspine_open_task", "paperspine_commit_milestone", "paperspine_authorize_materials") with exact field names ("grants[].include_paths", "skill_bridge.web_path") and concrete success conditions ("source=web_user", "user_confirmed=true", "readiness.ready=true"). It is not level 5 because most tool-call syntax is deferred to product-v1-workflow.md and the referenced update script is not present in this bundle (no scripts/ directory), leaving a few gaps.

4 / 5

Workflow Clarity

A clear 7-section scientific workflow is sequenced (anchor → research → contribution → writing/figures → render → review → delivery) with explicit validation checkpoints: "Only then begin the next dependent segment", mandatory read-back after "paperspine_commit_milestone", "Inspect every PDF page and the actual DOCX", and feedback loops ("If read-back fails, fix the concrete binding/sync error before claiming progression"; "After two unchanged attempts, change tactic"). Not level 5 because the density and interleaved policy digressions make several checkpoints implicit and the exact call sequence is delegated to references rather than stated inline.

4 / 5

Progressive Disclosure

Structure is good: every section signals one-level-deep markdown links to real files in references/ (all 34 referenced files verified to exist), organized by workflow stage with clear "Read X as applicable" navigation. It is not level 5 because the SKILL.md itself inlines substantial operational detail (update authority, Web-interruption policy, milestone mechanics) that belongs in the referenced playbooks, and it cites "scripts/paperspine_update.py" which is absent from the bundle.

4 / 5

Total

14

/

20

Passed

Description

55%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description states a clear, multi-part capability for evidence-bound paper production, but has no explicit "when to use" trigger guidance and lacks the synonyms (manuscript, journal, LaTeX, PDF/DOCX) users would naturally say. It is serviceable but below the quality of the good reference examples.

Suggestions

Add an explicit trigger clause, e.g. "Use when the user asks to write, revise, or deliver a research paper or manuscript, or mentions journals, venues, citations, or submission-ready PDF/DOCX outputs."

Include natural synonyms and extensions users would say — manuscript, thesis, preprint, LaTeX, .tex, .docx, references/citations — to broaden keyword coverage beyond the generic verbs.

Sharpen distinctiveness by naming the concrete artifacts (submission-oriented PDF, DOCX, editable LaTeX source) rather than only process verbs like "research, write, review, deliver".

DimensionReasoningScore

Specificity

"Research, write, review and deliver evidence-bound papers" names the domain and several concrete actions, with qualifiers like "editable outputs and same-task revision". It is not comprehensive — the actions are broad verbs with no coverage of figures, citations, or venue formatting — matching the 'several specific actions; minor gaps' anchor rather than the comprehensive level 5.

4 / 5

Completeness

The "what" is clear (research, write, review, deliver evidence-bound papers), but there is no "Use when..." clause or equivalent explicit trigger guidance, which caps completeness at 3 per the judging guidelines. It is not level 2 because the 'what' is concrete and multi-part, not vague.

3 / 5

Trigger Term Quality

"Research", "write", "review", "deliver", "papers" are natural terms, but common variations users would actually say are missing: "manuscript", "journal", "submission", "LaTeX", "citation", "thesis", or any file extensions (.tex, .pdf, .docx). This fits 'Some relevant keywords but missing common variations or synonyms', not level 4 which requires only a few natural terms missing.

3 / 5

Distinctiveness Conflict Risk

"Evidence-bound papers" with "same-task revision" gestures at a scientific-paper niche, but "Research, write, review" are generic verbs that overlap with general writing, document-editing, and code-review skills. 'Somewhat specific but could still overlap with similar skills' fits; it lacks the distinct file-type or artifact triggers that would justify level 4-5.

3 / 5

Total

13

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

referenced_paths_exist

Referenced path issues: 2 missing

Warning

Total

15

/

16

Passed

Repository
WUBING2023/PaperSpine
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.