CtrlK
BlogDocsLog inGet started
Tessl Logo

research-pipeline

Full end-to-end research pipeline: from a broad research direction through idea discovery, experiments, and review all the way to a polished paper PDF. Use when user says "全流程", "full pipeline", "从找idea到投稿", "end-to-end research", or wants the complete autonomous research lifecycle.

63

Quality

76%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./skills/research-pipeline/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

70%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered orchestrator skill with an exceptionally clear staged workflow, explicit gates, and acceptance criteria, plus mostly concrete invocations. Its weaknesses are redundancy (heartbeat doctrine and helper-resolution chains repeated inline) and a long monolithic body that inlines content the referenced shared files already carry.

Suggestions

State the external-cadence/heartbeat doctrine once (in the dedicated 'Overnight heartbeat' section) and reduce the opening blockquote to a two-line pointer to shared-references/external-cadence.md — the full doctrine currently appears twice.

Collapse the three repetitions of the canonical helper-resolution chain into one: define it once (or in a referenced integration-contract file) and refer to it from the heartbeat and resumable-runs sections.

Define the runtime variables ($ROOT, $RUN_ID, $STAGE, $N_NEW_FINDINGS, $CHOSEN_IDEA_TITLE) once at the top of the body, and write the run_state.py state-transition commands with their full script prefix so they are copy-paste executable.

DimensionReasoningScore

Conciseness

The body is dense operational instruction rather than concept explanation, but there is noticeable duplication: the external-cadence/heartbeat doctrine appears both in the opening blockquote and again in full in the "Overnight heartbeat" section, and the canonical helper-resolution chain (`.aris/tools → tools → $ARIS_REPO`) is stated three times, including a 7-line shell block. This fits the 3 anchor (mostly efficient but could be tightened) better than 2 (no pervasive padding) or 4 (only minor trims needed).

3 / 5

Actionability

Mostly executable guidance: exact slash-command invocations with argument passthrough (e.g. `/experiment-bridge "$CHOSEN_IDEA_TITLE" — code review: $CODE_REVIEW, base repo: $BASE_REPO`), concrete `run_state.py` invocations, and a filled-in report template. Minor gaps keep it below 5: state transitions are written in shorthand (`set <run_id> <phase> running` without the script prefix) and variables like `$ROOT`, `$RUN_ID`, `$STAGE`, `$N_NEW_FINDINGS` are used without being defined anywhere.

4 / 5

Workflow Clarity

Stages 1-5 are clearly sequenced with explicit gates, each gate's blocking vs non-blocking behavior is spelled out under AUTO_PROCEED, acceptance criteria are tabulated per phase, and there are feedback loops throughout (re-run failed stages, re-audit done-but-unaccepted stages, review rounds with an explicit STOP condition, bounded 3-attempt debugging). This matches the 5 anchor (clear sequence, explicit validation, feedback loops, checklists).

5 / 5

Progressive Disclosure

References are clearly signaled with markdown links (external-cadence.md, resumable-runs.md, output-versioning/manifest/language protocols), but no bundle files exist in the skill directory and the linked shared-references/templates are not resolvable here. Moreover, material that belongs in those referenced files — the full heartbeat doctrine and the multi-line shell helper-resolution chain — is duplicated inline in the ~390-line body. This fits the 3 anchor (references present, but content that should be separate is inline) rather than 4's 'appropriately placed' or 2's 'references buried'.

3 / 5

Total

15

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that explicitly states what the skill does and when to use it, with concrete bilingual trigger phrases in third person. It is specific and well-differentiated; only broader synonym coverage of natural user phrasings keeps it from the top of the scale.

DimensionReasoningScore

Specificity

The description names concrete pipeline actions — "from a broad research direction through idea discovery, experiments, and review all the way to a polished paper PDF" — listing several specific stages. It falls short of a 5 because the individual actions are stage names rather than fully enumerated capabilities, and short of 3 because it clearly covers the whole lifecycle, not just 1-2 actions.

4 / 5

Completeness

Both halves are explicit: the "what" is the full pipeline from research direction to polished paper PDF, and the "when" is a literal "Use when user says ..." clause with concrete trigger phrases. This matches the 5 anchor (explicit what AND when with concrete trigger phrases) and exceeds the 4 anchor where 'when' is present but less specific.

5 / 5

Trigger Term Quality

Triggers include natural phrases users would say: "全流程", "full pipeline", "从找idea到投稿", "end-to-end research", plus the paraphrase "wants the complete autonomous research lifecycle". Coverage is good with bilingual synonyms, but common single-stage variations a user might actually say (e.g. "write my paper", "run the whole pipeline", "投稿") are only partially represented, keeping it below the comprehensive 5 anchor.

4 / 5

Distinctiveness Conflict Risk

The pipeline framing ("full end-to-end research pipeline", "complete autonomous research lifecycle") carves a clear orchestrator niche distinct from its component skills (idea discovery, experiments, review, paper writing). Minor overlap risk remains: a user asking only for "end-to-end research" during a single sub-stage could plausibly trigger this skill instead of a component skill, so it is not the minimal-conflict 5.

4 / 5

Total

17

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

relative_links

Relative link issues: 6 suspicious

Warning

Total

13

/

16

Passed

Repository
wanshuiyin/Auto-claude-code-research-in-sleep
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.