CtrlK
BlogDocsLog inGet started
Tessl Logo

world-model-runtime

在重要任务上运行完整世界模型闭环的执行协议——恢复持久状态→建立竞争模型→显式预测→真实观察→修订模型→主动取证→决策→持久化。由 AGENTS.md World Model Gate 判 OFF/CORE/FULL 后加载;用于多步、高风险、跨会话、存在竞争解释或需要持久世界模型的真实任务,不用于简单一次性任务。

71

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

80%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, lean instruction-only protocol with concrete schemas, exact paths, and an explicit sequenced workflow with validation checkpoints. It is held to 4 rather than 5 by minor tightenability, instruction-shaped (non-copy-paste) examples, and a few policy-stated rather than enforced checkpoints.

DimensionReasoningScore

Conciseness

Dense, operational protocol prose with no padding and no explanation of concepts Claude already knows; each line earns its place, though a few section paragraphs (e.g. 各环要点) could be tightened further.

4 / 5

Actionability

Provides concrete schemas with required fields (Episode/Prediction), exact state paths, an explicit tool-risk grading ladder, and gating rules — mostly executable guidance, but as an instruction-shaped skill it stops short of fully copy-paste-ready examples.

4 / 5

Workflow Clarity

Defines an explicit ordered execution sequence (WM_ACTIVATE→STATE_RESTORE→…→STATE_PERSISTED) with 'sequence is acceptance' and validation checkpoints (guard rejects unbound predictions, prediction evaluation, canonical write threshold); a few checkpoints are policy-stated rather than machine-enforced.

4 / 5

Progressive Disclosure

Single self-contained SKILL.md with well-organized sections and no nested bundle references (references/scripts/assets absent); it exceeds the simple-skill line count but stays well-structured and navigable.

4 / 5

Total

16

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, specific description that answers what and when concretely, lists a comprehensive action chain, and carves out a distinct niche with an explicit negative boundary. Slightly technical trigger phrasing keeps it just short of a perfect trigger-term score.

DimensionReasoningScore

Specificity

Lists multiple concrete actions in sequence — '恢复持久状态→建立竞争模型→显式预测→真实观察→修订模型→主动取证→决策→持久化' — giving comprehensive coverage of what the protocol does.

5 / 5

Completeness

Explicitly answers both 'what' (the full closed-loop protocol with its eight steps) and 'when' (multi-step/high-stakes/cross-session/competing-interpretations, with an explicit 'not for simple one-off tasks' boundary).

5 / 5

Trigger Term Quality

Includes natural trigger phrases like '多步、高风险、跨会话、存在竞争解释' and the explicit negative '不用于简单一次性任务', but the terms lean technical/abstract rather than the everyday synonyms a user would naturally say.

4 / 5

Distinctiveness Conflict Risk

Occupies a clear niche (world-model runtime gated by the AGENTS.md World Model Gate) with explicit negative boundaries, so overlap with other skills is minimal.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
ooooooooooooooooooop/personal-ai
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.