CtrlK
BlogDocsLog inGet started
Tessl Logo

self-evolution

复盘已经发生的工作并沉淀改进,不是“我们来进化 X”的产品入口。 Use when: operator scope 发散偏离愿景、同类错误反复出现、SOP 流程缺口、有价值的知识/方法论值得沉淀。 Not for: 用户询问“能进化什么”、能力进化边界,或要求“我们来进化 X”(用 capability-evolution);日常 SOP 推进(正常执行);一次性个案 bug fix。 Output: Scope Guard Log 记录 / Evolution Proposal 提案 / Episode Card → Method/Skill 蒸馏 → Eval 验证。

66

Quality

79%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./cat-cafe-skills/self-evolution/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is dense, well-structured, and unusually actionable for a process-governance skill, with concrete templates, thresholds, and gated workflows for all three modes. Its weaknesses are token efficiency (dated changelog asides and internal-codename commentary) and progressive disclosure (a long monolithic file inlining protocols that would fit better in reference files).

Suggestions

Move the F167 理解偏差自记录 evidence protocol and the Mode C Eval Ledger spec (judge weights, case counts, gates) into reference files under references/ and keep one-line summaries with clearly signaled links in SKILL.md.

Remove the dated revision-history parenthetical ("2026-07-15 修订:...") and similar changelog commentary from the body, or relocate them to a deprecated/old-patterns section; they add tokens without changing current behavior.

Specify the failure path after the 30-day replay check (e.g., what happens when the same error class recurs) and how the Scope Guard success-rate metric adjusts trigger sensitivity, so those feedback loops are explicit rather than implicit.

DimensionReasoningScore

Conciseness

Mostly efficient — tables and terse imperative bullets throughout — but it includes unnecessary meta-commentary: the dated revision-history parenthetical ("2026-07-15 修订:旧版单次'笨猫'即当轮写档,与本节硬护栏 3...同节自相矛盾"), slogans ("被纠正不丢人,不记录才丢人"), and an internal-codename-laden aside ("挫败语气词后跟玩笑('笨猫哈哈哈')"). This matches 'mostly efficient but includes some unnecessary explanation or could be tightened', and the dated changelog content penalizes it further since it is not in an old-patterns section.

3 / 5

Actionability

The body gives copy-paste-ready material for the common cases: the exact Scope Guard log row format ("| {date} | {feat_id} | {signal_type} | {action_taken} | {outcome} | {agent} |"), a literal guardrail message template, the 5-slot Evolution Proposal template, a full markdown Case E{N} record template, and concrete eval gates with numbers (3-case smoke gate, 5-case promotion gate covering 3 case classes, "overall ≥ 3.5/5 AND boundary ≥ 4/5", judge weight table). For an instruction-only skill this is fully executable guidance.

5 / 5

Workflow Clarity

Multi-step processes are clearly sequenced with explicit checkpoints: the 6-step Mode B proposal flow includes "先闭环当前任务" and a "30 天验证" replay check, and Mode C has gated promotion (smoke/promotion gates, judge scoring, pass thresholds, "Judge 不能是创建知识的同一 agent"). Not 5 because some checkpoints are implicit — e.g., what to do when the 30-day replay check finds the error still recurring is left unspecified, and Mode A's "效果追踪" success-rate loop has no defined adjustment procedure.

4 / 5

Progressive Disclosure

Section structure is good, but this ~280-line body inlines several detailed protocols that belong in separate reference files — the F167 理解偏差自记录 evidence protocol (~35 lines), the full Eval Ledger spec with judge weights and gates, and the five-level maturity ladder — while deferring only once ("详见 ADR-015"). No references/ bundle exists, so template paths like `docs/evolution-proposals/TEMPLATE.md` and `evals/mode-c/TEMPLATE/` are repo-external pointers rather than a skill bundle; this matches 'some structure but content that should be separate is inline'. Not 2 because headers, tables, and the routing section make it navigable.

3 / 5

Total

15

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states what the skill does, gives explicit positive and negative triggers, names its outputs, and pre-empts confusion with the closely related capability-evolution skill. The only weakness is that trigger coverage relies on internal operator/SOP jargon rather than broader natural phrasings.

DimensionReasoningScore

Specificity

"复盘已经发生的工作并沉淀改进" names the function and the Output clause lists several concrete artifacts ("Scope Guard Log 记录 / Evolution Proposal 提案 / Episode Card → Method/Skill 蒸馏 → Eval 验证"), matching the 'several specific actions; minor gaps' anchor. Not 5 because the artifacts are named rather than the concrete actions themselves, leaving some coverage implicit.

4 / 5

Completeness

It explicitly answers both questions: what ("复盘已经发生的工作并沉淀改进") and when (explicit "Use when:" and "Not for:" clauses with concrete trigger phrases). Not below 5 because both are concrete and explicit, including negative triggers.

5 / 5

Trigger Term Quality

The Use when clause gives natural ecosystem phrases ("operator scope 发散偏离愿景、同类错误反复出现、SOP 流程缺口、有价值的知识/方法论值得沉淀") plus a Not for list, giving good keyword coverage. Not 5 because synonyms and common user phrasings beyond the operator/SOP vocabulary (e.g., plain-language variants a user might say) are not covered.

4 / 5

Distinctiveness Conflict Risk

The description opens with an explicit boundary ("不是'我们来进化 X'的产品入口") and the Not for clause routes overlapping intents to a named sibling skill ("用 capability-evolution"), plus it states the temporal dividing line (past-facing review vs. future-facing evolution). Clear niche with minimal conflict risk.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
zts212653/clowder-ai
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.