CtrlK
BlogDocsLog inGet started
Tessl Logo

loop-design-check

Design a goal-oriented agent loop or review one for failure modes: spinning, Goodhart-gaming the verifier, or running a wrong answer to completion. Covers machine-decidable goals, loop types, plan/build/judge skeletons, and runaway prevention; mechanism wiring lives in autonomous-loops. Use when designing, writing, or checking an agent loop. 中文触发:写 loop、设计 loop、做一个 loop、检查 loop 对不对、loop 体检、loop 会不会跑飞、可判定目标、五个崩法、plan build judge。

75

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body delivers high-value, opinionated design guidance with excellent sequencing, validation checkpoints, and a concrete worked example — everything in it is actionable and little of it is filler. Its weaknesses are minor: some rhetorical flourish that could be trimmed for token efficiency, and a monolithic single-file layout where a references/ bundle (e.g., the worked example or the failure-mode detail) would lighten the main file.

Suggestions

Trim rhetorical flourishes (e.g., "the system tends toward entropy, so maintain it", the aphoristic one-line close) to reclaim tokens without losing guidance.

Move the worked example and/or the failure-mode detail into a references/ file (e.g., references/worked-example.md) linked from the checklist section, keeping SKILL.md as a lean overview.

Tighten Step 4 (damping) into a concrete decision table — which damping to apply for servo vs. regulator, and default retry-cap values — instead of a single prose paragraph.

DimensionReasoningScore

Conciseness

The body is dense with non-obvious judgment content (the 4-condition gate, "Build may not edit the acceptance conditions to pass", the failure-mode table) rather than things Claude already knows, matching 'efficient'. It falls short of the lean anchor-5 because of rhetorical padding such as "the system tends toward entropy, so maintain it" and the aphoristic one-line close, which could be trimmed without losing guidance.

4 / 5

Actionability

For an instruction-only skill the guidance is fully actionable: a concrete 4-condition gate ("① the task repeats weekly or more ② verification can be automated…"), decision tables for loop type and skeleton, explicit rules ("Three failed retries → escalate to a human"), a five-row checklist with literal review questions, and a worked example showing the naive vs. fixed goal. This covers common cases concretely; score 4 would require missing key details, and none are missing.

5 / 5

Workflow Clarity

Action 1 is a clearly sequenced 5-step process with explicit checkpoints (the 4-condition veto gate, the Step-1 self-check "can they run one command and tell whether it's done?", the staged landing plan), and Action 2 is a per-row review checklist with explicit failure conditions and antibodies. Validation and feedback loops (retry caps, escalation, three-stage rollout) are present, matching the anchor-5 pattern of sequence + explicit validation + checklists.

5 / 5

Progressive Disclosure

Structure is good: clear section headers, a two-level 'use / don't use' scope statement, and well-signaled one-level pointers to the mechanism layer ("see `autonomous-loops` / `continuous-agent-loop`"). It sits below anchor 5 because the single ~137-line file inlines content that could be split into references (the worked example, the failure-mode/lineage notes), and there are no reference files at all to spread detail into.

4 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is a strong example: third-person, concise, explicit about both what the skill does and when to use it, with rich natural trigger terms in both English and Chinese and an explicit boundary against a sibling skill. It would both surface correctly in trigger matching and avoid conflict with the mechanism-layer skill.

DimensionReasoningScore

Specificity

Quotes like "Design a goal-oriented agent loop or review one for failure modes: spinning, Goodhart-gaming the verifier, or running a wrong answer to completion" list multiple specific concrete actions plus enumerated coverage ("machine-decidable goals, loop types, plan/build/judge skeletons, and runaway prevention"), matching the 'comprehensive coverage' anchor. It is not the level below (4) because there are no meaningful gaps in what it claims to cover.

5 / 5

Completeness

The 'what' is explicit (design and review agent loops, with covered topics and named failure modes) and the 'when' is explicit ("Use when designing, writing, or checking an agent loop") with concrete trigger phrases. This is exactly the anchor-5 example pattern (clear what AND when with concrete triggers); score 4 would require a weaker 'when', which is not the case.

5 / 5

Trigger Term Quality

"Use when designing, writing, or checking an agent loop" covers natural English triggers, and the Chinese trigger list ("写 loop、设计 loop、做一个 loop、检查 loop 对不对、loop 体检、loop 会不会跑飞、可判定目标、五个崩法、plan build judge") adds comprehensive synonyms users would actually say. This matches the 'comprehensive coverage of natural terms including synonyms' anchor; nothing common is missing.

5 / 5

Distinctiveness Conflict Risk

It carves a clear niche (goal definition and runaway prevention for agent loops) and explicitly fences off the sibling skill: "mechanism wiring lives in autonomous-loops." Triggers like "loop 会不会跑飞" and "检查 loop 对不对" are distinct from generic loop/timer keywords, so conflict risk is minimal, matching the 'clear niche with distinct triggers' anchor.

5 / 5

Total

20

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

Total

15

/

16

Passed

Repository
affaan-m/ECC
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.