CtrlK
BlogDocsLog inGet started
Tessl Logo

self-improvement-loops

This skill should be used when the harness, scaffold, workflow, or optimizer itself is the optimization target: recursive self-improvement (RSI) loops, meta-harnesses, self-improving harnesses that mine their own failures and propose bounded edits, evolutionary or population-based search over agent scaffolds, acceptance gates for self-modifying systems, and agentic context evolution where the mechanism that produces context is versioned and evolved. Route governance of a single autonomous loop (locked surfaces, durable logs, rollback, novelty gates, approval boundaries) to harness-engineering, measurement and quality-gate design to evaluation, judge design to advanced-evaluation, and remote sandbox infrastructure to hosted-agents.

68

Quality

83%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured, dense instruction skill with concrete operational guidance, clear validation-gated workflows, and clean progressive disclosure to a real reference file. Its only weakness is mild prose verbosity in places and design-altitude (rather than fully executable) treatment of some processes.

DimensionReasoningScore

Conciseness

Dense and information-rich, surfacing specific research systems and named claims Claude would not already know, with little padding about basics; a few prose passages could be tightened further, keeping it just short of fully lean.

4 / 5

Actionability

Provides concrete operational rules, a loop-readiness checklist, decision tables, an executable accept() gate function, and an archive layout; as an instruction-only design skill most guidance is specific and actionable, with only minor gaps in copy-paste-ready detail.

4 / 5

Workflow Clarity

Multi-stage processes are clearly sequenced (the three-stage failure-driven self-edit loop) with explicit validation checkpoints (two-split acceptance gate, programmatically re-verified locked regions, loop-readiness checklist), though some processes are described at design altitude rather than as a strict step-by-step runbook.

4 / 5

Progressive Disclosure

SKILL.md is a clear overview with well-signaled, one-level-deep references; the single referenced bundle file ./references/loop-design-evidence.md exists and is clearly labeled, and external resources are enumerated separately.

5 / 5

Total

17

/

20

Passed

Description

88%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, complete, and distinguishes itself via explicit routing to adjacent skills, with solid natural trigger terms. Its main weakness is a slightly technical register in the trigger phrasing rather than broad plain-language synonyms.

DimensionReasoningScore

Specificity

Lists multiple concrete capabilities ('mine their own failures and propose bounded edits', 'evolutionary or population-based search over agent scaffolds', 'acceptance gates for self-modifying systems', 'context evolution where the mechanism... is versioned and evolved'), giving comprehensive coverage of the domain.

5 / 5

Completeness

Explicitly answers 'when' ('This skill should be used when the harness, scaffold, workflow, or optimizer itself is the optimization target') and 'what' (the enumerated self-improvement loop families), plus concrete routing triggers.

5 / 5

Trigger Term Quality

Strong natural terms a user would say ('recursive self-improvement (RSI) loops', 'meta-harnesses', 'evolutionary search over agent scaffolds') with the RSI synonym included, though the phrasing leans technical rather than covering many plain-language variations.

4 / 5

Distinctiveness Conflict Risk

Has a clear niche and explicit routing to sibling skills ('Route governance... to harness-engineering', 'measurement and quality-gate design to evaluation', 'judge design to advanced-evaluation', 'remote sandbox infrastructure to hosted-agents'), with only minor residual overlap with harness-engineering given the inherently adjacent domain.

4 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
muratcankoylan/Agent-Skills-for-Context-Engineering
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.