CtrlK
BlogDocsLog inGet started
Tessl Logo

session-wrap-up

Review a completed, paused, or blocked coding-agent session to identify friction, mistakes, near misses, missing context, and opportunities to improve future work. Use when the user asks to wrap up, run a retrospective, explain what issues the agent faced, capture lessons learned, or recommend improvements to repository instructions, documentation, skills, tooling, tests, or workflows. Produce an evidence-based, propose-only report; do not apply follow-up changes unless the user asks.

75

Quality

92%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, actionable instruction-only skill with a clear sequenced workflow and explicit verification checkpoints. The main room for improvement is minor tightening of redundant guardrails and one worked example to make the report format fully concrete.

Suggestions

Add one short worked example showing a classified finding (category + what happened + impact) paired with a prioritized proposed improvement, to make the report format tangible.

Trim Guardrails entries that restate earlier guidance (e.g., the 'do not exaggerate' and 'do not hide errors behind tooling' points already surface in Interrogate/Classify) to improve token efficiency.

Consider folding the 'Verify the factual premise' paragraph and the 'cheap read-only checks' paragraph into a single explicit 'Verify current state' step in the sequence so the checkpoints read as numbered workflow steps rather than prose.

DimensionReasoningScore

Conciseness

The body is instruction-dense and assumes Claude's competence — no padded explanations of what a session or retrospective is. It earns a 4 rather than 5 because the Guardrails section partly restates earlier guidance (e.g., 'Do not exaggerate routine exploration into a systemic problem' and 'Do not hide the agent's own errors behind tooling or instruction complaints' echo points already made above), which could be trimmed. Not below 4: there is no verbosity explaining concepts Claude already knows.

4 / 5

Actionability

Guidance is concrete and specific: named finding categories, an exact report section list with per-section content, priority labels (now/soon/only if repeated), and specific checks ('inspect working-tree status and the changed-file summary'). It stops at 4 because, for an instruction-only skill, one worked example of a classified finding paired with a proposed improvement would make the format fully tangible. Not below 4: the guidance is already mostly executable, not abstract.

4 / 5

Workflow Clarity

A clear multi-step sequence (Interrogate → Classify → Recommend → Report) with explicit verification checkpoints: 'Verify the factual premise behind user questions', 'perform only the cheap read-only checks needed to confirm the current state', and reconciliation with recorded verification. Error-handling guidance is present ('State when relevant evidence is unavailable'; 'Say No material friction observed when that is the honest conclusion'). Not below 5: checkpoints are explicit, not merely implicit. The destructive/batch cap does not apply because the skill is explicitly propose-only.

5 / 5

Progressive Disclosure

A single self-contained SKILL.md with well-organized section headers (Interrogate, Classify, Recommend, Report Format, Guardrails) and no nested external references; no bundle files are present or needed. Not below 5: navigation via headers is clear and all content belongs inline as core guidance.

5 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: third-person voice, concrete action list, comprehensive natural trigger terms, and an explicit 'Use when' clause. Both the what and the when are answered clearly and specifically.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — 'identify friction, mistakes, near misses, missing context, and opportunities to improve future work' and 'Produce an evidence-based, propose-only report' — giving comprehensive coverage of what the skill does. Not below 5: coverage is not merely 'several actions with minor gaps'; it enumerates the full set of outputs the skill produces.

5 / 5

Completeness

Explicitly answers both 'what' ('Review a completed, paused, or blocked coding-agent session to identify ...') and 'when' ('Use when the user asks to wrap up, run a retrospective ...') with concrete trigger phrases. Not below 5: the when-clause is explicit and specific, not merely weakly implied.

5 / 5

Trigger Term Quality

Natural user-facing phrases are comprehensive and include synonyms: 'wrap up', 'run a retrospective', 'explain what issues the agent faced', 'capture lessons learned', and 'recommend improvements to ... documentation, skills, tooling, tests, or workflows'. Not below 5: these are exactly the terms a user would say, with common variations covered.

5 / 5

Distinctiveness Conflict Risk

Occupies a clear niche — session retrospectives / wrap-ups — with distinct triggers ('wrap up', 'retrospective', 'lessons learned') unlikely to fire for unrelated skills. Not below 5: overlap with general code-review or commit-message skills is minimal given the retrospective framing.

5 / 5

Total

20

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
chattocorp/chatto
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.