CtrlK
BlogDocsLog inGet started
Tessl Logo

retro-v2

Use when a branch has finished and the session-metrics card is all that is wanted — writes, commits and renders the card, and stops. The card-only half of `retro`, for sessions running close to their token budget. Where there is room to think, run `retro` instead — this one deliberately produces no findings and changes nothing.

72

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An unusually well-executed procedural skill: every command is literal and executable, destructive edges (PR body edits, directory-wide staging) have explicit guards, and all explanation is project-specific rather than filler. The minor weaknesses are repetition between the description and the opening sections, and rationale prose that could be split into a reference file.

DimensionReasoningScore

Conciseness

Nearly every paragraph carries non-obvious project constraints Claude could not know (the cross-repository `.closingIssuesReferences[0].number` bug, the private-tracker label rule, the guarded `git add`) — no padding on concepts Claude already knows. Not 5 because "When to run this instead of `retro`" and "Where this sits" repeat the description and state the deferred-analysis point twice, and a few rationale passages (e.g. the transcript-lifecycle paragraphs) could be trimmed.

4 / 5

Actionability

Fully executable, copy-paste-ready commands throughout: the three environment probes, `bun run --cwd tools/pr-metrics card -- --resolve-issue` with the markdown variant, the guarded `git add`/commit/push, the PR body append, and the idempotency `grep`. Common cases and failure modes (no `gh`, no network) are covered with concrete instructions. Not 4 because there is no gap — every step is a literal command with its purpose stated inline.

5 / 5

Workflow Clarity

Clear sequence (probe → write card → render onto PR → finish) with explicit validation checkpoints: the pre-flight network/`gh`/transcripts probe, the grep guard against stacking a second `<details>` block, the single-file staging rule, and recovery paths ("record which and carry on", "commit the card locally; it pushes with the next push"). The destructive risks the workflow does carry (PR body edit, commit/push) are each guarded. Not 4 because checkpoints and error-recovery guidance are present at every stage, not just most.

5 / 5

Progressive Disclosure

No bundle files exist, so this is a single-file skill with clean section headers and no nested references — nothing buried, nothing orphaned. Not 5 because at ~160 lines several long rationale blocks (the four rules in Step 1, the hook and transcript-lifecycle notes) are execution-adjacent rather than execution-critical and could live in a reference file, per the under-50-lines exception not applying here.

4 / 5

Total

18

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person, concise, with an explicit "Use when" trigger, concrete actions, and explicit disambiguation from the sibling `retro` skill. Its only weakness is that the card artifact's contents are referenced generically rather than enumerated.

DimensionReasoningScore

Specificity

Lists several concrete actions — "writes, commits and renders the card, and stops", "produces no findings and changes nothing" — but the card's contents (PR number, issue labels, spend) are only implied, leaving minor gaps in coverage. Not 3 because it goes beyond 1-2 actions; not 5 because the artifact itself is described generically as "the card" without its full payload.

4 / 5

Completeness

Explicitly answers both questions: what ("writes, commits and renders the card, and stops"; "produces no findings and changes nothing") and when ("Use when a branch has finished and the session-metrics card is all that is wanted", "for sessions running close to their token budget") with a concrete trigger clause. Voice is third person throughout, so no person penalty applies.

5 / 5

Trigger Term Quality

Good natural keyword coverage — "branch has finished", "session-metrics card", "token budget", "retro" — phrasing a user of this ecosystem would plausibly say. Not 5 because common variations like "wrap up the session" or "just the metrics" are absent; not 3 because multiple relevant natural terms are present, not just one or two generic ones.

4 / 5

Distinctiveness Conflict Risk

Occupies a clear niche and actively disambiguates against its only close sibling: "The card-only half of `retro`" and "Where there is room to think, run `retro` instead". A reader cannot mistake which skill to invoke; minimal conflict risk.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
englishstreetventures/osn
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.