CtrlK
BlogDocsLog inGet started
Tessl Logo

refit

Cross-session environment retrospective driven by OMC's own instrumentation — trace timelines, friction reports, logs, and plan notepads. Every finding lands on the surface that owns the fix (deterministic check, steering volume, tooling, or information access); nothing is written without user approval.

63

Quality

75%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/refit/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an exemplary instruction-only skill: concrete instruments, paths, and formats; an approval-gated workflow with a genuine prove-the-check-bites feedback loop; and a well-sectioned single-file layout that needs no bundle. The only room for improvement is trimming a handful of rhetorical asides to sharpen token efficiency.

DimensionReasoningScore

Conciseness

The body is dense with environment-specific information Claude cannot infer (instrument names, file paths, surface taxonomy, headless contract) and mostly earns its tokens. It sits at 4 rather than 5 because of a few rhetorical flourishes and rationale passages that could be trimmed — e.g. 'a guardrail nobody has seen fire is decoration, not protection' and the paragraph beginning 'The split rests on where enforcement pressure lives...' — and it never devolves into the level-3 pattern of explaining things Claude already knows.

4 / 5

Actionability

For an instruction-only skill, the guidance is fully concrete: exact instruments to read ('trace_timeline / trace_summary', 'omc session friction report'), exact paths ('.omc/logs/', '.omc/notepads/*/issues.md', '.omc/refit/pending-proposals.md'), a copy-ready proposal format ('finding → surface → intended change' plus one line of instrument evidence), and a precise verification procedure for new checks (run clean, demonstrate failing on a deliberate violation, revert). Per the code-vs-instruction scoring note, the absence of code is not penalized because the guidance is directly executable.

5 / 5

Workflow Clarity

The sequence is explicit and well-gated: read instruments first ('The survey starts from OMC's instruments, not from a checklist'), present findings ranked, 'Stop for user approval' with per-line veto, write approved findings to their surfaces with a pre-write discipline check, then record where each landed. Validation checkpoints are strong — the 'must be proven to bite before it counts as landed: run it clean once, then demonstrate it failing on a deliberately introduced violation, then revert the violation' loop is a textbook validate-fix-retry feedback loop — and the headless variant re-anchors the same approval gate. No destructive or batch write occurs without an explicit validation step, so no cap applies.

5 / 5

Progressive Disclosure

The skill has no bundle files (no references/, scripts/, or assets/ exist), and the ~46-line body is under the 50-line threshold with well-organized, clearly headed sections (Evidence first, Four fix-owner surfaces, Proposal and landing, Headless refit, Output). Per the simple-skill scoring note, progressive disclosure scores 5 on well-organized sections alone when no external references are needed, and nothing here would benefit from being split out.

5 / 5

Total

19

/

20

Passed

Description

58%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, jargon-grounded, and clearly distinct, but it is missing an explicit 'when to use' trigger clause, which caps completeness and leaves trigger-term coverage thin. Adding a 'Use when...' clause with natural user phrasing would move two dimensions up a level.

Suggestions

Add an explicit trigger clause, e.g. 'Use when the user asks for a retro, environment audit, or wants recurring friction from past sessions fixed.'

Include common user-facing synonyms ('retro', 'post-mortem', 'environment audit', 'setup improvements') alongside the OMC-specific jargon so the description matches how users actually phrase the request.

Lead with 1-2 concrete action verbs (e.g. 'Surveys OMC instrumentation and converts recorded friction into environment changes') so the capability reads as actions rather than a mechanism description.

DimensionReasoningScore

Specificity

The description names the domain ('Cross-session environment retrospective driven by OMC's own instrumentation') and enumerates several concrete elements: specific inputs ('trace timelines, friction reports, logs, and plan notepads') and the four fix surfaces ('deterministic check, steering volume, tooling, or information access'). It is not a level-5 comprehensive action list because the operating verbs are abstract ('retrospective', 'lands on the surface') rather than a full inventory of concrete actions; it is above level 3 because multiple specific capabilities are explicitly stated, not just 1-2.

4 / 5

Completeness

The 'what' is clearly stated (survey OMC instrumentation and route each finding to the surface that owns its fix, with nothing written without user approval), but there is no 'Use when...' clause or equivalent explicit trigger guidance, which caps this dimension at 3 per the judging guidelines. It is not a 2 because the 'what' is concrete and detailed, not vague.

3 / 5

Trigger Term Quality

There are some relevant keywords ('retrospective', 'logs', 'friction reports'), but the description is heavy on environment-specific jargon ('trace timelines', 'plan notepads', 'surfaces') and misses the natural phrases a user would say, such as 'environment audit', 'retro', 'post-mortem', or 'improve the setup'. It matches the level-3 anchor (some relevant keywords but missing common variations or synonyms) and not level 4, which requires broad natural-term coverage.

3 / 5

Distinctiveness Conflict Risk

The description carves out a clear niche (cross-session environment retrospective over OMC instrumentation) that is unlikely to collide with unrelated skills. It falls short of the level-5 anchor ('clear niche with distinct triggers') because the distinct triggers themselves are absent — the description relies on jargon specificity rather than explicit trigger phrases, leaving minor overlap risk with closely related retrospective/review skills.

4 / 5

Total

14

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
Yeachan-Heo/oh-my-claudecode
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.