CtrlK
BlogDocsLog inGet started
Tessl Logo

ce-compound-refresh

Refresh the repo's captured learnings against the current codebase. Use when auditing stale, overlapping, superseded, or drifted learnings; avoid general refactor, debugging, or code review unless the learnings store is explicit.

66

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered orchestrator body: a crisp sequenced pipeline with genuine error-recovery fallbacks, exact path-resolution and confirmation procedures, and exemplary one-level-deep reference structure. The costs are a compressed prose style with occasional rationale padding and validation details that are named in the body but defined only in the reference layer.

Suggestions

Trim the 'compounds value' rationale sentences (Worth lens and Discoverability Check sections) — they justify the skill's purpose rather than instruct the agent, and the body is otherwise pure rule-per-token.

Inline the two or three most safety-critical validation pre-checks (the auto-delete pre-checks from classify.md) as a short checklist in the Classify section, so the destructive-cap logic is verifiable from the body alone.

Rewrite the Worth lens intent-detection paragraph as a rule list (trigger words → action) matching the style of the other sections; its current nested conditionals cost re-reading.

DimensionReasoningScore

Conciseness

The body is dense and rule-per-sentence with no beginner-concept padding — e.g. "Candidates are the .md files under `<root>/solutions/`, excluding `README.md` and anything under `_archived/`" — assuming Claude's competence throughout. It sits below anchor 5 because of trimmable rhetorical rationale ("The store only compounds value if every doc can be trusted" and the similar closing flourish in Discoverability) and a somewhat convoluted Worth-lens paragraph that re-explains the confirm-before-investigate rule.

4 / 5

Actionability

Concrete, executable gating throughout: exact config-resolution steps ("Read `docs_root` from `<repo-root>/.compound-engineering/config.yaml` only (`<repo-root>` = `git rev-parse --show-toplevel`)"), a copy-paste numbered-options question block, and a verbatim subagent clause. It stops short of anchor 5 because the operational substance of most steps (classification criteria, per-doc fix procedures) lives one level deeper in the references — the body is an orchestrator that gates on reading them rather than containing the executable detail itself.

4 / 5

Workflow Clarity

A clearly sequenced pipeline (Mode → Worth lens → Artifact Root → Scope → Investigate → Classify → Execute → Vocabulary Capture → Report → Commit → Discoverability) with explicit error-recovery loops in the body: "A failed write is recorded as **recommended**, and the run continues", the git-failure fallback, and the blocking-question-or-numbered-options fallback. For a batch/destructive skill, anchor 5 requires explicit validation steps, but the body only names the validation machinery ("the auto-delete rule and its pre-checks", "unverifiable-is-not-false") and delegates its content to classify.md — a minor validation gap relative to anchor 5, comfortably above anchor 3.

4 / 5

Progressive Disclosure

Model progressive disclosure: an ~80-line overview that splits all detail across ten reference files, each verified to exist, each clearly signaled at the exact workflow point it is needed ("Read `references/classify.md` before assigning any of them"), with scripts and the assets template pushed out of the overview. Reference cross-links (e.g. per-action-flows.md → classify.md) target files also reachable directly from the body, keeping navigation one level deep and easy — matching anchor 5.

5 / 5

Total

17

/

20

Passed

Description

82%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete what-and-when with explicit trigger phrases and an unusually effective negative boundary that prevents misfires against similar-sounding refactor/review requests. The only notable gap is that the range of maintenance actions performed (consolidate, replace, delete, commit) is not surfaced in the description itself.

DimensionReasoningScore

Specificity

The description names the domain ("captured learnings" against the current codebase) and one core action ("Refresh") plus four concrete audit conditions ("stale, overlapping, superseded, or drifted"), but does not list the several distinct actions the skill actually performs (update, consolidate, replace, delete, report, commit). This matches anchor 3 (domain and 1-2 concrete actions, not comprehensive); it is below anchor 4 because the condition list enumerates what is checked, not multiple specific actions taken.

3 / 5

Completeness

It explicitly answers both parts: what ("Refresh the repo's captured learnings against the current codebase") and when ("Use when auditing stale, overlapping, superseded, or drifted learnings") with concrete trigger phrases, mirroring the anchor-5 good example structure. It is not anchor 4 because the 'when' clause is already explicit and trigger-phrase-specific rather than needing more precision.

5 / 5

Trigger Term Quality

Natural trigger terms are well covered — "auditing", "stale", "overlapping", "superseded", "drifted", "learnings" — and the negative boundary ("avoid general refactor, debugging, or code review") sharpens matching. It falls short of anchor 5 because common user phrasings like "outdated", "prune", or "clean up the solutions docs" are absent (clean-up intent is handled only inside the body).

4 / 5

Distinctiveness Conflict Risk

A clear niche (maintaining a captured-learnings store) with distinct triggers, plus an explicit de-confliction clause ("avoid general refactor, debugging, or code review unless the learnings store is explicit") that fences off its nearest competing skills. Minimal conflict risk, matching anchor 5.

5 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
EveryInc/compound-engineering-plugin
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.