CtrlK
BlogDocsLog inGet started
Tessl Logo

ralph

Self-referential loop until task completion with configurable verification reviewer

49

Quality

54%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/ralph/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

58%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an unusually rigorous, action-dense workflow with best-in-class validation gates and feedback loops, but it pays for that with heavy internal repetition and a monolithic structure that should be split into reference files. Moving the amendment-ledger spec, reconciliation config, and session caveats into a references/ directory and de-duplicating the feedback-gate rules would raise both conciseness and progressive disclosure without losing actionability.

Suggestions

De-duplicate the feedback-gate semantics: state the `omc ralph verify` baseline/exit-code rules once (the Precondition gate) and have Steps 1f, 4b, and 7.6 reference that statement instead of restating it; do the same for the Skill-vs-Agent invocation rule and the Step-7 anti-stop warning.

Extract the criterion-amendment ledger spec, the stale-state reconciliation JSON/check types, and the parallel-session caveats into a references/ directory (e.g. references/prd-amendments.md, references/reconciliation.md) and link them with one-level-deep pointers — this turns the monolith into a navigable overview.

Either bundle the cited docs (agent-tiers.md, company-context-interface.md, REFERENCE.md) or inline the minimal decision rules they contain, since currently a reader hits references that cannot be resolved.

DimensionReasoningScore

Conciseness

The ~340-line body restates the same rules multiple times: the feedback-gate/verify semantics appear in <Precondition_Feedback_Gate>, Step 1f, Step 4b, and Step 7.6; the "do not stop after Step 7 approval" anti-pattern is repeated in Step 7 and <Escalation_And_Stop_Conditions>; and the Skill-vs-Agent invocation rule is duplicated in Step 7.5 and <Tool_Usage>. Several padded, redundant sections match anchor 2 rather than the minor-trim anchor 4.

2 / 5

Actionability

Guidance is concrete and executable throughout: `omc ralph verify --write-baseline --session <sessionId>`, `Skill("oh-my-claudecode:ai-slop-cleaner")`, `Task(subagent_type="oh-my-claudecode:executor", model="haiku", ...)` with exit-code semantics, plus worked Good/Bad examples and a full amendment-ledger JSON shape. Not 5: cited materials like `docs/shared/agent-tiers.md`, `docs/company-context-interface.md`, and `docs/REFERENCE.md` are not part of the bundle, leaving gaps a reader cannot resolve.

4 / 5

Workflow Clarity

A clearly sequenced 10-step (plus 7.5/7.6) process with explicit validation at every checkpoint: a written feedback baseline, per-story verification via verify exit codes, reviewer sign-off gates, a mandatory post-deslop re-verification with rollback guidance, a rejection fix-and-re-verify loop, and a final checklist. This matches the anchor's validate → fix → retry feedback-loop standard exactly.

5 / 5

Progressive Disclosure

No bundle files exist (no references/, scripts/, or assets/ directories), and the SKILL.md is a ~30KB monolith: the reconciliation JSON schema, criterion-amendment ledger spec, parallel-session caveats, and background-execution rules are all inlined where separate reference files clearly belong, and the doc paths it does cite do not resolve in the bundle. This matches anchor 2 (content that clearly belongs in separate files is inlined) rather than 3, since there is effectively no working reference layer at all.

2 / 5

Total

13

/

20

Passed

Description

50%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is concise and names a distinctive concept, but it captures only a fragment of the skill's capability set and provides no trigger guidance for when to invoke it. The body contains excellent trigger phrases ("don't stop", "must complete", "keep going until done") that belong in the description.

Suggestions

Add a trigger clause to the description, e.g. "Use when the user says 'don't stop', 'keep going', 'must complete', or needs guaranteed verified completion" — this lifts completeness and trigger_term_quality simultaneously.

Enumerate 2-3 more concrete capabilities (PRD-driven story tracking, reviewer sign-off, post-review cleanup pass) instead of the abstract phrase "self-referential loop", which is jargon users would not say.

Include the skill's own name or a natural synonym ("ralph", persistence loop) so users and Claude can match invocations to the description.

DimensionReasoningScore

Specificity

"Self-referential loop until task completion" and "configurable verification reviewer" name the domain plus 1-2 concrete actions, but coverage is not comprehensive. It is above a bare domain label (2) yet does not list several specific actions such as PRD refinement or story tracking (4).

3 / 5

Completeness

The description gives a clear "what" (a persistence loop with reviewer verification) but contains no "Use when..." clause or equivalent explicit trigger guidance, which caps completeness at 3 per the judging guidelines. The body's <Use_When> triggers are not surfaced in the description itself.

3 / 5

Trigger Term Quality

Terms like "loop", "task completion", and "verification reviewer" are relevant, but the natural phrases a user would actually say for this skill ("keep going", "don't stop", "finish this", "persistence") appear only in the body, not the description. Missing common variations, so not 4; more than generic filler, so not 2.

3 / 5

Distinctiveness Conflict Risk

"Self-referential loop" carves a niche, but "task completion with verification" is generic enough to overlap with other completion/review skills, and no distinct trigger phrases are given. Somewhat specific but overlapping, matching anchor 3 rather than the mostly-distinct anchor 4.

3 / 5

Total

12

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
Yeachan-Heo/oh-my-claudecode
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.