CtrlK
BlogDocsLog inGet started
Tessl Logo

three-way-judge

Run an explicitly requested three-option workflow with three independent executors, one independent paired critic per option, bounded revision loops, and a blind evidence-backed judge. Use only when the user invokes $three-way-judge or explicitly requests this exact panel; this v1 skill does not run the nine-review deep mode.

70

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is an efficient, actionable protocol with strong progressive disclosure and explicit validation checkpoints. The main gap is that the end-to-end execution sequence is interleaved across multiple prose sections rather than consolidated into a single ordered checklist.

Suggestions

Consolidate the full execution sequence into one numbered checklist (dispatch lanes → critic gates → blind packet → judge → root verification), keeping the constraint sections as supporting rules.

Dedupe the independence/fingerprinting rules stated in both "Non-negotiable rules" and "Runtime behavior" so each appears once.

Add a short "Outputs" checklist mapping each returned artifact to the validation command that gates it, making the verify-then-report order explicit.

DimensionReasoningScore

Conciseness

The body is lean and assumes competence, but independence and fingerprinting concerns recur across "Non-negotiable rules" and "Runtime behavior", producing minor redundancy that could be tightened.

4 / 5

Actionability

Concrete executable commands with specific script paths and arguments (verify-evidence, validate-scorecard, validate-panel, build-blind-packet, verify-review-run) are provided, with only placeholder path tokens left to fill.

4 / 5

Workflow Clarity

Explicit validation gates and feedback loops (critic PASS, root verification, bounded improvement sweep, stalled-agent recovery) are present, but the execution sequence is distributed across prose rather than one ordered checklist.

4 / 5

Progressive Disclosure

A concise overview points via clearly signaled links to seven one-level-deep reference files and five scripts, all verified to exist, with detail appropriately split out of the main file.

5 / 5

Total

17

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, complete, and clearly gated to an explicit invocation, making it highly distinguishable. Trigger guidance is explicit though somewhat narrow in natural-synonym coverage.

DimensionReasoningScore

Specificity

Lists multiple concrete components — "three independent executors", "one independent paired critic per option", "bounded revision loops", "blind evidence-backed judge" — giving comprehensive coverage of what the workflow does.

5 / 5

Completeness

Clearly answers both "what" (three-option panel with executors/critics/judge) and "when" (explicit invocation trigger), with a concrete trigger phrase.

5 / 5

Trigger Term Quality

The explicit trigger "Use only when the user invokes $three-way-judge or explicitly requests this exact panel" is good, but natural-synonym coverage beyond the named command is limited.

4 / 5

Distinctiveness Conflict Risk

Heavily gated to a named invocation with an explicit boundary ("this v1 skill does not run the nine-review deep mode"), giving a clear niche with minimal conflict risk.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
stevenknowswhy/three-critic-judge-skill
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.