CtrlK
BlogDocsLog inGet started
Tessl Logo

edge-strategy-reviewer

Critically review strategy drafts from edge-strategy-designer for edge plausibility, overfitting risk, sample size adequacy, and execution realism. Use when strategy_drafts/*.yaml exists and needs quality gate before pipeline export. Outputs PASS/REVISE/REJECT verdicts with confidence scores.

74

Quality

93%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

90%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, fully executable skill body that keeps its token budget tight and pushes detail into real, well-signaled reference files. The residual gaps are minor: no explicit post-run interpretation/retry guidance, and the bias-checklist asset is not surfaced in the Resources section.

Suggestions

Add `assets/bias_checklist.yaml` to the Resources section (e.g., "`assets/bias_checklist.yaml` — default bias checklist for --bias-checklist; supply a custom YAML path to override") so the bundled asset is navigable like the two references.

Append one step to the Workflow or a short "After Running" note describing how to act on the output — e.g., read revision_instructions for REVISE drafts, return them to edge-strategy-designer, and re-review — to close the feedback loop.

In Verdict Logic, note what happens on unreadable/malformed draft YAML (skip vs. fail the run) so batch failures have a defined recovery path.

DimensionReasoningScore

Conciseness

The body is lean and operational: a criteria table, terse verdict logic, copy-paste commands, and a compact output example. It explains nothing Claude already knows, and the one long comment block (the bias-checklist flag) conveys non-obvious behavioral rules (PASS→REVISE downgrade, export-eligibility loss) rather than padding, so every token earns its place.

5 / 5

Actionability

The "Running the Script" section gives five complete, executable bash invocations covering the common cases (directory review, single draft, JSON + markdown summary, strict export, bias checklist), each with real flag names and paths. The YAML output example is concrete enough to interpret results without guesswork.

5 / 5

Workflow Clarity

The six-step workflow is clearly sequenced and the verdict logic supplies explicit decision checkpoints ("C1 or C2 severity=fail → immediate REJECT", "confidence >= 70 → PASS"), plus a REVISE path with revision instructions. It stops short of the 5 anchor because guidance on what to do after the script runs (interpreting findings, re-running after fixes, handling malformed input) is left implicit.

4 / 5

Progressive Disclosure

Good structure overall: the detailed C1-C8 scoring rubric and overfitting heuristics are correctly split into `references/review_criteria.md` and `references/overfitting_checklist.md`, both real files, clearly signaled one level deep in the Resources section. It misses the 5 anchor because the bundled bias checklist (assets/bias_checklist.yaml, invoked via --bias-checklist) is referenced only as "the bundled checklist" and is not listed in Resources with its path, leaving a minor navigation gap.

4 / 5

Total

18

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong third-person description that pairs a concrete capability list with an explicit, file-based trigger and equally concrete outputs. Its only weakness is slightly thin synonym coverage in trigger terms.

DimensionReasoningScore

Specificity

The description enumerates concrete review dimensions ("edge plausibility, overfitting risk, sample size adequacy, and execution realism") and concrete outputs ("PASS/REVISE/REJECT verdicts with confidence scores"), giving comprehensive, non-generic coverage of what the skill does. It is not merely naming a domain; every clause states a specific action or output.

5 / 5

Completeness

It explicitly answers both questions: what ("Critically review strategy drafts... Outputs PASS/REVISE/REJECT verdicts with confidence scores") and when ("Use when strategy_drafts/*.yaml exists and needs quality gate before pipeline export"), with a concrete file-based trigger phrase. Neither half is vague or merely implied.

5 / 5

Trigger Term Quality

Good natural keywords for the intended context: "strategy drafts", "strategy_drafts/*.yaml", "quality gate", "pipeline export", and the verdict terms. It falls just short of the 5 anchor because a few natural variations users might say (e.g., "validate a strategy", "review strategies", "critique strategy drafts") are absent and coverage relies on the file-path trigger.

4 / 5

Distinctiveness Conflict Risk

It occupies a clear niche tied to a specific upstream skill ("strategy drafts from edge-strategy-designer") and a specific artifact pattern ("strategy_drafts/*.yaml"), so the risk of it triggering for an unrelated skill is minimal. Unlike the 4 anchor's "minor overlap with closely related skills", no plausible sibling skill shares these triggers.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
tradermonty/claude-trading-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.