Content
82%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-engineered skill body: concise, accurately executable commands whose flags match the real script CLIs, a clearly sequenced workflow with a decision gate and iteration feedback loop, and a sensible split between the overview and real one-level-deep reference files. Remaining improvements are small: replace wildcard path placeholders with concrete examples, add an output-validation step (schema check or bundled tests), and annotate the reference files with when-to-read guidance.
Suggestions
Use concrete file paths instead of wildcard placeholders in the generate_pivots example (the script's --diagnosis takes a single path, so a shell glob would silently expand to the wrong file if multiple diagnoses exist).
Add a validation checkpoint after pivot generation, e.g. verifying output YAML against references/pivot_proposal_schema.md or running the bundled scripts/tests, to strengthen the workflow's feedback loop.
Annotate the Resources entries (or add inline signposts in the Workflow section) so each reference file states when to consult it, moving navigation from adequate to fully signaled.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is lean and assumes competence: every section (Overview, When to Use, Prerequisites, Output, Workflow, Quick Commands, Resources) is task-specific with zero filler, no explanation of concepts Claude already knows, and no padding. It matches the 'every token earns its place' anchor; nothing needed trimming to the level of 4. | 5 / 5 |
Actionability | The three Quick Commands are concrete, copy-paste-ready bash invocations whose flags match the actual argparse interfaces of detect_stagnation.py and generate_pivots.py (verified). Minor gaps keep it below 5: the --diagnosis and --strategy arguments use wildcard placeholders (reports/pivot_diagnosis_*.json) that must be filled in — and the script accepts a single diagnosis path, so a shell glob would silently pass only the last match — and there is no example of interpreting the diagnosis or ranking output. | 4 / 5 |
Workflow Clarity | The seven-step workflow is clearly sequenced, contains an explicit decision gate ("If stagnation detected, generate pivot proposals"), and closes a feedback loop by feeding selected pivots back into backtest-expert. It fits 'clear sequence with most checkpoints present; minor validation gaps': no step validates generated pivot YAML against the pivot proposal schema or mentions the bundled tests, and the file-generation steps produce batch outputs without a verify step. It is not capped at 3 because the operations are additive file generation, not destructive or destructive-batch changes. | 4 / 5 |
Progressive Disclosure | The SKILL.md is a clean overview with details correctly split into four real, one-level-deep reference files (stagnation_triggers, strategy_archetypes, pivot_techniques, pivot_proposal_schema — all verified present with substantive content) and two scripts. It sits at 4 rather than 5 because the Resources section lists reference paths bare, without one-line annotations or in-text signposts (e.g., workflow step 3 names the three pivot techniques but never says 'see references/pivot_techniques.md'), so navigation is good but not fully signaled. | 4 / 5 |
Total | 17 / 20 Passed |