Content
50%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is extremely strong on concrete, executable guidance — real scripts with full argument examples, output contracts, and hard-won domain pitfalls — but it is bloated: the same routing and set-operation rules are repeated 2-4 times, textbook interpretation tables are restated, two referenced reference files do not exist, and the CRITICAL section and the Analysis Workflow give conflicting DESeq2-vs-pydeseq2 instructions.
Suggestions
Deduplicate into single authoritative sections: the R-DESeq2-vs-pydeseq2 guidance (currently in CRITICAL #2, Analysis conventions, Known Limitations, and the dispersion subsection), the Venn union-denominator rule (currently three places), and the "also DE"/"uniquely DE" set-operation rules (each currently stated twice).
Move background Claude already knows out of SKILL.md — the Interpretation Framework thresholds table, Core Principles, Domain Reasoning normalization explainer, and Common Patterns table — into an existing reference file or drop them, cutting the ~500-line body substantially.
Fix the broken progressive-disclosure links by adding references/design_formula_guide.md and references/r_clusterprofiler_guide.md (or removing the links), and reconcile the Analysis Workflow's "Step 3: Run PyDESeq2" with the CRITICAL rule "Use R DESeq2, not pydeseq2" into one coherent execution track (also fixing the duplicated "3." item in the CRITICAL list).
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The ~500-line body has substantial duplication and background Claude already knows: the R-vs-pydeseq2 guidance appears in four places ("Use R DESeq2, not pydeseq2", "Analysis conventions", "Known Limitations", "Dispersion estimates: R DESeq2 vs pydeseq2 diverge"), the union-denominator rule appears three times, "Also DE in strain X" and "uniquely DE" rules each appear twice, and sections like "Core Principles" ("Data-first", "Statistical rigor") and the "Interpretation Framework" table (padj<0.05 = significant, LFC>1 = 2x) restate textbook knowledge. This is noticeably below the mostly-efficient anchor-3; not anchor 1 because the majority of the material (script contracts, notebook-filter pitfalls) is genuinely non-obvious and earns its tokens. | 2 / 5 |
Actionability | Highly concrete: full CLI invocations with arguments for all six primary scripts (e.g., the r_deseq2_wrapper.py multi-contrast example), parseable output-line formats, R code for dispersion queries, tu run commands, and decision tables mapping question phrasing to output lines. Falls short of 5 because the executable guidance is internally conflicted — the "Analysis Workflow" Step 3 directs to "Run PyDESeq2" and pydeseq2_workflow.md while the CRITICAL section commands "Use R DESeq2, not pydeseq2" — leaving the reader unsure which executable path is sanctioned. | 4 / 5 |
Workflow Clarity | A 7-step workflow exists (parse question → load/validate → inspect metadata → run → filter → dispersion → enrichment), each step linked to a reference file, with an Error Quick Reference table for recovery. But checkpoints are implicit rather than explicit, the CRITICAL "read the executed notebook FIRST" precedence rule and the primary-scripts track are not integrated into that workflow (they form a second, conflicting track), the CRITICAL numbered list has two items numbered "3", and the pydeseq2-vs-R contradiction breaks the single clear sequence — matching anchor 3 rather than 4. | 3 / 5 |
Progressive Disclosure | References are one level deep and each link carries a one-line description, and all 9 scripts exist — but 2 of the 12 referenced files are missing from references/ (design_formula_guide.md, referenced at the metadata step and in the file list; r_clusterprofiler_guide.md), so those links are broken. Additionally, hundreds of lines of edge-case heuristics (the entire "LOOK UP DON'T GUESS" and "Analysis conventions" sections) are inlined in SKILL.md rather than split into reference files, which is the anchor-3 pattern of content that should be separate living inline. | 3 / 5 |
Total | 12 / 20 Passed |