Research-brief + claim-record discipline for delegated or AI-assisted research work. Use when starting any substantive research task or long autonomous run (write the brief first), and when reporting results that will support a claim in a paper or decision (produce the claim record). Keeps faster execution from being confused with credible evidence.
Cheaper generation increases demand for scarce validation. The fix is two lightweight artifacts: a research brief that constrains what the system may do, and a claim record that constrains what may be reported. Applies to agents, workflows, sims, proof patches, data construction — anything delegated.
One short block, written before launching the work:
For each claim the work will support:
Proportionality: a routine task needs one short paragraph; heavier records only when branching is extensive, outputs will be reused, errors are consequential, or correction is costly.
The final decision is never "all checks green" — it is whether the evidence supports the proposed language. The options are: repair; narrower language; additional review; an exploratory/descriptive label; or decline to report. Never upgrade language beyond the evidence (associational ≠ causal; pointwise ≠ uniform; illustrated ≠ validated; imposed ≠ derived).
Reproducibility, implementation correctness, statistical performance, measurement validity, and identification/scope are different questions; evidence on one cannot answer another. Reproducible code may implement the wrong estimator; favorable simulations cannot establish an assumption; a correct estimate may answer the wrong question.
Every diagnosed failure becomes a durable artifact: a reusable test, a documented warning, or a memory entry stating the failure, why it happened, and the check that now prevents it. Preserve failed approaches and the reason they failed — a blocked route re-attempted without a new mechanism is waste.
verification-ladder.md — rung 5 (the ledger).claude/rules/orchestrator-protocol.md — RUN_CONFIG is the brief for a fan-out9d371f0
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.