Validate statistical research outputs for formulation quality, method-to- problem alignment, theory presence, experimental evidence, fair comparison, artifact completeness, and final-claim consistency.
57
66%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
Passed
No findings from the security scan
Fix and improve this skill with Tessl
tessl review fix ./external/agents/stat_research_agent/skills/stat-result-validator/SKILL.mdUse this skill after formulation, method proposal, theory, experimental evaluation, comparison, and result synthesis. It checks whether the final result is supported by a coherent statistical research chain.
Required for all topics:
progress/<TOPIC_ID>/step0_problem_formulation.md
progress/<TOPIC_ID>/step1_method_proposal.md
progress/<TOPIC_ID>/step2_theory_analysis.md
progress/<TOPIC_ID>/step3_experimental_evaluation.md
progress/<TOPIC_ID>/step4_comparison.md
progress/<TOPIC_ID>/step5_result_synthesis.md
progress/<TOPIC_ID>/step6_quality_audit.md
experiments/<TOPIC_ID>/config.yaml
experiments/<TOPIC_ID>/results/metrics.json
experiments/<TOPIC_ID>/results/run_manifest.json
experiments/<TOPIC_ID>/results/comparison_summary.md
experiments/<TOPIC_ID>/results/claim_verdicts.json
experiments/<TOPIC_ID>/report/paper.md
experiments/<TOPIC_ID>/README.mdAnalysis-specific source files and raw outputs are determined by the experiment
plan and should live under experiments/<TOPIC_ID>/src/ and
experiments/<TOPIC_ID>/results/.
The formulation must define:
Blocking failures:
Verify that:
Theory may be rigorous or partial, but it must be explicit.
Check for:
Blocking failures:
Check that:
run_manifest.json.Verify that:
Every final claim must be traceable to:
formulation -> method -> theory -> experiment -> comparisonUse:
PASS: formulation, theory, experiments, and comparisons support the claims.WARN: usable but has limitations that must be disclosed.FAIL: missing formulation, theory, evidence, or fair comparison prevents a
valid conclusion.e2e23c9
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.