>_ Evidence Strength Is Not Effect Size
Consider two hypothetical findings:
500,000 participants • Independent Replication
OR = 1.06 (Strong Evidence, Small Effect)180 participants • Single Pilot Study
OR = 2.10 (Weak Evidence, Large Effect)Finding B appears far more dramatic. But Finding A is vastly more trustworthy. A large effect from a small study is often an overestimation artifact (the Winner's Curse).
>_ The 7 Levels of Evidence Evaluation
Has the association been reproduced by independent research groups in new participant cohorts?
Does the study have adequate statistical power to detect small, polygenic common-variant effects?
Is the observed difference practically meaningful, beyond achieving P < 5 × 10⁻⁸?
OR 1.20 (95% CI 1.18–1.22) is precise; OR 1.20 (95% CI 0.88–1.65) spans neutral and is uncertain.
Was the trait objectively measured (e.g. polysomnography/wearables) or loosely self-reported?
Has the association been validated across diverse ancestries and reference populations?
Do experimental assays demonstrate altered gene expression, protein folding, or splicing?
>_ The GENOSTRIDE Evidence Scale
Important: GENOSTRIDE's "Strong Evidence" rating evaluates literature confidence for complex traits and nutrigenetics. It is NOT equivalent to an ACMG "Pathogenic" clinical classification.