Abstract
Genome-wide association studies (GWAS) have identified hundreds of associated loci across many common diseases. Most risk variants identified by GWAS will merely be tags for as-yet-unknown causal variants. It is therefore possible that identification of the causal variant, by fine mapping, will identify alleles with larger effects on genetic risk than those currently estimated from GWAS replication studies. We show that under plausible assumptions, whilst the majority of the per-allele relative risks (RR) estimated from GWAS data will be close to the true risk at the causal variant, some could be considerable underestimates. For example, for an estimated RR in the range 1.2-1.3, there is approximately a 38% chance that it exceeds 1.4 and a 10% chance that it is over 2. We show how these probabilities can vary depending on the true effects associated with low-frequency variants and on the minor allele frequency (MAF) of the most associated SNP. We investigate the consequences of the underestimation of effect sizes for predictions of an individual's disease risk and interpret our results for the design of fine mapping experiments. Although these effects mean that the amount of heritability explained by known GWAS loci is expected to be larger than current projections, this increase is likely to explain a relatively small amount of the so-called "missing" heritability.
MeSH Terms
Algorithms
Breast Neoplasms/genetics
Crohn Disease/genetics
Diabetes Mellitus, Type 2/genetics
Gene Frequency
Genetic Predisposition to Disease
Genome-Wide Association Study
Humans
Linkage Disequilibrium
Polymorphism, Single Nucleotide
Risk
Authors & Affiliations
4 authors, click to expand affiliations / ORCID
Spencer Chris
Wellcome Trust Centre for Human Genetics, University of Oxford, Oxford, UK.
[email protected]
Hechter Eliana
Vukcevic Damjan
Donnelly Peter
Conflict of Interest
The authors have declared that no competing interests exist.
References (26)
26 references, click to expand
-
The complex interplay among factors that influence allelic association.
Nat Rev Genet. 2004 Feb;5(2):89-100
PMID: 14735120
-
Linkage strategies for genetically complex traits. II. The power of affected relative pairs.
Am J Hum Genet. 1990 Feb;46(2):229-41
PMID: 2301393
-
A fine-scale map of recombination rates and hotspots across the human genome.
Science. 2005 Oct 14;310(5746):321-4
PMID: 16224025
-
Meta-analysis of genome-wide association data and large-scale replication identifies additional susceptibility loci for type 2 diabetes.
Nat Genet. 2008 May;40(5):638-45
PMID: 18372903
-
Family history and the risk of breast cancer: a systematic review and meta-analysis.
Int J Cancer. 1997 May 29;71(5):800-9
PMID: 9180149
-
Bayesian statistical methods for genetic association studies.
Nat Rev Genet. 2009 Oct;10(10):681-90
PMID: 19763151
-
Familial occurrence of inflammatory bowel disease.
N Engl J Med. 1991 Jan 10;324(2):84-8
PMID: 1984188
-
What can genome-wide association studies tell us about the genetics of common disease?
PLoS Genet. 2008 Feb;4(2):e33
PMID: 18454206
-
Genome-wide association defines more than 30 distinct susceptibility loci for Crohn's disease.
Nat Genet. 2008 Aug;40(8):955-62
PMID: 18587394
-
A haplotype map of the human genome.
Nature. 2005 Oct 27;437(7063):1299-320
PMID: 16255080
-
Polygenes, risk prediction, and targeted prevention of breast cancer.
N Engl J Med. 2008 Jun 26;358(26):2796-803
PMID: 18579814
-
A HapMap harvest of insights into the genetics of common disease.
J Clin Invest. 2008 May;118(5):1590-605
PMID: 18451988
-
The ENCODE (ENCyclopedia Of DNA Elements) Project.
Science. 2004 Oct 22;306(5696):636-40
PMID: 15499007
-
Prediction and interaction in complex disease genetics: experience in type 1 diabetes.
PLoS Genet. 2009 Jul;5(7):e1000540
PMID: 19584936
-
Risk of diabetes in siblings of index cases with Type 2 diabetes: implications for genetic studies.
Diabet Med. 2002 Jan;19(1):41-50
PMID: 11869302
-
Estimating the relative recurrence risk ratio using a global cross-ratio model.
Genet Epidemiol. 2003 Dec;25(4):293-302
PMID: 14639699
-
Risk ratio and rate ratio estimation in case-cohort designs: hypertension and cardiovascular mortality.
Stat Med. 1993 Sep 30;12(18):1733-45
PMID: 8248665
-
Bayes factors for genome-wide association studies: comparison with P-values.
Genet Epidemiol. 2009 Jan;33(1):79-86
PMID: 18642345
-
Are rare variants responsible for susceptibility to complex diseases?
Am J Hum Genet. 2001 Jul;69(1):124-37
PMID: 11404818
-
Genome-wide association studies: theoretical and practical concerns.
Nat Rev Genet. 2005 Feb;6(2):109-18
PMID: 15716907
-
Genome-wide association study of 14,000 cases of seven common diseases and 3,000 shared controls.
Nature. 2007 Jun 7;447(7145):661-78
PMID: 17554300
-
A second generation human haplotype map of over 3.1 million SNPs.
Nature. 2007 Oct 18;449(7164):851-61
PMID: 17943122
-
Personal genomes: The case of the missing heritability.
Nature. 2008 Nov 6;456(7218):18-21
PMID: 18987709
-
Overcoming the winner's curse: estimating penetrance parameters from case-control data.
Am J Hum Genet. 2007 Apr;80(4):605-15
PMID: 17357068
-
The International HapMap Project.
Nature. 2003 Dec 18;426(6968):789-96
PMID: 14685227
-
A new multipoint method for genome-wide association studies by imputation of genotypes.
Nat Genet. 2007 Jul;39(7):906-13
PMID: 17572673