Abstract
With the recent advances in high-throughput genotyping techniques, it is now possible to perform whole-genome association studies to fine map causal polymorphisms underlying important traits that influence susceptibility to human diseases and efficacy of drugs. Once a genome scan is completed the results can be sorted by the association statistic value. What is the probability that true positives will be encountered among the first most associated markers? When a particular polymorphism is found associated with the trait, there is a chance that it represents either a "true" or a "false" association (TA vs. FA). Setting appropriate significance thresholds has been considered to provide assurance of sufficient odds that the associations found to be significant are genuine. However, the problem with genome scans involving thousands of markers is that the statistic values of FAs can reach quite extreme magnitudes. In such situations, the distributions corresponding to TAs and the most extreme FAs become comparable and significance thresholds tend to penalize TAs and FAs in a similar fashion. When sorting between true and false associations, the "typical" place (i.e., rank) of TAs among the most significant outcomes becomes important, ordered by the association statistic value. The distribution of ranks that we study here allows calculation of several useful quantities. In particular, it gives the number of most significant markers needed for a follow-up study to guarantee that a true association is included with certain probability. This can be calculated conditionally on having applied a multiple-testing correction. Effects of multilocus (e.g., haplotype association) tests and impact of linkage disequilibrium on the distribution of ranks associated with TAs are evaluated and can be taken into account.
MeSH Terms
Computer Simulation
Genetic Diseases, Inborn/genetics
Genetic Markers/genetics
Genetic Predisposition to Disease
Genomics/methods
Haplotypes/genetics
Linkage Disequilibrium
Models, Genetic
Polymorphism, Genetic
Research Design
Chemicals
Genetic Markers
Authors & Affiliations
2 authors, click to expand affiliations / ORCID
Zaykin Dmitri V
National Institute of Environmental Health Sciences, National Institutes of Health, Research Triangle Park, NC 27709, USA. zaykind@niehs.nih.gov
Zhivotovsky Lev A
References (27)
27 references, click to expand
-
Significance levels in complex inheritance.
Am J Hum Genet. 1998 Mar;62(3):690-7
PMID: 9497238
-
Functional SNPs in the lymphotoxin-alpha gene that are associated with susceptibility to myocardial infarction.
Nat Genet. 2002 Dec;32(4):650-4
PMID: 12426569
-
Mapping mendelian factors underlying quantitative traits using RFLP linkage maps.
Genetics. 1989 Jan;121(1):185-99
PMID: 2563713
-
Bounds and normalization of the composite linkage disequilibrium coefficient.
Genet Epidemiol. 2004 Nov;27(3):252-7
PMID: 15389931
-
Is peak height sufficient?
Genet Epidemiol. 2001 May;20(4):403-8
PMID: 11319781
-
True and false positive peaks in genomewide scans: The long and the short of it.
Genet Epidemiol. 2001 May;20(4):409-14
PMID: 11319782
-
Power and efficiency of the TDT and case-control design for association scans.
Behav Genet. 2002 Mar;32(2):135-44
PMID: 12036111
-
Large-scale association study identifies ICAM gene region as breast and prostate cancer susceptibility locus.
Cancer Res. 2004 Dec 15;64(24):8906-10
PMID: 15604251
-
Effect of two- and three-locus linkage disequilibrium on the power to detect marker/phenotype associations.
Genetics. 2004 Oct;168(2):1029-40
PMID: 15514073
-
Two-stage designs for gene-disease association studies.
Biometrics. 2002 Mar;58(1):163-70
PMID: 11890312
-
Genome scans and candidate gene approaches in the study of common diseases and variable drug responses.
Trends Genet. 2003 Nov;19(11):615-22
PMID: 14585613
-
The Interaction of Selection and Linkage. I. General Considerations; Heterotic Models.
Genetics. 1964 Jan;49(1):49-67
PMID: 17248194
-
Using the false discovery rate approach in the genetic dissection of complex traits: a response to Weller et al.
Genetics. 2000 Apr;154(4):1917-8
PMID: 10950641
-
The replication requirement.
Nat Genet. 2001 Nov;29(3):244-5
PMID: 11687787
-
True and false positive peaks in genomewide scans: applications of length-biased sampling to linkage mapping.
Am J Hum Genet. 1997 Aug;61(2):430-8
PMID: 9311749
-
The future of genetic studies of complex human diseases.
Science. 1996 Sep 13;273(5281):1516-7
PMID: 8801636
-
Standardizing a composite measure of linkage disequilibrium.
Ann Hum Genet. 2004 May;68(Pt 3):234-9
PMID: 15180703
-
Testing association of statistically inferred haplotypes with discrete and continuous traits in samples of unrelated individuals.
Hum Hered. 2002;53(2):79-91
PMID: 12037407
-
On the advantage of haplotype analysis in the presence of multiple disease susceptibility alleles.
Genet Epidemiol. 2002 Oct;23(3):221-33
PMID: 12384975
-
Haplotype blocks and linkage disequilibrium in the human genome.
Nat Rev Genet. 2003 Aug;4(8):587-97
PMID: 12897771
-
Interval estimation of genetic susceptibility for retrospective case-control studies.
BMC Genet. 2004 May 11;5:9
PMID: 15137913
-
Truncated product method for combining P-values.
Genet Epidemiol. 2002 Feb;22(2):170-85
PMID: 11788962
-
Meta-analysis of genetic association studies supports a contribution of common variants to susceptibility to common disease.
Nat Genet. 2003 Feb;33(2):177-82
PMID: 12524541
-
Replication validity of genetic association studies.
Nat Genet. 2001 Nov;29(3):306-9
PMID: 11600885
-
Genetic dissection of complex traits: guidelines for interpreting and reporting linkage results.
Nat Genet. 1995 Nov;11(3):241-7
PMID: 7581446
-
Linkage disequilibrium mapping of complex disease: fantasy or reality?
Curr Opin Biotechnol. 1998 Dec;9(6):578-94
PMID: 9889136
-
Selection of genetic markers for association analyses, using linkage disequilibrium and haplotypes.
Am J Hum Genet. 2003 Jul;73(1):115-30
PMID: 12796855