Search PubMed⌕ Search

PubMed · 11788962

Truncated product method for combining P-values.

Abstract

We present a new procedure for combining P-values from a set of L hypothesis tests. Our procedure is to take the product of only those P-values less than some specified cut-off value and to evaluate the probability of such a product, or a smaller value, under the overall hypothesis that all L hypotheses are true. We give an explicit formulation for this P-value, and find by simulation that it can provide high power for detecting departures from the overall hypothesis. We extend the procedure to situations when tests are not independent. We present both real and simulated examples where the method is especially useful. These include exploratory analyses when L is large, such as genome-wide scans for marker-trait associations and meta-analytic applications that combine information from published studies, with potential for dealing with the "publication bias" phenomenon. Once the overall hypothesis is rejected, an adjustment procedure with strong family-wise error protection is available for smaller subsets of hypotheses, down to the individual tests.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

D V Zaykin, Lev A Zhivotovsky, P H Westfall, B S Weir. 2002. Truncated product method for combining P-values.. https://doi.org/10.1002/gepi.0042

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Statistical test to compare the linkage model and the admixture model based on central limit results.

In the Admixture Model, the probability that an individual carries a certain allele at a specific marker depends on the allele frequencies in K ancestral populations and the proportion of the individual's genome originating from these populations. The markers are assumed to be independent. The Linkage Model is a Hidden Markov Model that extends the Admixture Model by incorporating linkage between neighboring loci. We prove consistency and asymptotic normality of maximum likelihood estimators for the ancestry of individuals in the Linkage Model, complementing earlier results by (Pfaff et al., 2004; Pfaffelhuber and Rohde, 2022; Heinzel, 2025) for the Admixture Model. These results are used to prove that a statistical test that allows for model selection between the Admixture Model and the Linkage Model is an asymptotic level-α-test. Finally, we demonstrate the practical relevance of our results by applying the test to real-world data from The 1000 Genomes Project Consortium (2015).

Genetic Linkage↗

Haplotype thinking in lung disease.

To identify the genetic etiology of a disease of interest, disease-related characteristics (phenotypes) are often tested for association with genetic variants (genotypes). Although genetic association studies of single genetic variants have been widely performed, there has been increasing interest in studies of multiple adjacent genetic variants on one chromosome, known as a haplotype. In this review, we will provide background about the origin of haplotypes and why they can be useful in genetic studies; we will discuss approaches to determining haplotypes and performing haplotype-based genetic association studies; and we will compare single variant and haplotype-based approaches.

Genetic Linkage↗

Associations between DNA markers and resistance to diseases in sugarcane and effects of population substructure.

Association between markers and sugarcane diseases were investigated in a collection of 154 sugarcane clones, consisting of important ancestors or parents, and cultivars. 1,068 polymorphic AFLP and 141 SRR markers were scored across all clones. Data on the four most important diseases in the Australian sugarcane industry were obtained; these diseases being pachymetra root rot (Pachymetra chaunorhiza B.J. Croft & M.W. Dick), leaf scald (Xanthomonas albilineans Dowson), Fiji leaf gall (Fiji disease virus), and smut (Ustilago scitaminea H. & P. Sydow). By a simple regression analysis, association between markers and diseases could be readily detected. However, many of these associations were due to the effects of embedded population structure and random effects. After taking population structure into account, we found that 59% of the phenotypic variation in smut resistance ratings could be accounted for by 11 markers, 32% of variation for leaf scald and pachymetra root rot rating by 4 markers, and 26% of Fiji leaf gall by 5 markers. The results suggest that marker-trait associations can be readily detected in populations generated from modern sugarcane breeding programs. This may be due to special features of past sugarcane breeding programs leading to persistent linkage disequilibrium in modern parental populations.

Genetic Linkage↗