Search PubMed⌕ Search

PubMed · 2917462

Statistical evaluation of agreement between two methods for measuring a quantitative variable.

Abstract

Methodologic research is often concerned with determining whether two methods (procedures, laboratory instruments) can be used interchangeably for measuring some quantitative variable of interest. Logically, one method can be used as a surrogate of another provided the methods show high agreement on the measured results. Although the product-moment correlation (r) is often used as an indicator of agreement, this index is in fact inappropriate for this purpose. The intraclass correlation (r1) is the correct statistic for assessing agreement or consistency between two methods. Another criterion sometimes used for supporting interchangeability is the similarity of the mean measured results obtained by the two methods. However, similarity of means (aggregate agreement) does not necessarily indicate individual-subject agreement, and it is the latter that is the pre-requisite for interchangeability. On the other hand, a marked difference between two means (lack of aggregate agreement) does necessarily indicate lack of individual-subject agreement and therefore non-interchangeability. Herein we suggest that two methods for measuring a quantitative variable can be judged interchangeable provided all of the following conditions are met: first the methods must not exhibit marked additive or nonadditive systematic bias; second the difference between the two mean readings is not "statistically significant"; third, the lower limit of the 95% confidence interval of the intraclass correlation is at least 0.75. Statistical procedures to evaluate these conditions of interchangeability are described in detail. A computer program coded in SAS to carry out the procedures is listed in the Appendix. A similar program coded in DBASE III PLUS for the microcomputer is available upon request.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

J Lee, D Koh, C N Ong. 1989. Statistical evaluation of agreement between two methods for measuring a quantitative variable.. https://doi.org/10.1016/0010-4825(89)90036-x

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Molecular heterochrony and the evolution of sociality in bumblebees (Bombus terrestris).

Sibling care is a hallmark of social insects, but its evolution remains challenging to explain at the molecular level. The hypothesis that sibling care evolved from ancestral maternal care in primitively eusocial insects has been elaborated to involve heterochronic changes in gene expression. This elaboration leads to the prediction that workers in these species will show patterns of gene expression more similar to foundress queens, who express maternal care behaviour, than to established queens engaged solely in reproductive behaviour. We tested this idea in bumblebees (Bombus terrestris) using a microarray platform with approximately 4500 genes. Unlike the wasp Polistes metricus, in which support for the above prediction has been obtained, we found that patterns of brain gene expression in foundress and queen bumblebees were more similar to each other than to workers. Comparisons of differentially expressed genes derived from this study and gene lists from microarray studies in Polistes and the honeybee Apis mellifera yielded a shared set of genes involved in the regulation of related social behaviours across independent eusocial lineages. Together, these results suggest that multiple independent evolutions of eusociality in the insects might have involved different evolutionary routes, but nevertheless involved some similarities at the molecular level.

Analysis of Variance↗

Confidence intervals for the standardized effect arising in the comparison of two normal populations.

Confidence intervals for a standardized effect are derived after stabilizing the variance of the Welch t-statistic. Simulation studies demonstrate the viability of the resulting intervals for a wide range of parameter values and sample sizes as small as five. The methodology is extended to the combination of results from several studies, so as to obtain a confidence interval for a representative standardized effect for all the studies. The methods are illustrated on a recent meta-analytic study of systolic blood pressure reduction during a weight reducing regime, as well as the classical Mumford data on psychological intervention and hospital length of stay.

Analysis of Variance↗