Search PubMed⌕ Search

PubMed · 9333340

Regression estimator in ranked set sampling.

Abstract

Ranked set sampling (RSS) utilizes inexpensive auxiliary information about the ranking of the units in a sample to provide a more precise estimator of the population mean of the variable of interest Y, which is either difficult or expensive to measure. However, the ranking may not be perfect in most situations. In this paper, we assume that the ranking is done on the basis of a concomitant variable X. Regression-type RSS estimators of the population mean of Y will be proposed by utilizing this concomitant variable X in both the ranking process of the units and the estimation process when the population mean of X is known. When X has unknown mean, double sampling will be used to obtain an estimate for the population mean of X. It is found that when X and Y jointly follow a bivariate normal distribution, our proposed RSS regression estimator is more efficient than RSS and simple random sampling (SRS) naive estimators unless the correlation between X and Y is low (/rho/ < 0.4). Moreover, it is always superior to the regression estimator under SRS for all rho. When normality does not hold, this approach could still perform reasonably well as long as the shape of the distribution of the concomitant variable X is only slightly departed from symmetry. For heavily skewed distributions, a remedial measure will be suggested. An example of estimating the mean plutonium concentration in surface soil on the Nevada Test Site, Nevada, U.S.A., will be considered.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

P L Yu, K Lam. 1997. Regression estimator in ranked set sampling.. https://pubmed.ncbi.nlm.nih.gov/9333340/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Tutorial in biostatistics: competing risks and multi-state models.

Standard survival data measure the time span from some time origin until the occurrence of one type of event. If several types of events occur, a model describing progression to each of these competing risks is needed. Multi-state models generalize competing risks models by also describing transitions to intermediate events. Methods to analyze such models have been developed over the last two decades. Fortunately, most of the analyzes can be performed within the standard statistical packages, but may require some extra effort with respect to data preparation and programming. This tutorial aims to review statistical methods for the analysis of competing risks and multi-state models. Although some conceptual issues are covered, the emphasis is on practical issues like data preparation, estimation of the effect of covariates, and estimation of cumulative incidence functions and state and transition probabilities. Examples of analysis with standard software are shown.

Biometry↗

The role of education in biostatistical consulting.

Medical students, residents, postdoctoral fellows, and faculty commonly consult with biostatistical experts about study design and data analysis when conducting clinical research. The role of biostatistical training during these consultations is examined, and characterizations of the connections between biostatistical consultation and education are reviewed. The presence and kinds of teaching efforts during biostatistical consults at four academic research institutions over various periods of time between 1999 and 2005 (237 consultations in total) were recorded and are described. By site, 67, 70, 78, and 100 per cent of the consulting sessions included biostatistical training, with an overall 78 per cent (95 per cent CI: 73-83 per cent) of consultations including an educational component when all consultations were combined. Training covered a wide range of biostatistical topics. Seventy-five per cent of the consultations with faculty (120/161), 79 per cent with fellows and residents (31/39), and 100 per cent with medical students (10/10) included some degree of instruction in study design or statistical analysis topics. Results show that both the need and the opportunity exist for specialized biostatistical instruction during one-on-one sessions between a consulting biostatistician and physicians, medical students, and research staff. Academic researchers are ideally positioned to absorb this kind of training when they initiate a request for assistance with their own research project.

Biometry↗

Improving the quality of patient care using reliability measures: a classification tree approach.

This paper considers the application and interpretation of new reliability measures for a classification tree-based medical risk assessment tool. Following the construction of a classification tree reliability measures may then be used to provide an estimate of the precision of the classification and the probability in each terminal node of the classification tree. Identification of unreliable nodes (those that have low precision) in this application may indicate patient groups requiring closer monitoring or scenarios in which further information about the patient is required, thereby providing medical practitioners with an avenue for more informed decision making.

Biometry↗