Search PubMed⌕ Search

PubMed · 12148835

Complex sampling designs and statistical issues in secondary analysis.

Abstract

Conducting secondary analysis using large national survey data sets to answer pressing research questions is gaining acceptance in the nursing science community. There are, however, challenges confronted by researchers who wish to apply secondary analysis to large data sets due to the incorporation of complex sampling designs. This article presents sampling design issues inherent in many large national surveys and explains the rationale for applying sample and variance estimation weights when conducting statistical analyses. In addition, the rationale for using statistical software packages capable of analyzing data derived from complex sampling designs is described. Examples of differences in statistical outcomes with and without weights using Stata and SPSS are provided using data from the Medical Expenditure Panel Survey (MEPS). Based on the example analyses, the implications of the statistical outcome differences for study findings are discussed.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Shawn M Kneipp, Hossein N Yarandi. 2002. Complex sampling designs and statistical issues in secondary analysis.. https://doi.org/10.1177/019394502400446414

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Probability estimation when some observations are grouped.

This paper considers the use of additional questions for decreasing survey non-response rates and an approach for estimating a probability based on the results obtained. In a survey, the respondents are asked to answer an original question and follow-up questions, where the answers for the follow-up questions are grouped answers for the original question. For example, respondents are asked to provide an exact number of incidents, but in cases of 'Do not know' or 'Refuse' responses, they are subsequently asked to pick an answer from a less specific categorical scale. The new estimator obtains smaller variance asymptotically and does not depend on a distribution family. This method is applied to income questions in a survey regarding injury prevention and behaviours. Another application is survey data on intimate partner violence, where some amendments were applied for incorporating post-stratification weights and for using non-random grouping. For additional illustration, an example of parameter estimation on artificially generated data is presented.

Data Collection↗