Search PubMed⌕ Search

PubMed · 3755669

Design, development, and implementation of a data processing system for multiple controlled trials and epidemiologic studies.

Abstract

We were given the opportunity to design and implement a general data processing system to accommodate several different epidemiologic studies to be conducted by a new research group. A survey of 15 operating data centers was conducted in preparation for undertaking the design and development of our system. The results of the survey indicated that data processing activities can be classified, both conceptually and operationally, into three modules: data recording and data entry, data management, and data analysis, and that the data management functions were those amenable to generalization. Based on our survey and the varying needs of our studies, we selected a "mixed" hardware environment, using both a computer center mainframe and microcomputers. We created the systems using commercially available software, including a mainframe database manager and mainframe statistics packages, microcomputer data entry software, and a communications package to link the two environments. Our strategy was to buy software, when possible, rather than to build custom programs, and to let software tools govern hardware needs. Hardware independence, price, and functional capability directed our software choices, while hardware selection was constrained most importantly by available software, then by budget, by available computing resources, and finally by the marketplace. The system has been used successfully in three studies differing in design, size, data collection locale, and rate of data accrual.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

B S Hawkins, S W Singer. 1986. Design, development, and implementation of a data processing system for multiple controlled trials and epidemiologic studies.. https://doi.org/10.1016/0197-2456(86)90027-9

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Probability estimation when some observations are grouped.

This paper considers the use of additional questions for decreasing survey non-response rates and an approach for estimating a probability based on the results obtained. In a survey, the respondents are asked to answer an original question and follow-up questions, where the answers for the follow-up questions are grouped answers for the original question. For example, respondents are asked to provide an exact number of incidents, but in cases of 'Do not know' or 'Refuse' responses, they are subsequently asked to pick an answer from a less specific categorical scale. The new estimator obtains smaller variance asymptotically and does not depend on a distribution family. This method is applied to income questions in a survey regarding injury prevention and behaviours. Another application is survey data on intimate partner violence, where some amendments were applied for incorporating post-stratification weights and for using non-random grouping. For additional illustration, an example of parameter estimation on artificially generated data is presented.

Data Collection↗