Search PubMed⌕ Search

PubMed · 10984467

Exploring performance issues for a clinical database organized using an entity-attribute-value representation.

Abstract

BACKGROUND: The entity-attribute-value representation with classes and relationships (EAV/CR) provides a flexible and simple database schema to store heterogeneous biomedical data. In certain circumstances, however, the EAV/CR model is known to retrieve data less efficiently than conventionally based database schemas. OBJECTIVE: To perform a pilot study that systematically quantifies performance differences for database queries directed at real-world microbiology data modeled with EAV/CR and conventional representations, and to explore the relative merits of different EAV/CR query implementation strategies. METHODS: Clinical microbiology data obtained over a ten-year period were stored using both database models. Query execution times were compared for four clinically oriented attribute-centered and entity-centered queries operating under varying conditions of database size and system memory. The performance characteristics of three different EAV/CR query strategies were also examined. RESULTS: Performance was similar for entity-centered queries in the two database models. Performance in the EAV/CR model was approximately three to five times less efficient than its conventional counterpart for attribute-centered queries. The differences in query efficiency became slightly greater as database size increased, although they were reduced with the addition of system memory. The authors found that EAV/CR queries formulated using multiple, simple SQL statements executed in batch were more efficient than single, large SQL statements. CONCLUSION: This paper describes a pilot project to explore issues in and compare query performance for EAV/CR and conventional database representations. Although attribute-centered queries were less efficient in the EAV/CR model, these inefficiencies may be addressable, at least in part, by the use of more powerful hardware or more memory, or both.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

R S Chen, P Nadkarni, L Marenco, F Levin, J Erdos, P L Miller. Exploring performance issues for a clinical database organized using an entity-attribute-value representation.. https://doi.org/10.1136/jamia.2000.0070475

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

A randomised controlled trial of a patient based Diabetes Recall and Management System: the DREAM trial: a study protocol [ISRCTN32042030].

BACKGROUND: Whilst there is broad agreement on what constitutes high quality health care for people with diabetes, there is little consensus on the most efficient way of delivering it. Structured recall systems can improve the quality of care but the systems evaluated to date have been of limited sophistication and the evaluations have been carried out in small numbers of relatively unrepresentative settings. Hartlepool, Easington and Stockton currently operate a computerised diabetes register which has to date produced improvements in the quality of care but performance has now plateaued leaving substantial scope for further improvement. This study will evaluate the effectiveness and efficiency of an area wide 'extended' system incorporating a full structured recall and management system, actively involving patients and including clinical management prompts to primary care clinicians based on locally-adapted evidence based guidelines. METHODS: The study design is a two-armed cluster randomised controlled trial of 61 practices incorporating evaluations of the effectiveness of the system, its economic impact and its impact on patient wellbeing and functioning.

Database Management Systems↗

IDR: the ImmunoDeficiency Resource.

The ImmunoDeficiency Resource (IDR), freely available at http://www.uta.fi/imt/bioinfo/idr/, is a comprehensive knowledge base on immunodeficiencies. It is designed for different user groups such as researchers, physicians and nurses as well as patients and their families and the general public. Information on immunodeficiencies is stored as fact files, which are disease- and gene-based information resources. We have developed an inherited disease markup language (IDML) data model, which is designed for storing disease- and gene-specific data in extensible markup language (XML) format. The fact files written by the IDML can be used to present data in different contexts and platforms. All the information in the IDR is validated by expert curators.

Database Management Systems↗

The EcoCyc Database.

EcoCyc is an organism-specific pathway/genome database that describes the metabolic and signal-transduction pathways of Escherichia coli, its enzymes, its transport proteins and its mechanisms of transcriptional control of gene expression. EcoCyc is queried using the Pathway Tools graphical user interface, which provides a wide variety of query operations and visualization tools. EcoCyc is available at http://ecocyc.org/.

Database Management Systems↗