Search PubMed⌕ Search

PubMed · 10641964

Representing the UMLS as an object-oriented database: modeling issues and advantages.

Abstract

OBJECTIVE: The Unified Medical Language System (UMLS) combines many well-established authoritative medical informatics terminologies in one knowledge representation system. Such a resource is very valuable to the health care community and industry. However, the UMLS is very large and complex and poses serious comprehension problems for users and maintenance personnel. The authors present a representation to support the user's comprehension and navigation of the UMLS. DESIGN: An object-oriented database (OODB) representation is used to represent the two major components of the UMLS-the Metathesaurus and the Semantic Network-as a unified system. The semantic types of the Semantic Network are modeled as semantic type classes. Intersection classes are defined to model concepts of multiple semantic types, which are removed from the semantic type classes. RESULTS: The authors provide examples of how the intersection classes help expose omissions of concepts, highlight errors of semantic type classification, and uncover ambiguities of concepts in the UMLS. The resulting UMLS OODB schema is deeper and more refined than the Semantic Network, since intersection classes are introduced. The Metathesaurus is classified into more mutually exclusive, uniform sets of concepts. The schema improves the user's comprehension and navigation of the Metathesaurus. CONCLUSIONS: The UMLS OODB schema supports the user's comprehension and navigation of the Metathesaurus. It also helps expose and resolve modeling problems in the UMLS.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

H Gu, Y Perl, J Geller, M Halper, L M Liu, J J Cimino. Representing the UMLS as an object-oriented database: modeling issues and advantages.. https://doi.org/10.1136/jamia.2000.0070066

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

GI Epidemiology: databases for epidemiological studies.

BACKGROUND: Large databases are being increasingly used for examining the epidemiology and outcomes of digestive and liver disorders. The complexity and rigor of the methods used to conduct these studies are often underestimated. AIMS: For the most commonly used databases, we provide a brief description of the contents, highlight strengths and weakness, and provide links for more detailed information. We also present a systematic approach to utilizing large databases for addressing research questions, highlighting commonly encountered study design issues, as well as strategies for resolving these issues. CONCLUSIONS: 1. Research using large databases requires the same essential skills needed to conduct research studies using other data sources. These include a rigorous study design, expertise in analytic methods, and relevant research questions. 2. The completeness and accuracy of information contained in the database must be assessed. Methods for improving the quality and completeness of this information should be considered. 3. Despite similarities among large databases, gaining insight and experience into the structure and content of each database is essential. Key points *Large databases can be a powerful source of information to examine the clinical epidemiology and outcomes of digestive and liver disorders. * Research using large databases requires the same essential skills needed to conduct research studies using other data sources. These include a rigorous study design, expertise in analytic methods, and relevant research questions. * The completeness and accuracy of information contained in the database must be assessed. Methods for improving the quality and completeness of this information should be considered. * Despite similarities among large databases, gaining insight and experience into the structure and content of each database is essential. * Examples of commonly used large databases are presented with a synopsis of information contained in the database, as well as strengths and limitations of using the database for research.

Databases as Topic↗

Italian Rett database and biobank.

Rett syndrome is the second most common cause of severe mental retardation in females, with an incidence of approximately 1 out of 10,000 live female births. In addition to the classic form, a number of Rett variants have been described. MECP2 gene mutations are responsible for about 90% of classic cases and for a lower percentage of variant cases. Recently, CDKL5 mutations have been identified in the early onset seizures variant and other atypical Rett patients. While the high percentage of MECP2 mutations in classic patients supports the hypothesis of a single disease gene, the low frequency of mutated variant cases suggests genetic heterogeneity. Since 1998, we have performed clinical evaluation and molecular analysis of a large number of Italian Rett patients. The Italian Rett Syndrome (RTT) database has been developed to share data and samples of our RTT collection with the scientific community (http://www.biobank.unisi.it). This is the first RTT database that has been connected with a biobank. It allows the user to immediately visualize the list of available RTT samples and, using the "Search by" tool, to rapidly select those with specific clinical and molecular features. By contacting bank curators, users can request the samples of interest for their studies. This database encourages collaboration projects with clinicians and researchers from around the world and provides important resources that will help to better define the pathogenic mechanisms underlying Rett syndrome.

Databases as Topic↗

A database on treating drug addiction with traditional Chinese medicine.

AIMS: Traditional Chinese medicine (TCM) has been used to treat drug addiction for more than 160 years and valuable experiences have been accumulated with regard to patients' detoxification and rehabilitation. The aims of this project were (1) to establish a computerized, bilingual (Chinese-English) database on TCM for drug addiction; (2) to analyse the literature published in this field; and (3) to identify those Chinese herbs commonly used for drug addiction treatment. DESIGN: (1) Paper collection: related papers were collected through electronic databases and hand-searched materials; (2) data computerization: the Microsoft Access program and Delphi language were used as the major data management systems; (3) paper analysis: annual publications from 1989 to 2003 were classified and calculated; and (4) herbal analysis: the frequency of herbs used and herbal function categories were analysed. FINDINGS: (1) A special bilingual database that contained 340 works of professional literature, including 85 patent files on TCM for drug addiction, was established, in which more than 90% of the publications originated from mainland China; (2) the literature classification showed a significant increase in the number of publications on clinical and laboratory researches in this field over the past decade; (3) five functional categorizations of Chinese herbs and the 10 most frequently used Chinese herbs as well as three toxic herbs were identified from more than 200 herbs reported in 150 original research articles and 85 patent files. CONCLUSIONS: For the first time, the published data on TCM in the treatment of drug addiction were analysed systematically by using a new database. The results are invaluable for further laboratory and clinical studies to obtain more direct evidence.

Databases as Topic↗