Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “Database”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17Linked to original sources

A prototype Internet autopsy database. 1625 consecutive fetal and neonatal autopsy facesheets spanning 20 years.

OBJECTIVE: To demonstrate that cause-of-death statements can be generated by a computer algorithm from an autopsy database composed of diagnostic terms. DATA SOURCES: Over 49 000 autopsy facesheets contributed by over a dozen institutions were collected from a publicly accessible Internet autopsy database. This database is available at the following web site: http:@www.med.jhu.edu/pathology/iad.html STUDY SELECTION: To test the feasibility of creating and using a publicly available autopsy database, and to identify the technical and medicolegal problems that may arise with such a novel resource, a prototype study was designed by selecting autopsy facesheets from fetal and neonatal deaths. An algorithm was developed to determine the cause of death from the listing of anatomic diagnoses. DATA EXTRACTION: One thousand six hundred twenty-five fetal and neonatal autopsy facesheets were selected encompassing fetal and neonatal deaths occurring up to 28 days after birth. DATA SYNTHESIS: The algorithm determined causes of death from autopsy facesheet data in all cases. On review by an experienced pediatric pathologist, these automatically generated cause-of-death statements required no modification or only slight modification in over 90% of cases. CONCLUSIONS: A large multi-institutional autopsy database composed of demographic and diagnostic information has been deposited on the Internet. This information can be freely downloaded and used by any researcher without violating patient confidentiality. As a demonstration of one possible application of the database, fetal and neonatal autopsies generated cause-of-death statements using a computer algorithm. One can anticipate that the wealth of information contained in autopsy facesheets can be assembled into a database that will serve the public interest.

Algorithms↗

A regional perinatal database in southern Sweden--a basis for quality assurance in obstetrics and neonatology.

BACKGROUND: In order to ensure as few avoidable adverse outcomes of pregnancy as possible, it is necessary to continuously evaluate the quality of both obstetric and neonatal care. The eleven southernmost hospitals in Sweden have joined together in a project of developing a regional database, with special emphasis on rapid output of information in order to identify changing trends. METHODS: A regional computerized database has been developed, collecting variables and quality indicators agreed upon by all participants. Specific protocols have been designed for obstetric care, neonatal care and autopsy findings. All participating units transfer information on paper forms or via local computerized information systems. The regional database thus receives data on about 20,000 deliveries annually. RESULTS: Data collection started on September 1, 1994. The first results are due in March 1996, and thereafter on a regular basis every 3 months. Special methods for rapid analysis of raw data have been developed with the help of commercially available data analysis tools. CONCLUSIONS: It is possible to construct an information system with different computer platforms and different database tools at each participating facility, as long as the database systems are locally controllable. That a software is commercially available is no guarantee that data transfer to a central database is possible. Experience from participating sites also indicates that a specialized database is needed for registering obstetric data, as general computerized record-keeping systems are unable to cope with an event concerning more than one subject at a time.

Female↗

Prototype implementation of the integrated genomic database.

We aim to develop an open software system to handle human genome data. The system, called Integrated Genomic Database (IGD), will integrate information from many genomic databases and experimental resources into a comprehensive target-end database (IGD TED). Users will access front-end client systems (IGD FRED) to download data of interest to their computers and merge them with their own local data. FREDs will provide persistent storage of, and instant access to, retrieved data; a friendly graphical interface; tools for querying, browsing, analyzing, and editing local data; interface to external analysis; and tools for communicating with the outside world. The TED will be accessible over the network (online and offline) as a read-only resource for multiple clients. It collects data from major databases for nucleotide and protein sequences and structures, genome maps, experimental reagents, phenotypes, and bibliographic data, and sets of raw data produced at genome centers and laboratories. Beside character-based access via Gopher, WAIS, FTP, and several query language interfaces to the TED, we will develop a specialized front-end client, IGD FRED, with its own database manager, based on the ACEDB program. The FRED will support graphical display methods for sequence feature maps, chromosomal genetic and physical maps, and experimental objects like clone grids, etc. FRED will also provide an interface to important analysis software packages and tools for submitting data to external databases in their own format.

Computer Communication Networks↗

Up-to-date, and taxonomy-curated mcrA reference databases for methanogen community profiling.

The methyl-coenzyme M reductase subunit alpha gene (mcrA) is an important phylogenetic marker for high throughput ecological profiling of methanogenic archaea, central to industrial biological methane production and greenhouse gas emissions. Yet, dedicated reference databases predate current relevant NCBI sequence accumulation and archaeal taxonomic revision. We present three updated mcrA reference databases: (i) one derived from NCBI-catalogued methanogen genomes (1572 sequences); (ii) a database built by expansion of a previously published reference dataset, leveraging the NCBI nucleotide collection (27,942 sequences); (iii) a curated-taxonomy version of the latter. The updated amplicon databases provide a ∼ 3.5-fold sequence richness expansion, extend genus-level richness from 31 to 83 taxa, more than 4-fold species-level richness, and incorporate novel lineages compared with the previous reference dataset (e.g. Thermoplasmatota-encompassed). All databases were formatted to support analysis with relevant contemporary software pipelines and packages. Overall, the generated databases facilitate a highly improved characterization of methanogen diversity and ecology.

Archaea↗

Development, analysis, refinement, and utility of an interdisciplinary amyotrophic lateral sclerosis database.

The current status of evaluation and management provided by individual healthcare professionals (HCP) at amyotrophic lateral sclerosis (ALS) centers and clinics needs to be analyzed. This paper describes one ALS center's experiences with the development, analysis, refinement, and utility of an interdisciplinary, HCP-driven ALS database. The purpose and conceptual framework of the database, the general data that needed to be collected, and the types of reports that needed to be generated were determined, and, in collaboration with a computer programmer, data entry and database management systems were developed. Data were collected on 234 patients between September 1996 and August 1998, and were analyzed by a biostatistician. Based on review of the biostatistician's report and discussion of problems encountered with the systems, the database was then refined. Benefits of the database system included: systematization of data collection and reporting, reduction of redundant data collection by individuals, decreased variability of evaluation methods and management decisions from patient to patient, and increased availability of a variety of uniform patient information to assist team members in making care decisions. Ongoing refinement will ensure that this HCP-driven ALS database continues to be informative, practical and effective for decision-making and enhancing delivery of care.

Adult↗

Development of an animal genome database and its search system.

An animal genome database has been developed on a Unix workstation and maintained by a relational database management system. This database has focused on the comparative gene mapping between species to assist the mapping of the genes related to phenotypic traits in livestock. The linkage maps, cytogenetic maps, polymerase chain reaction primers of pig, cattle, mouse and human, and their references have been included in the database, and the correspondence among species have been stipulated in the database. In order to search the database effectively, the World Wide Web server (http://ws4.niai.affrc.go.jp/) and the electronic mail server system (e-mail: jgbase-mail@ niai.affrc.go.jp) have been developed on different Unix workstations. These servers are connected to the Internet.

Animals↗

Database-driven multi locus sequence typing (MLST) of bacterial pathogens.

MOTIVATION: Multi Locus Sequence Typing (MLST) is a newly developed typing method for bacteria based on the sequence determination of internal fragments of seven house-keeping genes. It has proved useful in characterizing and monitoring disease-causing and antibiotic resistant lineages of bacteria. The strength of this approach is that unlike data obtained using most other typing methods, sequence data are unambiguous, can be held on a central database and be queried through a web server. RESULTS: A database-driven software system (mlstdb) has been developed, which is used by public health laboratories and researchers globally to query their nucleotide sequence data against centrally held databases over the internet. The mlstdb system consists of a set of perl scripts for defining the database tables and generating the database management interface and dynamic web pages for querying the databases. AVAILABILITY: http://www.mlst.net.

Bacteria↗

BioMolQuest: integrated database-based retrieval of protein structural and functional information.

MOTIVATION: Information about a particular protein or protein family is usually distributed among multiple databases and often in more than one entry in each database. Retrieval and organization of this information can be a laborious task. This task is complicated even further by the existence of alternative terms for the same concept. RESULTS: The PDB, SWISS-PROT, ENZYME, and CATH databases have been imported into a combined relational database, BIOMOLQUEST: A powerful search engine has been built using this database as a back end. The search engine achieves significant improvements in query performance by automatically utilizing cross-references between the legacy databases. The results of the queries are presented in an organized, hierarchical way.

Abstracting and Indexing↗

The Database of Quantitative Cellular Signaling: management and analysis of chemical kinetic models of signaling networks.

MOTIVATION: Analysis of cellular signaling interactions is expected to pose an enormous informatics challenge, perhaps even larger than analyzing the genome. The complex networks arising from signaling processes are traditionally represented as block diagrams. A key step in the evolution toward a more quantitative understanding of signaling is to explicitly specify the kinetics of all chemical reaction steps in a pathway. Technical advances in proteomics and high-throughput protein interaction assays promise a flood of such quantitative data. While annotations, molecular information and pathway connectivity have been compiled in several databases, and there are several proposals for general cell model description languages, there is currently little experience with databases of chemical kinetics and reaction level models of signaling networks. RESULTS: The Database of Quantitative Cellular Signaling is a repository of models of signaling pathways. It is intended both to serve the growing field of chemical-reaction level simulation of signaling networks, and to anticipate issues in large-scale data management for signaling chemistry. AVAILABILITY: The Database of Quantitative Cellular Signaling is available at http://doqcs.ncbs.res.in. Links to the signaling model simulator, GENESIS/Kinetikit are at http://www.ncbs.res.in/~bhalla/kkit/index.html and are also provided from within the database. The database source code is available under the GNU Public License.

Abstracting and Indexing↗

MHCBN: a comprehensive database of MHC binding and non-binding peptides.

MHCBN is a comprehensive database of Major Histocompatibility Complex (MHC) binding and non-binding peptides compiled from published literature and existing databases. The latest version of the database has 19 777 entries including 17 129 MHC binders and 2648 MHC non-binders for more than 400 MHC molecules. The database has sequence and structure data of (a) source proteins of peptides and (b) MHC molecules. MHCBN has a number of web tools that include: (i) mapping of peptide on query sequence; (ii) search on any field; (iii) creation of data sets; and (iv) online data submission. The database also provides hypertext links to major databases like SWISS-PROT, PDB, IMGT/HLA-DB, GenBank and PUBMED.

Amino Acid Sequence↗

BioQuery: an object framework for building queries to biomedical databases.

SUMMARY: BioQuery is an application that helps scientists automate database searches. Users can build and store queries to public biomedical databases, and receive periodic updates on the results of those queries when new data is available. The application is implemented on a portable object framework that can provide database-searching capability to other applications. This framework is easily extensible, allowing users to develop plug-ins that provide access to new databases. BioQuery thus provides end-users with a complete database searching interface and updating service, and gives developers a toolkit to provide database-searching capability to their applications. AVAILABILITY: Free to all users: http://www.bioquery.org.

Biomedical Research↗

DAtA: database of Arabidopsis thaliana annotation.

The Database of Arabidopsis thaliana Annotation (D At A) was created to enable easy access to and analysis of all the Arabidopsis genome project annotation. The database was constructed using the completed A.thaliana genomic sequence data currently in GenBank. An automated annotation process was used to predict coding sequences for GenBank records that do not include annotation. D At A also contains protein motifs and protein similarities derived from searches of the proteins in D At A with motif databases and the non-redundant protein database. The database is routinely updated to include new GenBank submissions for Arabidopsis genomic sequences and new Blast and protein motif search results. A web interface to D At A allows coding sequences to be searched by name, comment, blast similarity or motif field. In addition, browse options present lists of either all the protein names or identified motifs present in the sequenced A.thaliana genome. The database can be accessed at http://baggage. stanford.edu/group/arabprotein/

Arabidopsis↗

PASS2: a semi-automated database of protein alignments organised as structural superfamilies.

PASS2 is a nearly automated version of CAMPASS and contains sequence alignments of proteins grouped at the level of superfamilies. This database has been created to fall in correspondence with SCOP database (1.53 release) and currently consists of 110 multi-member superfamilies and 613 superfamilies corresponding to single members. In multi-member superfamilies, protein chains with no more than 25% sequence identity have been considered for the alignment and hence the database aims to address sequence alignments which represent 26 219 protein domains under the SCOP 1.53 release. Structure-based sequence alignments have been obtained by COMPARER and the initial equivalences are provided automatically from a MALIGN alignment and subsequently augmented using STAMP4.0. The final sequence alignments have been annotated for the structural features using JOY4.0. Several interesting links are provided to other related databases and genome sequence relatives. Availability of reliable sequence alignments of distantly related proteins, despite poor sequence identity and single-member superfamilies, permit better sampling of structures in libraries for fold recognition of new sequences and for the understanding of protein structure-function relationships of individual superfamilies. The database can be queried by keywords and also by sequence search, interfaced by PSI-BLAST methods. Structure-annotated sequence alignments and several structural accessory files can be retrieved for all the superfamilies including the user-input sequence. The database can be accessed from http://www.ncbs.res.in/%7Efaculty/mini/campass/pass.html.

Amino Acid Sequence↗

Characteristics of the U.S. EPA's Office of Pesticide Programs' toxicity information databases.

The United States Environmental Protection Agency's Office of Pesticide Programs (OPP) requires that data from toxicity testing be submitted to the OPP to support the registration of pesticide chemicals. Once the toxicity data are submitted, they are entered into various toxicity databases. The studies are listed in an archival database to catalog and allow retrieval of the study for review. Reviews of toxicity studies are then placed into a separate database that can be retrieved to support a regulatory position. Toxicity information for health effects other than cancer and gene mutations from chronic exposure is reviewed through a reference dose (RfD) approach, and these decisions and supporting data are entered into an RfD database. Carcinogenicity data are reviewed by a peer review process, and these decisions are entered into a newly developed database to show the regulatory decision with supporting data. The mutagenicity data are reviewed and acceptable data are entered into the Genetic Activity Profile system to catalog and display the submitted information. These databases contain the information used for hazard evaluations as part of the OPP review of pesticide chemicals.

Animals↗

EXProt--a database for EXPerimentally verified Protein functions.

EXProt (database for EXPerimentally verified Protein functions) is a new non-redundant database containing protein sequences for which the function has been experimentally verified. It is a selection of 3976 entries from the Prokaryotes section of the EMBL Nucleotide Sequence Database, Release 66, and 375 entries from the Pseudomonas Community Annotation Project (PseudoCAP). The entries in EXProt all have a unique ID number and provide information about the organism, protein sequence, functional annotation, link to entry in original database, and if known, gene name and link to references in PubMed/Medline. The EXProt web page (http://www.cmbi.nl/EXProt) provides further details of the database and a link to a BLAST search (blastp & blastx) of the database. The EXProt entries are indexed in SRS (http://www.cmbi.nl/srs/) and can be searched by means of keywords. Authors can be reached by email (exprot(cmbi.kun.nl).

Amino Acid Sequence↗

Measuring use patterns of online journals and databases.

PURPOSE: This research sought to determine use of online biomedical journals and databases and to assess current user characteristics associated with the use of online resources in an academic health sciences center. SETTING: The Library of the Health Sciences-Peoria is a regional site of the University of Illinois at Chicago (UIC) Library with 350 print journals, more than 4,000 online journals, and multiple online databases. METHODOLOGY: A survey was designed to assess online journal use, print journal use, database use, computer literacy levels, and other library user characteristics. A survey was sent through campus mail to all (471) UIC Peoria faculty, residents, and students. RESULTS: Forty-one percent (188) of the surveys were returned. Ninety-eight percent of the students, faculty, and residents reported having convenient access to a computer connected to the Internet. While 53% of the users indicated they searched MEDLINE at least once a week, other databases showed much lower usage. Overall, 71% of respondents indicated a preference for online over print journals when possible. CONCLUSIONS: Users prefer online resources to print, and many choose to access these online resources remotely. Convenience and full-text availability appear to play roles in selecting online resources. The findings of this study suggest that databases without links to full text and online journal collections without links from bibliographic databases will have lower use. These findings have implications for collection development, promotion of library resources, and end-user training.

Computer Literacy↗

Non-sequence databases for biological activity and physicochemical properties.

A biological activity database and a physicochemical property database are described. They are intended to complement the protein sequence database of PIR-International. The Biological Activity Database and the Physicochemical Property Database contain information regarding the biological activity and the physicochemical properties of proteins, respectively. In addition they also provide information about wild-type molecules with which information concerning variant molecules may be compared. Data on artificial variant molecules are stored in the Artificial Variant Database which is described separately.

Amino Acid Sequence↗

A database model for studies of cocaine-dependent pregnant women and their families.

The database management functions for the Mothers Project are arranged into administrative and analytic task groups, and separate systems are devised for each. The task groups can be distinguished not only by differences in data structure but also by interface requirements. The administrative database system uses a relational database technology, whereas the analytic database system employs more traditional flat-file methods. Although the database management systems are complex, they are based on standard database practices, used in widely available software packages, and run on inexpensive desktop computing equipment.

Cocaine↗