Search PubMed⌕ Search

PubMed · 16455749

The LCB Data Warehouse.

Abstract

UNLABELLED: The Linnaeus Centre for Bioinformatics Data Warehouse (LCB-DWH) is a web-based infrastructure for reliable and secure microarray gene expression data management and analysis that provides an online service for the scientific community. The LCB-DWH is an effort towards a complete system for storage (using the BASE system), analysis and publication of microarray data. Important features of the system include: access to established methods within R/Bioconductor for data analysis, built-in connection to the Gene Ontology database and a scripting facility for automatic recording and re-play of all the steps of the analysis. The service is up and running on a high performance server. At present there are more than 150 registered users. AVAILABILITY: An open functional version is available at https://dw.lcb.uu.se/index.phtml?i_login=test. User accounts are created upon request. Additional facilities including plug-ins, user documentation and a password protected data storage system are available from http://www.lcb.uu.se/lcbdw.php

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Adam Ameur, Vladimir Yankovski, Stefan Enroth, Ola Spjuth, Jan Komorowski. 2006-02-02. The LCB Data Warehouse.. https://doi.org/10.1093/bioinformatics%2Fbtl036

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

NCBI GEO: mining tens of millions of expression profiles--database and tools update.

The Gene Expression Omnibus (GEO) repository at the National Center for Biotechnology Information (NCBI) archives and freely disseminates microarray and other forms of high-throughput data generated by the scientific community. The database has a minimum information about a microarray experiment (MIAME)-compliant infrastructure that captures fully annotated raw and processed data. Several data deposit options and formats are supported, including web forms, spreadsheets, XML and Simple Omnibus Format in Text (SOFT). In addition to data storage, a collection of user-friendly web-based interfaces and applications are available to help users effectively explore, visualize and download the thousands of experiments and tens of millions of gene expression patterns stored in GEO. This paper provides a summary of the GEO database structure and user facilities, and describes recent enhancements to database design, performance, submission format options, data query and retrieval utilities. GEO is accessible at http://www.ncbi.nlm.nih.gov/geo/

Computer Graphics↗

GO PaD: the Gene Ontology Partition Database.

Gene Ontology (GO) has been widely used to infer functional significance associated with sets of genes in order to automate discoveries within large-scale genetic studies. A level in GO's direct acyclic graph structure is often assumed to be indicative of its terms' specificities, although other work has suggested this assumption does not hold. Unfortunately, quantitative analysis of biological functions based on nodes at the same level (as is common in gene enrichment analysis tools) can lead to incorrect conclusions as well as missed discoveries due to inefficient use of available information. This paper addresses these using an informational theoretic approach encoded in the GO Partition Database that guarantees to maximize information for gene enrichment analysis. The GO Partition Database was designed to feature ontology partitions with GO terms of similar specificity. The GO partitions comprise varying numbers of nodes and present relevant information theoretic statistics, so researchers can choose to analyze datasets at arbitrary levels of specificity. The GO Partition Database, featuring GO partition sets for functional analysis of genes from human and 10 other commonly studied organisms with a total of 131,972 genes, is available on the internet at: bcl.med.harvard.edu/proj/gopart. The site also includes an online tutorial.

Computer Graphics↗

PROTCOM: searchable database of protein complexes enhanced with domain-domain structures.

The database of protein complexes (PROTCOM) is a compilation of known 3D structures of protein-protein complexes enriched with artificially created domain-domain structures using the available entries in the Protein Data Bank. The domain-domain structures are generated by parsing single chain structures into loosely connected domains and are important features of the database. The database (http://www.ces.clemson.edu/compbio/protcom) could be used for benchmarking purposes of the docking and other algorithms for predicting 3D structures of protein-protein complexes. The database can be utilized as a template database in the homology or threading methods for modeling the 3D structures of unknown protein-protein complexes. PROTCOM provides the scientific community with an integrated set of tools for browsing, searching, visualizing and downloading a pool of protein complexes. The user is given the option to select a subset of entries using a combination of up to 10 different criteria. As on July 2006 the database contains 1770 entries, each of which consists of the known 3D structures and additional relevant information that can be displayed either in text-only or in visual mode.

Computer Graphics↗