Search PubMed⌕ Search

PubMed · 15849016

Statistical analysis of microarray data.

Abstract

Microarrays promise dynamic snapshots of cell activity, but microarray results are unfortunately not straightforward to interpret. This article aims to distill the most useful practical results from the vast body of literature available on microarray data analysis. Topics covered include: experimental design issues, normalization, quality control, exploratory analysis, and tests for differential expression. Special attention is paid to the peculiarities of low-level analysis of Affymetrix chips, and the multiple testing problem in determining differential expression. The aim of this article is to provide useful answers to the most common practical issues in microarray data analysis. The main topics are pre-processing (normalization), and detecting differential expression. Subsidiary topics include experimental design, and exploratory analysis. Further discussion is found at the author's web page (http://discover.nci.nih.gov --> Notes on Microarray Data Analysis).

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Mark Reimers. 2005. Statistical analysis of microarray data.. https://doi.org/10.1080/13556210412331327795

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Clustering individuals using INMTD: a novel versatile multi-view embedding framework integrating omics and imaging data.

MOTIVATION: Combining omics and images can lead to a more comprehensive clustering of individuals than classic single-view approaches. Among the various approaches for multi-view clustering, nonnegative matrix tri-factorization (NMTF) and nonnegative Tucker decomposition (NTD) are advantageous in learning low-rank embeddings with promising interpretability. Besides, there is a need to handle unwanted drivers of clusterings (i.e. confounders). RESULTS: In this work, we introduce a novel multi-view clustering method based on NMTF and NTD, named INMTD, which integrates omics and 3D imaging data to derive unconfounded subgroups of individuals. According to the adjusted Rand index, INMTD outperformed other clustering methods on a synthetic dataset with known clusters. In the application to real-life facial-genomic data, INMTD generated biologically relevant embeddings for individuals, genetics, and facial morphology. By removing confounded embedding vectors, we derived an unconfounded clustering with better internal and external quality; the genetic and facial annotations of each derived subgroup highlighted distinctive characteristics. In conclusion, INMTD can effectively integrate omics data and 3D images for unconfounded clustering with biologically meaningful interpretation. AVAILABILITY AND IMPLEMENTATION: INMTD is freely available at https://github.com/ZuqiLi/INMTD.

Cluster Analysis↗

A graph spectral analysis of the structural similarity network of protein chains.

We present a simple method for the analysis of large networks based on their graph spectral properties. One of the advantages of this method is that it uses a single numerical computation to identify subclusters in a connected graph, which can significantly simplify the complexity involved in analyzing large graphs. This is illustrated using a network of protein chains constructed on the basis of their structural similarities. The large-scale network properties and the cluster and subcluster organization of the protein chain network are presented. We summarize the results of structural and functional analyses of the nodes present in these clusters and elucidate the implications of structural similarity in the protein chain universe.

Cluster Analysis↗