Search PubMed⌕ Search

PubMed · 16137321

TmaDB: a repository for tissue microarray data.

Abstract

BACKGROUND: Tissue microarray (TMA) technology has been developed to facilitate large, genome-scale molecular pathology studies. This technique provides a high-throughput method for analyzing a large cohort of clinical specimens in a single experiment thereby permitting the parallel analysis of molecular alterations (at the DNA, RNA, or protein level) in thousands of tissue specimens. As a vast quantity of data can be generated in a single TMA experiment a systematic approach is required for the storage and analysis of such data. DESCRIPTION: To analyse TMA output a relational database (known as TmaDB) has been developed to collate all aspects of information relating to TMAs. These data include the TMA construction protocol, experimental protocol and results from the various immunocytological and histochemical staining experiments including the scanned images for each of the TMA cores. Furthermore the database contains pathological information associated with each of the specimens on the TMA slide, the location of the various TMAs and the individual specimen blocks (from which cores were taken) in the laboratory and their current status i.e. if they can be sectioned into further slides or if they are exhausted. TmaDB has been designed to incorporate and extend many of the published common data elements and the XML format for TMA experiments and is therefore compatible with the TMA data exchange specifications developed by the Association for Pathology Informatics community. Finally the design of the database is made flexible such that TMA experiments from several types of cancer can be stored in a single database, which incorporates the national minimum data set required for pathology reports supported by the Royal College of Pathologists (UK). CONCLUSION: TmaDB will provide a comprehensive repository for TMA data such that a large number of results from the numerous immunostaining experiments can be efficiently compared for each of the TMA cores. This will allow a systematic, large-scale comparison of tumour samples to facilitate the identification of gene products of clinical importance such as therapeutic or prognostic markers. In addition this work will contribute to the establishment of a standard for reporting TMA data analogous to MIAME in the description of microarray data.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Archana Sharma-Oates, Philip Quirke, David R Westhead. 2005-09-01. TmaDB: a repository for tissue microarray data.. https://doi.org/10.1186/1471-2105-6-218

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Measurement of inter-rater agreement for transient events using Monte Carlo sampled permutations.

In this paper we demonstrate the adverse effect of serially observed data sequences containing transient events on the calculation of Cohen's kappa as an index of inter-rater agreement in the detection of these events. We develop and use a Monte-Carlo-based permutation technique to produce an empiric distribution of kappa in the presence of serial dependence. We find that the empiric confidence intervals for kappa tend to be wider than parametrically derived intervals and in the case of longer event lengths, are markedly so. We evaluate the effect of number and length of events, and further, describe and evaluate three permutation methods which match specific rating situations. Finally, we apply these techniques to the measurement of inter-rater agreement for sleep disordered breathing events, a transient event identified during nocturnal polysomnography, for which traditionally computed confidence intervals for kappa are incorrect.

Data Interpretation, Statistical↗

Non-normal path analysis in the presence of measurement error and missing data: a Bayesian analysis of nursing homes' structure and outcomes.

Path analytic models are useful tools in quantitative nursing research. They allow researchers to hypothesize causal inferential paths and test the significance of these paths both directly and indirectly through a mediating variable. A standard statistical method in the path analysis literature is to treat the variables as having a normal distribution and to estimate paths using several least squares regression equations. The parameters corresponding to the direct paths have point and interval estimates based on normal distribution theory. Indirect paths are a product of the direct path from the independent variable to the mediating variable and the direct path of the mediating variable to the dependent variable. However, in the case of non-normal distributions, the point and interval estimates of the indirect path become much more difficult to estimate. We address the issue of calculating indirect point and interval estimates in the case of non-normally distributed data. Our substantive application is a nursing home research problem in which the variables in the path analysis of interest involve variables with normal, Bernoulli, or Poisson distributions. Additionally, one of the Poisson variables is observed with error. This paper addresses estimating point and interval estimation of indirect paths for variables with non-normal distributions in the presence of missing data and measurement error. We handle these difficulties from a fully Bayesian point of view. We present our substantive path analysis motivated from a nursing home structure, process, and outcomes model. Our results focus on the impact job turnover in the nursing homes has on nursing home outcomes.

Data Interpretation, Statistical↗

Improving ecological inference using individual-level data.

In typical small-area studies of health and environment we wish to make inference on the relationship between individual-level quantities using aggregate, or ecological, data. Such ecological inference is often subject to bias and imprecision, due to the lack of individual-level information in the data. Conversely, individual-level survey data often have insufficient power to study small-area variations in health. Such problems can be reduced by supplementing the aggregate-level data with small samples of data from individuals within the areas, which directly link exposures and outcomes. We outline a hierarchical model framework for estimating individual-level associations using a combination of aggregate and individual data. We perform a comprehensive simulation study, under a variety of realistic conditions, to determine when aggregate data are sufficient for accurate inference, and when we also require individual-level information. Finally, we illustrate the methods in a case study investigating the relationship between limiting long-term illness, ethnicity and income in London.

Data Interpretation, Statistical↗