Identification of significant features in DNA microarray data
arXiv:1304.3838 · doi:10.1002/wics.1260
Abstract
DNA microarrays are a relatively new technology that can simultaneously measure the expression level of thousands of genes. They have become an important tool for a wide variety of biological experiments. One of the most common goals of DNA microarray experiments is to identify genes associated with biological processes of interest. Conventional statistical tests often produce poor results when applied to microarray data due to small sample sizes, noisy data, and correlation among the expression levels of the genes. Thus, novel statistical methods are needed to identify significant genes in DNA microarray experiments. This article discusses the challenges inherent in DNA microarray analysis and describes a series of statistical techniques that can be used to overcome these challenges. The problem of multiple hypothesis testing and its relation to microarray studies is also considered, along with several possible solutions.
35 pages, 6 figures. To be published in WIREs Computational Statistics
References in corpus (8)
- In silico prediction of protein-protein interactions in human macrophages
- On testing the significance of sets of genes
- Deriving chemosensitivity from cell lines: Forensic bioinformatics and reproducible research in high-throughput biology
- Microarrays, Empirical Bayes and the Two-Groups Model
- Random-set methods identify distinct aspects of the enrichment signal in gene-set analysis
- Robustness of multiple testing procedures against dependence
- Dependency and false discovery rate: Asymptotics
- Testing significance of features by lassoed principal components