2010Statistics in Biopharmaceutical ResearchRequires access

Multiple Comparisons in Microarray Data Analysis

Donghui Zhang, Li Liu

Open publisher page 0 citations

Abstract

Multiplicity is a challenging statistical issue in drug discovery, and a particular example is microarray study. The traditional approach of controlling of the family-wise error rate (FWER) is conservative when the number of tests is large. A more appropriate approach is to control the false discovery rate (FDR). Since the development of the Benjamini and Hochberg (BH) FDR procedure in 1995, many modifications have been proposed aimed at relaxing the requirement for independent test statistics or improving the power of the BH FDR procedure. Comparisons of these procedures in the current literature are not comprehensive and the conclusions on performances are inconsistent. The objectives of this article are three-fold: (a) to perform a more comprehensive comparison of extant multiple testing procedures using two real microarray datasets and various simulated data sets; (b) to explore potential reasons for the inconsistencies in published simulation results; and (c) to identify suitable FDR procedures under different scenarios according to covariance structure, percent of true null hypotheses among multiple tests, and sample size.

About this research paper

What this paper is about

Multiplicity is a challenging statistical issue in drug discovery, and a particular example is microarray study. The traditional approach of controlling of the family-wise error rate (FWER) is conservative when the number of tests is large. A more appropriate approach is to control the false discovery rate (FDR). Since the development of the Benjamini and Hochberg (BH) FDR procedure in 1995, many modifications have been proposed aimed at relaxing the requirement for independent test statistics or improving the power of the BH FDR procedure. Comparisons of these procedures in the current literature are not comprehensive and the conclusions on performances are inconsistent. The objectives of this article are three-fold: (a) to perform a more comprehensive comparison of extant multiple testing procedures using two real microarray datasets and various simulated data sets; (b) to explore potential reasons for the inconsistencies in published simulation results; and (c) to identify suitable FDR procedures under different scenarios according to covariance structure, percent of true null hypotheses among multiple tests, and sample size.

Why it matters

A significance statement is not available in the OpenAlex record.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Multiplicity is a challenging statistical issue in drug discovery, and a particular example is microarray study. The traditional approach of controlling of the family-wise error rate (FWER) is conservative when the number of tests is large. A more appropriate approach is to control the false discovery rate (FDR). Since the development of the Benjamini and Hochberg (BH) FDR procedure in 1995, many modifications have been proposed aimed at relaxing the requirement for independent test statistics or improving the power of the BH FDR procedure. Comparisons of these procedures in the current literature are not comprehensive and the conclusions on performances are inconsistent. The objectives of this article are three-fold: (a) to perform a more comprehensive comparison of extant multiple testing procedures using two real microarray datasets and various simulated data sets; (b) to explore potential reasons for the inconsistencies in published simulation results; and (c) to identify suitable FDR procedures under different scenarios according to covariance structure, percent of true null hypotheses among multiple tests, and sample size.

Key concepts: False discovery rate, Multiple comparisons problem, Statistical hypothesis testing, Sample size determination, Computer science, Type I and type II errors, Data mining, Extant taxon

Related papers

Back to paper searchBrowse research topicsOriginal source
Multiple Comparisons in Microarray Data Analysis — Research Paper | ScholarLens