2003Statistics in MedicineRequires access

Power comparison of two‐sided exact tests for association in 2 × 2 contingency tables using standard, mid p and randomized test versions

Stian Lydersen, Petter Laake

Open publisher page 23 citations

Abstract

Exact Pearson's chi square, likelihood ratio (LR), and Fisher's tests are obtained from the conditional distribution of its test statistic, given the row and column sums of the contingency table. The power and obtained significance level of the standard, mid p, and randomized versions of these tests are compared for two-sided tests in 2 x 2 tables, using binomial and multinomial sampling. The mid p type I error probabilities seldom exceed the nominal significance level. The mid p and randomized test versions have approximately the same power, and higher power than the standard test version. The power of the Pearson's chi square, LR and Fisher's test differ, and they differ in approximately the same way for standard, mid p and randomized test versions for any given set of parameters. There is no general ranking between the three tests. In many cases, Pearson's chi square and Fisher's tests have almost equal power, and higher power than LR. In a few cases, perhaps characterized by poorly balanced designs, LR performs best. Fisher's test seems to be slightly more robust even if the design is poor.

About this research paper

What this paper is about

Exact Pearson's chi square, likelihood ratio (LR), and Fisher's tests are obtained from the conditional distribution of its test statistic, given the row and column sums of the contingency table. The power and obtained significance level of the standard, mid p, and randomized versions of these tests are compared for two-sided tests in 2 x 2 tables, using binomial and multinomial sampling. The mid p type I error probabilities seldom exceed the nominal significance level. The mid p and randomized test versions have approximately the same power, and higher power than the standard test version. The power of the Pearson's chi square, LR and Fisher's test differ, and they differ in approximately the same way for standard, mid p and randomized test versions for any given set of parameters. There is no general ranking between the three tests. In many cases, Pearson's chi square and Fisher's tests have almost equal power, and higher power than LR. In a few cases, perhaps characterized by poorly balanced designs, LR performs best. Fisher's test seems to be slightly more robust even if the design is poor.

Why it matters

OpenAlex reports 23 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Exact Pearson's chi square, likelihood ratio (LR), and Fisher's tests are obtained from the conditional distribution of its test statistic, given the row and column sums of the contingency table. The power and obtained significance level of the standard, mid p, and randomized versions of these tests are compared for two-sided tests in 2 x 2 tables, using binomial and multinomial sampling. The mid p type I error probabilities seldom exceed the nominal significance level. The mid p and randomized test versions have approximately the same power, and higher power than the standard test version. The power of the Pearson's chi square, LR and Fisher's test differ, and they differ in approximately the same way for standard, mid p and randomized test versions for any given set of parameters. There is no general ranking between the three tests. In many cases, Pearson's chi square and Fisher's tests have almost equal power, and higher power than LR. In a few cases, perhaps characterized by poorly balanced designs, LR performs best. Fisher's test seems to be slightly more robust even if the design is poor.

Key concepts: Statistics, Contingency table, Mathematics, Exact test, Pearson's chi-squared test, Chi-square test, Type I and type II errors, Multinomial distribution

Related papers

Back to paper searchBrowse research topicsOriginal source
Power comparison of two‐sided exact tests for association in 2 × 2 contingency tables using standard, mid p and randomized test versions — Research Paper | ScholarLens