2017•World Journal of MethodologyOpen access

Predictive power of statistical significance

Thomas F Heston, Jackson M. King

Open full text 25 citations

Abstract

A statistically significant research finding should not be defined as a P -value of 0.05 or less, because this definition does not take into account study power.Statistical significance was originally defined by Fisher RA as a P -value of 0.05 or less.According to Fisher, any finding that is likely to occur by random variation no more than 1 in 20 times is considered significant.Neyman J and Pearson ES subsequently argued that Fisher's definition was incomplete.They proposed that statistical significance could only be determined by analyzing the chance of incorrectly considering a study finding was significant (a Type Ⅰ error) or incorrectly considering a study finding was insignificant (a Type Ⅱ error).Their definition of statistical significance is also incomplete because the error rates are considered separately, not together.A better definition of statistical significance is the positive predictive value of a P -value, which is equal to the power divided by the sum of power and the P -value.This definition is more complete and relevant than Fisher's or Neyman-Peason's definitions, because it takes into account both concepts of statistical significance.Using this definition, a statistically significant finding requires a P -value of 0.05 or less when the power is at least 95%, and a P -value of 0.032 or less when the power is 60%.To achieve statistical significance, P -values must be adjusted downward as the study power decreases.

Open-access reader

About this research paper

What this paper is about

A statistically significant research finding should not be defined as a P -value of 0.05 or less, because this definition does not take into account study power.Statistical significance was originally defined by Fisher RA as a P -value of 0.05 or less.According to Fisher, any finding that is likely to occur by random variation no more than 1 in 20 times is considered significant.Neyman J and Pearson ES subsequently argued that Fisher's definition was incomplete.They proposed that statistical significance could only be determined by analyzing the chance of incorrectly considering a study finding was significant (a Type Ⅰ error) or incorrectly considering a study finding was insignificant (a Type Ⅱ error).Their definition of statistical significance is also incomplete because the error rates are considered separately, not together.A better definition of statistical significance is the positive predictive value of a P -value, which is equal to the power divided by the sum of power and the P -value.This definition is more complete and relevant than Fisher's or Neyman-Peason's definitions, because it takes into account both concepts of statistical significance.Using this definition, a statistically significant finding requires a P -value of 0.05 or less when the power is at least 95%, and a P -value of 0.032 or less when the power is 60%.To achieve statistical significance, P -values must be adjusted downward as the study power decreases.

Why it matters

OpenAlex reports 25 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

A statistically significant research finding should not be defined as a P -value of 0.05 or less, because this definition does not take into account study power.Statistical significance was originally defined by Fisher RA as a P -value of 0.05 or less.According to Fisher, any finding that is likely to occur by random variation no more than 1 in 20 times is considered significant.Neyman J and Pearson ES subsequently argued that Fisher's definition was incomplete.They proposed that statistical significance could only be determined by analyzing the chance of incorrectly considering a study finding was significant (a Type Ⅰ error) or incorrectly considering a study finding was insignificant (a Type Ⅱ error).Their definition of statistical significance is also incomplete because the error rates are considered separately, not together.A better definition of statistical significance is the positive predictive value of a P -value, which is equal to the power divided by the sum of power and the P -value.This definition is more complete and relevant than Fisher's or Neyman-Peason's definitions, because it takes into account both concepts of statistical significance.Using this definition, a statistically significant finding requires a P -value of 0.05 or less when the power is at least 95%, and a P -value of 0.032 or less when the power is 60%.To achieve statistical significance, P -values must be adjusted downward as the study power decreases.

Key concepts: Statistical significance, Statistical power, p-value, Statistics, Type I and type II errors, Mathematics, Value (mathematics), Power (physics)

Related papers

Back to paper searchBrowse research topicsOriginal source
Predictive power of statistical significance — Research Paper | ScholarLens