Reports of the Death of Regression-Discontinuity Analysis are Greatly Exaggerated
Charles S. Reichardt, William M. K. Trochim, Joseph C. Cappelleri
Abstract
Charles S. Reichardt, William M. K. Trochim, Joseph C. Cappelleri
Abstract
Stanley (1991) argues that both random measurement error in the pretest and treatment-effect interactions bias the estimate of the treatment effect when multiple regression is used to analyze the data from a regression-discontinuity design (RDD). Stanley also argues that these biases are so severe that they should cause researchers to consider using statistical procedures other than regression analysis. The authors of the present article disagree. Curvilinearity in the regression of the posttest on pretest scores can be difficult to model, can bias the regression analysis of data from the RDD if not modeled correctly, and therefore should cause researchers to consider alternatives to regression analysis. If the regression surfaces are linear, however, unbiased estimates can be obtained easily via regression analysis, whether or not either random measurement error in the pretest or treatment-effect interactions are present. Improving upon regression analysis is a worthy goal but requires understanding just what are and are not the weaknesses of the method. In addressing these issues, this article elucidates some of the general principles that underlie the use of multiple regression to analyze data from the RDD quasi-experiment.
OpenAlex reports 25 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
Stanley (1991) argues that both random measurement error in the pretest and treatment-effect interactions bias the estimate of the treatment effect when multiple regression is used to analyze the data from a regression-discontinuity design (RDD). Stanley also argues that these biases are so severe that they should cause researchers to consider using statistical procedures other than regression analysis. The authors of the present article disagree. Curvilinearity in the regression of the posttest on pretest scores can be difficult to model, can bias the regression analysis of data from the RDD if not modeled correctly, and therefore should cause researchers to consider alternatives to regression analysis. If the regression surfaces are linear, however, unbiased estimates can be obtained easily via regression analysis, whether or not either random measurement error in the pretest or treatment-effect interactions are present. Improving upon regression analysis is a worthy goal but requires understanding just what are and are not the weaknesses of the method. In addressing these issues, this article elucidates some of the general principles that underlie the use of multiple regression to analyze data from the RDD quasi-experiment.
Key concepts: Regression discontinuity design, Regression analysis, Regression toward the mean, Regression diagnostic, Regression, Statistics, Linear regression, Cross-sectional regression