2024International Journal of Advanced Statistics and ProbabilityOpen access

On remedying the presence of heteroscedasticity in a multiple linear regression modelling

Emmanuel Uchenna Ohaegbulem, Victor Chijindu Iheaka

Open full text 2 citations

Abstract

This study demonstrated the very essence of remedying the presence of heteroscedasticity, where it existed, in regression modelling. Two different hypothetical data, Data A (the Original) and Data B (the Original), were used in this study for the purpose of illustration. The normality, multicollinearity and autocorrelation assumptions were satisfied, but the Breusch-Pagan test and the White test established the existence of heteroscedasticity in the two datasets. The estimated multiple linear regression model for Data A (the Original) was statistically significant with an R-square value of 0.976, an AIC value of 332.5929, and an SBC value of 347.2533; and the one for Data B (the Original) was also statistically significant with an R-square value of 0.553, an AIC value of 69.89669, and an SBC value of 82.15499. The Log-transformation was applied on the variables in Data A (the Original) and Data B (the Original) to give rise to new sets of data, Data A (Now with Heteroscedasticity Remedied) and Data B (Now with Heteroscedasticity Remedied); which equally satisfied the normality, multicollinearity and autocorrelation assumptions, and also satisfied that there were no existence of heteroscedasticity in the two datasets. Now, the estimated multiple linear regression model for Data A (Now with Heteroscedasticity Remedied) was statistically significant with an R-square value of 0.986, an AIC value of -135.021, and an SBC value of -120.361; and the estimated model for Data B (Now with Heteroscedasticity Remedied) was statistically significant with an R-square value of 0.624, an AIC value of -32.0801, and an SBC value of -19.8218. From the points of view of the values of the R-square (0.986>0.976 and 0.624>0.553), AIC (-135.021<332.5929 and -32.0801<69.89669) and SBC (-120.361<347.2533 and -19.8218<82.15499), it was evident that the estimated regression models for Data A (Now with Heteroscedasticity Remedied) and Data B (Now with Heteroscedasticity Remedied) were, respectively, better models when compared to the regression models for Data A (the Original) and Data B (the Original).

Open-access reader

About this research paper

What this paper is about

This study demonstrated the very essence of remedying the presence of heteroscedasticity, where it existed, in regression modelling. Two different hypothetical data, Data A (the Original) and Data B (the Original), were used in this study for the purpose of illustration. The normality, multicollinearity and autocorrelation assumptions were satisfied, but the Breusch-Pagan test and the White test established the existence of heteroscedasticity in the two datasets. The estimated multiple linear regression model for Data A (the Original) was statistically significant with an R-square value of 0.976, an AIC value of 332.5929, and an SBC value of 347.2533; and the one for Data B (the Original) was also statistically significant with an R-square value of 0.553, an AIC value of 69.89669, and an SBC value of 82.15499. The Log-transformation was applied on the variables in Data A (the Original) and Data B (the Original) to give rise to new sets of data, Data A (Now with Heteroscedasticity Remedied) and Data B (Now with Heteroscedasticity Remedied); which equally satisfied the normality, multicollinearity and autocorrelation assumptions, and also satisfied that there were no existence of heteroscedasticity in the two datasets. Now, the estimated multiple linear regression model for Data A (Now with Heteroscedasticity Remedied) was statistically significant with an R-square value of 0.986, an AIC value of -135.021, and an SBC value of -120.361; and the estimated model for Data B (Now with Heteroscedasticity Remedied) was statistically significant with an R-square value of 0.624, an AIC value of -32.0801, and an SBC value of -19.8218. From the points of view of the values of the R-square (0.986>0.976 and 0.624>0.553), AIC (-135.021<332.5929 and -32.0801<69.89669) and SBC (-120.361<347.2533 and -19.8218<82.15499), it was evident that the estimated regression models for Data A (Now with Heteroscedasticity Remedied) and Data B (Now with Heteroscedasticity Remedied) were, respectively, better models when compared to the regression models for Data A (the Original) and Data B (the Original).

Why it matters

OpenAlex reports 2 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

This study demonstrated the very essence of remedying the presence of heteroscedasticity, where it existed, in regression modelling. Two different hypothetical data, Data A (the Original) and Data B (the Original), were used in this study for the purpose of illustration. The normality, multicollinearity and autocorrelation assumptions were satisfied, but the Breusch-Pagan test and the White test established the existence of heteroscedasticity in the two datasets. The estimated multiple linear regression model for Data A (the Original) was statistically significant with an R-square value of 0.976, an AIC value of 332.5929, and an SBC value of 347.2533; and the one for Data B (the Original) was also statistically significant with an R-square value of 0.553, an AIC value of 69.89669, and an SBC value of 82.15499. The Log-transformation was applied on the variables in Data A (the Original) and Data B (the Original) to give rise to new sets of data, Data A (Now with Heteroscedasticity Remedied) and Data B (Now with Heteroscedasticity Remedied); which equally satisfied the normality, multicollinearity and autocorrelation assumptions, and also satisfied that there were no existence of heteroscedasticity in the two datasets. Now, the estimated multiple linear regression model for Data A (Now with Heteroscedasticity Remedied) was statistically significant with an R-square value of 0.986, an AIC value of -135.021, and an SBC value of -120.361; and the estimated model for Data B (Now with Heteroscedasticity Remedied) was statistically significant with an R-square value of 0.624, an AIC value of -32.0801, and an SBC value of -19.8218. From the points of view of the values of the R-square (0.986>0.976 and 0.624>0.553), AIC (-135.021<332.5929 and -32.0801<69.89669) and SBC (-120.361<347.2533 and -19.8218<82.15499), it was evident that the estimated regression models for Data A (Now with Heteroscedasticity Remedied) and Data B (Now with Heteroscedasticity Remedied) were, respectively, better models when compared to the regression models for Data A (the Original) and Data B (the Original).

Key concepts: Heteroscedasticity, Linear regression, Statistics, Econometrics, Linear model, Regression analysis, Regression, Mathematics

Related papers

Back to paper searchBrowse research topicsOriginal source
On remedying the presence of heteroscedasticity in a multiple linear regression modelling — Research Paper | ScholarLens