2018•DergiPark (Istanbul University)Open access

Comparative Study of Classical Test Theory and Item Response Theory Using Diagnostic Quantitative Economics Skill Test Item Analysis Results

Lydia İjeoma Eleje, Frederick Ekene Onah, C. C. Abanobi

Open full text 20 citations

Abstract

In this study a comparison was made on DQEST item parameters and test statistics results estimated using Classical Test Theory (CTT) approach and Item Response Theory (IRT) three parameter logistic model (3PLM) to find out the similarities and differences in the two frameworks. 517 randomly selected senior secondary three (SS3) economics students comprised the sample. Three research questions guided the study. Responses obtained from SS3 economics students in 50 multiple choice items of Diagnostic Quantitative Economics Skill Test (DQEST) were used for the analysis. DQEST items certified the unidimensionality, local independence and model-data fit assumptions. Then results from CTT and IRT analyses were compared. In terms of very difficult item and item that discriminate poorly, CTT were found not to be comparable with the 3PLM the most appropriate model for DQEST data. The calculated reliability value for CTT was found to be low when compared to that generated by 3PLM. Therefore, it could be concluded that there was disparity between CTT approach and 3 parameter IRT model in terms of item parameters and test statistics. Thus IRT model with the best data fit should be employed for an enhanced test validity and reliability.

About this research paper

What this paper is about

In this study a comparison was made on DQEST item parameters and test statistics results estimated using Classical Test Theory (CTT) approach and Item Response Theory (IRT) three parameter logistic model (3PLM) to find out the similarities and differences in the two frameworks. 517 randomly selected senior secondary three (SS3) economics students comprised the sample. Three research questions guided the study. Responses obtained from SS3 economics students in 50 multiple choice items of Diagnostic Quantitative Economics Skill Test (DQEST) were used for the analysis. DQEST items certified the unidimensionality, local independence and model-data fit assumptions. Then results from CTT and IRT analyses were compared. In terms of very difficult item and item that discriminate poorly, CTT were found not to be comparable with the 3PLM the most appropriate model for DQEST data. The calculated reliability value for CTT was found to be low when compared to that generated by 3PLM. Therefore, it could be concluded that there was disparity between CTT approach and 3 parameter IRT model in terms of item parameters and test statistics. Thus IRT model with the best data fit should be employed for an enhanced test validity and reliability.

Why it matters

OpenAlex reports 20 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

In this study a comparison was made on DQEST item parameters and test statistics results estimated using Classical Test Theory (CTT) approach and Item Response Theory (IRT) three parameter logistic model (3PLM) to find out the similarities and differences in the two frameworks. 517 randomly selected senior secondary three (SS3) economics students comprised the sample. Three research questions guided the study. Responses obtained from SS3 economics students in 50 multiple choice items of Diagnostic Quantitative Economics Skill Test (DQEST) were used for the analysis. DQEST items certified the unidimensionality, local independence and model-data fit assumptions. Then results from CTT and IRT analyses were compared. In terms of very difficult item and item that discriminate poorly, CTT were found not to be comparable with the 3PLM the most appropriate model for DQEST data. The calculated reliability value for CTT was found to be low when compared to that generated by 3PLM. Therefore, it could be concluded that there was disparity between CTT approach and 3 parameter IRT model in terms of item parameters and test statistics. Thus IRT model with the best data fit should be employed for an enhanced test validity and reliability.

Key concepts: Item response theory, Test (biology), Classical test theory, Econometrics, Test theory, Diagnostic test, Statistics, Economics

Related papers

Back to paper searchBrowse research topicsOriginal source
Comparative Study of Classical Test Theory and Item Response Theory Using Diagnostic Quantitative Economics Skill Test Item Analysis Results — Research Paper | ScholarLens