2018Chinese Journal of Medical EducationRequires access

Quality analysis of national English test for doctor of medicine based on two kinds of educational measurement theory

Zhang Quanhui

Open publisher page 0 citations

Abstract

Objective To analyze the quality of national English test for doctor of medicine from 2015 to 2017 using the Rasch model of item response theory (IRT) and the classical test theory (CTT), then compare the consistency of the results. Methods 40 607 papers of national English test for doctor of medicine from 2015 to 2017 were selected as the research materials. The reliabilities, difficulties and discriminations of the examinations were analyzed using the Rasch model of IRT and CTT respectively. Then the consistency of the results was analyzed. Results IRT analysis results showed that the parameters fit well with the Rasch model, the maximum information functions were all>25.00, the estimated errors were all 0.80, the difficulties were all between 0.31 and 0.63, and the discriminations were all between 0.12 and 0.29. The results of the two theoretical analyses were highly correlated. Conclusions The three years' tests can evaluate the application abilities of candidates' medical English well. In terms of parameter estimation and error analysis, Rasch model is more accurate than CTT, and Rasch model should be used more in the subsequent examination evaluations. Key words: National English test for doctor of medicine; Rasch model; Classical test theory

About this research paper

What this paper is about

Objective To analyze the quality of national English test for doctor of medicine from 2015 to 2017 using the Rasch model of item response theory (IRT) and the classical test theory (CTT), then compare the consistency of the results. Methods 40 607 papers of national English test for doctor of medicine from 2015 to 2017 were selected as the research materials. The reliabilities, difficulties and discriminations of the examinations were analyzed using the Rasch model of IRT and CTT respectively. Then the consistency of the results was analyzed. Results IRT analysis results showed that the parameters fit well with the Rasch model, the maximum information functions were all>25.00, the estimated errors were all 0.80, the difficulties were all between 0.31 and 0.63, and the discriminations were all between 0.12 and 0.29. The results of the two theoretical analyses were highly correlated. Conclusions The three years' tests can evaluate the application abilities of candidates' medical English well. In terms of parameter estimation and error analysis, Rasch model is more accurate than CTT, and Rasch model should be used more in the subsequent examination evaluations. Key words: National English test for doctor of medicine; Rasch model; Classical test theory

Why it matters

A significance statement is not available in the OpenAlex record.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Objective To analyze the quality of national English test for doctor of medicine from 2015 to 2017 using the Rasch model of item response theory (IRT) and the classical test theory (CTT), then compare the consistency of the results. Methods 40 607 papers of national English test for doctor of medicine from 2015 to 2017 were selected as the research materials. The reliabilities, difficulties and discriminations of the examinations were analyzed using the Rasch model of IRT and CTT respectively. Then the consistency of the results was analyzed. Results IRT analysis results showed that the parameters fit well with the Rasch model, the maximum information functions were all>25.00, the estimated errors were all 0.80, the difficulties were all between 0.31 and 0.63, and the discriminations were all between 0.12 and 0.29. The results of the two theoretical analyses were highly correlated. Conclusions The three years' tests can evaluate the application abilities of candidates' medical English well. In terms of parameter estimation and error analysis, Rasch model is more accurate than CTT, and Rasch model should be used more in the subsequent examination evaluations. Key words: National English test for doctor of medicine; Rasch model; Classical test theory

Key concepts: Rasch model, Classical test theory, Item response theory, Consistency (knowledge bases), Test (biology), Test theory, Polytomous Rasch model, Psychology

Related papers

Back to paper searchBrowse research topicsOriginal source
Quality analysis of national English test for doctor of medicine based on two kinds of educational measurement theory — Research Paper | ScholarLens