A comparative study of equating methods applied in standardized competence test for clinical medicine undergraduates

  • Zhang Quanhui ,
  • He Ju ,
  • Ren Jie ,
  • Zhang Ying ,
  • Lu Yan
Expand
  • 1Department of Information and Assessment, National Medicine Examination Center, Being 100097, China;
    2National Medicine Examination Center, Being 100097,China;
    3Institute of Language Testing and Talent Evaluation, Beijing Language and Culture University, Being 100083, China;
    4Department of Examination Management, National Medicine Examination Center, Being 100097, China;
    5Department of Development Research, National Medicine Examination Center, Being 100097, China

Received date: 2021-08-17

  Online published: 2022-06-29

Abstract

Objective This paper analyzes equating methods applied in Standardized Competence Test for undergraduates of clinical medicine based on classical test theory (CTT) and item response theory (IRT) in order to explore a more suitable equating method. Methods The research uses four equating methods based on the CTT and six equating methods based on the IRT.CTT equating methods include Tucker observation score linear equating method,Levine observation score linear equating method, equipercentile equating smoothing method and equating standard error equating unsmoothed method. While in the one-parameter model and two-parameter model of IRT, three calibration methods are used which are linking separate calibration, concurrent calibration and fixed Item Parameter Calibration. The stability of the 10 equating results is analyzed by the equating standard error. Results The results show that the equating standard error of CTT method is 0.7~1.6, while the equating standard error of IRT method is 0.2~0.6, IRT equating standard error is smaller than CTT equating method. Among four CTT equating methods, the equating standard error of Tucker observation score linear equating method is 0.7 as the smallest one, the error of equipercentile equating method is 1.6 as the largest one. Among six IRT equating methods, the result of one-parameter model is better than that of two-parameter model and the error of fixed item parameter calibration is the smallest one in one-parameter model, which the equating standard error is 0.2. Conclusions The fixed item parameter calibration in one-parameter model of IRT can be selected as the equating method of this test. Through equating, the score of year 2 is improved, and the eligibility criteria remain unchanged, which effectively achieves the score comparability and ensures the fairness of the test.

Cite this article

Zhang Quanhui , He Ju , Ren Jie , Zhang Ying , Lu Yan . A comparative study of equating methods applied in standardized competence test for clinical medicine undergraduates[J]. Chinese Journal of Medical Education, 2022 , 42(7) : 577 -580 . DOI: 10.3760/cma.j.cn115259-20210817-01034

References

[1] 杨志明.线性等值与等百分位等值的实施条件与步骤[J].教育测量与评价,2016(12):4-8,34. DOI:10.16518/j.cnki.emae.2016.12.002.
[2] 杨志明.学业水平考试事后等值的概念、条件与设计[J].教育测量与评价,2016,(11):4-8. DOI:10.16518/j.cnki.emae.2016.11.002.
[3] 张健,任杰.基于共同题非等组设计的等值结果评价标准研究综述[J].中国考试,2018(3):32-37. DOI:10.19360/j.cnki.11-3303/g4.2018.03.008.
[4] 范晓玲,欧阳淑兰,卢谢峰,等.高考数学试卷不同等值方法的比较研究[J].教育测量与评价,2018(10):47-55.DOI:10.16518/j.cnki.emae.2018.10.008.
[5] 王晓慧.等值问题在大规模考试中的研究——以上海市信息科技学业考等值研究为例[J].上海教育评估研究,2018,7(5):46-50.DOI:10.13794/j.cnki.shjee.2018.0075.
[6] 张飘.IRT框架下对含题组的语文阅读测验的等值分析[J].中国考试,2019(4):29-34. DOI:10.19360/j.cnki.11-3303/g4.2019.04.006.
[7] 庞海玉,康琳,刘雅茹.基于项目反应理论的老年医学知信行量表条目分析与评价[J].基础医学与临床,2019,39(8):1108-1113. DOI:10.3969/j.issn.1001-6325.2019.08.007.
[8] 庄然,郑淑园,田甜,等.项目反应理论在基础医学综合测试免疫学试题中的应用[J].细胞与分子免疫学杂志,2020,36(1):86-94.DOI:10.13423/j.cnki.cjcmi.008962.
Outlines

/