检索规则说明:AND代表“并且”;OR代表“或者”;NOT代表“不包含”;(注意必须大写,运算符两边需空一格)
检 索 范 例 :范例一: (K=图书馆学 OR K=情报学) AND A=范并思 范例二:J=计算机应用与软件 AND (U=C++ OR U=Basic) NOT M=Visual
作 者:丁树良[1] 吴锐[1] 张节兰[1] 熊建华[1]
机构地区:[1]江西师范大学计算机信息工程学院,南昌330022,江西师范大学鹰潭学院信息技术系,鹰潭335000
出 处:《心理学报》2008年第1期101-108,共8页Acta Psychologica Sinica
基 金:国家自然科学基金(60263005);江西省自然科学基金(0411021);省科技厅攻关项目;省教育厅科技项目;省高校人文社科研究项目(JY06201);全国教育考试科研规划项目(2006JKS3063);卫生部课题(JM20060070,KY200704)资助
摘 要:在项目反应理论框架下,根据已有文献提出了开发新的测验等值准则的方法,即许多准则都可以看成是通过对锚题上作答反应概率分布进行变换而导出。据此揭示了两个著名的等值准则——Haebara方法和Stock ing-Lord方法之间的联系,并且导出了一个新的等值准则——余弦等值准则。为了讨论余弦准则的行为表现,开展了一系列Monte-Carlo模拟研究。模拟结果表明,余弦准则在多级评分模型GPCM上表现比Haebara方法和Stock-ing--Lord方法都好,而对GRM和2PLM,其表现不如Haebara,但可以和Stock ing-Lord方法相提并论。这一发现提醒我们等值准则的选用是否恰当,不仅与等值系数所落的范围有关,而且还与项目反应函数(IRF)有更密切的关系。This paper, divided into two parts, discusses the following two issues: (1) the methodology of developing a new test equating criterion and (2) the behavior of a new test equating method, referred to as cosine criterion. Under the item response theory (IRT) and in light of the probability distribution of an examinee's response to some item, the fast part of this paper proposes the methodology derived from the published literature on some test equating criteria. Moreover, some test equating criteria could be regarded as certain functions of probability distributions. Based on this, a series of test equating approaches, such as the Haebara item characteristic curve equating method (Hcrit), Stocking - Lord test characteristic curve equating method (SLcrit), logcontract equating method, SQRT method, and weighted Haebara method, could be clearly illustrated. Further, the relationship between Hcrit and SLcrit was identified: if the mutual compensation of the responses to the anchor items is evident, then SLcrit is suitable, and if not, then Hcrit is more appropriate. In the second part of the paper, a new test equating criterion, known as cosine criterion (COScrit) was discussed as an example of the application of this methodology of the equating criteria. The results of the Monte Carlo study show that the behavior of the new criterion is better than that of Herit and SLcrit; this is evident when the data is fit to the generalized partial credit model (GPCM) in the sense that the root mean squared deviations (RMSDs) corresponding to the three criteria are compared. Further, the RMSD to COSerit is smaller and statistically significant. When the data is fit to the 2 - parameter logistic model 2PLM, or the graded response model (GRM),COSerit is comparable to SLcrit; in fact, it is considerably better than SLcrit, provided that the equating coefficient A is not smaller than 1.2. If, however, coefficient A is smaller than 12,an inverse result is observed. Nevertheless, COScrit is
分 类 号:B841[哲学宗教—基础心理学]
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在链接到云南高校图书馆文献保障联盟下载...
云南高校图书馆联盟文献共享服务平台 版权所有©
您的IP:216.73.216.38