IRT Test Equating: Relevant Issues and a Review of Recent Research

IRT Test Equating: Relevant Issues and a Review of Recent Research
复制标题

IRT 测试等同化:相关问题和近期研究回顾

DOI:
10.3102/00346543056004495
复制
发表时间:
1986
影响因子:
11.2
通讯作者:
R. Lissitz
R. Lissitz
中科院分区:
教育学1区
文献类型:
--
作者:
G. Skaggs;R. Lissitz

文献摘要

被引文献

相似文献

项目反应理论(IRT)在等值测验中的应用是近二十年来的一个研究热点。尽管研究量很大,但很难得出结论和做出概括,因为不同的研究使用了不同类型的测试,不同类型的样本,以及不同的方法来评估等同结果的准确性。本文的目的有三个:(a)回顾迄今为止的一些主要研究并综合其结果,(B)讨论哪些问题尚未回答,以及研究方法存在哪些问题,(c)为未来的研究提供方向。而早期的研究集中在比较等值方法和IRT模型,最近的研究已经解决了这样的统计问题,如等值的标准误差,参数稳定性和IRT模型的鲁棒性违反他们的假设。到目前为止,研究的一个主要发现是,期望一种等同方法为所有类型的测试提供最佳结果是不合理的。未来的研究必须确定条件,如多维性和测试内容,影响IRT等值。
The application of item response theory (IRT) methodology to test equating has been a research topic of considerable interest in the past 2 decades. Despite the volume of research, it has been difficult to draw conclusions and make generalizations because different studies have used different types of tests, different types of samples, and different methods for assessing the accuracy of equating results. The purpose of this paper is threefold: (a) to review some of the major studies thus far and synthesize their results, (b) to discuss what questions are as yet unanswered and what problems exist with research methodology, and (c) to provide direction for future research. Whereas earlier research focused on comparing equating methods and IRT models, recent research has addressed such statistical concerns as standard errors of equating, parameter stability, and robustness of IRT models to violations of their assumptions. A major finding from the research so far is that it is unreasonable to expect a single equating method to provide the best results for equating all types of tests. Future research must determine how conditions, such as multidimensionality and test content, affect IRT equating.