A comparison of computer adaptive tests (CATs) and short forms in terms of accuracy and number of items administrated using PROMIS profile

A comparison of computer adaptive tests (CATs) and short forms in terms of accuracy and number of items administrated using PROMIS profile
复制标题

DOI:
10.1007/s11136-019-02312-8
复制
发表时间:
2020-01-01
影响因子:
3.5
通讯作者:
Cella, David
Cella, David
中科院分区:
医学2区
文献类型:
--
作者:
Segawa, Eisuke;Schalet, Benjamin;Cella, David

文献摘要

被引文献

相似文献

目的在患者报告的结果测量信息系统(PROMIS)中,七个领域(身体功能,焦虑,抑郁,疲劳,睡眠障碍,社会功能和疼痛干扰)被打包在一起作为配置文件。这些域中的每一个也可以使用计算机自适应测试(CAT)或不同长度的简短形式(SF)(例如,4、6和8项)。我们比较了CAT与每个SF的准确性和管理的项目数。方法PROMIS量表采用项目反应理论(IRT)分级反应模型进行评分,以T分(均值50分,标准差10分)表示。我们从正态分布中模拟了10,000名受试者,症状量表的平均值为60,功能量表为40,每个领域的标准差为10。当标准误(SE)小于3.0时,我们认为受试者的评分是准确的。我们记录了准确分数的范围(准确范围)和管理的项目数量。结果CAT各领域的平均条目数为4.7个。CAT的准确范围比每个域中的所有SF更宽。CAT在将准确范围扩展到疲劳、身体功能和疼痛干扰的非常差的健康状况方面明显更好。大多数SF提供了相当宽的准确范围。结论相对于SFs,CAT提供了最广泛的准确范围,比SF4略多,比SF6和SF8少。大多数SF,特别是较长的SF,提供了相当宽的准确范围。
Purpose In the Patient-Reported Outcomes Measurement Information System (PROMIS), seven domains (Physical Function, Anxiety, Depression, Fatigue, Sleep Disturbance, Social Function, and Pain Interference) are packaged together as profiles. Each of these domains can also be assessed using computer adaptive tests (CATs) or short forms (SFs) of varying length (e.g., 4, 6, and 8 items). We compared the accuracy and number of items administrated of CAT versus each SF. Methods PROMIS instruments are scored using item response theory (IRT) with graded response model and reported as T scores (mean = 50, SD = 10). We simulated 10,000 subjects from the normal distribution with mean 60 for symptom scales and 40 for function scales, and standard deviation 10 in each domain. We considered a subject's score to be accurate when the standard error (SE) was less than 3.0. We recorded range of accurate scores (accurate range) and the number of items administrated. Results The average number of items administrated in CAT was 4.7 across all domains. The accurate range was wider for CAT compared to all SFs in each domain. CAT was notably better at extending the accurate range into very poor health for Fatigue, Physical Function, and Pain Interference. Most SFs provided reasonably wide accurate range. Conclusions Relative to SFs, CATs provided the widest accurate range, with slightly more items than SF4 and less than SF6 and SF8. Most SFs, especially longer ones, provided reasonably wide accurate range.