Measures of reliability in sports medicine and science

Measures of reliability in sports medicine and science
复制标题

DOI:
10.2165/00007256-200030010-00001
复制
发表时间:
2000-07-01
期刊:
影响因子:
9.8
通讯作者:
Hopkins, WG
Hopkins, WG
中科院分区:
医学1区
文献类型:
--
作者:
Hopkins, WG

文献摘要

被引文献

相似文献

可靠性是指在同一个体的重复试验中,测试、化验或其他测量的值的重复性。更好的可靠性意味着更高的单次测量精度和更好地跟踪研究或实际环境中测量的变化。信度的主要测量是受试者内的随机变化、平均数的系统变化和重测相关性。受试者内部变化的一种简单、适应性的形式是典型的(标准)测量误差:个人重复测量的标准偏差。对于运动医学和科学中的许多测量,典型的误差最好用变异系数(平均值的百分比)来表示。受试者内部差异的一种有偏见的、更有限的形式是一致的限度:两次试验之间个体测量结果变化的95%的可能范围。连续试验之间测量平均值的系统性变化代表了学习、动机或疲劳等影响;这些变化需要从受试者内部差异的估计中剔除。重测相关性很难解释,主要是因为它的值对参与者样本的异质性很敏感。可靠性的用途包括监测个体时的决策、测试或设备的比较、实验中样本大小的估计以及对治疗反应的个体差异大小的估计。估计可靠性的合理精确度需要大约50名研究参与者和至少3次试验。旨在评估测试或设备之间可靠性差异的研究需要复杂的设计和分析,而研究人员很少正确执行这些设计和分析。更广泛地了解可靠性,并采用典型误差作为可靠性的标准衡量标准,将改善对我们学科中的测试和设备的评估。
Reliability refers to the reproducibility of values of a test, assay or other measurement in repeated trials on the same individuals. Better reliability implies better precision of single measurements and better tracking of changes in measurements in research or practical settings. The main measures of reliability are within-subject random variation, systematic change in the mean, and retest correlation. A simple, adaptable form of within-subject variation is the typical (standard) error of measurement: the standard deviation of an individual's repeated measurements. For many measurements in sports medicine and science, the typical error is best expressed as a coefficient of variation (percentage of the mean). A biased, more limited form of within-subject variation is the limits of agreement: the 95% likely range of change of an individual's measurements between 2 trials. Systematic changes in the mean of a measure between consecutive trials represent such effects as learning, motivation or fatigue; these changes need to be eliminated from estimates of within-subject variation. Retest correlation is difficult to interpret, mainly because its value is sensitive to the heterogeneity of the sample of participants. Uses of reliability include decision-making when monitoring individuals, comparison of tests or equipment, estimation of sample size in experiments and estimation of the magnitude of individual differences in the response to a treatment. Reasonable precision for estimates of reliability requires approximately 50 study participants and at least 3 trials. Studies aimed at assessing variation in reliability between tests or equipment require complex designs and analyses that researchers seldom:perform correctly. A wider understanding of reliability and adoption of the typical error as the standard measure of reliability would improve the assessment of tests and equipment in our disciplines.