Accounting for overlap?: An application of Mezzich's κ statistic to test interrater reliability of interview data on parental accident and emergency attendance
Accounting for overlap?: An application of Mezzich's κ statistic to test interrater reliability of interview data on parental accident and emergency attendance
复制标题
DOI:
10.1046/j.1365-2648.2001.01718.x
复制
发表时间:
2001-03-01
影响因子:
3.8
通讯作者:
MacFaul, R
中科院分区:
文献类型:
--
作者:
Eccleston, P;Werneke, U;MacFaul, R
Study rationale. The number of interview studies with service users is rising because of growth in health services research. The level of agreement between multiple interview data coders requires statistical calculation to support results. Basic kappa statistics are often used hut this defends on having mutually exclusive data. Researchers should be aware that this is not valid when an interview word or paragraph can be coded into more than one category. The 'proportional overlap' kappa extension by Mezzich et al. (1981, Journal of Psychiatric Research 16, 29-39) has been investigated as an original solution.Objectives. To assess the level of agreement beyond chance between several raters of interview data by applying the 'proportional overlap' kappa statistic by Mezzich et al. to verbal interview data. The clinical area investigated was child attendance at an Accident and Emergency Department, where parental attendance experiences have been under-explored.Methods. Two researchers using a coding schedule coded a random sample of interview transcripts. These data were applied to Mezzich's procedure; coder 1 notes that a paragraph refers to category A and B but coder 2 notes A, B and C. The total agreement overlap in this case was 0.66 because two actual agreements out of three possible agreements were made. This was repeated fur each paragraph and divided by the number of coding pairs. All agreement values were summed then subsequently divided by the total number of paragraphs to get P-o (total number of observed agreements) and by the total number of coding pairs to get P-e (total number of agreements by chance alone). P-o and P-e were used in the basic IC formula to assess interview coding reliability.Results. The overall mean P-o was 0.61, the mean P-e was 0.32, with a kappa score of 0.43; a moderate level of agreement which was: statistically significant (t = 4.8, P < 0.001, d.f. = 23).Conclusion. Mezzich's procedure may be applied to interview data to calculate agreement levels between several coders.