Kappa coefficients in medical research

被引:396
作者
Kraemer, HC
Periyakoil, VS
Noda, A
机构
[1] Stanford Univ, Dept Psychiat & Behav Sci, Sch Med, Stanford, CA 94305 USA
[2] VA Palo Alto Hlth Care Syst, Palo Alto, CA USA
关键词
kappa; reliability; validity; consensus;
D O I
10.1002/sim.1180
中图分类号
Q [生物科学];
学科分类号
07 ; 0710 ; 09 ;
摘要
Kappa coefficients are measures of correlation between categorical variables often used as reliability or validity coefficients. We recapitulate development and definitions of the K (categories) by M (ratings) kappas (K x M), discuss what they are well- or ill-designed to do, and summarize where kappas now stand with regard to their application in medical research. The 2 x M (M greater than or equal to 2) intraclass kappa seems the ideal measure of binary reliability; a 2 x 2 weighted kappa is an excellent choice, though not a unique one, as a validity measure. For both the intraclass and weighted kappas, we address continuing problems with kappas. There are serious problems with using the K x M intraclass (K > 2) or the various K x M weighted kappas for K > 2 or M > 2 in any context, either because they convey incomplete and possibly misleading information, or because other approaches are preferable to their use. We illustrate the use of the recommended kappas with applications in medical research. Copyright (C) 2002 John Wiley Sons, Ltd.
引用
收藏
页码:2109 / 2129
页数:21
相关论文
共 68 条
[1]   MAXIMUM-LIKELIHOOD-ESTIMATION OF AGREEMENT IN THE CONSTANT PREDICTIVE PROBABILITY MODEL, AND ITS RELATION TO COHEN KAPPA [J].
AICKIN, M .
BIOMETRICS, 1990, 46 (02) :293-302
[2]  
[Anonymous], 1972, The dependability of behaviourial measurements: Theory of generalzsability for scores and profiles
[3]   Beyond kappa: A review of interrater agreement measures [J].
Banerjee, M .
CANADIAN JOURNAL OF STATISTICS-REVUE CANADIENNE DE STATISTIQUE, 1999, 27 (01) :3-23
[4]   METHODS AND THEORY OF RELIABILITY [J].
BARTKO, JJ ;
CARPENTER, WT .
JOURNAL OF NERVOUS AND MENTAL DISEASE, 1976, 163 (05) :307-317
[5]  
Blackman NJM, 2000, STAT MED, V19, P723, DOI 10.1002/(SICI)1097-0258(20000315)19:5<723::AID-SIM379>3.0.CO
[6]  
2-A
[7]   2X2 KAPPA-COEFFICIENTS - MEASURES OF AGREEMENT OR ASSOCIATION [J].
BLOCH, DA ;
KRAEMER, HC .
BIOMETRICS, 1989, 45 (01) :269-287
[8]   Hypothesis testing and effect size estimation in clinical trials [J].
Borenstein, M .
ANNALS OF ALLERGY ASTHMA & IMMUNOLOGY, 1997, 78 (01) :5-11
[9]  
Breiman L., 1984, BIOMETRICS, DOI DOI 10.2307/2530946
[10]   COEFFICIENT KAPPA - SOME USES, MISUSES, AND ALTERNATIVES [J].
BRENNAN, RL ;
PREDIGER, DJ .
EDUCATIONAL AND PSYCHOLOGICAL MEASUREMENT, 1981, 41 (03) :687-699