Large-Sample Variance of Fleiss Generalized Kappa

被引:13
|
作者
Gwet, Kilem L. [1 ]
机构
[1] AgreeStat Analyt, POB 2696, Gaithersburg, MD 20886 USA
关键词
Fleiss kappa; Cohen kappa; interrater reliability; Gwet AC1;
D O I
10.1177/0013164420973080
中图分类号
G44 [教育心理学];
学科分类号
0402 ; 040202 ;
摘要
Cohen's kappa coefficient was originally proposed for two raters only, and it later extended to an arbitrarily large number of raters to become what is known as Fleiss' generalized kappa. Fleiss' generalized kappa and its large-sample variance are still widely used by researchers and were implemented in several software packages, including, among others, SPSS and the R package "rel." The purpose of this article is to show that the large-sample variance of Fleiss' generalized kappa is systematically being misused, is invalid as a precision measure for kappa, and cannot be used for constructing confidence intervals. A general-purpose variance expression is proposed, which can be used in any statistical inference procedure. A Monte-Carlo experiment is presented, showing the validity of the new variance estimation procedure.
引用
收藏
页码:781 / 790
页数:10
相关论文
共 50 条