login

Beyond kappa: A review of interrater agreement measures

Canadian Journal of StatisticsPublished 1 March 1999
Mousumi Banerjee, Michelle C. Capozzoli, Laura A. McSweeney, Debajyoti Sinha
Citations979
SJR quartileQ2
SJR score0.59
SNIP0.91

Abstract

Abstract In 1960, Cohen introduced the kappa coefficient to measure chance‐corrected nominal scale agreement between two raters. Since then, numerous extensions and generalizations of this interrater agreement measure have been proposed in the literature. This paper reviews and critiques various approaches to the study of interrater agreement, for which the relevant data comprise either nominal or ordinal categorical ratings from multiple raters. It presents a comprehensive compilation of the main statistical approaches to this problem, descriptions and characterizations of the underlying models, and discussions of related statistical methodologies for estimation and confidence‐interval construction. The emphasis is on various practical scenarios and designs that underlie the development of these measures, and the interrelationships between them.

Keywords

Decision Sciences