Abstract
In recent years, the kappa coefficient of agreement has become the de facto standard for evaluating intercoder agreement for tagging tasks. In this squib, we highlight issues that affect κ and that the community has largely neglected. First, we discuss the assumptions underlying different computations of the expected agreement component of κ. Second, we discuss how prevalence and bias affect the κ measure.
Cite
CITATION STYLE
APA
Eugenio, B. D., & Glass, M. (2004). The Kappa Statistic: A Second Look. Computational Linguistics, 30(1), 95–101. https://doi.org/10.1162/089120104773633402
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.
Already have an account? Sign in
Sign up for free