Skip to content

The Measurement of Observer Agreement for Categorical Data

John Richard Landis, Gary G. Koch

Biometrics · 1977 · 81,688 citationsOpen access

Abstract

This paper presents a general statistical methodology for the analysis of multivariate categorical data arising from observer reliability studies. The procedure essentially involves the construction of functions of the observed proportions which are directed at the extent to which the observers agree among themselves and the construction of test statistics for hypotheses involving these functions. Tests for interobserver bias are presented in terms of first-order marginal homogeneity and measures of interobserver agreement are developed as generalized kappa-type statistics. These procedures are illustrated with a clinical diagnosis example from the epidemiological literature.

Cite this paper

Landis, J. R., & Koch, G. G. (1977). The measurement of observer agreement for categorical data. Biometrics, 33(1), 159. https://doi.org/10.2307/2529310

Read it with every claim anchored

Add this paper to a project, ask questions of it, and get answers that point to the exact passage.

Start free
  1. A Guideline of Selecting and Reporting Intraclass Correlation Coefficients for Reliability Research2016
  2. Intraclass correlations: Uses in assessing rater reliability.1979
  3. The Strengthening the Reporting of Observational Studies in Epidemiology (STROBE) Statement: Guidelines for Reporting Observational Studies2007
  4. Meta-analysis of Observational Studies in Epidemiology A Proposal for Reporting2000
  5. The Strengthening the Reporting of Observational Studies in Epidemiology (STROBE) Statement: Guidelines for Reporting Observational Studies2007

Metadata from OpenAlex (CC0). Citations are generated from the published record.