2012•Biochemia MedicaOpen access

Interrater reliability: the kappa statistic

Marry L. McHugh

Open full text 19,388 citations

Abstract

Many situations in the healthcare industry rely on multiple people to collect research or clinical laboratory data .The question of consistency, or agreement among the individuals collecting data immediately arises due to the variability among human observers .Well-designed research studies must therefore include procedures that measure agreement among the various data collectors .Study designs typically involve training the data collectors, and measuring the extent to which they record the same scores for the same phenomena .Perfect agreement is seldom achieved, and confidence in study results is partly a function of the amount of disagreement, or error introduced into the study from inconsistency among the data collectors .The extent of agreement among data collectors is called, "interrater reliability" .Interrater reliability is a concern to one degree or another in most large studies due to the fact that multiple people collecting data may experience and interpret the phenomena of interest differently .

Open-access reader

About this research paper

What this paper is about

Many situations in the healthcare industry rely on multiple people to collect research or clinical laboratory data .The question of consistency, or agreement among the individuals collecting data immediately arises due to the variability among human observers .Well-designed research studies must therefore include procedures that measure agreement among the various data collectors .Study designs typically involve training the data collectors, and measuring the extent to which they record the same scores for the same phenomena .Perfect agreement is seldom achieved, and confidence in study results is partly a function of the amount of disagreement, or error introduced into the study from inconsistency among the data collectors .The extent of agreement among data collectors is called, "interrater reliability" .Interrater reliability is a concern to one degree or another in most large studies due to the fact that multiple people collecting data may experience and interpret the phenomena of interest differently .

Why it matters

OpenAlex reports 19388 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Many situations in the healthcare industry rely on multiple people to collect research or clinical laboratory data .The question of consistency, or agreement among the individuals collecting data immediately arises due to the variability among human observers .Well-designed research studies must therefore include procedures that measure agreement among the various data collectors .Study designs typically involve training the data collectors, and measuring the extent to which they record the same scores for the same phenomena .Perfect agreement is seldom achieved, and confidence in study results is partly a function of the amount of disagreement, or error introduced into the study from inconsistency among the data collectors .The extent of agreement among data collectors is called, "interrater reliability" .Interrater reliability is a concern to one degree or another in most large studies due to the fact that multiple people collecting data may experience and interpret the phenomena of interest differently .

Key concepts: Inter-rater reliability, Kappa, Cohen's kappa, Statistics, Statistic, Reliability (semiconductor), Mathematics, Test (biology)

Related papers

Back to paper searchBrowse research topicsOriginal source
Interrater reliability: the kappa statistic — Research Paper | ScholarLens