AIExplainer
Machine Learning Intermediate 2 min read

What is inter-rater agreement?

A measure of how much two or more raters agree on their assessments or ratings

Inter-rater agreement is a statistical method used to evaluate the consistency of ratings or assessments made by different people, often used in research and evaluation studies to ensure that results are reliable and consistent

Imagine two movie critics watching the same film and giving it a rating out of 10 - if they both give it a similar rating, there is high inter-rater agreement, but if one gives it a 5 and the other a 9, there is low inter-rater agreement

A hospital might use inter-rater agreement to evaluate the consistency of diagnoses made by different doctors, to ensure that patients are receiving accurate and reliable care

Inter-rater agreement is used in various fields such as psychology, education, and healthcare to assess the reliability of measurements, evaluations, and assessments made by different raters or observers

One common misconception is that high inter-rater agreement always means that the ratings are accurate, when in fact it only measures the consistency of the ratings, not their accuracy

The concept of inter-rater agreement has been around since the early 20th century, but it wasn't until the 1970s and 1980s that statistical methods for measuring it were developed

inter-rater reliability inter-observer agreement rater consistency

Three products for different needs — explore what’s relevant to you.