What Determines Inter-Coder Agreement in Manual Annotations? A Meta-Analytic Investigation-Reference-Cited by-同舟云学术

What Determines Inter-Coder Agreement in Manual Annotations? A Meta-Analytic Investigation

Published:2011-12 Issue:4 Volume:37 Page:699-725
ISSN:0891-2017
Container-title:Computational Linguistics
language:en
Short-container-title:Computational Linguistics

Author:

Bayerl Petra Saskia¹,Paul Karsten Ingmar²

Affiliation:

1. Erasmus University

2. University of Erlangen-Nuremberg

Abstract

Recent discussions of annotator agreement have mostly centered around its calculation and interpretation, and the correct choice of indices. Although these discussions are important, they only consider the “back-end” of the story, namely, what to do once the data are collected. Just as important in our opinion is to know how agreement is reached in the first place and what factors influence coder agreement as part of the annotation process or setting, as this knowledge can provide concrete guidelines for the planning and set-up of annotation projects. To investigate whether there are factors that consistently impact annotator agreement we conducted a meta-analytic investigation of annotation studies reporting agreement percentages. Our meta-analysis synthesized factors reported in 96 annotation studies from three domains (word-sense disambiguation, prosodic transcriptions, and phonetic transcriptions) and was based on a total of 346 agreement indices. Our analysis identified seven factors that influence reported agreement values: annotation domain, number of categories in a coding scheme, number of annotators in a project, whether annotators received training, the intensity of annotator training, the annotation purpose, and the method used for the calculation of percentage agreements. Based on our results we develop practical recommendations for the assessment, interpretation, calculation, and reporting of coder agreement. We also briefly discuss theoretical implications for the concept of annotation quality.

Publisher

MIT Press - Journals

Subject

Artificial Intelligence,Computer Science Applications,Linguistics and Language,Language and Linguistics

Link

https://www.mitpressjournals.org/doi/pdf/10.1162/COLI_a_00074

Reference28 articles.

1. Inter-Coder Agreement for Computational Linguistics

2. From Annotator Agreement to Noise Models

3. How meta-analysis increases statistical power.

4. Evaluating Discourse and Dialogue Coding Schemes

Cited by 63 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Language-based machine perception: linguistic perspectives on the compilation of captioning datasets;Digital Scholarship in the Humanities;2024-06-21

2. Taming our Wild Data;Dutch Journal of Applied Linguistics;2024-03-26

3. Extracting informational cues between initial coin offering projects and the public;Psychology & Marketing;2024-02-14

4. The Database of Constructions with Lexical Repetitions “RepLeCon” and Inter-Annotator Agreement;Springer Geography;2024

5. Subject Cataloging by Norwegian Cataloging Agencies;Cataloging & Classification Quarterly;2023-12-29