Eliciting Confidence for Improving Crowdsourced Audio Annotations-Reference-Cited by-同舟云学术

Eliciting Confidence for Improving Crowdsourced Audio Annotations

Published:2022-03-30 Issue:CSCW1 Volume:6 Page:1-25
ISSN:2573-0142
Container-title:Proceedings of the ACM on Human-Computer Interaction
language:en
Short-container-title:Proc. ACM Hum.-Comput. Interact.

Author:

Méndez Méndez Ana Elisa¹,Cartwright Mark²,Bello Juan Pablo¹,Nov Oded¹

Affiliation:

1. New York University, New York, NY, USA

2. New Jersey Institute of Technology & New York University, Newark, NJ, USA

Abstract

In this work we explore confidence elicitation methods for crowdsourcing "soft" labels, e.g., probability estimates, to reduce the annotation costs for domains with ambiguous data. Machine learning research has shown that such "soft" labels are more informative and can reduce the data requirements when training supervised machine learning models. By reducing the number of required labels, we can reduce the costs of slow annotation processes such as audio annotation. In our experiments we evaluated three confidence elicitation methods: 1) "No Confidence" elicitation, 2) "Simple Confidence" elicitation, and 3) "Betting" mechanism for confidence elicitation, at both individual (i.e., per participant) and aggregate (i.e., crowd) levels. In addition, we evaluated the interaction between confidence elicitation methods, annotation types (binary, probability, and z-score derived probability), and "soft" versus "hard" (i.e., binarized) aggregate labels. Our results show that both confidence elicitation mechanisms result in higher annotation quality than the "No Confidence" mechanism for binary annotations at both participant and recording levels. In addition, when aggregating labels at the recording level, results indicate that we can achieve comparable results to those with 10-participant aggregate annotations using fewer annotators if we aggregate "soft" labels instead of "hard" labels. These results suggest that for binary audio annotation using a confidence elicitation mechanism and aggregating continuous labels we can obtain higher annotation quality, more informative labels, with quality differences more pronounced with fewer participants. Finally, we propose a way of integrating these confidence elicitation methods into a two-stage, multi-label annotation pipeline.

Funder

National Science Foundation

Publisher

Association for Computing Machinery (ACM)

Subject

Computer Networks and Communications,Human-Computer Interaction,Social Sciences (miscellaneous)

Link

https://dl.acm.org/doi/pdf/10.1145/3512935

Reference79 articles.

1. Active Learning and Crowd-Sourcing for Machine Translation;Ambati Vamshi;LREC,2010

2. Truth Is a Lie: Crowd Truth and the Seven Myths of Human Annotation

3. Immediate Feedback and Opportunity to Revise Answers to Open-Ended Questions

4. Effects of feedback elaboration and feedback timing during computer-based practice in mathematics problem solving

5. Bayesian Aggregation of Categorical Distributions with Applications in Crowdsourcing

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. RCTD: Reputation-Constrained Truth Discovery in Sybil Attack Crowdsourcing Environment;Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining;2024-08-24