Directions in abusive language training data, a systematic review: Garbage in, garbage out-Reference-Cited by-同舟云学术

Directions in abusive language training data, a systematic review: Garbage in, garbage out

Published:2020-12-28 Issue:12 Volume:15 Page:e0243300
ISSN:1932-6203
Container-title:PLOS ONE
language:en
Short-container-title:PLoS ONE

Author:

Vidgen Bertie,Derczynski Leon^ORCID

Abstract

Data-driven and machine learning based approaches for detecting, categorising and measuring abusive content such as hate speech and harassment have gained traction due to their scalability, robustness and increasingly high performance. Making effective detection systems for abusive content relies on having the right training datasets, reflecting a widely accepted mantra in computer science: Garbage In, Garbage Out. However, creating training datasets which are large, varied, theoretically-informed and that minimize biases is difficult, laborious and requires deep expertise. This paper systematically reviews 63 publicly available training datasets which have been created to train abusive language classifiers. It also reports on creation of a dedicated website for cataloguing abusive language data hatespeechdata.com. We discuss the challenges and opportunities of open science in this field, and argue that although more dataset sharing would bring many benefits it also poses social and ethical risks which need careful consideration. Finally, we provide evidence-based recommendations for practitioners creating new abusive content training datasets.

Publisher

Public Library of Science (PLoS)

Subject

Multidisciplinary

Reference143 articles.

1. Williams M. Hatred behind the scenes: a report on the rise of online hate speech. London: Mishcon de Reya; 2019.

2. HM Government. Online Harms White Paper. London: Department of Digital, Culture, Media and Society; 2019.

3. Zampieri M, Malmasi S, Nakov P, Rosenthal S, Farra N, Kumar R. SemEval-2019 Task 6: Identifying and Categorizing Offensive Language in Social Media (OffensEval). In: Proceedings of the 13th International Workshop on Semantic Evaluation; 2019. p. 75–86.

4. Cyberhate on social media in the aftermath of Woolwich: a case study in computational criminology and big data;M Williams;British Journal of Criminology,2016

Cited by 91 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Using Explainable AI (XAI) for Identification of Subjectivity in Hate Speech Annotations for Low-Resource Languages;4th International Workshop on OPEN CHALLENGES IN ONLINE SOCIAL NETWORKS;2024-09-10

2. From hate to harmony: Leveraging large language models for safer speech in times of COVID-19 crisis;Heliyon;2024-08

3. The Role of Gender in Hate Speech Targeting Politicians: Evidence from Finnish Twitter;International Journal of Politics, Culture, and Society;2024-06-13

4. The Unseen Targets of Hate: A Systematic Review of Hateful Communication Datasets;Social Science Computer Review;2024-06-13

5. User-Centric Modeling of Online Hate Through the Lens of Psycholinguistic Patterns and Behaviors in Social Media;IEEE Transactions on Computational Social Systems;2024-06