A machine learning approach to detect potentially harmful and protective suicide-related content in broadcast media-Reference-Cited by-同舟云学术

A machine learning approach to detect potentially harmful and protective suicide-related content in broadcast media

Published:2024-05-14 Issue:5 Volume:19 Page:e0300917
ISSN:1932-6203
Container-title:PLOS ONE
language:en
Short-container-title:PLoS ONE

Author:

Metzler Hannah,Baginski Hubert,Garcia David^ORCID,Niederkrotenthaler Thomas^ORCID

Abstract

Suicide-related media content has preventive or harmful effects depending on the specific content. Proactive media screening for suicide prevention is hampered by the scarcity of machine learning approaches to detect specific characteristics in news reports. This study applied machine learning to label large quantities of broadcast (TV and radio) media data according to media recommendations reporting suicide. We manually labeled 2519 English transcripts from 44 broadcast sources in Oregon and Washington, USA, published between April 2019 and March 2020. We conducted a content analysis of media reports regarding content characteristics. We trained a benchmark of machine learning models including a majority classifier, approaches based on word frequency (TF-IDF with a linear SVM) and a deep learning model (BERT). We applied these models to a selection of more simple (e.g., focus on a suicide death), and subsequently to putatively more complex tasks (e.g., determining the main focus of a text from 14 categories). Tf-idf with SVM and BERT were clearly better than the naive majority classifier for all characteristics. In a test dataset not used during model training, F1-scores (i.e., the harmonic mean of precision and recall) ranged from 0.90 for celebrity suicide down to 0.58 for the identification of the main focus of the media item. Model performance depended strongly on the number of training samples available, and much less on assumed difficulty of the classification task. This study demonstrates that machine learning models can achieve very satisfactory results for classifying suicide-related broadcast media content, including multi-class characteristics, as long as enough training samples are available. The developed models enable future large-scale screening and investigations of broadcast media.

Funder

Vibrant Emotional Health

Vienna Science and Technology Fund

Publisher

Public Library of Science (PLoS)

Reference32 articles.

1. Suicide data.;World Health Organization;Published,2021

2. Association between suicide reporting in the media and suicide: systematic review and meta-analysis;T Niederkrotenthaler;BMJ,2020

3. The influence of suggestion on suicide: substantive and theoretical implications of the Werther effect.;DP Phillips;Am Sociol Rev,1974