Pediatric Injury Surveillance From Uncoded Emergency Department Admission Records in Italy: Machine Learning–Based Text-Mining Approach-Reference-Cited by-同舟云学术

Pediatric Injury Surveillance From Uncoded Emergency Department Admission Records in Italy: Machine Learning–Based Text-Mining Approach

Published:2023-07-12 Issue: Volume:9 Page:e44467
ISSN:2369-2960
Container-title:JMIR Public Health and Surveillance
language:en
Short-container-title:JMIR Public Health Surveill

Author:

Azzolina Danila^ORCID,Bressan Silvia^ORCID,Lorenzoni Giulia^ORCID,Baldan Giulia Andrea^ORCID,Bartolotta Patrizia^ORCID,Scognamiglio Federico^ORCID,Francavilla Andrea^ORCID,Lanera Corrado^ORCID,Da Dalt Liviana^ORCID,Gregori Dario^ORCID

Abstract

Background Unintentional injury is the leading cause of death in young children. Emergency department (ED) diagnoses are a useful source of information for injury epidemiological surveillance purposes. However, ED data collection systems often use free-text fields to report patient diagnoses. Machine learning techniques (MLTs) are powerful tools for automatic text classification. The MLT system is useful to improve injury surveillance by speeding up the manual free-text coding tasks of ED diagnoses. Objective This research aims to develop a tool for automatic free-text classification of ED diagnoses to automatically identify injury cases. The automatic classification system also serves for epidemiological purposes to identify the burden of pediatric injuries in Padua, a large province in the Veneto region in the Northeast Italy. Methods The study includes 283,468 pediatric admissions between 2007 and 2018 to the Padova University Hospital ED, a large referral center in Northern Italy. Each record reports a diagnosis by free text. The records are standard tools for reporting patient diagnoses. An expert pediatrician manually classified a randomly extracted sample of approximately 40,000 diagnoses. This study sample served as the gold standard to train an MLT classifier. After preprocessing, a document-term matrix was created. The machine learning classifiers, including decision tree, random forest, gradient boosting method (GBM), and support vector machine (SVM), were tuned by 4-fold cross-validation. The injury diagnoses were classified into 3 hierarchical classification tasks, as follows: injury versus noninjury (task A), intentional versus unintentional injury (task B), and type of unintentional injury (task C), according to the World Health Organization classification of injuries. Results The SVM classifier achieved the highest performance accuracy (94.14%) in classifying injury versus noninjury cases (task A). The GBM method produced the best results (92% accuracy) for the unintentional and intentional injury classification task (task B). The highest accuracy for the unintentional injury subclassification (task C) was achieved by the SVM classifier. The SVM, random forest, and GBM algorithms performed similarly against the gold standard across different tasks. Conclusions This study shows that MLTs are promising techniques for improving epidemiological surveillance, allowing for the automatic classification of pediatric ED free-text diagnoses. The MLTs revealed a suitable classification performance, especially for general injuries and intentional injury classification. This automatic classification could facilitate the epidemiological surveillance of pediatric injuries by also reducing the health professionals’ efforts in manually classifying diagnoses for research purposes.

Publisher

JMIR Publications Inc.

Subject

Public Health, Environmental and Occupational Health,Health Informatics

Reference36 articles.

1. Our Shrinking Globe

2. Unintentional Injuries in Childhood: Analysis and Recommendations

3. Unintentional Injuries in Pediatrics

4. WHOWorld report on child injury prevention2019-02-04https://www.who.int/publications/i/item/9789241563574

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Use of a Large Language Model to Identify and Classify Injuries With Free-Text Emergency Department Data;JAMA Network Open;2024-05-28

2. Development of machine learning-based predictors for early diagnosis of hepatocellular carcinoma;Scientific Reports;2024-03-04