Text-based predictions of COVID-19 diagnosis from self-reported chemosensory descriptions-Reference-Cited by-同舟云学术

Text-based predictions of COVID-19 diagnosis from self-reported chemosensory descriptions

Published:2023-07-27 Issue:1 Volume:3 Page:
ISSN:2730-664X
Container-title:Communications Medicine
language:en
Short-container-title:Commun Med

Author:

Li Hongyang,Gerkin Richard C.,Bakke Alyssa^ORCID,Norel Raquel^ORCID,Cecchi Guillermo^ORCID,Laudamiel Christophe,Niv Masha Y.^ORCID,Ohla Kathrin,Hayes John E.,Parma Valentina,Meyer Pablo^ORCID

Abstract

Abstract Background There is a prevailing view that humans’ capacity to use language to characterize sensations like odors or tastes is poor, providing an unreliable source of information. Methods Here, we developed a machine learning method based on Natural Language Processing (NLP) using Large Language Models (LLM) to predict COVID-19 diagnosis solely based on text descriptions of acute changes in chemosensation, i.e., smell, taste and chemesthesis, caused by the disease. The dataset of more than 1500 subjects was obtained from survey responses early in the COVID-19 pandemic, in Spring 2020. Results When predicting COVID-19 diagnosis, our NLP model performs comparably (AUC ROC ~ 0.65) to models based on self-reported changes in function collected via quantitative rating scales. Further, our NLP model could attribute importance of words when performing the prediction; sentiment and descriptive words such as “smell”, “taste”, “sense”, had strong contributions to the predictions. In addition, adjectives describing specific tastes or smells such as “salty”, “sweet”, “spicy”, and “sour” also contributed considerably to predictions. Conclusions Our results show that the description of perceptual symptoms caused by a viral infection can be used to fine-tune an LLM model to correctly predict and interpret the diagnostic status of a subject. In the future, similar models may have utility for patient verbatims from online health portals or electronic health records.

Publisher

Springer Science and Business Media LLC

Subject

General Medicine

Link

https://www.nature.com/articles/s43856-023-00334-5.pdf

Reference41 articles.

1. Koleck, T. A., Dreisbach, C., Bourne, P. E. & Bakken, S. Natural language processing of symptoms documented in free-text narratives of electronic health records: a systematic review. J. Am. Med. Inform. Assoc. 26, 364–379 (2019).

2. Cook, B. L. et al. Novel use of natural language processing (NLP) to predict suicidal ideation and psychiatric symptoms in a text-based mental health intervention in Madrid. Comput. Math. Methods Med. 2016, 8708434 (2016).

3. Velupillai, S. et al. Using clinical Natural Language Processing for health outcomes research: overview and actionable suggestions for future advances. J. Biomed. Inform. 88, 11–19 (2018).

4. Dey, S. et al. Human-centered explainability for life sciences, healthcare, and medical informatics. Patterns Prejudice 3, 100493 (2022).

5. Moein, S. T. et al. Smell dysfunction: a biomarker for COVID-19. Int. Forum Allergy Rhinol. 10, 944–950 (2020).

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Large Language Models in Healthcare and Medical Domain: A Review;Informatics;2024-08-07

2. Recent Advances in Large Language Models for Healthcare;BioMedInformatics;2024-04-16

3. DR-GPT: a large language model for medical report analysis of diabetic retinopathy patients;2024-01-17