Abstract
Background
A novel disease poses special challenges for informatics solutions. Biomedical informatics relies for the most part on structured data, which require a preexisting data or knowledge model; however, novel diseases do not have preexisting knowledge models. In an emergent epidemic, language processing can enable rapid conversion of unstructured text to a novel knowledge model. However, although this idea has often been suggested, no opportunity has arisen to actually test it in real time. The current coronavirus disease (COVID-19) pandemic presents such an opportunity.
Objective
The aim of this study was to evaluate the added value of information from clinical text in response to emergent diseases using natural language processing (NLP).
Methods
We explored the effects of long-term treatment by calcium channel blockers on the outcomes of COVID-19 infection in patients with high blood pressure during in-patient hospital stays using two sources of information: data available strictly from structured electronic health records (EHRs) and data available through structured EHRs and text mining.
Results
In this multicenter study involving 39 hospitals, text mining increased the statistical power sufficiently to change a negative result for an adjusted hazard ratio to a positive one. Compared to the baseline structured data, the number of patients available for inclusion in the study increased by 2.95 times, the amount of available information on medications increased by 7.2 times, and the amount of additional phenotypic information increased by 11.9 times.
Conclusions
In our study, use of calcium channel blockers was associated with decreased in-hospital mortality in patients with COVID-19 infection. This finding was obtained by quickly adapting an NLP pipeline to the domain of the novel disease; the adapted pipeline still performed sufficiently to extract useful information. When that information was used to supplement existing structured data, the sample size could be increased sufficiently to see treatment effects that were not previously statistically detectable.
Reference16 articles.
1. ChapmanWDowlingJIvanovOGestelandPOlszewskiREspinoJWagnerMEvaluating natural language processing applications applied to outbreak and disease surveillanceProceedings of 36th symposium on the interface: computing science and statistics 2004200436th Symposium on the Interface: Computing Science and Statistics 2004May 26-29, 2004Baltimore, MD
2. Comparison of Natural Language Processing Biosurveillance Methods for Identifying Influenza From Encounter Notes
3. Calcium channel blocker amlodipine besylate is associated with reduced case fatality rate of COVID-19 patients with hypertension
4. DevlinJChangMLeeKToutanovaKBERT: Pre-training of Deep Bidirectional Transformers for Language UnderstandingarXivcs201810102018-11-17http://arxiv.org/abs/1810.04805
Cited by
63 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献