Retrieval-Based Diagnostic Decision Support (Preprint)

Author:

Abdullahi Tassallah AminaORCID,Mercurio LauraORCID,Singh RitambharaORCID,Eickhoff CarstenORCID

Abstract

BACKGROUND

Diagnostic errors pose significant health risks and contribute to patient mortality. With the growing accessibility of electronic health records, machine learning models offer a promising avenue for enhancing diagnosis quality. Current research has primarily focused on a limited set of diseases with ample training data, neglecting diagnostic scenarios with limited data availability.

OBJECTIVE

This study aims to develop an information retrieval (IR) based framework that accommodates data sparsity to facilitate broader diagnostic decision support.

METHODS

We present an IR-based diagnostic decision support framework called CliniqIR. It employs clinical text records, the Unified Medical Language System (UMLS) Metathesaurus, and 33M PubMed abstracts to classify a broad spectrum of diagnoses independent of training data availability. We compare CliniqIR's performance to pre-trained clinical transformer models (like ClinicalBERT) in supervised and zero-shot settings. Subsequently, we combine the strength of supervised fine-tuned ClinicalBERT and CliniqIR to build an ensemble framework that delivers state-of-the-art diagnostic predictions.

RESULTS

CliniqIR returns the correct diagnosis for a DC3 case among its top-3 predictions, on average, on a rare disease dataset (DC3) with no training data. On the MIMIC-III dataset, CliniqIR outperforms ClinicalBERT in predicting diagnoses with fewer than five training samples by an average Mean Reciprocal Rank (MRR) of 9%. In a zero-shot setting, where no specific training was conducted, CliniqIR also outperforms the pre-trained transformer models by an MRR of 10%. Furthermore, our ensemble framework surpassed the individual constituent models by a minimum of 8% in MRR.

CONCLUSIONS

Our experiments highlight the importance of IR in leveraging unstructured knowledge resources to identify infrequently encountered diagnoses. In addition, our ensemble framework benefits from combining the complementary strengths of the supervised and retrieval-based models to diagnose a broad spectrum of diseases.

Publisher

JMIR Publications Inc.

Cited by 1 articles. 订阅此论文施引文献 订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3