symptomcheckR: an R package for analyzing and visualizing symptom checker performance-Reference-Cited by-同舟云学术

symptomcheckR: an R package for analyzing and visualizing symptom checker performance

Published:2024-02-06 Issue: Volume: Page:
ISSN:
Container-title:
language:
Short-container-title:

Author:

Kopka Marvin^ORCID,Feufel Markus A.^ORCID

Abstract

AbstractBackgroundA major stream of research on symptom checkers aims at evaluating the technology’spredictive accuracy, but apart from general trends, the results are marked by high variability. Several authors suggest that this variability might in part be due to different assessment methods and a lack of standardization. To improve the reliability of symptom checker evaluation studies, several approaches have been suggested, including standardizing input procedures, the generation of test vignettes, and the assignment of gold standard solutions for these vignettes. Recently, we suggested a third approach––test-theoretic metrics for standardized performance reporting–– to allow systematic and comprehensive comparisons of symptom checker performance. However, calculating these metrics is time-consuming and error prone, which could hamper the use and effectiveness of these metrics.ResultsWe developed the R package symptomcheckR as an open-source software to assist researchers in calculating standard metrics to evaluate symptom checker performance individually and comparatively and produce publicationready figures. These metrics include accuracy (by triage level), safety of advice (i.e., rate of correct or overtriage), comprehensiveness (i.e., how many cases could be entered or were assessed), inclination to overtriage (i.e., how risk-averse a symptom checker is) and a capability comparison score (i.e., a score correcting for case difficulty and comprehensiveness that enables a fair and reliable comparison of different symptom checkers). Each metric can be obtained using a single command and visualized with another command. For the analysis of individual or the comparison of multiple symptom checkers, single commands can be used to produce a comprehensive performance profile that complements the standard focus on accuracy with additional metrics that reveal strengths and weaknesses of symptom checkers.ConclusionsOur package supports ongoing efforts to improve the quality of vignette-based symptom checker evaluation studies by means of standardized methods. Specifically, with our package, adhering to reporting standards and metrics becomes easier, simple, and time efficient. Ultimately, this may help users gain a more systematic understanding of the strengths and limitations of symptom checkers for different use cases (e.g., all-purpose symptom checkers for general medicine versus symptom checkers that aim at improving triage in emergency departments), which can improve patient safety and resource allocation.

Publisher

Cold Spring Harbor Laboratory

Reference35 articles.

1. Examining the impact of a symptom assessment application on patient-physician interaction among self-referred walk-in patients in the emergency depart-ment (AKUSYM): study protocol for a multi-center, randomized controlled, parallel-group superiority trial;Trials,2022

2. The diagnostic and triage accuracy of digital and online symptom checker tools: a systematic review;npj Digit Med,2022

3. Triage and Diagnostic Accuracy of Online Symptom Checkers: Systematic Review;J Med Internet Res,2023

4. A scoping review on the use and usefulness of online symptom checkers and triage systems: How to proceed?;Front Med,2023

5. Impact of NHS 111 Online on the NHS 111 telephone service and urgent care system: a mixed-methods study;Health Serv Deliv Res,2021

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Evaluating self-triage accuracy of laypeople, symptom-assessment apps, and large language models: A framework for case vignette development using a representative design approach (RepVig);2024-04-03