Error Rates of Data Processing Methods in Clinical Research: A Systematic Review and Meta-Analysis

Author:

Garza Maryam Y.1,Williams Tremaine1,Ounpraseuth Songthip1,Hu Zhuopei1,Lee Jeannette1,Snowden Jessica1,Walden Anita C.2,Simon Alan E.3,Devlin Lori A.4,Young Leslie W.5,Zozus Meredith N.6

Affiliation:

1. University of Arkansas for Medical Sciences

2. University of Colorado Denver, Anschutz Medical Campus

3. National Institutes of Health

4. University of Louisville

5. University of Vermont

6. The University of Texas Health Science Center at San Antonio

Abstract

Abstract Background: Over the last 30 years, empirical assessments of data accuracy in clinical research have been reported in the literature. Although there have been articles summarizing results reported in multiple papers, there has been little synthesis of these results. Further, although notable exceptions exist, little evidence has been obtained regarding the relative accuracy of different data processing methods. Methods: A systematic review of the literature was performed to identify clinical research studies that evaluated the quality of data obtained from data processing methods typically used in clinical research (e.g., medical record abstraction, optical scanning, single-data entry, and double-data entry). A total of 93 papers meeting our inclusion criteria were categorized according to their data processing methods. Quantitative information on data accuracy was abstracted from the articles and pooled. Meta-analysis of single proportions based on an inverse variance method and generalized linear mixed model approach of studies from the literature were used to derive an overall estimate of error rates across data processing methods for comparison. Results: Review of the literature indicated that the accuracy associated with data processing methods varies widely, with error rates ranging from 2 errors per 10,000 fields to 2,784errors per 10,000 fields. The medical record abstraction process for data acquisition in clinical research was associated with both high and highly variable error rates, with a variability of 3 orders of magnitude in accuracy (70 – 2,784 errors per 10,000 fields). Error rates for data processed with optical methods were comparable to data processed using single-data entry (2 – 358 vs. 4 – 650 per 10,000 fields, respectively). In comparison, double-data entry was associated with the lowest error rates (4 – 33 per 10,000 fields). Conclusions: Data processing and cleaning methods may explain a significant amount of the variability in data accuracy.

Publisher

Research Square Platform LLC

Cited by 1 articles. 订阅此论文施引文献 订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3