Affiliation:
1. Department of Computer Information Systems, College of Computer Science and Information Technology, Imam Abdulrahman Bin Faisal University, P.O. Box 1982, Dammam 31441, Saudi Arabia
Abstract
Data preprocessing moves the data from raw to ready for analysis. Data resulting from fraud compromises the quality of the data and the resulting analysis. It can exist in datasets such that it goes undetected since it is included in the analysis. This study proposed a process for measuring the effect of fraudulent data during data preparation and its possible influence on quality. The five-step process begins with identifying the business rules related to the business process(s) affected by fraud and their associated quality dimensions. This is followed by measuring the business rules in the specified timeframe, detecting fraudulent data, cleaning them, and measuring their quality after cleaning. The process was implemented in the case of occupational fraud within a hospital context and the illegal issuance of underserved sick leave. The aim of the application is to identify the quality dimensions that are influenced by the injected fraudulent data and how these dimensions are affected. This study agrees with the existing literature and confirms its effects on timeliness, coherence, believability, and interpretability. However, this did not show any effect on consistency. Further studies are needed to arrive at a generalizable list of the quality dimensions that fraud can affect.
Subject
Information Systems and Management,Computer Science Applications,Information Systems
Reference22 articles.
1. Gee, S. (2014). Fraud and Fraud Detection: A Data Analytics Approach, Wiley.
2. Detecting Data Quality Issues in Clinical Trials: Current Practices and Recommendations;Knepper;Ther. Innov. Regul. Sci.,2016
3. The Challenge of Quality Evaluation in Fraud Detection;Puentes;HAL,2018
4. Allen, M., and Cervo, D. (2015). Multi-Domain Master Data Management, Elsevier.
5. DAMA-DMBOK (2015). The DAMA Guide to the Data Management Body of Knowledge, Technics Publications, LLC. [2nd ed.].