A study of dealing class imbalance problem with machine learning methods for code smell severity detection using PCA-based feature selection technique-Reference-Cited by-同舟云学术

A study of dealing class imbalance problem with machine learning methods for code smell severity detection using PCA-based feature selection technique

Published:2023-09-27 Issue:1 Volume:13 Page:
ISSN:2045-2322
Container-title:Scientific Reports
language:en
Short-container-title:Sci Rep

Author:

Rao Rajwant Singh,Dewangan Seema,Mishra Alok,Gupta Manjari

Abstract

AbstractDetecting code smells may be highly helpful for reducing maintenance costs and raising source code quality. Code smells facilitate developers or researchers to understand several types of design flaws. Code smells with high severity can cause significant problems for the software and may cause challenges for the system's maintainability. It is quite essential to assess the severity of the code smells detected in software, as it prioritizes refactoring efforts. The class imbalance problem also further enhances the difficulties in code smell severity detection. In this study, four code smell severity datasets (Data class, God class, Feature envy, and Long method) are selected to detect code smell severity. In this work, an effort is made to address the issue of class imbalance, for which, the Synthetic Minority Oversampling Technique (SMOTE) class balancing technique is applied. Each dataset's relevant features are chosen using a feature selection technique based on principal component analysis. The severity of code smells is determined using five machine learning techniques: K-nearest neighbor, Random forest, Decision tree, Multi-layer Perceptron, and Logistic Regression. This study obtained the 0.99 severity accuracy score with the Random forest and Decision tree approach with the Long method code smell. The model's performance is compared based on its accuracy and three other performance measurements (Precision, Recall, and F-measure) to estimate severity classification models. The impact of performance is also compared and presented with and without applying SMOTE. The results obtained in the study are promising and can be beneficial for paving the way for further studies in this area.

Funder

NTNU Norwegian University of Science and Technology

Publisher

Springer Science and Business Media LLC

Subject

Multidisciplinary

Link

https://www.nature.com/articles/s41598-023-43380-8.pdf

Reference44 articles.

1. Ogheneovo, E. On the relationship between software complexity and maintenance costs. J. Comput. Commun. 2, 1–16. https://doi.org/10.4236/jcc.2014.214001 (2014).

2. Wiegers, K. & Beatty, J. Software Requirements (Pearson Education, London, 2013).

3. Chung, L. & do, P. L. J. C. S. On non-functional requirements in software engineering. in Conceptual Modeling: Foundations and Applications (Lecture Notes in Computer Science) (eds. Borgida, A. T., Chaudhri, V., Giorgini, P. & YuE, E. S.) 363–379 (Springer, Cham, 2009).

4. Mishra, D. & Mishra, A. Simplified software inspection process in compliance with international standards. Comput. Stand. Interfaces 31(4), 763–771 (2009).

5. Yu, L. & Mishra, A. Risk analysis of global software development and proposed solutions. Automatika 51(1), 89–98 (2010).

Cited by 9 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. EvaluateXAI: A framework to evaluate the reliability and consistency of rule-based XAI techniques for software analytics tasks;Journal of Systems and Software;2024-11

2. Dynamic stacking ensemble for cross-language code smell detection;PeerJ Computer Science;2024-08-15

3. Semantic and traditional feature fusion for software defect prediction using hybrid deep learning model;Scientific Reports;2024-07-01

4. Improving accuracy of code smells detection using machine learning with data balancing techniques;The Journal of Supercomputing;2024-06-05

5. Indian Sign Language Interpretation using Skeletal Features and LSTM Networks;2024 IEEE International Conference on Interdisciplinary Approaches in Technology and Management for Social Innovation (IATMSI);2024-03-14