A COMPARATIVE STUDY ON PERFORMANCE OF BASIC AND ENSEMBLE CLASSIFIERS WITH VARIOUS DATASETS-Reference-Cited by-同舟云学术

A COMPARATIVE STUDY ON PERFORMANCE OF BASIC AND ENSEMBLE CLASSIFIERS WITH VARIOUS DATASETS

Published:2023-03-31 Issue:1 Volume:19 Page:107-132
ISSN:2353-6977
Container-title:Applied Computer Science
language:
Short-container-title:acs

Author:

Gunakala Archana^ORCID,Shahid Afzal Hussain^ORCID

Abstract

Classification plays a critical role in machine learning (ML) systems for processing images, text and high -dimensional data. Predicting class labels from training data is the primary goal of classification. An optimal model for a particular classification problem is chosen on the basis of the model's performance and execution time. This paper compares and analyses the performance of basic as well as ensemble classifiers utilizing 10 -fold cross validation and also discusses their essential concepts, advantages, and disadvantages. In this study five basic classifiers namely Naïve Bayes (NB), Multi-layer Perceptron (MLP), Support Vector Machine (SVM), Decision Tree (DT), and Random Forest (RF) and the ensemble of all the five classifiers along with few more combinations are compared with five University of California Irvine (UCI) ML Repository datasets and a Diabetes Health Indicators dataset from kaggle repository. To analyze and compare the performance of classifiers, evaluation metrics like Accuracy, Recall, Precision, Area Under Curve (AUC) and F-Score are used. Experimental results showed that SVM performs best on two out of the six datasets (Diabetes Health Indicators and waveform), RF performs best for Arrhythmia, Sonar, Tic-tac-toe datasets, and the best ensemble combination is found to be DT+SVM+RF on Ionosphere dataset having respective accuracies 72.58%, 90.38%, 81.63%, 73.59%, 94.78% and 94.01% and the proposed ensemble combinations outperformed over the conventional models for few datasets.

Publisher

Politechnika Lubelska

Subject

Artificial Intelligence,Industrial and Manufacturing Engineering,Computer Science Applications,Economics, Econometrics and Finance (miscellaneous),Mechanical Engineering,Biomedical Engineering,Information Systems,Control and Systems Engineering

Reference40 articles.

1. Alshayeji, M. H., Ellethy, H., Abed, S., & Gupta, R. (2022). Computer-aided detection of breast cancer on the Wisconsin dataset: An artificial neural networks approach. Biomedical Signal Processing and Control, 71(PA), 103141. https://doi.org/10.1016/j.bspc.2021.103141

2. Alshdaifat, E., Al-hassan, M., & Aloqaily, A. (2021). Effective heterogeneous ensemble classification: An alternative approach for selecting base classifiers. ICT Express, 7(3), 342–349. https://doi.org/10.1016/j.icte.2020.11.005

3. Baumann, P., Hochbaum, D. S., & Yang, Y. T. (2019). A comparative study of the leading machine learning techniques and two new optimization algorithms. European Journal of Operational Research, 272(3), 1041–1057. https://doi.org/10.1016/j.ejor.2018.07.009

4. bin Basir, M. A., & binti Ahmad, F. (2017). New Feature Selection Model Based Ensemble Rule Classifiers Method for Dataset Classification. International Journal of Artificial Intelligence & Applications, 8(2), 37–43. https://doi.org/10.5121/ijaia.2017.8204

5. Chandrika, Divya, C., Gowramma, G. S., & Varun, C. R. (2018). A comparative analysis on evaluation of classification algorithms based on ionospheric data. International Journal of Computer Sciences and Engineering, 6(5), 636–640. https://doi.org/10.26438/ijcse/v6i5.636640

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Comparison of Machine Learning Tree-Based Algorithms to Predict Future Paratuberculosis ELISA Results Using Repeat Milk Tests;Animals;2024-04-05

2. Comparative Study of Multiple Algorithms Classification for Land Use and Land Cover Change Detection and its Impact on Local Climate of Mardan District, Pakistan;2024