Ensemble Model for Educational Data Mining Based on Synthetic Minority Oversampling Technique

Author:

R Manoharan1ORCID,Stalin M.Subi2,Loganathan Ganesh Babu3,s Deepa4,k Venkateswaran5

Affiliation:

1. Apollo Engineering College

2. Arignar Anna Institute of Science and Technology

3. Tishk International University

4. Kongu Engineering College

5. St Joseph Engineering College

Abstract

Abstract Data mining in the classroom is a well-known field that involves data mining concepts, statistical analysis, and machine learning concepts, all of which are applied to educational data. These EDM processed data are frequently used to analyse various aspects of the business and process model. Existing and traditional process models involve the use of traditional statistical techniques to process data, which necessitates a significant amount of manual intervention for data modelling and pre-processing. To address the issues raised above, this paper proposes a novel technique that combines a machine learning model with statistical approaches. This machine learning combination combines various classifiers such as Decision Tree Logistic Regression, Random Forest, Multiplayer Perceptron, K Nearest Neighbor, Decision Tree, and so on. Because of the limited data availability, the information used in the observation is highly imbalanced. As a result, the above-mentioned technique is combined with the universally benchmarked model known as Synthetic Minority Oversampling Technique (SMOTE) to discuss issues related to class imbalance. In addition, the performance evaluation is statistically performed to demonstrate the efficacy of the suggested strategies. A technological college in India obtained the main student data collection, which included information on 6,807 students with characteristics. A synthetic minority oversampling approach sensor is used to handle the imbalanced data set. The model is calibrated using eight methodologies, which are then evaluated to determine the dimensions that will help produce the best suitable model to categorise a student based on his achievements.

Publisher

Research Square Platform LLC

Reference24 articles.

1. Sapkota N, Alsadoon A, Prasad PWC, Elchouemi A, Singh AK (2019) "Data Summarization Using Clustering and Classification: Spectral Clustering Combined with k-Means Using NFPH," International Conference on Machine Learning, Big Data, Cloud and Parallel Computing (COMITCon), Faridabad, India, 2019, pp. 146–151, doi: 10.1109/COMITCon.2019.8862218

2. Li H, Lu Q (2017) "K-CV parameter optimization method in the application of SVM classification data," 2017 IEEE 2nd International Conference on Big Data Analysis (ICBDA), Beijing, pp. 25–29, doi: 10.1109/ICBDA.2017.8078838

3. Chandra S, Kaur M (2015) "Creation of an Adaptive Classifier to enhance the classification accuracy of existing classification algorithms in the field of Medical Data Mining," 2nd International Conference on Computing for Sustainable Global Development (INDIACom), New Delhi, 2015, pp. 376–381

4. Okfalisa I, Gazalba, Mustakim, Reza NGI, "Comparative analysis of k-nearest neighbor and modified k-nearest neighbor algorithm for data classification," 2017 2nd International conferences on Information Technology, Information Systems and, Engineering E (2017) (ICITISEE), Yogyakarta, pp. 294–298, doi: 10.1109/ICITISEE.2017.8285514

5. Pristyanto Y, Pratama I, Nugraha AF (2018) "Data level approach for imbalanced class handling on educational data mining multiclass classification," International Conference on Information and Communications Technology (ICOIACT), Yogyakarta, 2018, pp. 310–314, doi: 10.1109/ICOIACT.2018.8350792

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3