A Two-Stage Voting-Boosting Technique for Ensemble Learning in Social Network Sentiment Classification-Reference-Cited by-同舟云学术

A Two-Stage Voting-Boosting Technique for Ensemble Learning in Social Network Sentiment Classification

Published:2023-03-24 Issue:4 Volume:25 Page:555
ISSN:1099-4300
Container-title:Entropy
language:en
Short-container-title:Entropy

Author:

Cui Su¹^ORCID,Han Yiliang¹^ORCID,Duan Yifei²,Li Yu¹,Zhu Shuaishuai¹,Song Chaoyue¹

Affiliation:

1. Department of Electronic Information, Engineering University of Chinese People’s Armed Police Force, Xi’an 710086, China

2. Department of Computer and Information Technology, University of Pennsylvania, Philadelphia, PA 19019, USA

Abstract

In recent years, social network sentiment classification has been extensively researched and applied in various fields, such as opinion monitoring, market analysis, and commodity feedback. The ensemble approach has achieved remarkable results in sentiment classification tasks due to its superior performance. The primary reason behind the success of ensemble methods is the enhanced diversity of the base classifiers. The boosting method employs a sequential ensemble structure to construct diverse data while also utilizing erroneous data by assigning higher weights to misclassified samples in the next training round. However, this method tends to use a sequential ensemble structure, resulting in a long computation time. Conversely, the voting method employs a concurrent ensemble structure to reduce computation time but neglects the utilization of erroneous data. To address this issue, this study combines the advantages of voting and boosting methods and proposes a new two-stage voting boosting (2SVB) concurrent ensemble learning method for social network sentiment classification. This novel method not only establishes a concurrent ensemble framework to decrease computation time but also optimizes the utilization of erroneous data and enhances ensemble performance. To optimize the utilization of erroneous data, a two-stage training approach is implemented. Stage-1 training is performed on the datasets by employing a 3-fold cross-segmentation approach. Stage-2 training is carried out on datasets that have been augmented with the erroneous data predicted by stage 1. To augment the diversity of base classifiers, the training stage employs five pre-trained deep learning (PDL) models with heterogeneous pre-training frameworks as base classifiers. To reduce the computation time, a two-stage concurrent ensemble framework was established. The experimental results demonstrate that the proposed method achieves an F1 score of 0.8942 on the coronavirus tweet sentiment dataset, surpassing other comparable ensemble methods.

Funder

National Natural Science Foundation of China

Innovative Research Team in Engineering University of People’s Armed Police, China

Natural Science Foundation of Shaanxi Province, China

National Social Science Fund of China

Publisher

MDPI AG

Subject

General Physics and Astronomy

Link

https://www.mdpi.com/1099-4300/25/4/555/pdf

Reference70 articles.

1. Microblog sentiment analysis with emoticon space model;Jiang;J. Comput. Sci. Technol.,2015

2. Twitter sentiment classification for measuring public health concerns;Ji;Soc. Netw. Anal. Min.,2015

3. Hu, X., Tang, L., Tang, J., and Liu, H. (2013, January 4–8). Exploiting social relations for sentiment analysis in microblogging. Proceedings of the 6th ACM International Conference on Web Search and Data Mining, Rome, Italy.

4. Social media contents based sentiment analysis and prediction system;Yoo;Expert Syst. Appl.,2018

5. Monitoring the public opinion about the vaccination topic from tweets analysis;Ducange;Expert Syst. Appl.,2019

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Imbalanced ensemble learning leveraging a novel data-level diversity metric;Pattern Recognition;2025-01

2. FMC: Multimodal fake news detection based on multi-granularity feature fusion and contrastive learning;Alexandria Engineering Journal;2024-12

3. SARD: Fake news detection based on CLIP contrastive learning and multimodal semantic alignment;Journal of King Saud University - Computer and Information Sciences;2024-10

4. A Model Ensemble Approach for Conversational Quadruple Extraction;Natural Language Processing and Chinese Computing;2023