Feature selection for sentiment analysis using hybrid multiobjective evolutionary algorithm

Author:

Gul Rimsha1,Bashir Maryam1

Affiliation:

1. FAST School of Computing, National University of Computer and Emerging Sciences, Lahore, Pakistan

Abstract

As the volume of data continues to grow, the significance of text classification is on the rise. This vast amount of data majorly exists in the form of texts. Effective data preparation is essential to extract sentiment data from this vast amount of text, as irrelevant and redundant information can impede valuable insights. Feature selection is an important step in the data preparation phase as it eliminates irrelevant and insignificant features from the huge features set. There exist a large body of work related to feature selection for image processing but limited research is done for text data. While some studies recognize the significance of feature selection in text classification, but there is still need for more efficient sentiment analysis models that optimize feature selection and reduce computational. This manuscript aims to bridge these gaps by introducing a hybrid multi-objective evolutionary algorithm as a feature selection mechanism, combining the power of multiple objectives and evolutionary processes. The approach combines two feature selection techniques within a binary classification model: a filter method, Information Gain (IG), and an evolutionary wrapper method, Binary Multi-Objective Grey Wolf Optimizer (BMOGWO). Experimental evaluations are conducted across six diverse datasets. It achieves a reduction of over 90 percent in feature size while improving accuracy by nearly nine percent. These results showcase the model’s efficiency in terms of computational time and its efficacy in terms of higher classification accuracy which improves sentiment analysis performance. This improvement can be beneficial for various applications, including recommendation systems, reviews analysis, and public opinion observation. However, it’s crucial to acknowledge certain limitations of this study. These encompass the need for broader classifier evaluation, and scalability considerations with larger datasets. These identified limitations serve as directions for future research and the enhancement of the proposed approach.

Publisher

IOS Press

Reference33 articles.

1. Migdal Miki , How big data empowers organizations to work smarter, not harder, https://www.forbes.com/sites/forbestechcouncil//08/23/how-big-data-empowers-organizations-to-work-smarter-not-?sh=41ba2045532f.

2. Reviewer credibility and sentiment analysis based user profilemodelling for online product recommendation;Hu;IEEE Access,2020

3. Xu X. , What are customers commenting on, and how is their satisfaction affected? examining online reviews in the on-demand food service context, Decis. Support Syst 142(113467) (2021).

4. A system for real-time twitter sentiment analysis of us presidential election cycle;Wang;Proc. ACL Syst. Demonstrations,2012

5. A survey on sentiment analysis challenges;Hussein;J. King Saud Univ.-Eng. Sci.,2018

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3