Multi-class Twitter sentiment classification with emojis

Author:

Li Mengdi,Ch’ng Eugene,Chong Alain Yee Loong,See Simon

Abstract

Purpose Recently, various Twitter Sentiment Analysis (TSA) techniques have been developed, but little has paid attention to the microblogging feature – emojis, and few works have been conducted on the multi-class sentiment analysis of tweets. The purpose of this paper is to consider the popularity of emojis on Twitter and investigate the feasibility of an emoji training heuristic for multi-class sentiment classification of tweets. Tweets from the “2016 Orlando nightclub shooting” were used as a source of study. Besides, this study also aims to demonstrate how mapping can contribute to interpreting sentiments. Design/methodology/approach The authors presented a methodological framework to collect, pre-process, analyse and map public Twitter postings related to the shooting. The authors designed and implemented an emoji training heuristic, which automatically prepares the training data set, a feature needed in Big Data research. The authors improved upon the previous framework by advancing the pre-processing techniques, enhancing feature engineering and optimising the classification models. The authors constructed the sentiment model with a logistic regression classifier and selected features. Finally, the authors presented how to visualise citizen sentiments on maps dynamically using Mapbox. Findings The sentiment model constructed with the automatically annotated training sets using an emoji approach and selected features performs well in classifying tweets into five different sentiment classes, with a macro-averaged F-measure of 0.635, a macro-averaged accuracy of 0.689 and the MAEM of 0.530. Compared to those experimental results in related works, the results are satisfactory, indicating the model is effective and the proposed emoji training heuristic is useful and feasible in multi-class TSA. The maps authors created, provide a much easier-to-understand visual representation of the data, and make it more efficient to monitor citizen sentiments and distributions. Originality/value This work appears to be the first to conduct multi-class sentiment classification on Twitter with automatic annotation of training sets using emojis. Little attention has been paid to applying TSA to monitor the public’s attitudes towards terror attacks and country’s gun policies, the authors consider this work to be a pioneering work. Besides, the authors have introduced a new data set of 2016 Orlando Shooting tweets, which will be made available for other researchers to mine the public’s political opinions about gun policies.

Publisher

Emerald

Subject

Industrial and Manufacturing Engineering,Strategy and Management,Computer Science Applications,Industrial relations,Management Information Systems

Reference35 articles.

1. Sentiment analysis of Twitter data,2011

2. The effects of emoji in sentiment analysis;International Journal of Computer and Electrical Engineering,2017

3. TwiSE at semeval-2016 task 4: Twitter sentiment classification,2016

4. Robust sentiment detection on Twitter from biased and noisy data,2010

5. On using Twitter to monitor political sentiment and predict election results,2011

Cited by 37 articles. 订阅此论文施引文献 订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献

1. What multimodal components, tools, dataset and focus of emotion are used in the current research of multimodal emotion: a systematic literature review;Cogent Social Sciences;2024-07-16

2. Analysing Protest-Related Tweets: An Evaluation of Techniques by the Open Source Intelligence Team;Lecture Notes in Networks and Systems;2024

3. Embracing emojis: Bridging the gap in workplace technology adoption and elevating communication effectiveness;Journal of Information Technology Teaching Cases;2023-12-11

4. Uncovering Customer Issues in E-Commerce: Sentiment Analysis and Topic Modeling Approach;2023 6th International Conference on Information and Communications Technology (ICOIACT);2023-11-10

5. The Role of Emojis in Sentiment Analysis of Financial Microblogs;2023 Fourth International Conference on Intelligent Data Science Technologies and Applications (IDSTA);2023-10-24

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3