A Heterogeneous Ensemble Learning Framework for Spam Detection in Social Networks with Imbalanced Data-Reference-Cited by-同舟云学术

A Heterogeneous Ensemble Learning Framework for Spam Detection in Social Networks with Imbalanced Data

Published:2020-01-31 Issue:3 Volume:10 Page:936
ISSN:2076-3417
Container-title:Applied Sciences
language:en
Short-container-title:Applied Sciences

Author:

Zhao Chensu^ORCID,Xin Yang,Li Xuefeng^ORCID,Yang Yixian,Chen Yuling

Abstract

The popularity of social networks provides people with many conveniences, but their rapid growth has also attracted many attackers. In recent years, the malicious behavior of social network spammers has seriously threatened the information security of ordinary users. To reduce this threat, many researchers have mined the behavior characteristics of spammers and have obtained good results by applying machine learning algorithms to identify spammers in social networks. However, most of these studies overlook class imbalance situations that exist in real world data. In this paper, we propose a heterogeneous stacking-based ensemble learning framework to ameliorate the impact of class imbalance on spam detection in social networks. The proposed framework consists of two main components, a base module and a combining module. In the base module, we adopt six different base classifiers and utilize this classifier diversity to construct new ensemble input members. In the combination module, we introduce cost sensitive learning into deep neural network training. By setting different costs for misclassification and dynamically adjusting the weights of the prediction results of the base classifiers, we can integrate the input members and aggregate the classification results. The experimental results show that our framework effectively improves the spam detection rate on imbalanced datasets.

Publisher

MDPI AG

Subject

Fluid Flow and Transfer Processes,Computer Science Applications,Process Chemistry and Technology,General Engineering,Instrumentation,General Materials Science

Link

https://www.mdpi.com/2076-3417/10/3/936/pdf

Reference47 articles.

1. Malicious accounts: Dark of the social networks

2. @spam: the underground on 140 characters or less;Grier,2010

3. Social turing tests: Crowdsourcing sybil detection;Wang,2013

4. Design and Evaluation of a Real-Time URL Spam Filtering Service;Thomas,2011

5. Understanding and combating link farming in the twitter social network;Ghosh,2012

Cited by 54 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Ab‐HIDS: An anomaly‐based host intrusion detection system using frequency of N‐gram system call features and ensemble learning for containerized environment;Concurrency and Computation: Practice and Experience;2024-08-06

2. Advancing Email Spam Classification using Machine Learning and Deep Learning Techniques;Engineering, Technology & Applied Science Research;2024-08-02

3. A Collaborative Learning Technique for Improved Email Security;VFAST Transactions on Software Engineering;2024-06-30

4. Scalable Learning Framework for Detecting New Types of Twitter Spam with Misuse and Anomaly Detection;Sensors;2024-04-02

5. A Fake Profile Detection Model Using Multistage Stacked Ensemble Classification;Proceedings of Engineering and Technology Innovation;2024-02-29