Graph neural networks and cross-protocol analysis for detecting malicious IP addresses-Reference-Cited by-同舟云学术

Graph neural networks and cross-protocol analysis for detecting malicious IP addresses

Published:2022-09-14 Issue: Volume: Page:
ISSN:2199-4536
Container-title:Complex & Intelligent Systems
language:en
Short-container-title:Complex Intell. Syst.

Author:

Huang Yonghong^ORCID,Negrete Joanna,Wagener John,Fralick Celeste,Rodriguez Armando,Peterson Eric,Wosotowsky Adam

Abstract

AbstractAn internet protocol (IP) address is the foundation of the Internet, allowing connectivity between people, servers, Internet of Things, and services across the globe. Knowing what is connecting to what and where connections are initiated is crucial to accurately assess a company’s or individual’s security posture. IP reputation assessment can be quite complex because of the numerous services that may be hosted on that IP address. For example, an IP might be serving millions of websites from millions of different companies like web hosting companies often do, or it could be a large email system sending and receiving emails for millions of independent entities. The heterogeneous nature of an IP address typically makes it challenging to interpret the security risk. To make matters worse, adversaries understand this complexity and leverage the ambiguous nature of the IP reputation to exploit further unsuspecting Internet users or devices connected to the Internet. In addition, traditional techniques like dirty-listing cannot react quickly enough to changes in the security climate, nor can they scale large enough to detect new exploits that may be created and disappear in minutes. In this paper, we introduce the use of cross-protocol analysis and graph neural networks (GNNs) in semi-supervised learning to address the speed and scalability of assessing IP reputation. In the cross-protocol supervised approach, we combine features from the web, email, and domain name system (DNS) protocols to identify ones which are the most useful in discriminating suspicious and benign IPs. In our second experiment, we leverage the most discriminant features and incorporate them into the graph as nodes’ features. We use GNNs to pass messages from node to node, propagating the signal to the neighbors while also gaining the benefit of having the originating nodes being influenced by neighboring nodes. Thanks to the relational graph structure we can use only a small portion of labeled data and train the algorithm in a semi-supervised approach. Our dataset represents real-world data that is sparse and only contain a small percentage of IPs with verified clean or suspicious labels but are connected. The experimental results demonstrate that the system can achieve

$$85.28\%$$

85.28%accuracy in detecting malicious IP addresses at scale with only

$$5\%$$

5%of labeled data.

Publisher

Springer Science and Business Media LLC

Subject

Computational Mathematics,Engineering (miscellaneous),Information Systems,Artificial Intelligence

Link

https://link.springer.com/content/pdf/10.1007/s40747-022-00838-y.pdf

Reference47 articles.

1. Statista (2021) Number of internet users worldwide from 2005 to 2018 (in millions) [Online]. https://www.statista.com/statistics/617136/digital-population-worldwide/

2. Levine L (2008) DNS blacklists and whitelists [Online]. https://tools.ietf.org/html/draft-irtf-asrg-dnsbl-08

3. DNSWL (2017) Whitelisting DKIM-signed domains [Online]. https://www.dnswl.org/

4. Greylisting (2016) Greylisting explained [Online]. https://www.greylisting.org/

5. Berkeley Security Information Office (2019) Aggressive IP Distribution List [Online]. https://security.berkeley.edu/services/aggressive-ip-distribution-aid-list/

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Transfer and online learning for IP maliciousness prediction in a concept drift scenario;Wireless Networks;2024-01-28

2. An explainable intrusion detection system based on feature importance;2023 IEEE 12th International Conference on Cloud Networking (CloudNet);2023-11-01

3. Graph Neural Networks for Intelligent Modelling in Network Management and Orchestration: A Survey on Communications;Electronics;2022-10-19