Classification of Malicious URLs Using Machine Learning-Reference-Cited by-同舟云学术

Classification of Malicious URLs Using Machine Learning

Published:2023-09-08 Issue:18 Volume:23 Page:7760
ISSN:1424-8220
Container-title:Sensors
language:en
Short-container-title:Sensors

Author:

Abad Shayan¹,Gholamy Hassan¹,Aslani Mohammad¹

Affiliation:

1. Department of Computer and Geo-Spatial Sciences, University of Gävle, 801 76 Gävle, Sweden

Abstract

Amid the rapid proliferation of thousands of new websites daily, distinguishing safe ones from potentially harmful ones has become an increasingly complex task. These websites often collect user data, and, without adequate cybersecurity measures such as the efficient detection and classification of malicious URLs, users’ sensitive information could be compromised. This study aims to develop models based on machine learning algorithms for the efficient identification and classification of malicious URLs, contributing to enhanced cybersecurity. Within this context, this study leverages support vector machines (SVMs), random forests (RFs), decision trees (DTs), and k-nearest neighbors (KNNs) in combination with Bayesian optimization to accurately classify URLs. To improve computational efficiency, instance selection methods are employed, including data reduction based on locality-sensitive hashing (DRLSH), border point extraction based on locality-sensitive hashing (BPLSH), and random selection. The results show the effectiveness of RFs in delivering high precision, recall, and F1 scores, with SVMs also providing competitive performance at the expense of increased training time. The results also emphasize the substantial impact of the instance selection method on the performance of these models, indicating its significance in the machine learning pipeline for malicious URL classification.

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Biochemistry,Instrumentation,Atomic and Molecular Physics, and Optics,Analytical Chemistry

Link

https://www.mdpi.com/1424-8220/23/18/7760/pdf

Reference22 articles.

1. A Novel Approach for Phishing URLs Detection Using Lexical Based Machine Learning in a Real-Time Environment;Gupta;Comput. Commun.,2021

2. Cybersecurity;Veale;Internet Policy Rev.,2020

3. Phishing URL Detection Using Machine Learning Methods;Ahammad;Adv. Eng. Softw.,2022

4. Phorecasting Phishing Attacks: A New Approach for Predicting the Appearance of Phishing Websites;Wardman;Cyber-Secur. Digit.,2016

5. Security Busters: Web Browser Security vs. Rogue Sites;Virvilis;Comput. Secur.,2015

Cited by 6 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. AI-enabled approach for enhancing obfuscated malware detection: a hybrid ensemble learning with combined feature selection techniques;International Journal of System Assurance Engineering and Management;2024-03-28

2. Detection of phishing addresses and pages with a data set balancing approach by generative adversarial network (GAN) and convolutional neural network (CNN) optimized with swarm intelligence;Concurrency and Computation: Practice and Experience;2024-01-29

3. Developing a Model to Detect Malicious URLs using Different Classification Algorithms;2024 21st Learning and Technology Conference (L&T);2024-01-15

4. Efficient Classification of Malicious URLs: M-BERT—A Modified BERT Variant for Enhanced Semantic Understanding;IEEE Access;2024

5. WebGuardML: Safeguarding Users with Malicious URL Detection Using Machine Learning;2023 26th International Conference on Computer and Information Technology (ICCIT);2023-12-13