JOURNAL ARTICLE

Classification of Malicious URLs Using Machine Learning

S. AbadHassan GholamyMohammad Aslani

Year: 2023 Journal:   Sensors Vol: 23 (18)Pages: 7760-7760   Publisher: Multidisciplinary Digital Publishing Institute

Abstract

Amid the rapid proliferation of thousands of new websites daily, distinguishing safe ones from potentially harmful ones has become an increasingly complex task. These websites often collect user data, and, without adequate cybersecurity measures such as the efficient detection and classification of malicious URLs, users’ sensitive information could be compromised. This study aims to develop models based on machine learning algorithms for the efficient identification and classification of malicious URLs, contributing to enhanced cybersecurity. Within this context, this study leverages support vector machines (SVMs), random forests (RFs), decision trees (DTs), and k-nearest neighbors (KNNs) in combination with Bayesian optimization to accurately classify URLs. To improve computational efficiency, instance selection methods are employed, including data reduction based on locality-sensitive hashing (DRLSH), border point extraction based on locality-sensitive hashing (BPLSH), and random selection. The results show the effectiveness of RFs in delivering high precision, recall, and F1 scores, with SVMs also providing competitive performance at the expense of increased training time. The results also emphasize the substantial impact of the instance selection method on the performance of these models, indicating its significance in the machine learning pipeline for malicious URL classification.

Keywords:
Computer science Support vector machine Machine learning Random forest Locality-sensitive hashing Artificial intelligence Pipeline (software) Identification (biology) Context (archaeology) Data mining Task (project management) Hash function Computer security Hash table Engineering

Metrics

34
Cited By
21.03
FWCI (Field Weighted Citation Impact)
22
Refs
0.99
Citation Normalized Percentile
Is in top 1%
Is in top 10%

Citation History

Topics

Spam and Phishing Detection
Physical Sciences →  Computer Science →  Information Systems
Advanced Malware Detection Techniques
Physical Sciences →  Computer Science →  Signal Processing
Web Data Mining and Analysis
Physical Sciences →  Computer Science →  Information Systems

Related Documents

JOURNAL ARTICLE

Detecting Malicious URLs Using Machine Learning

Munkhbayar Bat-ErdeneOyuntungalag BuyanturDensmaa BatbayarByambadorj DondogmegdJuhee Choi

Journal:   Journal of Electrical Engineering and Technology Year: 2025 Vol: 20 (6)Pages: 4537-4550
JOURNAL ARTICLE

Detection of malicious URLs using machine learning

Nuria Reyes-DortaPino Caballero‐GilCarlos Rosa-Remedios

Journal:   Wireless Networks Year: 2024 Vol: 30 (9)Pages: 7543-7560
JOURNAL ARTICLE

Recognition of Malicious URLs Using Machine Learning

Sweety GargMs. Sayed Shifa Mohd Imran

Journal:   INTERANTIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING AND MANAGEMENT Year: 2024 Vol: 08 (008)Pages: 1-4
© 2026 ScienceGate Book Chapters — All rights reserved.