Towards Lightweight URL-Based Phishing Detection

Open Access

13 June 2021

journal article
research article
Published by MDPI AG in Future Internet

Vol. 13 (6), 154
https://doi.org/10.3390/fi13060154

Abstract

Nowadays, the majority of everyday computing devices, irrespective of their size and operating system, allow access to information and online services through web browsers. However, the pervasiveness of web browsing in our daily life does not come without security risks. This widespread practice of web browsing in combination with web users’ low situational awareness against cyber attacks, exposes them to a variety of threats, such as phishing, malware and profiling. Phishing attacks can compromise a target, individual or enterprise, through social interaction alone. Moreover, in the current threat landscape phishing attacks typically serve as an attack vector or initial step in a more complex campaign. To make matters worse, past work has demonstrated the inability of denylists, which are the default phishing countermeasure, to protect users from the dynamic nature of phishing URLs. In this context, our work uses supervised machine learning to block phishing attacks, based on a novel combination of features that are extracted solely from the URL. We evaluate our performance over time with a dataset which consists of active phishing attacks and compare it with Google Safe Browsing (GSB), i.e., the default security control in most popular web browsers. We find that our work outperforms GSB in all of our experiments, as well as performs well even against phishing URLs which are active one year after our model’s training.

Keywords

This publication has 21 references indexed in Scilit:

New rule-based phishing detection method
Expert Systems with Applications, 2016
A novel approach to protect against phishing attacks at client side using auto-updated white-list
EURASIP Journal on Information Security, 2016
A random forest guided tour
TEST, 2016
Security Busters: Web browser security vs. rogue sites
Computers & Security, 2015
Intelligent rule‐based phishing websites classification
IET Information Security, 2014
Adaptive regularization of weight vectors
Machine Learning, 2013
CANTINA+
ACM Transactions on Information and System Security, 2011
Greedy function approximation: A gradient boosting machine.
The Annals of Statistics, 2001
ANFIS: adaptive-network-based fuzzy inference system
IEEE Transactions on Systems, Man, and Cybernetics, 1993
Multilayer feedforward networks are universal approximators
Neural Networks, 1989

Cited by 25 articles