作者
Rizka Purwanto, Arindam Pal, Alan Blair, Sanjay Jha
发表日期
2022
期刊
IEEE Transactions on Information Forensics and Security
简介
In this paper, we propose a feature-free method for detecting phishing websites using the Normalized Compression Distance (NCD), a parameter-free similarity measure which computes the similarity of two websites by compressing them, thus eliminating the need to perform any feature extraction. It also removes any dependence on a specific set of website features. This method examines the HTML of webpages and computes their similarity with known phishing websites, in order to classify them. We use the Furthest Point First algorithm to perform phishing prototype extractions, in order to select instances that are representative of a cluster of phishing webpages. We also introduce the use of an incremental learning algorithm as a framework for continuous and adaptive detection without extracting new features when concept drift occurs. On a large dataset, our proposed method significantly outperforms previous …
引用总数
学术搜索中的文章
RW Purwanto, A Pal, A Blair, S Jha - IEEE Transactions on Information Forensics and …, 2022