CANTINA+-Reference-Cited by-同舟云学术

CANTINA+

Published:2011-09 Issue:2 Volume:14 Page:1-28
ISSN:1094-9224
Container-title:ACM Transactions on Information and System Security
language:en
Short-container-title:ACM Trans. Inf. Syst. Secur.

Author:

Xiang Guang¹,Hong Jason¹,Rose Carolyn P.¹,Cranor Lorrie¹

Affiliation:

1. Carnegie Mellon University

Abstract

Phishing is a plague in cyberspace. Typically, phish detection methods either use human-verified URL blacklists or exploit Web page features via machine learning techniques. However, the former is frail in terms of new phish, and the latter suffers from the scarcity of effective features and the high false positive rate (FP). To alleviate those problems, we propose a layered anti-phishing solution that aims at (1) exploiting the expressiveness of a rich set of features with machine learning to achieve a high true positive rate (TP) on novel phish, and (2) limiting the FP to a low level via filtering algorithms. Specifically, we proposed CANTINA+, the most comprehensive feature-based approach in the literature including eight novel features, which exploits the HTML Document Object Model (DOM), search engines and third party services with machine learning techniques to detect phish. Moreover, we designed two filters to help reduce FP and achieve runtime speedup. The first is a near-duplicate phish detector that uses hashing to catch highly similar phish. The second is a login form filter, which directly classifies Web pages with no identified login form as legitimate. We extensively evaluated CANTINA+ with two methods on a diverse spectrum of corpora with 8118 phish and 4883 legitimate Web pages. In the randomized evaluation, CANTINA+ achieved over 92% TP on unique testing phish and over 99% TP on near-duplicate testing phish, and about 0.4% FP with 10% training phish. In the time-based evaluation, CANTINA+ also achieved over 92% TP on unique testing phish, over 99% TP on near-duplicate testing phish, and about 1.4% FP under 20% training phish with a two-week sliding window. Capable of achieving 0.4% FP and over 92% TP, our CANTINA+ has been demonstrated to be a competitive anti-phishing solution.

Funder

Army Research Office

Division of Computing and Communication Foundations

Publisher

Association for Computing Machinery (ACM)

Subject

Safety, Risk, Reliability and Quality,General Computer Science

Link

https://dl.acm.org/doi/pdf/10.1145/2019599.2019606

Reference28 articles.

1. 3sharp report. 2006. Gone phishing: Evaluating anti-phishing tools for windows. http://www.3sharp.com/projects/antiphishing/gone-phishing.pdf. 3sharp report . 2006. Gone phishing: Evaluating anti-phishing tools for windows. http://www.3sharp.com/projects/antiphishing/gone-phishing.pdf.

2. Detecting visually similar Web pages

Cited by 310 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Enhanced Malicious Traffic Detection in Encrypted Communication Using TLS Features and a Multi-class Classifier Ensemble;Journal of Network and Systems Management;2024-08-10

2. A new perspective exploration of machine learning algorithms for defending Side-Channel attacks;Multimedia Tools and Applications;2024-07-02

3. Understanding Characteristics of Phishing Reports from Experts and Non-Experts on Twitter;IEICE Transactions on Information and Systems;2024-07-01

4. APFormer: Anti-Phishing Transformer for Website-Phishing Detection Via Joint Feature Learning;2024 International Conference on Engineering & Computing Technologies (ICECT);2024-05-23

5. Phishing URL detection generalisation using Unsupervised Domain Adaptation;Computer Networks;2024-05