Online supervised spam filter evaluation-Reference-Cited by-同舟云学术

Online supervised spam filter evaluation

Published:2007-07 Issue:3 Volume:25 Page:11
ISSN:1046-8188
Container-title:ACM Transactions on Information Systems
language:en
Short-container-title:ACM Trans. Inf. Syst.

Author:

Cormack Gordon V.¹,Lynam Thomas R.¹

Affiliation:

1. University of Waterloo, Waterloo, ON, Canada

Abstract

Eleven variants of six widely used open-source spam filters are tested on a chronological sequence of 49086 e-mail messages received by an individual from August 2003 through March 2004. Our approach differs from those previously reported in that the test set is large, comprises uncensored raw messages, and is presented to each filter sequentially with incremental feedback. Misclassification rates and Receiver Operating Characteristic Curve measurements are reported, with statistical confidence intervals. Quantitative results indicate that content-based filters can eliminate 98% of spam while incurring 0.1% legitimate email loss. Qualitative results indicate that the risk of loss depends on the nature of the message, and that messages likely to be lost may be those that are less critical. More generally, our methodology has been encapsulated in a free software toolkit, which may used to conduct similar experiments.

Publisher

Association for Computing Machinery (ACM)

Subject

Computer Science Applications,General Business, Management and Accounting,Information Systems

Link

https://dl.acm.org/doi/pdf/10.1145/1247715.1247717

Reference46 articles.

1. Agresti A. 1996. An Introduction to Categorical Data Analysis. Wiley New York. Agresti A. 1996. An Introduction to Categorical Data Analysis. Wiley New York.

2. Burton B. 2002. Spamprobe---A fast Bayesian spam filter. http://spamprobe.sourceforge.net. Burton B. 2002. Spamprobe---A fast Bayesian spam filter. http://spamprobe.sourceforge.net.

Cited by 54 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Evaluation of Classification Algorithms for Effective Spam Email Detection Using Spam Email Dataset;2024

2. Efficient e-mail spam filtering approach combining Logistic Regression model and Orthogonal Atomic Orbital Search algorithm;Applied Soft Computing;2023-09

3. Machine Learning Approaches for Text Mining and Spam E-mail Filtering: Industry 4.0 Perspective;Artificial Intelligence and Data Science in Recommendation System: Current Trends, Technologies and Applications;2023-08-14

4. A Novel Fuzzy-Logic-Based Multi-Criteria Metric for Performance Evaluation of Spam Email Detection Algorithms;Applied Sciences;2022-07-12

5. Classifier Calibration: with application to threat scores in cybersecurity;IEEE Transactions on Dependable and Secure Computing;2022