Document Type
Conference Paper
Rights
Available under a Creative Commons Attribution Non-Commercial Share Alike 4.0 International Licence
Disciplines
1.2 COMPUTER AND INFORMATION SCIENCE
Abstract
Because of the changing nature of spam, a spam filtering system that uses machine learning will need to be dynamic. This suggests that a case-based (memory-based) approach may work well. Case-Based Reasoning (CBR) is a lazy approach to machine learning where induction is delayed to run time. This means that the case base can be updated continuously and new training data is immediately available to the induction process. In this paper we present a detailed description of such a system called ECUE and evaluate design decisions concerning the case representation. We compare its performance with an alternative system that uses Na¨ıve Bayes. We find that there is little to choose between the two alternatives in cross-validation tests on data sets. However, ECUE does appear to have some advantages in tracking concept drift over time.
DOI
https://doi.org/10.1007/978-3-540-28631-8_11
Recommended Citation
Delany, Sarah Jane and Cunningham, Padraig: An analysis of case-base editing in a spam filtering systems. Advances in Case-Based Reasoning (Proceedings of the 7th. European Conferenc e on Case Based Reasoning, ECCBR-04), edited by P.Funk and Gonzales Calero , LNAI 3155, pp.128-141.
Publication Details
In Advances in Case-Based Reasoning (Proceedings of the 7th. European Conferenc e on Case Based Reasoning, ECCBR-04), edited by P.Funk and Gonzales Calero , LNAI 3155, pp.128-141. Available from the publisher here
doi:10.1007/978-3-540-28631-8_11