Stacking classifiers for anti-spam filtering of e-mail

Published: 01 Jan 2001, Last Modified: 20 May 2025CoRR 2001EveryoneRevisionsBibTeXCC BY-SA 4.0
Abstract: We evaluate empirically a scheme for combining classifiers, known as stacked generalization, in the context of anti-spam filtering, a novel cost-sensitive application of text categorization. Unsolicited commercial e-mail, or "spam", floods mailboxes, causing frustration, wasting bandwidth, and exposing minors to unsuitable content. Using a public corpus, we show that stacking can improve the efficiency of automatically induced anti-spam filters, and that such filters can be used in real-life applications.
Loading