Naive Rule Induction For Text ClassificationBased On Key-phrases
Nikitas Ν. Karanikolas, Christos Skourlas · WIT transactions on information and communication technologies · 2005
In this paper, we focus on the induction of naive rules for classifying text documents. An algorithm is briefly described for the creation of key-phrases from a given set of documents and these key-phrases are organized and used as features for the automatic classification of new documents. An Authority list of key-phrases is specified by the algorithm containing key-phrases that are frequent within the documents of only one or few classes in the training set. In this framework, this last property permitted us the creation of naive rules that measure the similarity of new documents with the existing classes.