A Corpus-Independent Feature Set for Style-Based Text Categorization

Moshe Koppel, Navot Akiva, Ido Dagan · 2003

We suggest a corpus-independent feature set appropriate for style-based text categorization problems. To achieve this, we introduce a new measure on linguistic features, called stability, which captures the extent to which a language element, such as a word or syntactic construct, is replaceable by semantically equivalent elements. This measure may be

Read the paper · More papers on PaperTik