Improving classification of tweets using word-word co-occurrence information from a large external corpus

Hugo L. Hammer, Anis Yazidi, Aleksander Bai, Paal Einar Engelstad · 2016

Classifying tweets is an intrinsically hard task as tweets are short messages which makes traditional bags of words based approach inefficient. In fact, bags of words approaches ignores relationships between important terms that do not co-occur literally.

Read the paper · More papers on PaperTik