Research on Micro-Blog Hot Topics Mining Model on Sentence Constituents

Long Xue Xiao · Information Sciences · 2015

Because the traditional clustering analysis is not applicable to short text, this article selectsthe sentence similarity computing method based on component to calculate similarity between short texts.We obtain sentence constituents by parsing, and choose the words constitute parts of the sentence as keywords. Then we calculate the semantic similarity between key words based on the Hownet. The similaritybetween the texts can calculate by weighted summing the semantic similarity between key words. Accord-ing to this, we can construct the text similarity matrix and do clustering analysis on it. At last, we can minethe hot topics of micro-blogs. Finally, the experiment proved the feasibility of the proposed method.

Read the paper · More papers on PaperTik