Automatic identification and labeling of non-clauses based on part of speech

Qiong Li · Journal of Changchun Institute of Technology · 2011

In order to build a finishing compound-sentence corpus for Chinese Information Process,automatic word segmentation and POS tagging work should be completed first of all.On this basis,automatic classification and labeling of levels and relationship between clauses should be conducted.We can use the POS information to develop a set of rules to achieve some non-clause of automatic identification and labeling,but also can build a phrase library,which includes the phrase language fragments.

Read the paper · More papers on PaperTik