Preliminary study on Kazak Part-of-Speech automatic tagging
Yan Liu · Computer Engineering and Applications Journal · 2008
Part-of-Speech tagging is playing a key role in many such information processing.Kazak,as one of the minority languages and characters being universally applied or used in Xinjiang,some basic problems in natural language treatment become the problems to be solved urgently.The thesis analyzes the configuration of Kazak morpheme characteristics.Based on the completement of one-level tagging of the dictionary,it adopts statistical methods,gaining model training parameter under the bi-gram HMM,and adopting the Viterbi algorithm to complete the Part-of-Speech tagging based on the statistical method.Finally adopting the Kazak language regular storehouse in revising parts of speech.The thesis finally compares and tests the methods of pure use of statistics and that of giving first place to statistical methods and assists the methods being amended with regulation.And final result indicates that the latter method enhances the correctness rate in arrangement.