Identification of Chinese Names Based on Statistics

Huang De · Zhongwen xinxi xuebao · 2001

Identification of Chinese names is one of important techniques to improve the accuracy of automatic word segmentation. This paper proposes an effective model based on statistics to identify Chinese names. It establishes rewards punishment mechanism and supervised learning mechanism, and presents the reliability for the word segmentation in the model. The experiments show that the precision and recall rate respectively reach 95.97% and 95.52% by close test, while the precision and recall rate are 92.37% and 88.62% by open test.

Read the paper · More papers on PaperTik