Research of intelligent word segmentation and information retrieval

Xiaofei Li, Xusheng Xie · 2010

Chinese information retrieval process is somewhat different from the English information retrieval process. In consideration of the existing problems and difficulties of Chinese language information processing, Hibernate search was introduced to exploit information retrieval engine in this paper. A Chinese language analyzer based on the word stock was adopted to process Chinese language information, therefore this analyzer could advance with the times by updating the word stock at any time. However, ambiguity errors caused by the Chinese language analyzer always interfered with the degree of accuracy of the result. During the time of information retrieval, a secondary word segmentation algorithm was used in order to improve Chinese language information retrieval precision. The result list given in this paper had shown that the intelligent Chinese segmentation algorithm had improved the system performance well.

Read the paper · More papers on PaperTik