Text Information Extraction Based on Hidden Markov Model
Zhiping Chen · Acta Simulata Systematica Sinica · 2004
Text information extraction is an important method of processing large quantity of text. The application of hidden Markov model to information extraction is a relatively new research topic. A new algorithm based on hidden Markov Model is proposed for text information extraction. The algorithm makes use of the information of format and list separators to segment text, and then combines hidden Markov model for text information extraction. The simulation results show that the new algorithm exceeds the original one that hasnt segment text into blocks in precision and recall.