A Document Content Extraction Model Using Keyword Correlation Analysis.
Jiang‐Liang Hou, Chuan-An Chan · 2003
Owing to the drastic development of the information and Internet technologies, large amount of information and documents can be easily accessed through the electronic network. In addition to the efficiency of document acquisition, another typical issue for document management is the document content extraction. In order to provide the critical contents of a document to the knowledge requester, a thesaurus indicating the keyword correlation is required for accurate content extraction. This paper presents an algorithm for