Research on the Information Extraction Based on Unknown Data Sources
Junhua Chen · Computer Engineering and Science · 2008
This paper introduces the technology of wrapper and clustering.When we do not know their data sources,we will segment the web pages containing tags and search the data section we care without artificial interference.Finally we take advantage of matching and indexing in order to extract information and put them into databases.By studying the second search and data mining,we can search and extract data so as to offer individualized information without knowing the data source.