XML: URL Data Set Creation for Future Web Mining Research Avenues
Krishna Murthy. A, Suresha · Computational intelligence · 2012
The rapid expansion of the internet has made web a popular place for disseminating and collecting information and also it opens many research topics on varies research fields. Since last few years, several attempts have been made on Web based research particularly based on HTML web pages because of its more availability. So that many Research Data sets have created and few of them are made available on Web. But W3 consortium stated that, HTML does not provide a better description of semantic structure of the web page contents. To overcome this draw back Web developers started to develop Web page(s) on XML, Flash kind of new technologies [1]. It makes a way for new Research methods. This article mainly focuses on Data Set creation on XML Web pages by using Sequential search, Link Extraction and string based classification methods for future research avenues on XML Web pages.