Design of a Framework for Knowledge Based Web Page Ranking
PremSagar Sharma, Sharma A.K, Pankaj Kumar Garg · International Journal of Engineering and Technology · 2017
Web is growing exponentially.The search mechanisms need to provide relevant and high quality web pages that too in short time to the internet user.The standard search engines utilize the link structure of the web to measure the quality of Web pages.Wherein it has been observed that the some less popular and lowly ranked but significantly important web pages remains missing.In this paper a framework for knowledge based web page ranking is being presented.It provides relevant and quality information in desirable time with the help of a proxy server.This framework exploits the content of the web to measure the quality of web pages.Keywords -Introduction, web page ranking algorithms, proposed framework for knowledge base web page raking.1. INTRODUCTION Increasing popularity of internet has rendered World Wide Web, a rich collection of hypertext documents belonging to different domains.Web based information retrieval system called search engine, though has made things easy for information seeker but still it does not provide guarantee about the correctness of the information provided to the user.Many times the information is not precise.Information retrieval systemprovides the information to the user based on certain retrieval criterion.For instance, it may search the web for identifying documents which contain information on a given subject.Due to the large size of the WWW, it is very common that a large number of documents get identified related to a particular domain.Therefore to help guide users towards finding the best matching documents, a ranking mechanism is employed by the search engine.Common methods for ranking are either based on relevancy where the documents are ordered from most relevant to least relevant or on the basis ofpopularity where documents are ordered from being most popular to least popular.It is important to understand that the term popularity is normally the result of link analysis and not user feedback.A web search engine typically consists of a ranking System thatmeasure the importance of Web Pages (discussed in sec-1.2). Web MiningIn web mining, the techniques of data mining are used to automatically discover and extract information from Web documents and web services.With a view to extract something useful out of the Web.The following tasks are generally performed for this purpose: 1) Resource finding: Useful resources are retrieved from the web documents i.e. we extract the data which are accessible on the web either through online or offline mode.2) Information selection and pre-processing: Specific information is selected automatically and pre processing of that information is carried out with a view to data cleaning, normalization, feature extraction etc. 3) Transformation: The original retrieved data is transformed into informationby rejuvenation of stop words to obtain the necessary representation for finding phrases in training mass.4) Generalization:The general patterns present in individual web sites or across multiple sites are found by generalization.Machine learning and data mining techniques are employed for this purpose.5) Analysis:Validation and interpretation of the mined patterns is done in phase of analysis.It has got an important role for pattern mining.The human being plays an important role for knowledge discovery technique on the web.1.2 Web Mining Taxonomy Web mining is categorized into three different types as shown in Fig.-1[13].