SEPHWIR: Search Engine Parsing for Hidden Web Information Retrieval

Manpreet Singh Sehgal, Sachin Kumar Gupta, Twinkle Sehgal · 2022

The world wide web consists of web pages and databases from which web pages can be generated on demand. It is presumed that the information stored in the databases is of better quality than the one published onto already created static webpages. The web search engine architectures are tuned to access static webpages and index and rank them in their results against the query. This results into the accessibility to the kind of information that is not of better quality. This paper talks about the approach to parse such results of the search engines and find the entry points to the databases to fetch the better-quality information.

Read the paper · More papers on PaperTik