Web Crawler Design Issues: A Review

Deepika, Ashutosh Dixit · International Journal of Managment, IT and Engineering · 2012

The large size and the dynamic nature of the Web increase the need for updating Web based information retrieval systems. Crawlers facilitate the process by following the hyperlinks in Web pages to automatically download a partial snapshot of the Web. While some systems rely on crawlers that exhaustively crawl the Web, others focus on topic specific collections. In present paper the various types of crawlers are discussed. The paper also discusses several web crawler design issues along with their solutions.

Read the paper · More papers on PaperTik