Focused Web Crawlers on Domain-Specific Retrieval Systems

Ika Oktavia Suzanti, Fakhrur Razi, Husni Husni, Eka Mala Sari Rochman, Nurhayati Fitriani · IOP Conference Series Materials Science and Engineering · 2021

Abstract The need for a large and growing internet has formed a new culture that symbolizes widespread dissemination of knowledge, information and data. Everyday hundreds of pages as well as new information were added. deleted and modified. Search engines is a finding information tool. First step of search engines is web crawler, a process crawling webpage to obtain information about its content. Focused web crawler is one of web crawler that scrapping certain relevant webpages to given topic and ignores others that have nothing to do with. Indonesia is an archipelago that number of island reached 17.491. where it makes Indonesia rich of different kinds of travel, culture and food that is characteristic of each region. The amount of information available does not rule out possibility that there are some food recipes do not include staples used so that web crawlers are needed to find out the ingredients. With this application, it is expected to facilitate the search for information on various processed recipes with meat-based ingredients from all regions in Indonesia.

Read the paper · More papers on PaperTik