Gusti Ahmad Fanshuri Alfarisy, Fitra A. Bachtiar
Crawlers are commonly used to traverse and collect all public webs that are connected through links. The general crawlers could not be used for crawling or collecting web pages with a particular topic such as food recipe. This paper, propose focused web crawler for Indonesian food recipes using simple classification based on the analysis of Indonesian recipes available on the internet, providing priority levels of a link through anchor text and URLs, and restricting the traverse by the depth. The focused crawler is tested on 4 different query to collect 100 recipes each. The results show that focused web crawler provide higher relevance of 81.75 % than general crawler that uses breath first with 16.00 % relevance. Furthermore, with the same amount of time, focused web crawler is able to collect more relevant web page than the general crawler. Therefore, the proposed crawler can collect recipes on the web based on user query effectively. © 2017 IEEE.
Faculty of Computer Science, Brawijaya University, Malang, Indonesia