A WEB SYLLABUS CRAWLER AND ITS EFFICIENCY EVALUATION
Yoshihiro Matsunaga, Shintaro Yamada, Eisuke Ito, Sachio Hirokawa · Institutional Repositories DataBase (IRDB) · 2003
Many educational material are available on the web. Typical examples are syllabus pages of universities. A university and a department of the university usually provides a complete list of the courses which they offer. Therefore, syllabus data are not only typical but outstanding educational material with respect to quality and quantity. We are developing a syllabus database of Japanese universities. As the first step of the project, we developed a crawler that gathers syllabus pages on the Web. This paper describes a crawling mechanism which utilizes the characteristic keywords of syllabus pages. The efficiency evaluation is presented compared with general purpose crawlers.