A Metasearch Engine Using the Measure of Uniquenes

Yamashita Yoko, Shinichi Morishita, Yoshifumi Masunaga · 2003

Abstract The rapid growth of data available in the WWW has been demanding a wide variety of search engines that facili-tate the task of mining informative web pages. However, divergence among ranking strategies of various search engines mayyield serious disagreement among ranks even against the same query, motivating the development of meta-search engines thataggregate multiple ranks into a single list. Although there have been a lot of studies on aggregation procedures, most of themare likely to make little of niche and unique web pages that occupy high positions in a couple of ranks but low places in theothers. In this paper, we propose three different kind of measures to define such unique and niche pages. Tests uncover thatweb pages ranked high according to these measures are typically fresh pages that are not listed in Yahoo!. Key words Web and Internet, information fusion, information retrieval 1. 背景と目的 近年, WWW(World Wide Web) 上のデータ量は爆発的な勢いで増加している. この膨大なデータの中から自分の必要とする情報を, リンクだけを辿って見つけ出すのは困難になってきており, 自分の必要とする情報のキーワードのみから関係するページを検索するサーチエンジンの重要性がますます高まっている. 現在, 多くのユーザに使用されているサーチエンジン数は 10を超えている. これらのサーチエンジンはそれぞれ, データ収集の方式, 索引付けの方式, ランキング方式の戦略が異なっている .そのため, 同一検索キーワードを問い合わせても, サーチエンジン毎に検索結果には違いが生じる. それぞれのサーチエンジンは, 他との差別化を図り独自性のある検索結果を提供している. そこで複数の検索結果を統合することで, 偏りの少ない総合的なランキング結果を返すメタサーチエンジン[3][6] の研究・開発が盛んになっている. また, WWW 上のデータ量が膨大なた

Read the paper · More papers on PaperTik