On the Parallel Implementation of Sparse Matrix Information Retrieval Engine
Ankit Kumar Jain, Nazli Goharian · 2002
We demonstrate a parallel implementation of a sparse matrix information retrieval engine. We use a shared nothing PC cluster. We perform our experiments with TREC disk 4 and 5 data, a NIST 2 Gigabytes standard benchmark text collection on 2, 4, 6, 8, 10, 12 and 14 processing nodes with different queries. We compare the results with the results of sequential inverted index, a conventional and common indexing and query processing method. The experimental results are promising and show a significant speedup.