Scatter/Gather: a cluster-based approach to browsing large document collections

Douglass R. Cutting, David R. Karger, Jan Ole Pedersen, John W. Tukey · 1992

Document clustering has not been well received as an information retrieval tool. Objections to its use fall into two main categories: first, that clustering is too slow for large corpora (with running time often quadratic in the number of documents); and second, that clustering does not appreciably improve retrieval.

Read the paper · More papers on PaperTik