N-gram-based text filtering for TREC-2

William B. Cavnar · Text REtrieval Conference · 1993

Most text retrieval and filtering systems depend heavily on the accuracy of the text they process. In other words, the various mechanismms that they use depend on every word in the queries being correctly and completely spelled. To get around this limitation, our experimental text filtering system uses N-gram-based matching for document retrieval and routing tasks. The systems's first application was for the TREC-2 retrieval and routing task. Its performace on this task was promising, pointing the way for several types of enhancements, both for speed and effectiveness

Read the paper · More papers on PaperTik