Inverted index search in data mining

Miloš Ilić, Petar Spalević, Mladen Đ. Veinović · 2014

Data mining has its origins in various disciplines. Two most important data mining disciplines are statistics and machine learning. Data mining is a process of finding new, useful knowledge from data using different techniques. These techniques provide faster and better search for large amounts of data. Inverted index is structure that can be used in data mining process. That is a sorted list of words, with the list of corresponding documents attached to each word. Authors explored inverted index structure for a big corpus of documents. For that purpose, authors created application that use inverted index structure. Application uses open source library named Lucene.

Read the paper · More papers on PaperTik