TextProc: a natural language processing framework

Janez Brezovnik, Milan Ojsteršek · 2010

Our implementation of a natural language processing framework (called TextProc) is described in this paper. We start with a general overview of the framework and continue with detailed description of its parts. Actual language processing is implemented as software plug-ins. Plug-ins can be put together into processes that perform a practical natural processing function. One such process is plagiarism detection, which is explained in detail. The process for plagiarism detection is actually used in digital library of University of Maribor and the integration of digital library with TextProc is also briefly described. At the end of this paper some ideas for future development are given.

Read the paper · More papers on PaperTik