Building a syntactic rules-based stemmer to improve search effectiveness for arabic language

Walid Cherif, Abdellah Madani, Mohamed Kissi · 2014

Nowadays, The world is experiencing a huge growth in the volume of exchanged texts, which makes some of it untapped. Text Mining is the set of techniques that analyze these large masses of information, extract relations that can be unknown beforehand and provide solutions that help decision making. In this sense, stemming is a common requirement of these techniques. It includes reducing different grammatical forms of a word and bringing them to a common base form. In what follows, we will discuss these treatment methods for arabic text, show their limits and provide new algorithm to improve them.

Read the paper · More papers on PaperTik