The performance of text similarity algorithms

Didik Dwi Prasetya, Aji Prasetya Wibawa, Tsukasa Hirashima · International Journal of Advances in Intelligent Informatics · 2018

Text similarity measurement compares text with available references to indicate the degree of similarity between those objects. There have been many studies of text similarity and resulting in various approaches and algorithms. This paper investigates four majors text similarity measurements which include String-based, Corpus-based, Knowledge-based, and Hybrid similarities. The results of the investigation showed that the semantic similarity approach is more rational in finding substantial relationship between texts.

Read the paper · More papers on PaperTik