Using the length of the speech to measure the opinion
Luigi Lancieri, Eric Leprêtre · 2013
This article describes an automated technique that allows to differentiate texts expressing a positive or a negative opinion. The basic principle is based on the observation that positive texts are statistically shorter than negative ones. From this observation of the psycholinguistic human behavior, we derive a heuristic that is employed to generate connoted lexicons with a low level of prior knowledge. The lexicon is then used to compute the level of opinion of an unknown text. Our primary goal is to reduce the need of the human implication (domain and language) in the generation of the lexicon in order to have a process with the highest possible autonomy. The resulting adaptability would represent an advantage with free or approximate expression commonly found in social networks environment.