Keyword Extraction Through Contextual Semantic Analysis of Documents
Terry Ruas, William I. Grosky · 2017
Keywords in a text are often used to suggest the main concepts being discussed and help to index them. However, most of the traditional approaches make use of techniques that rely on analyzing just the syntactic aspect of texts, ignoring the meaning they convey and more importantly, the semantic effect of one word over another (context). This paper explores two alternative approaches to extract concept-terms based on semantic features embedded in textual documents extracted from Wikipedia. The first approach extends the concept of Word Sense Disambiguation (WSD), and the second approach enhances the theory behind traditional lexical chains. These applied techniques also consider distinct levels of abstraction with respect to the meanings of words, in addition to the context in which they appear. Our initial results show that both techniques are robust and can extract the main concepts in a document without human intervention or supervision.