quanteda: Quantitative Analysis of Textual Data

Kenneth R. Benoit, Kohei Watanabe, Haiyan Wang, Paul Nulty, Adam Obeng, Stefan C Müller, Akitaka Matsuo, Will Lowe · 2015

A fast, flexible, and comprehensive framework for quantitative text analysis in R. Provides functionality for corpus management, creating and manipulating tokens and n-grams, exploring keywords in context, forming and manipulating sparse matrices of documents by features and feature co-occurrences, analyzing keywords, computing feature similarities and distances, applying content dictionaries, applying supervised and unsupervised machine learning, visually representing text and text analyses, and more.

Read the paper · More papers on PaperTik