Statistical Corpus Exploitation

Hermann L. Moisl · Oxford University Press eBooks · 2013

The aim of this chapter is to encourage corpus linguists to use quantitative and more specifically statistical methods in analysing large digital electronic corpora, focusing in particular on cluster analysis. The first part of the discussion motivates the use of cluster analysis in corpus linguistics, the second gives an outline account of data creation and clustering with reference to the Newcastle Electronic Corpus of Tyneside English, and the third is a selective literature review.

Read the paper · More papers on PaperTik