Statistical Corpus Exploitation
Hermann L. Moisl · Oxford University Press eBooks · 2013
The aim of this chapter is to encourage corpus linguists to use quantitative and more specifically statistical methods in analysing large digital electronic corpora, focusing in particular on cluster analysis. The first part of the discussion motivates the use of cluster analysis in corpus linguistics, the second gives an outline account of data creation and clustering with reference to the Newcastle Electronic Corpus of Tyneside English, and the third is a selective literature review.