Corpus data in usage-based linguistics
Stefan Τh. Gries · Human cognitive processing · 2011
The use of corpus data in cognitive linguistics brings with it a host of methodological problems. One concerns the degree of granularity that provides the most insightful results. The present study investigates two granularity issues – different inflectional forms and (register-)based corpus parts. First, I compare the results of a lemma-based corpus analysis of an English argument structure construction to an inflectional-form-based corpus analysis to determine whether the two approaches result in different suggestions concerning the semantics of the construction at issue. Second, I outline how to determine whether data from different corpus parts/registers result in different semantic generalizations of the same construction and how relevant corpus distinctions can be determined in an objective bottom-up manner.