Corpus Linguistics and Information Structure Research

Anke Lüdeling, Julia Ritz, Manfred Stede, Amir Zeldes · Oxford University Press eBooks · 2015

Abstract This chapter describes the contributions that Corpus Linguistics (the study of linguistic phenomena by means of systematically exploiting collections of naturally-occurring linguistic data) can make to IS research. It discusses issues of designing a corpus that can serve as a basis for qualitative or quantitative studies, and then turns to the central issue of data annotation: what corpora are available that have been annotated with IS-related annotations, and how can such annotations be evaluated? In case a corpus does not have direct IS annotation, can other types of annotations, especially in the form of multi-layer annotation, be used as indirect evidence for the presence of IS phenomena? Next, the present state of the art in automatic IS annotation (by means of techniques from computational linguistics) is sketched, and finally, several sample studies that exploit IS annotations are introduced briefly.

Read the paper · More papers on PaperTik