Corpus-based research into language. In honour of Jan Aarts

N.H.J. Oostdijk, P.J.M. de Haan · Radboud Repository (Radboud University) · 1994

Flor AARTS: A tribute to Jan Aarts. Nelleke OOSTDIJK and Pieter de HAAN: Introduction. PART I: THE ENCODING AND TAGGING OF CORPORA. Stig JOHANSSON: Continuity and change in the encoding of computer corpora. Sidney GREENBAUM and Ni YIBIN: Tagging the British ICE Corpus: English word classes. Geoffrey LEECH, Roger GARSIDE, and Michael BRYANT: The large-scale grammatical tagging of text: Experience with the British National Corpus. Willem MEIJS: Computerized lexicons and theoretical models. Louise GUTHRIE, Joe GUTHRIE, and Jim COWIE: Resolving lexical ambiguity. PART II: PARSING AND DATABASES. Ted BRISCOE: Prospects for practical parsing of unrestricted text: Robust statistical parsing techniques. Fred KARLSSON: Robust parsing of unconstrained text. Clive SOUTER and Eric ATWELL: Using parsed corpora: A review of current practice. Ezra BLACK: An experiment in customizing the Lancaster Treebank. Geoffrey SAMPSON: SUSANNE: A Domesday Book of English grammar. William GALE and Kenneth CHURCH: What is wrong with adding one? PART III: LINGUISTIC EXPLORATION OF THE DATA. Douglas BIBER and Edward FINEGAN: Intra-textual variation within medical research articles. Bengt ALTENBERG: On the functions of such in spoken and written English. Anna-Brita STENSTROM and Jan SVARTVIK: Imparsable speech: Repeats and other nonfluencies in spoken English. References. List of contributors.

Read the paper · More papers on PaperTik