Adding semantic role annotation to a corpus of written Dutch
Paola Monachesi, Gerwert Stevens, Jantine Trapman · 2007
We present an approach to automatic semantic role labeling (SRL) carried out in the context of the Dutch Language Corpus Initiative (D-Coi) project. Adapting earlier research which has mainly focused on English to the Dutch situation poses an interesting challenge especially because there is no semantically annotated Dutch corpus available that can be used as training data. Our automatic SRL approach consists of three steps: bootstrapping from a syntactically annotated corpus by means of a rule-based tagger developed for this purpose, manual correction on the basis of the Prop-Bank guidelines which have been adapted to Dutch and training a machine learning system on the manually corrected data.