CoNLL-X SharedTask: Multi-lingual Dependency Parsing
Johan Hall, Jens Nilsson · KTH Publication Database DiVA (KTH Royal Institute of Technology) · 2006
The goal of this report is to summarize our experiments and present the final result of our participation in the CoNLL-X Shared Task 2006. The topic of this year's shared task was multi-lingual dependency parsing. The organizers have prepared 13 existing dependency treebanks so that they all comply to the same markup format. The training and test data for the languages differ in size, granularity and quality, but they have tried to even out differences in the markup format. No additional information is allowed to be used besides the provided training data, forcing the parser to be fully automatic and data-driven. Ideally, the same parser should be trainable for all languages, possibly by adjusting parameters. The main goal is to assign labeled dependency structure for all languages on held out test data, approximately 5 000 tokens for each language. The main metric for comparison of the different parsers of the participants is therefore labeled attachment score.