A multilayer annotated corpus for Turkish
Olcay Taner Yıldız, Koray Ak, Gökhan Ercan, Ozan Topsakal, Cengiz Asmazoglu · 2018
In this paper, we present the first multilayer annotated corpus for Turkish, which is a low-resourced agglutinative language. Our dataset consists of 9,600 sentences translated from the Penn Treebank Corpus. Annotated layers contain syntactic and semantic information including morphological disambiguation of words, named entity annotation, shallow parse, sense annotation, and semantic role label annotation.