Dari Dataset for Part-of-Speech

Ghezal Ahmad Jan Zia · DepositOnce · 2020

This dataset is related to the task of part-of-speech tagging on the Dari language. It will be usable for many tasks of Natural Language processing on Dari text. The size of the dataset is 12K and it is annotated manually. The tagset used in this dataset is the Universal Tagger.

Read the paper · More papers on PaperTik