Dari Dataset for Part-of-Speech
Ghezal Ahmad Jan Zia · DepositOnce · 2020
This dataset is related to the task of part-of-speech tagging on the Dari language. It will be usable for many tasks of Natural Language processing on Dari text. The size of the dataset is 12K and it is annotated manually. The tagset used in this dataset is the Universal Tagger.