TEI P5 as a Text Encoding Standard for Multilevel Corpus Annotation

Piotr Bański, Adam Przepiórkowski · 2010

The need for text encoding standards for language resources (LRs) is widely acknowledged: within the International Standards Organization (ISO) Technical Committee 37 / Subcommittee 4 (TC 37 / SC 4), work in this area has been going on since the early 2000s, and working groups devoted to this issue have been set up in two current panEuropean projects, CLARIN (http://www.clari n.eu/) and FLaReNet (http://www.flarenet.e u/). It is obvious that standards are necessary for the interoperability of tools and for the facilitation of data exchange between projects, but they are also needed within projects, especially where multiple partners and multiple levels of linguistic data are involved.

Read the paper · More papers on PaperTik