Parallel writing in East Asian languages and its representation in metadata in light of the DCMI abstract model

Akira Miyazawa · International Conference on Dublin Core and Metadata Applications · 2007

This paper discusses the parallel writing tradition in East Asian languages and its representation in metadata. Parallel writing systems in these languages do not use the same scripts, but they all share common scheme and have well-established tradition in bibliographic data. Their data representation in the MARC bibliographic format is handled in variety of ways. Even in the metadata world, representation of parallel writing shows some inconsistencies. It is therefore desirable to establish new common way of representation. For this purpose, this paper discusses the class of the represented values in terms of the DCMI Abstract Model (DCAM). In the case of properties such as Title, it is possible to see the associated value as literal, but for parallel writing, it is more appropriate to see such value as a of words. Accordingly, parallel writing can be represented as multiple value strings associated with value of the class sequence of words. Even so, one remaining problem is that the language tags used in the value string language cannot also specify writing systems. Enumeration of the types of writing systems in various languages and registration with RFC 4646 would be required in order to express this information in DCAM value string languages.

Read the paper · More papers on PaperTik