Parallel Text Collections at Linguistic Data Consortium

Xiaoyi Ma · 1999

The Linguistic Data Consortium (LDC) is an open consortium of universities, companies and government research laboratories. It creates, collects and distributes speech and text databases, lexicons, and other resources for research and development purposes. This paper describes past and current work on creation of parallel text corpora, and reviews existing and upcoming collections at LDC.

Read the paper · More papers on PaperTik