Cross language text categorization by acquiring multilingual domain models from comparable corpora

Alfio Gliozzo, Carlo Strapparava · 2005

In a multilingual scenario, the classical monolingual text categorization problem can be reformulated as a cross language TC task, in which we have to cope with two or more languages (e.g. English and Italian). In this setting, the system is trained using labeled examples in a source language (e.g. English), and it classifies documents in a different target language (e.g. Italian).

Read the paper · More papers on PaperTik