Does Neural Machine Translation Benefit from Larger Context?

Sébastien Jean, Stanislas Lauly, Orhan Fırat, Kyunghyun Cho · arXiv (Cornell University) · 2017

We propose a neural machine translation architecture that models the surrounding text in addition to the source sentence. These models lead to better performance, both in terms of general translation quality and pronoun prediction, when trained on small corpora, although this improvement largely disappears when trained with a larger corpus. We also discover that attention-based neural machine translation is well suited for pronoun prediction and compares favorably with other approaches that were specifically designed for this task.

Read the paper · More papers on PaperTik