DCU System Report on the WMT 2017 Multi-modal Machine Translation Task
Iacer Calixto, Koel Dutta Chowdhury, Qun Liu · 2017
We report experiments with multi-modal neural machine translation models that incorporate global visual features in different parts of the encoder and decoder, and use the VGG19 network to extract features for all images.In our experiments, we explore both different strategies to include global image features and also how ensembling different models at inference time impact translations.Our submissions ranked 3rd best for translating from English into French, always improving considerably over an neural machine translation baseline across all language pair evaluated, e.g. an increase of 7.0-9.2ME-TEOR points.