Neural Machine Translation with Phrase-Level Universal Visual Representations

03/19/2022
by   Qingkai Fang, et al.
0

Multimodal machine translation (MMT) aims to improve neural machine translation (NMT) with additional visual information, but most existing MMT methods require paired input of source sentence and image, which makes them suffer from shortage of sentence-image pairs. In this paper, we propose a phrase-level retrieval-based method for MMT to get visual information for the source input from existing sentence-image data sets so that MMT can break the limitation of paired sentence-image input. Our method performs retrieval at the phrase level and hence learns visual information from pairs of source phrase and grounded region, which can mitigate data sparsity. Furthermore, our method employs the conditional variational auto-encoder to learn visual representations which can filter redundant visual information and only retain visual information related to the phrase. Experiments show that the proposed method significantly outperforms strong baselines on multiple MMT datasets, especially when the textual context is limited.

READ FULL TEXT

page 2

page 3

page 7

research
09/21/2020

Generative Imagination Elevates Machine Translation

There are thousands of languages on earth, but visual perception is shar...
research
07/26/2022

Multimodal Neural Machine Translation with Search Engine Based Image Retrieval

Recently, numbers of works shows that the performance of neural machine ...
research
06/06/2016

Neural Machine Translation with External Phrase Memory

In this paper, we propose phraseNet, a neural machine translator with a ...
research
08/20/2019

Phrase Localization Without Paired Training Examples

Localizing phrases in images is an important part of image understanding...
research
04/07/2020

Towards Multimodal Simultaneous Neural Machine Translation

Simultaneous translation involves translating a sentence before the spea...
research
05/31/2022

VALHALLA: Visual Hallucination for Machine Translation

Designing better machine translation systems by considering auxiliary in...
research
02/28/2019

Non-Parametric Adaptation for Neural Machine Translation

Neural Networks trained with gradient descent are known to be susceptibl...

Please sign up or login with your details

Forgot password? Click here to reset