Unsupervised Multi-Modal Neural Machine Translation

descriptionPublicationkeyboard_double_arrow_right Article , Preprint , Conference object 01 Jun 2019Embargo end date: 01 Jan 2018Publisher:IEEEJournal:2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Authors: Yuanhang Su; Kai Fan 0002; Nguyen Bach; C.-C. Jay Kuo; Fei Huang 0002;

doi: 10.1109/cvpr.2019.01073 , 10.48550/arxiv.1811.11365

arXiv: 1811.11365

Unsupervised Multi-Modal Neural Machine Translation

- Summary
- Subjects
- Metrics

Abstract

Unsupervised neural machine translation (UNMT) has recently achieved remarkable results with only large monolingual corpora in each language. However, the uncertainty of associating target with source sentences makes UNMT theoretically an ill-posed problem. This work investigates the possibility of utilizing images for disambiguation to improve the performance of UNMT. Our assumption is intuitively based on the invariant property of image, i.e., the description of the same visual content by different languages should be approximately similar. We propose an unsupervised multi-modal machine translation (UMNMT) framework based on the language translation cycle consistency loss conditional on the image, targeting to learn the bidirectional multi-modal translation simultaneously. Through an alternate training between multi-modal and uni-modal, our inference model can translate with or without the image. On the widely used Multi30K dataset, the experimental results of our approach are significantly better than those of the text-only UNMT on the 2016 test dataset.

Accepted to CVPR 2019

Related Organizations

University of California System
United States
Alibaba Group (China)
China (People's Republic of)
University of Southern California
United States

Keywords

FOS: Computer and information sciences, Computer Science - Computation and Language, Computer Vision and Pattern Recognition (cs.CV), Computer Science - Computer Vision and Pattern Recognition, Computation and Language (cs.CL)

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	21
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Top 10%
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Top 10%
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Top 10%

Found an issue? Give us feedback

21

Top 10%

Green

Fields of Science

engineering and technology

electrical engineering, electronic engineering, information engineering

Fields of Science

engineering and technology

electrical engineering, electronic engineering, information engineering