LVP-M3:多语种多语种多式机器翻译的有语言意识的视觉快速 (LVP-M3: Language-aware Visual Prompt for Multilingual Multimodal Machine Translation)

Multimodal Machine Translation (MMT) focuses on enhancing text-only translation with visual features, which has attracted considerable attention from both natural language processing and computer vision communities. Recent advances still struggle to train a separate model for each language pair, which is costly and unaffordable when the number of languages increases in the real world. In other words, the multilingual multimodal machine translation (Multilingual MMT) task has not been investigated, which aims to handle the aforementioned issues by providing a shared semantic space for multiple languages. Besides, the image modality has no language boundaries, which is superior to bridging the semantic gap between languages. To this end, we first propose the Multilingual MMT task by establishing two new Multilingual MMT benchmark datasets covering seven languages. Then, an effective baseline LVP-M3 using visual prompts is proposed to support translations between different languages, which includes three stages (token encoding, language-aware visual prompt generation, and language translation). Extensive experimental results on our constructed benchmark datasets demonstrate the effectiveness of LVP-M3 method for Multilingual MMT.

翻译：多式机器翻译(MMT)工作的重点是加强具有视觉特征的只读文本翻译,这吸引了自然语言处理和计算机视觉界的极大关注。最近的进展仍然是难以为每种语文分别培训一个模型,当现实世界语言数量增加时,这种模型成本很高,负担不起。换句话说,多语种多式联运机器翻译(多语种MMT)任务尚未调查,其目的是通过为多种语言提供一个共同的语义空间来处理上述问题。此外,图像模式没有语言界限,这优于弥合语言之间的语义差距。为此目的,我们首先提出多语种MMMT任务,方法是建立两套新的多语种MMMT基准数据集,涵盖七种语言。然后,建议使用视觉提示的有效基线LVP-M3,以支持不同语言之间的翻译,其中包括三个阶段(对调、语言有觉识的视觉生成和语言翻译)。我们构建的基准数据集的广泛实验结果显示多语种MMMMT方法的有效性。

相关内容

Machine Translation

关注 210

机器翻译（Machine Translation）涵盖计算语言学和语言工程的所有分支，包含多语言方面。特色论文涵盖理论，描述或计算方面的任何下列主题:双语和多语语料库的编写和使用，计算机辅助语言教学，非罗马字符集的计算含义，连接主义翻译方法，对比语言学等。官网地址：http://dblp.uni-trier.de/db/journals/mt/

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

专知会员服务

115+阅读 · 2020年4月5日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日