Letz Translate: 卢森堡语低资源机器翻译</s> (Letz Translate: Low-Resource Machine Translation for Luxembourgish)

from arxiv, The associated model is published on HuggingFace: https://huggingface.co/etamin/Letz-Translate-OPUS-LB-EN The Dictionary used in this paper is available in Github: https://github.com/Etamin/Ltz_dictionary

Natural language processing of Low-Resource Languages (LRL) is often challenged by the lack of data. Therefore, achieving accurate machine translation (MT) in a low-resource environment is a real problem that requires practical solutions. Research in multilingual models have shown that some LRLs can be handled with such models. However, their large size and computational needs make their use in constrained environments (e.g., mobile/IoT devices or limited/old servers) impractical. In this paper, we address this problem by leveraging the power of large multilingual MT models using knowledge distillation. Knowledge distillation can transfer knowledge from a large and complex teacher model to a simpler and smaller student model without losing much in performance. We also make use of high-resource languages that are related or share the same linguistic root as the target LRL. For our evaluation, we consider Luxembourgish as the LRL that shares some roots and properties with German. We build multiple resource-efficient models based on German, knowledge distillation from the multilingual No Language Left Behind (NLLB) model, and pseudo-translation. We find that our efficient models are more than 30\% faster and perform only 4\% lower compared to the large state-of-the-art NLLB model.

翻译：低源语言(LLL)的自然语言处理往往因缺乏数据而面临挑战。因此,在低资源环境中实现准确的机器翻译(MT)是一个实际问题,需要实际的解决办法。多语言模型的研究表明,有些LLL可以使用这种模型。但是,它们的庞大规模和计算需要使得它们在有限的环境中使用(例如移动/互联网设备或有限的/老服务器)不切实际。在本文中,我们通过利用大型多语言MT模型的力量,利用知识蒸馏来解决这一问题。知识蒸馏可以把知识从一个大型和复杂的教师模型转移到一个更简单和较小的学生模型,而不会在性能方面损失很多。我们还利用与目标LLLLL有关系或具有相同语言根基的高资源语言语言。我们的评价认为,卢森堡是与德国分享一些根基和特性的LLLLS。我们根据德语建立多种资源效率模型,从多语言的“不留语言后面”模型(NLLLB)中知识蒸馏,以及假转换。我们发现,我们高效的模型比NLLLLV要快30以上,并且只进行更低的模型。</s>

相关内容

MoDELS

关注 43

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

NLP必读经典文献100篇

专知会员服务

124+阅读 · 2020年9月8日