EASE:实体 -- -- 实体 -- -- 违反认知学习收押刑期 (EASE: Entity-Aware Contrastive Learning of Sentence Embedding)

We present EASE, a novel method for learning sentence embeddings via contrastive learning between sentences and their related entities. The advantage of using entity supervision is twofold: (1) entities have been shown to be a strong indicator of text semantics and thus should provide rich training signals for sentence embeddings; (2) entities are defined independently of languages and thus offer useful cross-lingual alignment supervision. We evaluate EASE against other unsupervised models both in monolingual and multilingual settings. We show that EASE exhibits competitive or better performance in English semantic textual similarity (STS) and short text clustering (STC) tasks and it significantly outperforms baseline methods in multilingual settings on a variety of tasks. Our source code, pre-trained models, and newly constructed multilingual STC dataset are available at https://github.com/studio-ousia/ease.

翻译：我们提出EASE,这是通过对判决及其相关实体的对比性学习而嵌入判决的一种新颖方法,使用实体监督的好处有两个方面:(1) 实体已证明是文字语义的有力指标,因此应当为判决嵌入提供丰富的培训信号;(2) 实体的定义独立于语言,从而提供有用的跨语言协调监督;我们对照单一语言和多语言环境中其他不受监督的模式,对EASE进行评估;我们表明,EASE在英语语义文本相似性和短文本组别(STS)任务中表现出竞争性或更好的表现,在多种任务中大大优于多语种环境中的基线方法。我们的源代码、预先培训的模式和新建的多语种STC数据集可在https://github.com/studio-ousia/sease查阅。

相关内容

EASE

关注 0

软件工程评估（Evaluation and Assessment in Software Engineering，EASE）会议是一个国际领先的会议场所，学术界和实践者可以在此展示和讨论他们对基于证据的软件工程的研究及其对软件实践的影响。第23届EASE将于2019年4月在丹麦哥本哈根举行，由哥本哈根IT大学主办。EASE 2019欢迎向不同领域提交高质量的研究报告：完整的研究论文、短篇论文和手工艺品、新兴成果和愿景、行业轨迹、博士研讨会、海报。官网链接：https://ease2019.org/

【IJCAJ 2019】多视角知识图谱嵌入的实体对齐，Multi-view Knowledge Graph Embedding for Entity Alignment

专知会员服务

59+阅读 · 2020年6月30日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

【AAAI2020论文】概念结构化嵌入医疗文本表示（Learning Conceptual-Contextual Embeddings for Medical Text）

专知会员服务

49+阅读 · 2019年11月15日