组成,注意,还是两者兼而有之? (Composition, Attention, or Both?) - 专知论文

会员服务 ·

0

泛函 · Attention · 组合性 · 泛化理论 · MoDELS ·

2022 年 12 月 14 日

Composition, Attention, or Both?

翻译：组成,注意,还是两者兼而有之?

Ryo Yoshida,Yohei Oseki

from arxiv, Accepted by Findings of EMNLP 2022

In this paper, we propose a novel architecture called Composition Attention Grammars (CAGs) that recursively compose subtrees into a single vector representation with a composition function, and selectively attend to previous structural information with a self-attention mechanism. We investigate whether these components -- the composition function and the self-attention mechanism -- can both induce human-like syntactic generalization. Specifically, we train language models (LMs) with and without these two components with the model sizes carefully controlled, and evaluate their syntactic generalization performance against six test circuits on the SyntaxGym benchmark. The results demonstrated that the composition function and the self-attention mechanism both play an important role to make LMs more human-like, and closer inspection of linguistic phenomenon implied that the composition function allowed syntactic features, but not semantic features, to percolate into subtree representations.

翻译：在本文中,我们提出一个名为“组成注意语法”的新结构,将亚树重新组成成一个具有组成功能的单一矢量代表,有选择地以自我注意机制关注先前的结构信息。我们调查这些组成部分 -- -- 组成功能和自我注意机制 -- -- 是否既能诱发类似人的同义法的概括化。具体地说,我们用和没有这两个组成部分的模型来训练语言模型(LMs),并仔细控制这两个模型的大小,对照语权基准上的六个测试电路来评价其综合概括性表现。结果显示,组成功能和自我注意机制都发挥了重要作用,使LMs更像人一样,更密切地检查语言现象意味着,组成功能允许合成特征,但非语权特征,进入子树木的表述。

0

相关内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

20篇「ACL2020」最新论文抢先看！看自然语言处理2020在研究什么？

20篇「ACL2020」最新论文抢先看！看自然语言处理2020在研究什么？

专知会员服务

97+阅读 · 2020年4月10日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

开放知识图谱

1+阅读 · 2022年4月4日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

miR-5591靶向AGER/ROS/JNK抑制MSCs氧化应激损伤在糖尿病创面修复中的作用及机制

国家自然科学基金

0+阅读 · 2015年12月31日

c-MET信号通路与肝癌耐药机制的研究

国家自然科学基金

0+阅读 · 2014年12月31日

MKP-4调节ERK信号通路在肝细胞癌发生发展中的意义

国家自然科学基金

0+阅读 · 2013年12月31日

PI3K/AKT/mTOR通路对上皮细胞间质化的调控在胃癌化疗耐药中的作用及其机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

南蛇藤提取物靶向PI3K/Akt/mTOR信号通路抑制肝癌早期转移的作用及机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

从海马-杏仁核神经元PI3K/Akt/mTOR信号通路与细胞骨架关系探讨抑郁症发病机制

国家自然科学基金

0+阅读 · 2012年12月31日

Fibulin-5/β1-integrin 信号通路在醛固酮诱导血管平滑肌细胞凋亡中的作用

国家自然科学基金

0+阅读 · 2012年12月31日

片仔癀调控microRNA抑制结肠癌上皮间质转化的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

TLR4信号通路介导DFMG抗AS作用机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

mTOR 催化抑制剂抗肿瘤作用机制的研究

国家自然科学基金

0+阅读 · 2011年12月31日

Initialisation from lattice Boltzmann to multi-step Finite Difference methods: modified equations and discrete observability

Arxiv

0+阅读 · 2023年2月15日

Contrastive Multimodal Learning for Emergence of Graphical Sensory-Motor Communication

Arxiv

0+阅读 · 2023年2月14日

simpleKT: A Simple But Tough-to-Beat Baseline for Knowledge Tracing

Arxiv

0+阅读 · 2023年2月14日

An Empirical Bayes Approach for Constructing the Confidence Intervals of Clonality and Entropy

Arxiv

0+阅读 · 2023年2月14日

Cross-Layer Retrospective Retrieving via Layer Attention

Arxiv

0+阅读 · 2023年2月10日

Graph Ordering Attention Networks

Arxiv

12+阅读 · 2022年11月21日

Deep Neural Network Based Relation Extraction: An Overview

Arxiv

14+阅读 · 2021年1月6日

DAGCN: Dual Attention Graph Convolutional Networks

Arxiv

16+阅读 · 2019年4月4日

Bilinear Attention Networks

Arxiv

11+阅读 · 2018年5月21日

Learning to Count Objects in Natural Images for Visual Question Answering

Arxiv

12+阅读 · 2018年2月15日

VIP会员

文章信息

相关主题

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

20篇「ACL2020」最新论文抢先看！看自然语言处理2020在研究什么？

20篇「ACL2020」最新论文抢先看！看自然语言处理2020在研究什么？

专知会员服务

97+阅读 · 2020年4月10日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

操作系统智能体：基于多模态大模型（MLLM）的通用计算设备智能体综述

《美国太空军系统全生命周期建模、仿真与分析效能提升方案》最新84页报告

【博士论文】推进数据高效的深度学习：非参数 Transformer、主动测试与上下文学习

自主人工智能：未来战争是否将是自主化的？

相关资讯

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

开放知识图谱

1+阅读 · 2022年4月4日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

相关论文

Initialisation from lattice Boltzmann to multi-step Finite Difference methods: modified equations and discrete observability

Arxiv

0+阅读 · 2023年2月15日

Contrastive Multimodal Learning for Emergence of Graphical Sensory-Motor Communication

Arxiv

0+阅读 · 2023年2月14日

simpleKT: A Simple But Tough-to-Beat Baseline for Knowledge Tracing

Arxiv

0+阅读 · 2023年2月14日

An Empirical Bayes Approach for Constructing the Confidence Intervals of Clonality and Entropy

Arxiv

0+阅读 · 2023年2月14日

Cross-Layer Retrospective Retrieving via Layer Attention

Arxiv

0+阅读 · 2023年2月10日

Graph Ordering Attention Networks

Arxiv

12+阅读 · 2022年11月21日

Deep Neural Network Based Relation Extraction: An Overview

Arxiv

14+阅读 · 2021年1月6日

DAGCN: Dual Attention Graph Convolutional Networks

Arxiv

16+阅读 · 2019年4月4日

Bilinear Attention Networks

Arxiv

11+阅读 · 2018年5月21日

Learning to Count Objects in Natural Images for Visual Question Answering

Arxiv

12+阅读 · 2018年2月15日

相关基金

miR-5591靶向AGER/ROS/JNK抑制MSCs氧化应激损伤在糖尿病创面修复中的作用及机制

国家自然科学基金

0+阅读 · 2015年12月31日

c-MET信号通路与肝癌耐药机制的研究

国家自然科学基金

0+阅读 · 2014年12月31日

MKP-4调节ERK信号通路在肝细胞癌发生发展中的意义

国家自然科学基金

0+阅读 · 2013年12月31日

PI3K/AKT/mTOR通路对上皮细胞间质化的调控在胃癌化疗耐药中的作用及其机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

南蛇藤提取物靶向PI3K/Akt/mTOR信号通路抑制肝癌早期转移的作用及机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

从海马-杏仁核神经元PI3K/Akt/mTOR信号通路与细胞骨架关系探讨抑郁症发病机制

国家自然科学基金

0+阅读 · 2012年12月31日

Fibulin-5/β1-integrin 信号通路在醛固酮诱导血管平滑肌细胞凋亡中的作用

国家自然科学基金

0+阅读 · 2012年12月31日

片仔癀调控microRNA抑制结肠癌上皮间质转化的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

TLR4信号通路介导DFMG抗AS作用机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

mTOR 催化抑制剂抗肿瘤作用机制的研究

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员