流动:以共识为基础的对话评价框架,利用分部法流程 (FlowEval: A Consensus-Based Dialogue Evaluation Framework Using Segment Act Flows) - 专知论文

会员服务 ·

0

任务对话系统 · Extensibility · 数据集 · Better · 基准 ·

2022 年 11 月 3 日

FlowEval: A Consensus-Based Dialogue Evaluation Framework Using Segment Act Flows

翻译：流动:以共识为基础的对话评价框架,利用分部法流程

Jianqiao Zhao,Yanyang Li,Wanyu Du,Yangfeng Ji,Dong Yu,Michael R. Lyu,Liwei Wang

from arxiv, EMNLP 2022 camera-ready version

Despite recent progress in open-domain dialogue evaluation, how to develop automatic metrics remains an open problem. We explore the potential of dialogue evaluation featuring dialog act information, which was hardly explicitly modeled in previous methods. However, defined at the utterance level in general, dialog act is of coarse granularity, as an utterance can contain multiple segments possessing different functions. Hence, we propose segment act, an extension of dialog act from utterance level to segment level, and crowdsource a large-scale dataset for it. To utilize segment act flows, sequences of segment acts, for evaluation, we develop the first consensus-based dialogue evaluation framework, FlowEval. This framework provides a reference-free approach for dialog evaluation by finding pseudo-references. Extensive experiments against strong baselines on three benchmark datasets demonstrate the effectiveness and other desirable characteristics of our FlowEval, pointing out a potential path for better dialogue evaluation.

翻译：尽管在开放域对话评价方面最近取得了进展,但如何制定自动指标仍是一个尚未解决的问题。我们探索对话评价的潜力,其特点是对话行为信息,在以前的方法中,这种信息很少被明确效仿。然而,一般而言,在发言一级,对话行为是粗粗的颗粒,因为一种言论可以包含具有不同功能的多个部分。因此,我们提议采取分部分行动,将对话行为从发声级别扩大到分层一级,并为它提供大规模数据集。为了评估,我们开发了第一个基于共识的对话评价框架,即RlowEval。这个框架为对话评价提供了一种无参考方法,通过寻找假参照。针对三个基准数据集的强大基线进行的广泛实验,显示了我们流动值的有效性和其他可取特征,指出了改进对话评价的潜在途径。

0

相关内容

任务对话系统

任务对话系统

20篇「ACL2020」最新论文抢先看！看自然语言处理2020在研究什么？

20篇「ACL2020」最新论文抢先看！看自然语言处理2020在研究什么？

专知会员服务

97+阅读 · 2020年4月10日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【论文】图上的表示学习综述

【论文】图上的表示学习综述

机器学习研究会

15+阅读 · 2017年9月24日

EADIA调节抑癌基因DCC凋亡通路的分子机制研究

国家自然科学基金

0+阅读 · 2016年12月31日

IRAK-M在肾结核皮、髓质差异性发病中的作用和机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

CoFe2O4/BaSrTiO3复合势垒多铁隧道结的制备及隧穿特性研究

国家自然科学基金

0+阅读 · 2013年12月31日

Reticulon-1介导的内质网应激在糖尿病肾病发病机制中的作用

国家自然科学基金

0+阅读 · 2013年12月31日

柑橘黄龙病亚洲种病原( Cadidatus Liberibacter assiaticus)重组抗体的研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于低毒性Mn: ZnS 量子点的活体肿瘤靶向荧光成像

国家自然科学基金

0+阅读 · 2012年12月31日

加工番茄可溶性固形物含量的全基因组关联分析与连锁作图

国家自然科学基金

0+阅读 · 2011年12月31日

基于弱相互作用的分子尺度器件电荷传输机理与性能调控

国家自然科学基金

0+阅读 · 2011年12月31日

“#32473;体-受体”#22411;稀土发光纳米粒子的制备和荧光调控

国家自然科学基金

0+阅读 · 2009年12月31日

量子点生物效应的热化学研究

国家自然科学基金

0+阅读 · 2008年12月31日

What Makes for Good Tokenizers in Vision Transformer?

What Makes for Good Tokenizers in Vision Transformer?

Arxiv

0+阅读 · 2022年12月21日

Exploring Consistency in Cross-Domain Transformer for Domain Adaptive Semantic Segmentation

Arxiv

0+阅读 · 2022年12月21日

Tracing and Removing Data Errors in Natural Language Generation Datasets

Arxiv

0+阅读 · 2022年12月21日

MoralDial: A Framework to Train and Evaluate Moral Dialogue Systems via Constructing Moral Discussions

Arxiv

0+阅读 · 2022年12月21日

mFACE: Multilingual Summarization with Factual Consistency Evaluation

Arxiv

0+阅读 · 2022年12月20日

Evaluation for Change

Arxiv

0+阅读 · 2022年12月20日

Joint Spatio-Temporal Modeling for the Semantic Change Detection in Remote Sensing Images

Arxiv

0+阅读 · 2022年12月17日

Natural Language Descriptions of Deep Visual Features

Arxiv

12+阅读 · 2022年1月26日

Exploiting Fine-grained Face Forgery Clues via Progressive Enhancement Learning

Arxiv

12+阅读 · 2021年12月28日

Reasoning in Dialog: Improving Response Generation by Context Reading Comprehension

Arxiv

12+阅读 · 2020年12月14日

VIP会员

文章信息

相关主题

任务对话系统

相关VIP内容

20篇「ACL2020」最新论文抢先看！看自然语言处理2020在研究什么？

20篇「ACL2020」最新论文抢先看！看自然语言处理2020在研究什么？

专知会员服务

97+阅读 · 2020年4月10日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

【NeurIPS 2025】稳定电影度量：面向专业视频生成的结构化分类与评测体系

战场AI决策支持系统

【博士论文】面向排序与扩散模型的安全、高效与鲁棒强化学习

面向 AI 生成图像的安全与鲁棒水印：全面综述

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【论文】图上的表示学习综述

【论文】图上的表示学习综述

机器学习研究会

15+阅读 · 2017年9月24日

相关论文

What Makes for Good Tokenizers in Vision Transformer?

What Makes for Good Tokenizers in Vision Transformer?

Arxiv

0+阅读 · 2022年12月21日

Exploring Consistency in Cross-Domain Transformer for Domain Adaptive Semantic Segmentation

Arxiv

0+阅读 · 2022年12月21日

Tracing and Removing Data Errors in Natural Language Generation Datasets

Arxiv

0+阅读 · 2022年12月21日

MoralDial: A Framework to Train and Evaluate Moral Dialogue Systems via Constructing Moral Discussions

Arxiv

0+阅读 · 2022年12月21日

mFACE: Multilingual Summarization with Factual Consistency Evaluation

Arxiv

0+阅读 · 2022年12月20日

Evaluation for Change

Arxiv

0+阅读 · 2022年12月20日

Joint Spatio-Temporal Modeling for the Semantic Change Detection in Remote Sensing Images

Arxiv

0+阅读 · 2022年12月17日

Natural Language Descriptions of Deep Visual Features

Arxiv

12+阅读 · 2022年1月26日

Exploiting Fine-grained Face Forgery Clues via Progressive Enhancement Learning

Arxiv

12+阅读 · 2021年12月28日

Reasoning in Dialog: Improving Response Generation by Context Reading Comprehension

Arxiv

12+阅读 · 2020年12月14日

相关基金

EADIA调节抑癌基因DCC凋亡通路的分子机制研究

国家自然科学基金

0+阅读 · 2016年12月31日

IRAK-M在肾结核皮、髓质差异性发病中的作用和机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

CoFe2O4/BaSrTiO3复合势垒多铁隧道结的制备及隧穿特性研究

国家自然科学基金

0+阅读 · 2013年12月31日

Reticulon-1介导的内质网应激在糖尿病肾病发病机制中的作用

国家自然科学基金

0+阅读 · 2013年12月31日

柑橘黄龙病亚洲种病原( Cadidatus Liberibacter assiaticus)重组抗体的研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于低毒性Mn: ZnS 量子点的活体肿瘤靶向荧光成像

国家自然科学基金

0+阅读 · 2012年12月31日

加工番茄可溶性固形物含量的全基因组关联分析与连锁作图

国家自然科学基金

0+阅读 · 2011年12月31日

基于弱相互作用的分子尺度器件电荷传输机理与性能调控

国家自然科学基金

0+阅读 · 2011年12月31日

“#32473;体-受体”#22411;稀土发光纳米粒子的制备和荧光调控

国家自然科学基金

0+阅读 · 2009年12月31日

量子点生物效应的热化学研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员