以能源为基础的解释性案文建模示范模式 (Latent Diffusion Energy-Based Model for Interpretable Text Modeling) - 专知论文

会员服务 ·

0

Learning · MoDELS · 潜在 · MCMC · INFORMS ·

2022 年 6 月 13 日

Latent Diffusion Energy-Based Model for Interpretable Text Modeling

翻译：以能源为基础的解释性案文建模示范模式

Peiyu Yu,Sirui Xie,Xiaojian Ma,Baoxiong Jia,Bo Pang,Ruigi Gao,Yixin Zhu,Song-Chun Zhu,Ying Nian Wu

from arxiv, ICML 2022

Latent space Energy-Based Models (EBMs), also known as energy-based priors, have drawn growing interests in generative modeling. Fueled by its flexibility in the formulation and strong modeling power of the latent space, recent works built upon it have made interesting attempts aiming at the interpretability of text modeling. However, latent space EBMs also inherit some flaws from EBMs in data space; the degenerate MCMC sampling quality in practice can lead to poor generation quality and instability in training, especially on data with complex latent structures. Inspired by the recent efforts that leverage diffusion recovery likelihood learning as a cure for the sampling issue, we introduce a novel symbiosis between the diffusion models and latent space EBMs in a variational learning framework, coined as the latent diffusion energy-based model. We develop a geometric clustering-based regularization jointly with the information bottleneck to further improve the quality of the learned latent space. Experiments on several challenging tasks demonstrate the superior performance of our model on interpretable text modeling over strong counterparts.

翻译：深层空间以能源为基础的模型(EBM)也被称为以能源为基础的前身,在基因模型方面引起了越来越多的兴趣。由于在潜在空间的构思方面的灵活性和强大的建模能力,最近基于这一模型的工程做出了令人感兴趣的尝试,目的是解释文本模型的可解释性;然而,潜层空间EBM也继承了数据空间EBM的一些缺陷;实践中的低劣MCMC取样质量会导致培训质量差和不稳定,特别是复杂潜质结构数据的培训。由于最近努力利用扩散回收可能性学习作为取样问题的解药,我们把扩散模型与潜在空间EBM之间的新型共生关系引入一个变异学习框架中,作为潜在的扩散能源基模型。我们与信息瓶颈一起开发了基于几何集群的正规化,以进一步提高所学过的潜在空间的质量。关于若干具有挑战性的任务的实验表明,我们关于可解释的文本模型优于强大的对应方。

0

相关内容

Learning

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

对比学习简述

专知会员服务

90+阅读 · 2021年6月29日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

2019年机器学习框架回顾

2019年机器学习框架回顾

专知会员服务

36+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

开放知识图谱

1+阅读 · 2022年4月4日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Industry Talk2

【ICIG2021】Latest News & Announcements of the Industry Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年7月29日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

vae 相关论文表示学习 1

vae 相关论文表示学习 1

CreateAMind

12+阅读 · 2018年9月6日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

胶质瘤侵袭过程中DNMT1沉默miR-134与ERK信号通路自激活的表观新机制

国家自然科学基金

0+阅读 · 2015年12月31日

长链非编码RNA CAR intergenic 10在细胞衰老中的作用和机制

国家自然科学基金

1+阅读 · 2013年12月31日

具有共格/半共格界面关系的Cu2O-TiO2异质结的制备及其形成机理研究

国家自然科学基金

0+阅读 · 2013年12月31日

以紫外光固化双亲共聚物为功能性软模板聚合水溶性PEDOT导电材料的研究

国家自然科学基金

0+阅读 · 2013年12月31日

棉铃虫性信息素腺体ACCase基因的克隆及功能分析

国家自然科学基金

0+阅读 · 2013年12月31日

基于脆弱性的大气颗粒物重金属健康风险研究

国家自然科学基金

0+阅读 · 2013年12月31日

非系统性创业风险的识别和控制机制：基于认知视角的实证研究

国家自然科学基金

0+阅读 · 2012年12月31日

城市大气颗粒物重金属污染特征及健康风险评估

国家自然科学基金

0+阅读 · 2012年12月31日

计及分布式发电的配电网自愈控制研究

国家自然科学基金

0+阅读 · 2012年12月31日

Reality-based Interaction用户界面模型和评估方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

Interpretable Time Series Clustering Using Local Explanations

Arxiv

0+阅读 · 2022年8月1日

Diffusion-Based Representation Learning

Arxiv

0+阅读 · 2022年8月1日

Composable Text Control Operations in Latent Space with Ordinary Differential Equations

Arxiv

0+阅读 · 2022年8月1日

Inter-model Interpretability: Self-supervised Models as a Case Study

Arxiv

0+阅读 · 2022年7月31日

Learning an Interpretable Model for Driver Behavior Prediction with Inductive Biases

Arxiv

0+阅读 · 2022年7月31日

INSightR-Net: Interpretable Neural Network for Regression using Similarity-based Comparisons to Prototypical Examples

Arxiv

0+阅读 · 2022年7月31日

The Causal Learning of Retail Delinquency

Arxiv

14+阅读 · 2020年12月17日

Latent Relation Language Models

Arxiv

21+阅读 · 2019年8月21日

Learning with Interpretable Structure from RNN

Arxiv

19+阅读 · 2018年10月25日

An Interpretable Reasoning Network for Multi-Relation Question Answering

Arxiv

13+阅读 · 2018年6月1日

VIP会员

文章信息

相关主题

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

对比学习简述

专知会员服务

90+阅读 · 2021年6月29日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

2019年机器学习框架回顾

2019年机器学习框架回顾

专知会员服务

36+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【CMU博士论文】数据驱动决策中的激励、信息与不确定性

DGP双粒度提示框架：图增强大模型助力欺诈检测

【ICCV2025】ESSENTIAL：用于视频类增量学习的情景记忆与语义记忆整合

唯快不破：大型语言模型高效架构综述

相关资讯

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

开放知识图谱

1+阅读 · 2022年4月4日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Industry Talk2

【ICIG2021】Latest News & Announcements of the Industry Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年7月29日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

vae 相关论文表示学习 1

vae 相关论文表示学习 1

CreateAMind

12+阅读 · 2018年9月6日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

相关论文

Interpretable Time Series Clustering Using Local Explanations

Arxiv

0+阅读 · 2022年8月1日

Diffusion-Based Representation Learning

Arxiv

0+阅读 · 2022年8月1日

Composable Text Control Operations in Latent Space with Ordinary Differential Equations

Arxiv

0+阅读 · 2022年8月1日

Inter-model Interpretability: Self-supervised Models as a Case Study

Arxiv

0+阅读 · 2022年7月31日

Learning an Interpretable Model for Driver Behavior Prediction with Inductive Biases

Arxiv

0+阅读 · 2022年7月31日

INSightR-Net: Interpretable Neural Network for Regression using Similarity-based Comparisons to Prototypical Examples

Arxiv

0+阅读 · 2022年7月31日

The Causal Learning of Retail Delinquency

Arxiv

14+阅读 · 2020年12月17日

Latent Relation Language Models

Arxiv

21+阅读 · 2019年8月21日

Learning with Interpretable Structure from RNN

Arxiv

19+阅读 · 2018年10月25日

An Interpretable Reasoning Network for Multi-Relation Question Answering

Arxiv

13+阅读 · 2018年6月1日

相关基金

胶质瘤侵袭过程中DNMT1沉默miR-134与ERK信号通路自激活的表观新机制

国家自然科学基金

0+阅读 · 2015年12月31日

长链非编码RNA CAR intergenic 10在细胞衰老中的作用和机制

国家自然科学基金

1+阅读 · 2013年12月31日

具有共格/半共格界面关系的Cu2O-TiO2异质结的制备及其形成机理研究

国家自然科学基金

0+阅读 · 2013年12月31日

以紫外光固化双亲共聚物为功能性软模板聚合水溶性PEDOT导电材料的研究

国家自然科学基金

0+阅读 · 2013年12月31日

棉铃虫性信息素腺体ACCase基因的克隆及功能分析

国家自然科学基金

0+阅读 · 2013年12月31日

基于脆弱性的大气颗粒物重金属健康风险研究

国家自然科学基金

0+阅读 · 2013年12月31日

非系统性创业风险的识别和控制机制：基于认知视角的实证研究

国家自然科学基金

0+阅读 · 2012年12月31日

城市大气颗粒物重金属污染特征及健康风险评估

国家自然科学基金

0+阅读 · 2012年12月31日

计及分布式发电的配电网自愈控制研究

国家自然科学基金

0+阅读 · 2012年12月31日

Reality-based Interaction用户界面模型和评估方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员