半监督的参数生成深端变量模型 (Deep Latent Variable Models for Semi-supervised Paraphrase Generation) - 专知论文

会员服务 ·

0

潜变量/隐变量 · MoDELS · 潜在 · Performer · 监督模型 ·

2023 年 1 月 5 日

Deep Latent Variable Models for Semi-supervised Paraphrase Generation

翻译：半监督的参数生成深端变量模型

Jialin Yu,Alexandra I. Cristea,Anoushka Harit,Zhongtian Sun,Olanrewaju Tahir Aduragba,Lei Shi,Noura Al Moubayed

This paper explores deep latent variable models for semi-supervised paraphrase generation, where the missing target pair is modelled as a latent paraphrase sequence. We present a novel unsupervised model named variational sequence auto-encoding reconstruction (VSAR), which performs latent sequence inference given an observed text. To leverage information from text pairs, we introduce a supervised model named dual directional learning (DDL). Combining VSAR with DDL (DDL+VSAR) enables us to conduct semi-supervised learning; however, the combined model suffers from a cold-start problem. To combat this issue, we propose to deal with better weight initialisation, leading to a two-stage training scheme named knowledge reinforced training. Our empirical evaluations suggest that the combined model yields competitive performance against the state-of-the-art supervised baselines on complete data. Furthermore, in scenarios where only a fraction of the labelled pairs are available, our combined model consistently outperforms the strong supervised model baseline (DDL and Transformer) by a significant margin.

翻译：本文探索了半监督参数生成的深潜潜变数模型, 缺少的目标对是一个潜在的参数序列。我们提出了一个新的未经监督的模型, 名为变异序列自动编码重建( VSAR), 用于对观察到的文本进行潜在序列推断。为了利用文本对的信息, 我们引入了一个名为双向定向学习( DDL)的监管模型。将VSAR与DDL( DDL+VSAR)相结合, 使我们能够进行半监督学习; 但是, 合并模型存在一个冷却的启动问题。为了解决这一问题, 我们提议处理更好的权重初始化, 导致一个称为知识强化培训的两阶段培训计划。我们的经验评估表明, 合并模型能产生与全数据最新监管基线的竞争性性能。此外, 在只有一小部分贴有标签的对子( DDL和变压器) 能够持续以显著的幅度超越强大的监管模型基线( DDL和变压器) 。

0

相关内容

潜变量/隐变量

潜变量/隐变量

NeurlPS 2022 | 自然语言处理相关论文分类整理

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

75+阅读 · 2022年6月28日

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【代码资源】GAN | 七份最热GAN文章及代码分享（Github 1000+Stars）

【代码资源】GAN | 七份最热GAN文章及代码分享（Github 1000+Stars）

专知

13+阅读 · 2018年6月24日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

稀土硫氧化物上转换荧光探针的一步合成与生物成像研究

国家自然科学基金

0+阅读 · 2015年12月31日

S3AGA样本（Spitzer-SDSS Spectral Atlas of Galaxies and AGNs)及其AGN研究

国家自然科学基金

0+阅读 · 2014年12月31日

用LAMOST的巡天数据搜索和研究激变变星

国家自然科学基金

0+阅读 · 2014年12月31日

染料包埋的核壳结构YAG:Ce3+/SiO2荧光粉的制备与发光性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

原子层沉积稀土氧化物和硅酸盐纳米复合薄膜硅基MOS电致发光器件的研究

国家自然科学基金

0+阅读 · 2012年12月31日

ZnSe:Mn/ZnSe/PMMA纳米复合薄膜在脉冲强磁场下的物性研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于SERS编码的Capase探针激活效应的研究

国家自然科学基金

0+阅读 · 2011年12月31日

电光晶体中亚微米周期畴结构的PFM极化制备及动力学演化研究

国家自然科学基金

0+阅读 · 2009年12月31日

高光度blazar的甚高能伽马射线辐射研究

国家自然科学基金

0+阅读 · 2009年12月31日

基于碱土复合氧化物的LED用单一多色光转换材料研究

国家自然科学基金

0+阅读 · 2008年12月31日

Are We Really Making Much Progress? Bag-of-Words vs. Sequence vs. Graph vs. Hierarchy for Single- and Multi-Label Text Classification

Arxiv

0+阅读 · 2023年3月3日

Learning Better Masking for Better Language Model Pre-training

Arxiv

0+阅读 · 2023年3月3日

Mixture of Soft Prompts for Controllable Data Generation

Arxiv

0+阅读 · 2023年3月2日

Integrated Parameter-Efficient Tuning for General-Purpose Audio Models

Arxiv

0+阅读 · 2023年3月2日

STUNT: Few-shot Tabular Learning with Self-generated Tasks from Unlabeled Tables

Arxiv

0+阅读 · 2023年3月2日

Distance-based Weight Transfer for Fine-tuning from Near-field to Far-field Speaker Verification

Arxiv

0+阅读 · 2023年3月1日

Analog Bits: Generating Discrete Data using Diffusion Models with Self-Conditioning

Arxiv

0+阅读 · 2023年3月1日

Pre-training Text Representations as Meta Learning

Arxiv

13+阅读 · 2020年4月12日

MAD-GAN: Multivariate Anomaly Detection for Time Series Data with Generative Adversarial Networks

MAD-GAN: Multivariate Anomaly Detection for Time Series Data with Generative Adversarial Networks

Arxiv

15+阅读 · 2019年1月15日

Exploring Models and Data for Remote Sensing Image Caption Generation

Arxiv

14+阅读 · 2017年12月21日

VIP会员

文章信息

相关主题

潜变量/隐变量

相关VIP内容

NeurlPS 2022 | 自然语言处理相关论文分类整理

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

75+阅读 · 2022年6月28日

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

GPT-5如何对齐？从硬性拒绝到安全完成：走向以输出为中心的安全训练

【伯克利博士论文】超越人类监督的视觉智能

【ICCV2025】SO(3) 上连续非保守动力系统的预测

2025年中国数据要素行业发展研究报告

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【代码资源】GAN | 七份最热GAN文章及代码分享（Github 1000+Stars）

【代码资源】GAN | 七份最热GAN文章及代码分享（Github 1000+Stars）

专知

13+阅读 · 2018年6月24日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

相关论文

Are We Really Making Much Progress? Bag-of-Words vs. Sequence vs. Graph vs. Hierarchy for Single- and Multi-Label Text Classification

Arxiv

0+阅读 · 2023年3月3日

Learning Better Masking for Better Language Model Pre-training

Arxiv

0+阅读 · 2023年3月3日

Mixture of Soft Prompts for Controllable Data Generation

Arxiv

0+阅读 · 2023年3月2日

Integrated Parameter-Efficient Tuning for General-Purpose Audio Models

Arxiv

0+阅读 · 2023年3月2日

STUNT: Few-shot Tabular Learning with Self-generated Tasks from Unlabeled Tables

Arxiv

0+阅读 · 2023年3月2日

Distance-based Weight Transfer for Fine-tuning from Near-field to Far-field Speaker Verification

Arxiv

0+阅读 · 2023年3月1日

Analog Bits: Generating Discrete Data using Diffusion Models with Self-Conditioning

Arxiv

0+阅读 · 2023年3月1日

Pre-training Text Representations as Meta Learning

Arxiv

13+阅读 · 2020年4月12日

MAD-GAN: Multivariate Anomaly Detection for Time Series Data with Generative Adversarial Networks

MAD-GAN: Multivariate Anomaly Detection for Time Series Data with Generative Adversarial Networks

Arxiv

15+阅读 · 2019年1月15日

Exploring Models and Data for Remote Sensing Image Caption Generation

Arxiv

14+阅读 · 2017年12月21日

相关基金

稀土硫氧化物上转换荧光探针的一步合成与生物成像研究

国家自然科学基金

0+阅读 · 2015年12月31日

S3AGA样本（Spitzer-SDSS Spectral Atlas of Galaxies and AGNs)及其AGN研究

国家自然科学基金

0+阅读 · 2014年12月31日

用LAMOST的巡天数据搜索和研究激变变星

国家自然科学基金

0+阅读 · 2014年12月31日

染料包埋的核壳结构YAG:Ce3+/SiO2荧光粉的制备与发光性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

原子层沉积稀土氧化物和硅酸盐纳米复合薄膜硅基MOS电致发光器件的研究

国家自然科学基金

0+阅读 · 2012年12月31日

ZnSe:Mn/ZnSe/PMMA纳米复合薄膜在脉冲强磁场下的物性研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于SERS编码的Capase探针激活效应的研究

国家自然科学基金

0+阅读 · 2011年12月31日

电光晶体中亚微米周期畴结构的PFM极化制备及动力学演化研究

国家自然科学基金

0+阅读 · 2009年12月31日

高光度blazar的甚高能伽马射线辐射研究

国家自然科学基金

0+阅读 · 2009年12月31日

基于碱土复合氧化物的LED用单一多色光转换材料研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员