可垂直高山修复度量计: 重新审视 Gumbel- Softmax (Invertible Gaussian Reparameterization: Revisiting the Gumbel-Softmax) - 专知论文

会员服务 ·

0

再参数化/重参数化 · 求逆 · 单纯形 · 泛函 · Continuity ·

2022 年 8 月 29 日

Invertible Gaussian Reparameterization: Revisiting the Gumbel-Softmax

翻译：可垂直高山修复度量计: 重新审视 Gumbel- Softmax

Andres Potapczynski,Gabriel Loaiza-Ganem,John P. Cunningham

from arxiv, Accepted at NeurIPS 2020

The Gumbel-Softmax is a continuous distribution over the simplex that is often used as a relaxation of discrete distributions. Because it can be readily interpreted and easily reparameterized, it enjoys widespread use. We propose a modular and more flexible family of reparameterizable distributions where Gaussian noise is transformed into a one-hot approximation through an invertible function. This invertible function is composed of a modified softmax and can incorporate diverse transformations that serve different specific purposes. For example, the stick-breaking procedure allows us to extend the reparameterization trick to distributions with countably infinite support, thus enabling the use of our distribution along nonparametric models, or normalizing flows let us increase the flexibility of the distribution. Our construction enjoys theoretical advantages over the Gumbel-Softmax, such as closed form KL, and significantly outperforms it in a variety of experiments. Our code is available at https://github.com/cunningham-lab/igr.

翻译：Gumbel- Softmax 是一个持续分布的简单符号, 通常用作离散分布的松散。因为它可以很容易地解释, 容易地进行重新校准, 它被广泛使用。我们提议一个模块化的、更灵活的组合, 包括可重新校准分布, 使高斯噪音通过一个不可逆的功能转换成一热近似。这个不可逆的函数由修改的软体轴组成, 并可以包含各种不同的特殊用途的变形。例如, 棍棒破碎程序允许我们把重新校准的把戏扩展至分布, 并有相当无限的支持, 从而使我们能够使用非对称模型的分布, 或正常化的流让我们增加分布的灵活性。我们的建筑在理论上优于 Gumbel- Softmax, 例如封闭式 KL, 并在各种实验中大大超越了它。我们的代码可以在 https://github.com/ cunningham-lab/ igr 上查阅。

0

相关内容

再参数化/重参数化

再参数化/重参数化

【2022新书】高效深度学习，Efficient Deep Learning Book

【2022新书】高效深度学习，Efficient Deep Learning Book

专知会员服务

126+阅读 · 2022年4月21日

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

NLP必读经典文献100篇

专知会员服务

124+阅读 · 2020年9月8日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

最新BERT相关论文清单，BERT-related Papers

最新BERT相关论文清单，BERT-related Papers

专知会员服务

53+阅读 · 2019年9月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新5篇度量学习（Metric Learning）相关论文—人脸验证、BIER、自适应图卷积、注意力机制、单次学习

【论文推荐】最新5篇度量学习（Metric Learning）相关论文—人脸验证、BIER、自适应图卷积、注意力机制、单次学习

专知

17+阅读 · 2018年2月11日

Capsule Networks解析

Capsule Networks解析

机器学习研究会

11+阅读 · 2017年11月12日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

基于VIA族和IB族杂质深能级的硅亚带隙光谱响应机理研究

国家自然科学基金

0+阅读 · 2015年12月31日

Sestrin2/AMPK信号通路调控新生鼠缺氧缺血脑损伤细胞自噬的新机制

国家自然科学基金

0+阅读 · 2015年12月31日

寡层过渡金属二硫族化合物电子结构的角分辨光电子能谱研究

国家自然科学基金

0+阅读 · 2015年12月31日

混凝土Weibull统计尺寸效应理论模型改进研究

国家自然科学基金

0+阅读 · 2013年12月31日

二阶随机微分方程的Runge-Kutta方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

激光诱导荧光光谱研究二氧化硫和硫化氢离子低电子态结构

国家自然科学基金

0+阅读 · 2012年12月31日

快裂变颈部发射的同位旋效应与亚饱和密区对称能的约束

国家自然科学基金

0+阅读 · 2012年12月31日

基于高熵效应的镁铝异质金属焊接组织演变及界面反应机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

新型高速大容量长距离光纤频域传输机理和方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

新型高稳定全光纤NICE-OHMS色散光谱技术研究

国家自然科学基金

0+阅读 · 2009年12月31日

Learning from Few Samples: Transformation-Invariant SVMs with Composition and Locality at Multiple Scales

Arxiv

0+阅读 · 2022年10月16日

Improved Robust Algorithms for Learning with Discriminative Feature Feedback

Arxiv

0+阅读 · 2022年10月16日

Robust Flow-based Conformal Inference (FCI) with Statistical Guarantee

Arxiv

0+阅读 · 2022年10月15日

Causal Discovery in Heterogeneous Environments Under the Sparse Mechanism Shift Hypothesis

Arxiv

0+阅读 · 2022年10月15日

Scalable Stochastic Parametric Verification with Stochastic Variational Smoothed Model Checking

Arxiv

0+阅读 · 2022年10月14日

Privacy-Preserving and Lossless Distributed Estimation of High-Dimensional Generalized Additive Mixed Models

Arxiv

0+阅读 · 2022年10月14日

Neural Implicit Representations for Physical Parameter Inference from a Single Video

Arxiv

0+阅读 · 2022年10月14日

Nonlinear approximation of high-dimensional anisotropic analytic functions

Arxiv

0+阅读 · 2022年10月13日

The Causal Learning of Retail Delinquency

Arxiv

15+阅读 · 2020年12月17日

Additive Margin Softmax for Face Verification

Arxiv

11+阅读 · 2018年1月18日

VIP会员

文章信息

相关主题

再参数化/重参数化

相关VIP内容

【2022新书】高效深度学习，Efficient Deep Learning Book

【2022新书】高效深度学习，Efficient Deep Learning Book

专知会员服务

126+阅读 · 2022年4月21日

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

NLP必读经典文献100篇

专知会员服务

124+阅读 · 2020年9月8日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

最新BERT相关论文清单，BERT-related Papers

最新BERT相关论文清单，BERT-related Papers

专知会员服务

53+阅读 · 2019年9月29日

热门VIP内容

开通专知VIP会员享更多权益服务

网络科学赋能人工智能: 现状与展望

【NeurIPS2025教程】解释人工智能模型：可解释人工智能、数据中心人工智能与机制可解释性的方法与机遇

人工智能赋能作战行动：以俄乌战争为例

【ETHZ博士论文】表征学习在推进深度学习中的作用：效率、可扩展性与推理

相关资讯

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新5篇度量学习（Metric Learning）相关论文—人脸验证、BIER、自适应图卷积、注意力机制、单次学习

【论文推荐】最新5篇度量学习（Metric Learning）相关论文—人脸验证、BIER、自适应图卷积、注意力机制、单次学习

专知

17+阅读 · 2018年2月11日

Capsule Networks解析

Capsule Networks解析

机器学习研究会

11+阅读 · 2017年11月12日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

相关论文

Learning from Few Samples: Transformation-Invariant SVMs with Composition and Locality at Multiple Scales

Arxiv

0+阅读 · 2022年10月16日

Improved Robust Algorithms for Learning with Discriminative Feature Feedback

Arxiv

0+阅读 · 2022年10月16日

Robust Flow-based Conformal Inference (FCI) with Statistical Guarantee

Arxiv

0+阅读 · 2022年10月15日

Causal Discovery in Heterogeneous Environments Under the Sparse Mechanism Shift Hypothesis

Arxiv

0+阅读 · 2022年10月15日

Scalable Stochastic Parametric Verification with Stochastic Variational Smoothed Model Checking

Arxiv

0+阅读 · 2022年10月14日

Privacy-Preserving and Lossless Distributed Estimation of High-Dimensional Generalized Additive Mixed Models

Arxiv

0+阅读 · 2022年10月14日

Neural Implicit Representations for Physical Parameter Inference from a Single Video

Arxiv

0+阅读 · 2022年10月14日

Nonlinear approximation of high-dimensional anisotropic analytic functions

Arxiv

0+阅读 · 2022年10月13日

The Causal Learning of Retail Delinquency

Arxiv

15+阅读 · 2020年12月17日

Additive Margin Softmax for Face Verification

Arxiv

11+阅读 · 2018年1月18日

相关基金

基于VIA族和IB族杂质深能级的硅亚带隙光谱响应机理研究

国家自然科学基金

0+阅读 · 2015年12月31日

Sestrin2/AMPK信号通路调控新生鼠缺氧缺血脑损伤细胞自噬的新机制

国家自然科学基金

0+阅读 · 2015年12月31日

寡层过渡金属二硫族化合物电子结构的角分辨光电子能谱研究

国家自然科学基金

0+阅读 · 2015年12月31日

混凝土Weibull统计尺寸效应理论模型改进研究

国家自然科学基金

0+阅读 · 2013年12月31日

二阶随机微分方程的Runge-Kutta方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

激光诱导荧光光谱研究二氧化硫和硫化氢离子低电子态结构

国家自然科学基金

0+阅读 · 2012年12月31日

快裂变颈部发射的同位旋效应与亚饱和密区对称能的约束

国家自然科学基金

0+阅读 · 2012年12月31日

基于高熵效应的镁铝异质金属焊接组织演变及界面反应机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

新型高速大容量长距离光纤频域传输机理和方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

新型高稳定全光纤NICE-OHMS色散光谱技术研究

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员