整形 SVD 共变调制和内部分解 (Orthogonal SVD Covariance Conditioning and Latent Disentanglement) - 专知论文

会员服务 ·

0

正交 · 奇异值分解 · HTTPS · 潜在 · 泛化理论 ·

2022 年 12 月 11 日

Orthogonal SVD Covariance Conditioning and Latent Disentanglement

翻译：整形 SVD 共变调制和内部分解

Yue Song,Nicu Sebe,Wei Wang

from arxiv, Accepted by IEEE T-PAMI. arXiv admin note: substantial text overlap with arXiv:2207.02119

Inserting an SVD meta-layer into neural networks is prone to make the covariance ill-conditioned, which could harm the model in the training stability and generalization abilities. In this paper, we systematically study how to improve the covariance conditioning by enforcing orthogonality to the Pre-SVD layer. Existing orthogonal treatments on the weights are first investigated. However, these techniques can improve the conditioning but would hurt the performance. To avoid such a side effect, we propose the Nearest Orthogonal Gradient (NOG) and Optimal Learning Rate (OLR). The effectiveness of our methods is validated in two applications: decorrelated Batch Normalization (BN) and Global Covariance Pooling (GCP). Extensive experiments on visual recognition demonstrate that our methods can simultaneously improve covariance conditioning and generalization. The combinations with orthogonal weight can further boost the performance. Moreover, we show that our orthogonality techniques can benefit generative models for better latent disentanglement through a series of experiments on various benchmarks. Code is available at: \href{https://github.com/KingJamesSong/OrthoImproveCond}{https://github.com/KingJamesSong/OrthoImproveCond}.

翻译：在神经网络中插入 SVD 元层时, 容易使共变错误, 这会损害培训稳定性和一般化能力的模式。在本文中, 我们系统地研究如何通过对 SVD 前层强制执行正数调整来改进共变调节。首先是调查重量的现有正数处理方法。但是, 这些技术可以改进调试, 但是会损害性能。为了避免这种副作用, 我们提议了近半正数梯度和最佳学习率( OLR ) 。我们的方法的有效性在两种应用中得到验证: 与调理学相关的批次正常化( BN) 和全球共变异组合( GCP ) 。关于视觉识别的广泛实验表明, 我们的方法可以同时改善共变调和一般化。与正数的组合可以进一步提升性能。此外, 我们的正数法技术可以通过一系列基准实验, 有利于更潜在分解模型。代码可以在以下网址上找到: Oror Batch/ Ormabrefr@Khobus/ Ormabs@ Gomes.

0

相关内容

NeurlPS 2022 | 自然语言处理相关论文分类整理

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

最新BERT相关论文清单，BERT-related Papers

最新BERT相关论文清单，BERT-related Papers

专知会员服务

53+阅读 · 2019年9月29日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

ICLR2019最佳论文出炉

ICLR2019最佳论文出炉

专知

12+阅读 · 2019年5月6日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

vae 相关论文表示学习 1

vae 相关论文表示学习 1

CreateAMind

12+阅读 · 2018年9月6日

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

专知

23+阅读 · 2018年1月18日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

采用pinball loss的MEE算法研究

国家自然科学基金

1+阅读 · 2013年12月31日

Partial Spread Bent函数与Bent-Negabent函数的构造及密码学性质研究

国家自然科学基金

0+阅读 · 2013年12月31日

联合显著性检测和对象分割的算法研究

国家自然科学基金

0+阅读 · 2013年12月31日

Kronheimer-Nakajima quiver 模空间与有理曲面

国家自然科学基金

1+阅读 · 2013年12月31日

柯萨奇病毒 A16 实心/空心颗粒抗原性差异的结构基础

国家自然科学基金

0+阅读 · 2013年12月31日

中国BTV-1与BTV-4毒株Seg2重排致病毒变异机制的研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于list-mode数据的快速SART真3D PET断层重建算法的研究

国家自然科学基金

0+阅读 · 2011年12月31日

远红外新型Te基硫系玻璃研制及相关性质研究

国家自然科学基金

0+阅读 · 2009年12月31日

CD73在动脉粥样硬化斑块破裂中的作用及其机制研究

国家自然科学基金

0+阅读 · 2009年12月31日

高分辨大视野的多针孔SPECT成像研究

国家自然科学基金

0+阅读 · 2008年12月31日

Learning a Gaussian Mixture Model from Imperfect Training Data for Robust Channel Estimation

Arxiv

0+阅读 · 2023年2月13日

Dimension Reduction and MARS

Arxiv

0+阅读 · 2023年2月11日

Verifying Generalization in Deep Learning

Arxiv

0+阅读 · 2023年2月11日

Data-Efficient Structured Pruning via Submodular Optimization

Arxiv

0+阅读 · 2023年2月10日

Out-of-Distribution Representation Learning for Time Series Classification

Arxiv

1+阅读 · 2023年2月10日

Model-Based Regression Adjustment with Model-Free Covariates for Network Interference

Arxiv

0+阅读 · 2023年2月10日

Trading Information between Latents in Hierarchical Variational Autoencoders

Arxiv

0+阅读 · 2023年2月9日

Neural Architecture Search without Training

Neural Architecture Search without Training

Arxiv

10+阅读 · 2021年6月11日

Adversarial Mutual Information for Text Generation

Adversarial Mutual Information for Text Generation

Arxiv

13+阅读 · 2020年6月30日

How to train your MAML

Arxiv

26+阅读 · 2019年3月5日

VIP会员

文章信息

相关主题

奇异值分解

相关VIP内容

NeurlPS 2022 | 自然语言处理相关论文分类整理

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

最新BERT相关论文清单，BERT-related Papers

最新BERT相关论文清单，BERT-related Papers

专知会员服务

53+阅读 · 2019年9月29日

热门VIP内容

开通专知VIP会员享更多权益服务

《使用量化测量将传感器节点关联到融合中心的算法设计》171页

军事前沿模型

提升军事训练能力的最佳人工智能模拟工具

《社交媒体信息作战》最新48页技术报告

相关资讯

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

ICLR2019最佳论文出炉

ICLR2019最佳论文出炉

专知

12+阅读 · 2019年5月6日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

vae 相关论文表示学习 1

vae 相关论文表示学习 1

CreateAMind

12+阅读 · 2018年9月6日

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

专知

23+阅读 · 2018年1月18日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

相关论文

Learning a Gaussian Mixture Model from Imperfect Training Data for Robust Channel Estimation

Arxiv

0+阅读 · 2023年2月13日

Dimension Reduction and MARS

Arxiv

0+阅读 · 2023年2月11日

Verifying Generalization in Deep Learning

Arxiv

0+阅读 · 2023年2月11日

Data-Efficient Structured Pruning via Submodular Optimization

Arxiv

0+阅读 · 2023年2月10日

Out-of-Distribution Representation Learning for Time Series Classification

Arxiv

1+阅读 · 2023年2月10日

Model-Based Regression Adjustment with Model-Free Covariates for Network Interference

Arxiv

0+阅读 · 2023年2月10日

Trading Information between Latents in Hierarchical Variational Autoencoders

Arxiv

0+阅读 · 2023年2月9日

Neural Architecture Search without Training

Neural Architecture Search without Training

Arxiv

10+阅读 · 2021年6月11日

Adversarial Mutual Information for Text Generation

Adversarial Mutual Information for Text Generation

Arxiv

13+阅读 · 2020年6月30日

How to train your MAML

Arxiv

26+阅读 · 2019年3月5日

相关基金

采用pinball loss的MEE算法研究

国家自然科学基金

1+阅读 · 2013年12月31日

Partial Spread Bent函数与Bent-Negabent函数的构造及密码学性质研究

国家自然科学基金

0+阅读 · 2013年12月31日

联合显著性检测和对象分割的算法研究

国家自然科学基金

0+阅读 · 2013年12月31日

Kronheimer-Nakajima quiver 模空间与有理曲面

国家自然科学基金

1+阅读 · 2013年12月31日

柯萨奇病毒 A16 实心/空心颗粒抗原性差异的结构基础

国家自然科学基金

0+阅读 · 2013年12月31日

中国BTV-1与BTV-4毒株Seg2重排致病毒变异机制的研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于list-mode数据的快速SART真3D PET断层重建算法的研究

国家自然科学基金

0+阅读 · 2011年12月31日

远红外新型Te基硫系玻璃研制及相关性质研究

国家自然科学基金

0+阅读 · 2009年12月31日

CD73在动脉粥样硬化斑块破裂中的作用及其机制研究

国家自然科学基金

0+阅读 · 2009年12月31日

高分辨大视野的多针孔SPECT成像研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员