双级优化框架,以便能够进行随机和全球差异减少算法 (A framework for bilevel optimization that enables stochastic and global variance reduction algorithms) - 专知论文

会员服务 ·

0

方差减小 · 方差 · 线性的 · 优化器 · 价值函数 ·

2022 年 9 月 12 日

A framework for bilevel optimization that enables stochastic and global variance reduction algorithms

翻译：双级优化框架,以便能够进行随机和全球差异减少算法

Mathieu Dagréou,Pierre Ablin,Samuel Vaiter,Thomas Moreau

Bilevel optimization, the problem of minimizing a value function which involves the arg-minimum of another function, appears in many areas of machine learning. In a large scale empirical risk minimization setting where the number of samples is huge, it is crucial to develop stochastic methods, which only use a few samples at a time to progress. However, computing the gradient of the value function involves solving a linear system, which makes it difficult to derive unbiased stochastic estimates. To overcome this problem we introduce a novel framework, in which the solution of the inner problem, the solution of the linear system, and the main variable evolve at the same time. These directions are written as a sum, making it straightforward to derive unbiased estimates. The simplicity of our approach allows us to develop global variance reduction algorithms, where the dynamics of all variables is subject to variance reduction. We demonstrate that SABA, an adaptation of the celebrated SAGA algorithm in our framework, has $O(\frac1T)$ convergence rate, and that it achieves linear convergence under Polyak-Lojasciewicz assumption. This is the first stochastic algorithm for bilevel optimization that verifies either of these properties. Numerical experiments validate the usefulness of our method.

翻译：双层优化双层优化, 将包含另一个函数最小化的值函数最小化的问题, 出现在机器学习的许多领域。在大规模实验风险最小化的大规模实验中, 样本数量巨大, 开发随机分析方法至关重要。然而, 计算值函数的梯度需要解决线性系统, 这使得难以得出公正的随机估计值。要克服这个问题, 我们引入了一个新颖的框架, 解决内部问题、线性系统解决方案和主要变量同时演变。这些方向是写成一个总和, 直截了当地得出不偏直的估计数。我们方法的简单性使我们能够开发全球差异减少算法, 所有的变量的动态都会降低差异。我们证明, SABA, 是我们框架中值得庆祝的SAGA算法的调整, 具有$( orc1T) 的汇合率, 并在Polyak- Lojaciewicz假设下实现线性融合。这是用于验证这些属性的双层优化方法的首个高级算法。

0

相关内容

方差减小

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

深度学习优化算法，73页ppt，Optimization Algorithms on Deep Learning

深度学习优化算法，73页ppt，Optimization Algorithms on Deep Learning

专知会员服务

135+阅读 · 2021年6月16日

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

【ICCV 2019 Toturial】Global Optimization for Geometric Understanding with Provable Guarantees（具有可证明保证的几何理解的全局优化）

【ICCV 2019 Toturial】Global Optimization for Geometric Understanding with Provable Guarantees（具有可证明保证的几何理解的全局优化）

专知会员服务

18+阅读 · 2019年11月1日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

集值优化问题的逼近解及二阶最优性条件

国家自然科学基金

0+阅读 · 2014年12月31日

小麦太谷核不育基因Ms2的图位克隆

国家自然科学基金

0+阅读 · 2014年12月31日

用于肿瘤标志物检测的稀土掺杂超小MF2(M=Ca,Sr,Ba)纳米荧光探针及其发光物理

国家自然科学基金

0+阅读 · 2013年12月31日

基于证据推理算法的建筑用能行为理论模型研究

国家自然科学基金

0+阅读 · 2013年12月31日

精神分裂症记忆障碍的脑网络组学研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于联合决策与估计的高频超视距雷达信息处理与融合

国家自然科学基金

3+阅读 · 2011年12月31日

miR-140在肿瘤转移中的作用及机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于无约束凸优化的多尺度动态图像分割方法研究

国家自然科学基金

0+阅读 · 2009年12月31日

急性淋巴细胞白血病（ALL）逃逸NK细胞杀伤的机制研究

国家自然科学基金

0+阅读 · 2008年12月31日

Sorcin蛋白在胃癌耐药细胞中的相互作用网络研究

国家自然科学基金

0+阅读 · 2008年12月31日

Targeted active learning for probabilistic models

Arxiv

0+阅读 · 2022年10月21日

Fast and numerically stable particle-based online additive smoothing: the AdaSmooth algorithm

Arxiv

0+阅读 · 2022年10月21日

Efficient Submodular Optimization under Noise: Local Search is Robust

Arxiv

0+阅读 · 2022年10月21日

Stochastic Adaptive Activation Function

Arxiv

0+阅读 · 2022年10月21日

Exact Inference for Stochastic Epidemic Models via Uniformly Ergodic Block Sampling

Arxiv

0+阅读 · 2022年10月20日

Efficient variational approximations for state space models

Arxiv

0+阅读 · 2022年10月20日

No-Regret Dynamics in the Fenchel Game: A Unified Framework for Algorithmic Convex Optimization

Arxiv

0+阅读 · 2022年10月20日

Bring Your Own Algorithm for Optimal Differentially Private Stochastic Minimax Optimization

Arxiv

0+阅读 · 2022年10月19日

Hyper-Parameter Optimization: A Review of Algorithms and Applications

Hyper-Parameter Optimization: A Review of Algorithms and Applications

Arxiv

16+阅读 · 2020年3月12日

Optimization for deep learning: theory and algorithms

Optimization for deep learning: theory and algorithms

Arxiv

106+阅读 · 2019年12月19日

VIP会员

文章信息

相关主题

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

深度学习优化算法，73页ppt，Optimization Algorithms on Deep Learning

深度学习优化算法，73页ppt，Optimization Algorithms on Deep Learning

专知会员服务

135+阅读 · 2021年6月16日

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

【ICCV 2019 Toturial】Global Optimization for Geometric Understanding with Provable Guarantees（具有可证明保证的几何理解的全局优化）

【ICCV 2019 Toturial】Global Optimization for Geometric Understanding with Provable Guarantees（具有可证明保证的几何理解的全局优化）

专知会员服务

18+阅读 · 2019年11月1日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《物联网（IoT）中的无人机通信高效控制》135页

《在GNSS信号降级环境中利用共识实现无人机集群稳健协调》

中程单向攻击无人机的战略意义：俄乌战争启示

《面向无人机集群的避障动态传感器覆盖算法》最新38页

相关资讯

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

相关论文

Targeted active learning for probabilistic models

Arxiv

0+阅读 · 2022年10月21日

Fast and numerically stable particle-based online additive smoothing: the AdaSmooth algorithm

Arxiv

0+阅读 · 2022年10月21日

Efficient Submodular Optimization under Noise: Local Search is Robust

Arxiv

0+阅读 · 2022年10月21日

Stochastic Adaptive Activation Function

Arxiv

0+阅读 · 2022年10月21日

Exact Inference for Stochastic Epidemic Models via Uniformly Ergodic Block Sampling

Arxiv

0+阅读 · 2022年10月20日

Efficient variational approximations for state space models

Arxiv

0+阅读 · 2022年10月20日

No-Regret Dynamics in the Fenchel Game: A Unified Framework for Algorithmic Convex Optimization

Arxiv

0+阅读 · 2022年10月20日

Bring Your Own Algorithm for Optimal Differentially Private Stochastic Minimax Optimization

Arxiv

0+阅读 · 2022年10月19日

Hyper-Parameter Optimization: A Review of Algorithms and Applications

Hyper-Parameter Optimization: A Review of Algorithms and Applications

Arxiv

16+阅读 · 2020年3月12日

Optimization for deep learning: theory and algorithms

Optimization for deep learning: theory and algorithms

Arxiv

106+阅读 · 2019年12月19日

相关基金

集值优化问题的逼近解及二阶最优性条件

国家自然科学基金

0+阅读 · 2014年12月31日

小麦太谷核不育基因Ms2的图位克隆

国家自然科学基金

0+阅读 · 2014年12月31日

用于肿瘤标志物检测的稀土掺杂超小MF2(M=Ca,Sr,Ba)纳米荧光探针及其发光物理

国家自然科学基金

0+阅读 · 2013年12月31日

基于证据推理算法的建筑用能行为理论模型研究

国家自然科学基金

0+阅读 · 2013年12月31日

精神分裂症记忆障碍的脑网络组学研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于联合决策与估计的高频超视距雷达信息处理与融合

国家自然科学基金

3+阅读 · 2011年12月31日

miR-140在肿瘤转移中的作用及机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于无约束凸优化的多尺度动态图像分割方法研究

国家自然科学基金

0+阅读 · 2009年12月31日

急性淋巴细胞白血病（ALL）逃逸NK细胞杀伤的机制研究

国家自然科学基金

0+阅读 · 2008年12月31日

Sorcin蛋白在胃癌耐药细胞中的相互作用网络研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员