我们能否在全球范围优化交叉校准损失? (Can we globally optimize cross-validation loss? Quasiconvexity in ridge regression) - 专知论文

会员服务 ·

0

岭回归 · 全局优化 · 优化器 · 损失 · CASE ·

2021 年 7 月 19 日

Can we globally optimize cross-validation loss? Quasiconvexity in ridge regression

翻译：我们能否在全球范围优化交叉校准损失?

William T. Stephenson,Zachary Frangella,Madeleine Udell,Tamara Broderick

from arxiv, 20 pages, 6 figures

Models like LASSO and ridge regression are extensively used in practice due to their interpretability, ease of use, and strong theoretical guarantees. Cross-validation (CV) is widely used for hyperparameter tuning in these models, but do practical optimization methods minimize the true out-of-sample loss? A recent line of research promises to show that the optimum of the CV loss matches the optimum of the out-of-sample loss (possibly after simple corrections). It remains to show how tractable it is to minimize the CV loss. In the present paper, we show that, in the case of ridge regression, the CV loss may fail to be quasiconvex and thus may have multiple local optima. We can guarantee that the CV loss is quasiconvex in at least one case: when the spectrum of the covariate matrix is nearly flat and the noise in the observed responses is not too high. More generally, we show that quasiconvexity status is independent of many properties of the observed data (response norm, covariate-matrix right singular vectors and singular-value scaling) and has a complex dependence on the few that remain. We empirically confirm our theory using simulated experiments.

翻译：LASSO和山脊回归等模型由于其可解释性、使用方便性以及强有力的理论保障而在实践中被广泛使用。交叉校准(CV)在这些模型中被广泛用于超参数调制,但实际优化方法可以最大限度地减少真实的体外损失?最近的一系列研究承诺表明,CV损失的最佳性与SAPSO和山脊回归的最佳性相匹配(在简单校正之后可能存在)。仍需表明,将CV损失减少到最小程度是多少。在本文件中,我们表明,在山脊回归的情况下,CV损失可能无法成为准电离子,因此可能具有多个局部的奥地性。我们可以保证,CV损失至少在一种情况下是准电流:当COV矩阵的频谱接近平和观察到的应对措施的噪音不高时。更一般地说,我们表明,准电解状态与观察到的数据的许多特性是独立的(对应规范、正变量对准正对准的正向向向向量和奇向值理论),我们使用模拟的实验仍然具有复杂的依赖性。

0

相关内容

岭回归

【经典书】计算最优传输，209页pdf，Computational Optimal Transport

【经典书】计算最优传输，209页pdf，Computational Optimal Transport

专知会员服务

75+阅读 · 2021年1月10日

【干货书】机器学习速查手册，135页pdf

【干货书】机器学习速查手册，135页pdf

专知会员服务

127+阅读 · 2020年11月20日

【伯克利-Ke Li】学习优化，74页ppt，Learning to Optimize

【伯克利-Ke Li】学习优化，74页ppt，Learning to Optimize

专知会员服务

41+阅读 · 2020年7月23日

【NLP模型压缩方法综述】《A Survey of Methods for Model Compression in NLP》by Madison May

【NLP模型压缩方法综述】《A Survey of Methods for Model Compression in NLP》by Madison May

专知会员服务

43+阅读 · 2020年4月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【SIGIR2018】五篇对抗训练文章

【SIGIR2018】五篇对抗训练文章

专知

12+阅读 · 2018年7月9日

【论文推荐】最新十篇度量学习相关论文—可量化表示、非线性度量学习、在线深度量学习、大间隔最近邻、判别深度度量、域自适应

【论文推荐】最新十篇度量学习相关论文—可量化表示、非线性度量学习、在线深度量学习、大间隔最近邻、判别深度度量、域自适应

专知

12+阅读 · 2018年5月18日

Hierarchical Disentangled Representations

Hierarchical Disentangled Representations

CreateAMind

4+阅读 · 2018年4月15日

【推荐】决策树/随机森林深入解析

【推荐】决策树/随机森林深入解析

机器学习研究会

5+阅读 · 2017年9月21日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

最佳实践：深度学习用于自然语言处理（三）

最佳实践：深度学习用于自然语言处理（三）

待字闺中

3+阅读 · 2017年8月20日

【学习】Hierarchical Softmax

【学习】Hierarchical Softmax

机器学习研究会

4+阅读 · 2017年8月6日

Auto-Encoding GAN

Auto-Encoding GAN

CreateAMind

7+阅读 · 2017年8月4日

PKLM: A flexible MCAR test using Classification

PKLM: A flexible MCAR test using Classification

Arxiv

0+阅读 · 2021年9月21日

The Canny-Emiris conjecture for the sparse resultant

Arxiv

0+阅读 · 2021年9月21日

Sharp global convergence guarantees for iterative nonconvex optimization: A Gaussian process perspective

Arxiv

0+阅读 · 2021年9月20日

Revisiting the Characteristics of Stochastic Gradient Noise and Dynamics

Arxiv

0+阅读 · 2021年9月20日

Deep Quantile Regression for Uncertainty Estimation in Unsupervised and Supervised Lesion Detection

Arxiv

0+阅读 · 2021年9月20日

A New Non-parametric Test for Multivariate Paired Data and Pair Matching

Arxiv

0+阅读 · 2021年9月19日

Online Multiobjective Minimax Optimization and Applications

Arxiv

0+阅读 · 2021年9月18日

Regression Discontinuity Design with Potentially Many Covariates

Arxiv

0+阅读 · 2021年9月17日

Adaptive Ridge-Penalized Functional Local Linear Regression

Arxiv

0+阅读 · 2021年9月17日

On the Implicit Assumptions of GANs

Arxiv

6+阅读 · 2018年11月29日

VIP会员

文章信息

相关主题

相关VIP内容

【经典书】计算最优传输，209页pdf，Computational Optimal Transport

【经典书】计算最优传输，209页pdf，Computational Optimal Transport

专知会员服务

75+阅读 · 2021年1月10日

【干货书】机器学习速查手册，135页pdf

【干货书】机器学习速查手册，135页pdf

专知会员服务

127+阅读 · 2020年11月20日

【伯克利-Ke Li】学习优化，74页ppt，Learning to Optimize

【伯克利-Ke Li】学习优化，74页ppt，Learning to Optimize

专知会员服务

41+阅读 · 2020年7月23日

【NLP模型压缩方法综述】《A Survey of Methods for Model Compression in NLP》by Madison May

【NLP模型压缩方法综述】《A Survey of Methods for Model Compression in NLP》by Madison May

专知会员服务

43+阅读 · 2020年4月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【ACMMM2025教程】打击网络虚假信息视频：特征分析、检测与防范，170页ppt

海军无人系统：海上作战的演进而非革命

Nature 子刊 | SciToolAgent:知识图谱引导的科学工具智能体

多媒体顶会ACM Multimedia 2025各大奖项揭晓！格拉斯哥大学等获最佳论文，中科院自动化所等获最佳学生论文

相关资讯

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【SIGIR2018】五篇对抗训练文章

【SIGIR2018】五篇对抗训练文章

专知

12+阅读 · 2018年7月9日

【论文推荐】最新十篇度量学习相关论文—可量化表示、非线性度量学习、在线深度量学习、大间隔最近邻、判别深度度量、域自适应

【论文推荐】最新十篇度量学习相关论文—可量化表示、非线性度量学习、在线深度量学习、大间隔最近邻、判别深度度量、域自适应

专知

12+阅读 · 2018年5月18日

Hierarchical Disentangled Representations

Hierarchical Disentangled Representations

CreateAMind

4+阅读 · 2018年4月15日

【推荐】决策树/随机森林深入解析

【推荐】决策树/随机森林深入解析

机器学习研究会

5+阅读 · 2017年9月21日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

最佳实践：深度学习用于自然语言处理（三）

最佳实践：深度学习用于自然语言处理（三）

待字闺中

3+阅读 · 2017年8月20日

【学习】Hierarchical Softmax

【学习】Hierarchical Softmax

机器学习研究会

4+阅读 · 2017年8月6日

Auto-Encoding GAN

Auto-Encoding GAN

CreateAMind

7+阅读 · 2017年8月4日

相关论文

PKLM: A flexible MCAR test using Classification

PKLM: A flexible MCAR test using Classification

Arxiv

0+阅读 · 2021年9月21日

The Canny-Emiris conjecture for the sparse resultant

Arxiv

0+阅读 · 2021年9月21日

Sharp global convergence guarantees for iterative nonconvex optimization: A Gaussian process perspective

Arxiv

0+阅读 · 2021年9月20日

Revisiting the Characteristics of Stochastic Gradient Noise and Dynamics

Arxiv

0+阅读 · 2021年9月20日

Deep Quantile Regression for Uncertainty Estimation in Unsupervised and Supervised Lesion Detection

Arxiv

0+阅读 · 2021年9月20日

A New Non-parametric Test for Multivariate Paired Data and Pair Matching

Arxiv

0+阅读 · 2021年9月19日

Online Multiobjective Minimax Optimization and Applications

Arxiv

0+阅读 · 2021年9月18日

Regression Discontinuity Design with Potentially Many Covariates

Arxiv

0+阅读 · 2021年9月17日

Adaptive Ridge-Penalized Functional Local Linear Regression

Arxiv

0+阅读 · 2021年9月17日

On the Implicit Assumptions of GANs

Arxiv

6+阅读 · 2018年11月29日

微信扫码咨询专知VIP会员