交叉验证:它估计什么以及它做得如何? (Cross-validation: what does it estimate and how well does it do it?) - 专知论文

会员服务 ·

0

估计/估计量 · 交叉验证 · 方差 · MoDELS · 数据拆分 ·

2022 年 6 月 10 日

Cross-validation: what does it estimate and how well does it do it?

翻译：交叉验证:它估计什么以及它做得如何?

Stephen Bates,Trevor Hastie,Robert Tibshirani

Cross-validation is a widely-used technique to estimate prediction error, but its behavior is complex and not fully understood. Ideally, one would like to think that cross-validation estimates the prediction error for the model at hand, fit to the training data. We prove that this is not the case for the linear model fit by ordinary least squares; rather it estimates the average prediction error of models fit on other unseen training sets drawn from the same population. We further show that this phenomenon occurs for most popular estimates of prediction error, including data splitting, bootstrapping, and Mallow's Cp. Next, the standard confidence intervals for prediction error derived from cross-validation may have coverage far below the desired level. Because each data point is used for both training and testing, there are correlations among the measured accuracies for each fold, and so the usual estimate of variance is too small. We introduce a nested cross-validation scheme to estimate this variance more accurately, and we show empirically that this modification leads to intervals with approximately correct coverage in many examples where traditional cross-validation intervals fail.

翻译：交叉校准是用来估计预测误差的一种广泛使用的方法,但其行为是复杂和不完全理解的。理想的是,人们会认为交叉校准估计手头模型的预测误差,符合培训数据。我们证明线性模型适合普通最小方格的情况并非如此; 而它估计适合同一人群的其他无形培训组的模型的平均预测误差。我们进一步表明,这种现象发生在大多数流行的预测误差估计中,包括数据分离、制靴和Mallow's Cp。下一步, 交叉校准产生的预测误差标准信任期的覆盖面可能远远低于理想水平。由于每个数据点用于培训和测试,每个折数的测量误差都有关联性,因此通常的差异估计太小。我们采用了嵌套的交叉校准办法来更准确地估计这一差异,我们从经验上表明,这种修改导致间隔,在传统的交叉校准间隔期失败的许多例子中,这种间隔范围大致正确。

0

相关内容

估计/估计量

估计/估计量

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

UC.Berkeley CS189讲义教材:《机器学习全面指南》，185页pdf

专知会员服务

162+阅读 · 2020年1月16日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

IEEE TII Call For Papers

IEEE TII Call For Papers

CCF多媒体专委会

3+阅读 · 2022年3月24日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

内源性逆转录病毒在小鼠胚胎干细胞中的转录抑制机制

国家自然科学基金

0+阅读 · 2016年12月31日

血红素加氧酶1抑制猪繁殖与呼吸综合征病毒复制的分子机制

国家自然科学基金

0+阅读 · 2014年12月31日

肾癌中KLF4转录调控细胞外基质蛋白fibulin-1的分子机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

PGC-1α调节骨骼肌脂肪酸代谢和胰岛素抵抗的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

CD226分子调控血小板功能参与动脉粥样硬化疾病的机制

国家自然科学基金

0+阅读 · 2012年12月31日

PIM-1信号通路在非小细胞肺癌EGFR-TKI获得性耐药中的作用及其分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

P2-HNF4α促进肝癌细胞增殖的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

LncRNAs在非小细胞肺癌EGFR-TKIs耐药中的作用及分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

肝细胞癌血管生成拟态的分子机制研究

国家自然科学基金

0+阅读 · 2009年12月31日

附睾蛋白酶抑制剂(EPPIN)基因转录调控的分子机理

国家自然科学基金

0+阅读 · 2009年12月31日

Reconstruction of inhomogeneous media by iterative reconstruction algorithm with learned projector

Reconstruction of inhomogeneous media by iterative reconstruction algorithm with learned projector

Arxiv

0+阅读 · 2022年7月26日

Variance estimation in graphs with the fused lasso

Arxiv

0+阅读 · 2022年7月26日

Differentially Private Estimation via Statistical Depth

Arxiv

0+阅读 · 2022年7月26日

When does SGD favor flat minima? A quantitative characterization via linear stability

Arxiv

0+阅读 · 2022年7月25日

Estimating Extreme Value Index by Subsampling for Massive Datasets with Heavy-Tailed Distributions

Arxiv

0+阅读 · 2022年7月25日

An Exploration of How Training Set Composition Bias in Machine Learning Affects Identifying Rare Objects

Arxiv

0+阅读 · 2022年7月25日

Finite-sample bias-correction factors for the median absolute deviation based on the Harrell-Davis quantile estimator and its trimmed modification

Arxiv

0+阅读 · 2022年7月25日

Correcting Model Bias with Sparse Implicit Processes

Arxiv

0+阅读 · 2022年7月21日

The Causal Learning of Retail Delinquency

Arxiv

14+阅读 · 2020年12月17日

Which Knowledge Graph Is Best for Me?

Arxiv

11+阅读 · 2018年9月28日

VIP会员

文章信息

相关主题

估计/估计量

相关VIP内容

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

UC.Berkeley CS189讲义教材:《机器学习全面指南》，185页pdf

专知会员服务

162+阅读 · 2020年1月16日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

大语言模型基准综述

《自适应训练辅助系统概念导论及其在空战指挥官加速培训中的应用》125页

【剑桥博士论文】多智能体学习中的神经多样性

以色列-伊朗空战：短暂而激烈冲突的启示

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

IEEE TII Call For Papers

IEEE TII Call For Papers

CCF多媒体专委会

3+阅读 · 2022年3月24日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

Reconstruction of inhomogeneous media by iterative reconstruction algorithm with learned projector

Reconstruction of inhomogeneous media by iterative reconstruction algorithm with learned projector

Arxiv

0+阅读 · 2022年7月26日

Variance estimation in graphs with the fused lasso

Arxiv

0+阅读 · 2022年7月26日

Differentially Private Estimation via Statistical Depth

Arxiv

0+阅读 · 2022年7月26日

When does SGD favor flat minima? A quantitative characterization via linear stability

Arxiv

0+阅读 · 2022年7月25日

Estimating Extreme Value Index by Subsampling for Massive Datasets with Heavy-Tailed Distributions

Arxiv

0+阅读 · 2022年7月25日

An Exploration of How Training Set Composition Bias in Machine Learning Affects Identifying Rare Objects

Arxiv

0+阅读 · 2022年7月25日

Finite-sample bias-correction factors for the median absolute deviation based on the Harrell-Davis quantile estimator and its trimmed modification

Arxiv

0+阅读 · 2022年7月25日

Correcting Model Bias with Sparse Implicit Processes

Arxiv

0+阅读 · 2022年7月21日

The Causal Learning of Retail Delinquency

Arxiv

14+阅读 · 2020年12月17日

Which Knowledge Graph Is Best for Me?

Arxiv

11+阅读 · 2018年9月28日

相关基金

内源性逆转录病毒在小鼠胚胎干细胞中的转录抑制机制

国家自然科学基金

0+阅读 · 2016年12月31日

血红素加氧酶1抑制猪繁殖与呼吸综合征病毒复制的分子机制

国家自然科学基金

0+阅读 · 2014年12月31日

肾癌中KLF4转录调控细胞外基质蛋白fibulin-1的分子机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

PGC-1α调节骨骼肌脂肪酸代谢和胰岛素抵抗的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

CD226分子调控血小板功能参与动脉粥样硬化疾病的机制

国家自然科学基金

0+阅读 · 2012年12月31日

PIM-1信号通路在非小细胞肺癌EGFR-TKI获得性耐药中的作用及其分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

P2-HNF4α促进肝癌细胞增殖的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

LncRNAs在非小细胞肺癌EGFR-TKIs耐药中的作用及分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

肝细胞癌血管生成拟态的分子机制研究

国家自然科学基金

0+阅读 · 2009年12月31日

附睾蛋白酶抑制剂(EPPIN)基因转录调控的分子机理

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员