有条件预测性绩效措施的低信任度 (Multiplicity-adjusted bootstrap tilting lower confidence bounds for conditional prediction performance measures) - 专知论文

会员服务 ·

0

Performer · 置信度 · MoDELS · 性能度量 · 自助法/自举法 ·

2022 年 10 月 24 日

Multiplicity-adjusted bootstrap tilting lower confidence bounds for conditional prediction performance measures

翻译：有条件预测性绩效措施的低信任度

Pascal Rink,Werner Brannath

In machine learning, the selection of a promising model from a potentially large number of competing models and the assessment of its generalization performance are critical tasks that need careful consideration. Typically, model selection and evaluation are strictly separated endeavors, splitting the sample at hand into a training, validation, and evaluation set, and only compute a single confidence interval for the prediction performance of the final selected model. We however propose an algorithm how to compute valid lower confidence bounds for multiple models that have been selected based on their prediction performances in the evaluation set by interpreting the selection problem as a simultaneous inference problem. We use bootstrap tilting and a maxT-type multiplicity correction. The approach is universally applicable for any combination of prediction models, any model selection strategy, and any prediction performance measure that accepts weights. We conducted various simulation experiments which show that our proposed approach yields lower confidence bounds that are at least comparably good as bounds from standard approaches, and that reliably reach the nominal coverage probability. In addition, especially when sample size is small, our proposed approach yields better performing prediction models than the default selection of only one model for evaluation does.

翻译：在机器学习中,从众多可能相互竞争的模型中选择一个有希望的模式,并评估其总体性能,这些都是需要认真考虑的关键任务。通常,模型选择和评价是严格分开的努力,将手头的样本分成一个培训、鉴定和评价组,并且只计算最后选定的模型预测性能的单一信任间隔。然而,我们提出一个算法,如何根据多个模型的预测性能来计算其有效的较低信任界限,这些模型是根据其在评估中的预测性能而选定的,通过将选择问题解释为同时发生的推论问题。我们使用靴套倾斜和最大T型多重校正。该方法普遍适用于预测性模型的任何组合、任何模型选择战略和任何接受重量的预测性业绩计量。我们进行了各种模拟实验,这些实验表明,我们拟议的方法产生的信任界限比标准方法的界限要低,而且至少比得上好,而且可靠地达到名义覆盖概率。此外,在样本规模小的情况下,我们提议的方法比默认选择的评价模型的准确性要好。

0

相关内容

Performer

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

多尺度多场应力耦合致密砂岩体积改造裂缝评价模型研究

国家自然科学基金

0+阅读 · 2015年12月31日

基于WorldView-3和OP-ELM的矿化蚀变提取方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

含有缺失值的纵向数据回归模型的稳健推断

国家自然科学基金

3+阅读 · 2012年12月31日

实时安全关键系统的建模、仿真与验证

国家自然科学基金

1+阅读 · 2012年12月31日

云计算环境中身份基海量数据分布式PDP的研究

国家自然科学基金

0+阅读 · 2012年12月31日

LncRNAs在非小细胞肺癌EGFR-TKIs耐药中的作用及分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

低自旋磁各向异性过渡金属配合物的合成与应用研究

国家自然科学基金

0+阅读 · 2011年12月31日

退化k-Hessian方程解的正则性研究

国家自然科学基金

0+阅读 · 2011年12月31日

帕金森病发病机制中的DNA甲基化作用研究

国家自然科学基金

0+阅读 · 2009年12月31日

EMCD和ALCHEMI研究单个DMS纳米结构的铁磁性内禀属性

国家自然科学基金

0+阅读 · 2009年12月31日

On the Robustness of Normalizing Flows for Inverse Problems in Imaging

Arxiv

0+阅读 · 2022年12月8日

Mind the Gap: Measuring Generalization Performance Across Multiple Objectives

Arxiv

0+阅读 · 2022年12月8日

On the Global Solution of Soft k-Means

Arxiv

0+阅读 · 2022年12月7日

An integrated approach to test for missing not at random

Arxiv

0+阅读 · 2022年12月7日

A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification

Arxiv

0+阅读 · 2022年12月7日

Root-finding Approaches for Computing Conformal Prediction Set

Arxiv

0+阅读 · 2022年12月7日

Calibration and generalizability of probabilistic models on low-data chemical datasets with DIONYSUS

Arxiv

0+阅读 · 2022年12月6日

Learning to Bound Counterfactual Inference in Structural Causal Models from Observational and Randomised Data

Arxiv

0+阅读 · 2022年12月6日

On regression and classification with possibly missing response variables in the data

Arxiv

0+阅读 · 2022年12月6日

Concentration of measure bounds for matrix-variate data with missing values

Arxiv

0+阅读 · 2022年12月6日

VIP会员

文章信息

相关主题

自助法/自举法

相关VIP内容

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【CMU博士论文】以人为中心的强化学习

任务规划与地形分析：现代复杂环境作战导航体系

认知优势：人工智能在国家安全决策中的核心作用

大模型赋能的具身智能：决策与具身学习综述

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

On the Robustness of Normalizing Flows for Inverse Problems in Imaging

Arxiv

0+阅读 · 2022年12月8日

Mind the Gap: Measuring Generalization Performance Across Multiple Objectives

Arxiv

0+阅读 · 2022年12月8日

On the Global Solution of Soft k-Means

Arxiv

0+阅读 · 2022年12月7日

An integrated approach to test for missing not at random

Arxiv

0+阅读 · 2022年12月7日

A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification

Arxiv

0+阅读 · 2022年12月7日

Root-finding Approaches for Computing Conformal Prediction Set

Arxiv

0+阅读 · 2022年12月7日

Calibration and generalizability of probabilistic models on low-data chemical datasets with DIONYSUS

Arxiv

0+阅读 · 2022年12月6日

Learning to Bound Counterfactual Inference in Structural Causal Models from Observational and Randomised Data

Arxiv

0+阅读 · 2022年12月6日

On regression and classification with possibly missing response variables in the data

Arxiv

0+阅读 · 2022年12月6日

Concentration of measure bounds for matrix-variate data with missing values

Arxiv

0+阅读 · 2022年12月6日

相关基金

多尺度多场应力耦合致密砂岩体积改造裂缝评价模型研究

国家自然科学基金

0+阅读 · 2015年12月31日

基于WorldView-3和OP-ELM的矿化蚀变提取方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

含有缺失值的纵向数据回归模型的稳健推断

国家自然科学基金

3+阅读 · 2012年12月31日

实时安全关键系统的建模、仿真与验证

国家自然科学基金

1+阅读 · 2012年12月31日

云计算环境中身份基海量数据分布式PDP的研究

国家自然科学基金

0+阅读 · 2012年12月31日

LncRNAs在非小细胞肺癌EGFR-TKIs耐药中的作用及分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

低自旋磁各向异性过渡金属配合物的合成与应用研究

国家自然科学基金

0+阅读 · 2011年12月31日

退化k-Hessian方程解的正则性研究

国家自然科学基金

0+阅读 · 2011年12月31日

帕金森病发病机制中的DNA甲基化作用研究

国家自然科学基金

0+阅读 · 2009年12月31日

EMCD和ALCHEMI研究单个DMS纳米结构的铁磁性内禀属性

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员