Translated title: 深度学习模型的可靠性改进：通过模型间潜在协议提高可靠性 (Great Models Think Alike: Improving Model Reliability via Inter-Model Latent Agreement) - 专知论文

会员服务 ·

0

潜在 · 高可靠 · 高可靠性 · 连贯性 · 模型预测 ·

2023 年 5 月 2 日

Great Models Think Alike: Improving Model Reliability via Inter-Model Latent Agreement

翻译：Translated title: 深度学习模型的可靠性改进：通过模型间潜在协议提高可靠性

Ailin Deng,Miao Xiong,Bryan Hooi

from arxiv, ICML 2023

Reliable application of machine learning is of primary importance to the practical deployment of deep learning methods. A fundamental challenge is that models are often unreliable due to overconfidence. In this paper, we estimate a model's reliability by measuring \emph{the agreement between its latent space, and the latent space of a foundation model}. However, it is challenging to measure the agreement between two different latent spaces due to their incoherence, \eg, arbitrary rotations and different dimensionality. To overcome this incoherence issue, we design a \emph{neighborhood agreement measure} between latent spaces and find that this agreement is surprisingly well-correlated with the reliability of a model's predictions. Further, we show that fusing neighborhood agreement into a model's predictive confidence in a post-hoc way significantly improves its reliability. Theoretical analysis and extensive experiments on failure detection across various datasets verify the effectiveness of our method on both in-distribution and out-of-distribution settings.

翻译：Translated Abstract: 机器学习的可靠应用对于深度学习方法的实际部署至关重要。一个基本的挑战是模型经常因为过度自信而不可靠。在本文中，我们通过测量模型的潜在空间与基础模型的潜在空间之间的协议来估计模型的可靠性。然而，由于这些潜在空间的不连贯性，如任意旋转和不同的维度，所以测量两个不同潜在空间之间的协议是具有挑战性的。为了克服这种不连贯性，我们设计了一种潜在空间间邻域协议（neighborhood agreement measure）方法，并发现潜在空间之间的这种协议与模型预测的可靠性有惊人的相关性。此外，我们证明将邻域协议融入模型预测的置信度中，可以极大地提高模型的可靠性。理论分析和在各种数据集上的故障检测的大量实验证明了我们方法在分布内和分布外情况下的有效性。

0

相关内容

如何提升深度学习可靠性？DeepMind研究科学家Stutz博士论文《理解改进深度学习中的鲁棒性和不确定性估计》，291页pdf

如何提升深度学习可靠性？DeepMind研究科学家Stutz博士论文《理解改进深度学习中的鲁棒性和不确定性估计》，291页pdf

专知会员服务

37+阅读 · 2022年10月21日

【Hugging Face】使用自定义数据集微调语义分割模型，Fine-Tune a Semantic Segmentation Model with a Custom Dataset

【Hugging Face】使用自定义数据集微调语义分割模型，Fine-Tune a Semantic Segmentation Model with a Custom Dataset

专知会员服务

21+阅读 · 2022年3月18日

【开放书】卡耐基梅隆大学Elaine Shi 教授《Foundations of Distributed Consensus and Blockchains（分布式共识和区块链的基础）》150页pdf

【开放书】卡耐基梅隆大学Elaine Shi 教授《Foundations of Distributed Consensus and Blockchains（分布式共识和区块链的基础）》150页pdf

专知会员服务

30+阅读 · 2022年2月22日

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

专知会员服务

104+阅读 · 2022年2月10日

【斯坦福大学AI】BERT, ELMo， & GPT-2:上下文化的单词表示是怎样的?

专知会员服务

35+阅读 · 2020年3月28日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

167+阅读 · 2020年3月18日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

灾难性遗忘问题新视角：迁移-干扰平衡

灾难性遗忘问题新视角：迁移-干扰平衡

CreateAMind

17+阅读 · 2019年7月6日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

【Awesome】最全的机器学习可解释性资料（machine-learning-interpretability）

【Awesome】最全的机器学习可解释性资料（machine-learning-interpretability）

专知

29+阅读 · 2019年3月1日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

利用动态深度学习预测金融时间序列基于Python

利用动态深度学习预测金融时间序列基于Python

量化投资与机器学习

18+阅读 · 2018年10月30日

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

专知

20+阅读 · 2018年6月29日

【论文推荐】最新七篇图像描述生成相关论文—CNN+CNN、对抗样本、显著性和上下文注意力、条件生成对抗网络、风格化

【论文推荐】最新七篇图像描述生成相关论文—CNN+CNN、对抗样本、显著性和上下文注意力、条件生成对抗网络、风格化

专知

25+阅读 · 2018年5月28日

高维积分波动率矩阵的估计及其在资产投资中的应用

国家自然科学基金

0+阅读 · 2015年12月31日

具脉冲影响的Van der Pol方程的复杂动力学行为研究

国家自然科学基金

0+阅读 · 2015年12月31日

高维回归模型的预测稳定性研究

国家自然科学基金

3+阅读 · 2015年12月31日

对偶Auslander转置及其诱导模类的同调性质研究

国家自然科学基金

0+阅读 · 2015年12月31日

几种新型稀土金属间化合物中位错性质的高压效应

国家自然科学基金

0+阅读 · 2013年12月31日

Calderon问题和边界刚性问题

国家自然科学基金

0+阅读 · 2013年12月31日

面向多用户MIMO干扰的混合型无线Mesh网络自由度上限及可达方案研究

国家自然科学基金

0+阅读 · 2013年12月31日

冰云辐射性质参数化对东亚夏季风模拟的影响研究

国家自然科学基金

0+阅读 · 2012年12月31日

软件可靠性测试的数学模型研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于信道Time/Power度量指标的TOA测距误差模型及其应用研究

国家自然科学基金

0+阅读 · 2011年12月31日

Uncertainty Quantification via Spatial-Temporal Tweedie Model for Zero-inflated and Long-tail Travel Demand Prediction

Arxiv

0+阅读 · 2023年6月16日

Finite state verifiers with both private and public coins

Arxiv

0+阅读 · 2023年6月15日

On the Feasibility of Cross-Task Transfer with Model-Based Reinforcement Learning

Arxiv

0+阅读 · 2023年6月15日

Improving Training Stability for Multitask Ranking Models in Recommender Systems

Arxiv

0+阅读 · 2023年6月15日

Blind identification of Ambisonic reduced room impulse response

Arxiv

0+阅读 · 2023年6月15日

Some observations on the distribution of order statistics under simple-random-sampling without replacement

Arxiv

0+阅读 · 2023年6月14日

WizardCoder: Empowering Code Large Language Models with Evol-Instruct

Arxiv

0+阅读 · 2023年6月14日

Pretraining Language Models with Human Preferences

Arxiv

0+阅读 · 2023年6月14日

On the Robustness of Latent Diffusion Models

Arxiv

0+阅读 · 2023年6月14日

Improving the Generalizability of Trajectory Prediction Models with Frenet-Based Domain Normalization

Arxiv

0+阅读 · 2023年6月14日

VIP会员

文章信息

相关主题

相关VIP内容

如何提升深度学习可靠性？DeepMind研究科学家Stutz博士论文《理解改进深度学习中的鲁棒性和不确定性估计》，291页pdf

如何提升深度学习可靠性？DeepMind研究科学家Stutz博士论文《理解改进深度学习中的鲁棒性和不确定性估计》，291页pdf

专知会员服务

37+阅读 · 2022年10月21日

【Hugging Face】使用自定义数据集微调语义分割模型，Fine-Tune a Semantic Segmentation Model with a Custom Dataset

【Hugging Face】使用自定义数据集微调语义分割模型，Fine-Tune a Semantic Segmentation Model with a Custom Dataset

专知会员服务

21+阅读 · 2022年3月18日

【开放书】卡耐基梅隆大学Elaine Shi 教授《Foundations of Distributed Consensus and Blockchains（分布式共识和区块链的基础）》150页pdf

【开放书】卡耐基梅隆大学Elaine Shi 教授《Foundations of Distributed Consensus and Blockchains（分布式共识和区块链的基础）》150页pdf

专知会员服务

30+阅读 · 2022年2月22日

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

专知会员服务

104+阅读 · 2022年2月10日

【斯坦福大学AI】BERT, ELMo， & GPT-2:上下文化的单词表示是怎样的?

专知会员服务

35+阅读 · 2020年3月28日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

167+阅读 · 2020年3月18日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

Deep Research（深度研究）：系统性综述

《革新战术战场空间能力：反无人机系统》报告

【普林斯顿博士论文】用于语音的生成式通用模型

螺旋式开发作为战略资产：美军启示

相关资讯

灾难性遗忘问题新视角：迁移-干扰平衡

灾难性遗忘问题新视角：迁移-干扰平衡

CreateAMind

17+阅读 · 2019年7月6日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

【Awesome】最全的机器学习可解释性资料（machine-learning-interpretability）

【Awesome】最全的机器学习可解释性资料（machine-learning-interpretability）

专知

29+阅读 · 2019年3月1日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

利用动态深度学习预测金融时间序列基于Python

利用动态深度学习预测金融时间序列基于Python

量化投资与机器学习

18+阅读 · 2018年10月30日

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

专知

20+阅读 · 2018年6月29日

【论文推荐】最新七篇图像描述生成相关论文—CNN+CNN、对抗样本、显著性和上下文注意力、条件生成对抗网络、风格化

【论文推荐】最新七篇图像描述生成相关论文—CNN+CNN、对抗样本、显著性和上下文注意力、条件生成对抗网络、风格化

专知

25+阅读 · 2018年5月28日

相关论文

Uncertainty Quantification via Spatial-Temporal Tweedie Model for Zero-inflated and Long-tail Travel Demand Prediction

Arxiv

0+阅读 · 2023年6月16日

Finite state verifiers with both private and public coins

Arxiv

0+阅读 · 2023年6月15日

On the Feasibility of Cross-Task Transfer with Model-Based Reinforcement Learning

Arxiv

0+阅读 · 2023年6月15日

Improving Training Stability for Multitask Ranking Models in Recommender Systems

Arxiv

0+阅读 · 2023年6月15日

Blind identification of Ambisonic reduced room impulse response

Arxiv

0+阅读 · 2023年6月15日

Some observations on the distribution of order statistics under simple-random-sampling without replacement

Arxiv

0+阅读 · 2023年6月14日

WizardCoder: Empowering Code Large Language Models with Evol-Instruct

Arxiv

0+阅读 · 2023年6月14日

Pretraining Language Models with Human Preferences

Arxiv

0+阅读 · 2023年6月14日

On the Robustness of Latent Diffusion Models

Arxiv

0+阅读 · 2023年6月14日

Improving the Generalizability of Trajectory Prediction Models with Frenet-Based Domain Normalization

Arxiv

0+阅读 · 2023年6月14日

相关基金

高维积分波动率矩阵的估计及其在资产投资中的应用

国家自然科学基金

0+阅读 · 2015年12月31日

具脉冲影响的Van der Pol方程的复杂动力学行为研究

国家自然科学基金

0+阅读 · 2015年12月31日

高维回归模型的预测稳定性研究

国家自然科学基金

3+阅读 · 2015年12月31日

对偶Auslander转置及其诱导模类的同调性质研究

国家自然科学基金

0+阅读 · 2015年12月31日

几种新型稀土金属间化合物中位错性质的高压效应

国家自然科学基金

0+阅读 · 2013年12月31日

Calderon问题和边界刚性问题

国家自然科学基金

0+阅读 · 2013年12月31日

面向多用户MIMO干扰的混合型无线Mesh网络自由度上限及可达方案研究

国家自然科学基金

0+阅读 · 2013年12月31日

冰云辐射性质参数化对东亚夏季风模拟的影响研究

国家自然科学基金

0+阅读 · 2012年12月31日

软件可靠性测试的数学模型研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于信道Time/Power度量指标的TOA测距误差模型及其应用研究

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员