ENLODI: 积极一致培训综合逻辑差异阻断 (ELODI: Ensemble Logit Difference Inhibition for Positive-Congruent Training) - 专知论文

会员服务 ·

0

可约的 · 对数几率 · MoDELS · 参考模型 · 蒸馏 ·

2022 年 5 月 13 日

ELODI: Ensemble Logit Difference Inhibition for Positive-Congruent Training

翻译：ENLODI: 积极一致培训综合逻辑差异阻断

Yue Zhao,Yantao Shen,Yuanjun Xiong,Shuo Yang,Wei Xia,Zhuowen Tu,Bernt Schiele,Stefano Soatto

from arxiv, Tech report

Negative flips are errors introduced in a classification system when a legacy model is replaced with a new one. Existing methods to reduce the negative flip rate (NFR) either do so at the expense of overall accuracy using model distillation, or use ensembles, which multiply inference cost prohibitively. We present a method to train a classification system that achieves paragon performance in both error rate and NFR, at the inference cost of a single model. Our method introduces a generalized distillation objective, Logit Difference Inhibition (LDI), that penalizes changes in the logits between the new and old model, without forcing them to coincide as in ordinary distillation. LDI affords the model flexibility to reduce error rate along with NFR. The method uses a homogeneous ensemble as the reference model for LDI, hence the name Ensemble LDI, or ELODI. The reference model can then be substituted with a single model at inference time. The method leverages the observation that negative flips are typically not close to the decision boundary, but often exhibit large deviations in the distance among their logits, which are reduced by ELODI.

翻译：负翻转是指在以新的模式取代遗留模式时,在分类系统中引入错误。现有的降低负翻转率(NFR)的方法要么以使用模型蒸馏的总体准确性为代价,要么以使用模型蒸馏的总体准确性为代价来降低负翻转率(NFR),或者使用模型组合,这种组合将极高的推论成本乘以极高的推论成本。我们提出了一个方法,以单一模型的推论成本来培训一个在错误率和NFR中都达到参数性能的分类系统。我们的方法引入了一个普遍的蒸馏目标,即Logit差异触发(LDI),它惩罚新式和旧模型之间对日志的修改,而不会迫使它们与普通的蒸馏同步。LDI提供了模型的灵活性,以降低误差率和NFR。该方法使用同质共性共性共性词作为LDI的参考模型,因此名称为Esemble LDI,或ELODI。然后在推论时可以用一个单一模型取代。这种方法利用这一观察,即负翻通常不接近决定边界,但往往显示其日志之间的偏差很大。

0

相关内容

可约的

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

专知会员服务

44+阅读 · 2020年11月2日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

IEEE ICKG 2022: Call for Papers

IEEE ICKG 2022: Call for Papers

机器学习与推荐算法

3+阅读 · 2022年3月30日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

强化学习三篇论文避免遗忘等

强化学习三篇论文避免遗忘等

CreateAMind

20+阅读 · 2019年5月24日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

专知

20+阅读 · 2018年6月29日

基于自主学习的Ad hoc Agent序贯决策研究

国家自然科学基金

45+阅读 · 2015年12月31日

PARP1通路抑制分子RNF146调控星形胶质细胞凋亡在AD中的作用研究

国家自然科学基金

0+阅读 · 2014年12月31日

Intraflagellar Transport运输纤毛蛋白的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

米诺环素对脑出血后神经细胞凋亡和坏死的影响

国家自然科学基金

0+阅读 · 2012年12月31日

Wnt/β-catenin和 Hedgehog信号通路互作在骨关节中的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

MAPK通路在气道高反应性发生中对G-蛋白偶联受体的调控机制

国家自然科学基金

0+阅读 · 2012年12月31日

MeCP2-PTEN调控神经干细胞增殖分化影响孤独症发生的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

miR146a抑制Smad4对骨髓间充质干细胞成骨分化调控的研究

国家自然科学基金

0+阅读 · 2009年12月31日

TR3相互作用新蛋白机理研究

国家自然科学基金

1+阅读 · 2008年12月31日

高性能壳聚糖纳米微囊介导双基因共转染ADSCs的成骨研究

国家自然科学基金

0+阅读 · 2008年12月31日

Identifiability of Label Noise Transition Matrix

Identifiability of Label Noise Transition Matrix

Arxiv

0+阅读 · 2022年7月4日

Necessary and Sufficient Condition for the Existence of Zero-Determinant Strategies in Repeated Games

Arxiv

0+阅读 · 2022年7月4日

Eliciting and Learning with Soft Labels from Every Annotator

Arxiv

0+阅读 · 2022年7月2日

Action-modulated midbrain dopamine activity arises from distributed control policies

Arxiv

0+阅读 · 2022年7月1日

An Enumeration Algorithm for Binary Coprime Polynomials with Nonzero Constant Term

Arxiv

0+阅读 · 2022年7月1日

Improved Generalization Bounds for Adversarially Robust Learning

Arxiv

0+阅读 · 2022年7月1日

Self-Training of Handwritten Word Recognition for Synthetic-to-Real Adaptation

Self-Training of Handwritten Word Recognition for Synthetic-to-Real Adaptation

Arxiv

0+阅读 · 2022年6月30日

Consensus Function from an $L_p^q-$norm Regularization Term for its Use as Adaptive Activation Functions in Neural Networks

Arxiv

0+阅读 · 2022年6月30日

Instance-level loss based multiple-instance learning framework for acoustic scene classification

Arxiv

0+阅读 · 2022年6月30日

Improving Event Causality Identification via Self-Supervised Representation Learning on External Causal Statement

Arxiv

15+阅读 · 2021年6月3日

VIP会员

文章信息

相关主题

相关VIP内容

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

专知会员服务

44+阅读 · 2020年11月2日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

NeurIPS 2025 | 自动化所新作速览（一）

大型语言模型（LLM）赋能的知识图谱构建：综述

NeurIPS 2025 | 自动化所新作速览（二）

领域特定文本分类中的预训练语言模型新进展：系统综述

相关资讯

IEEE ICKG 2022: Call for Papers

IEEE ICKG 2022: Call for Papers

机器学习与推荐算法

3+阅读 · 2022年3月30日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

强化学习三篇论文避免遗忘等

强化学习三篇论文避免遗忘等

CreateAMind

20+阅读 · 2019年5月24日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

专知

20+阅读 · 2018年6月29日

相关论文

Identifiability of Label Noise Transition Matrix

Identifiability of Label Noise Transition Matrix

Arxiv

0+阅读 · 2022年7月4日

Necessary and Sufficient Condition for the Existence of Zero-Determinant Strategies in Repeated Games

Arxiv

0+阅读 · 2022年7月4日

Eliciting and Learning with Soft Labels from Every Annotator

Arxiv

0+阅读 · 2022年7月2日

Action-modulated midbrain dopamine activity arises from distributed control policies

Arxiv

0+阅读 · 2022年7月1日

An Enumeration Algorithm for Binary Coprime Polynomials with Nonzero Constant Term

Arxiv

0+阅读 · 2022年7月1日

Improved Generalization Bounds for Adversarially Robust Learning

Arxiv

0+阅读 · 2022年7月1日

Self-Training of Handwritten Word Recognition for Synthetic-to-Real Adaptation

Self-Training of Handwritten Word Recognition for Synthetic-to-Real Adaptation

Arxiv

0+阅读 · 2022年6月30日

Consensus Function from an $L_p^q-$norm Regularization Term for its Use as Adaptive Activation Functions in Neural Networks

Arxiv

0+阅读 · 2022年6月30日

Instance-level loss based multiple-instance learning framework for acoustic scene classification

Arxiv

0+阅读 · 2022年6月30日

Improving Event Causality Identification via Self-Supervised Representation Learning on External Causal Statement

Arxiv

15+阅读 · 2021年6月3日

相关基金

基于自主学习的Ad hoc Agent序贯决策研究

国家自然科学基金

45+阅读 · 2015年12月31日

PARP1通路抑制分子RNF146调控星形胶质细胞凋亡在AD中的作用研究

国家自然科学基金

0+阅读 · 2014年12月31日

Intraflagellar Transport运输纤毛蛋白的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

米诺环素对脑出血后神经细胞凋亡和坏死的影响

国家自然科学基金

0+阅读 · 2012年12月31日

Wnt/β-catenin和 Hedgehog信号通路互作在骨关节中的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

MAPK通路在气道高反应性发生中对G-蛋白偶联受体的调控机制

国家自然科学基金

0+阅读 · 2012年12月31日

MeCP2-PTEN调控神经干细胞增殖分化影响孤独症发生的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

miR146a抑制Smad4对骨髓间充质干细胞成骨分化调控的研究

国家自然科学基金

0+阅读 · 2009年12月31日

TR3相互作用新蛋白机理研究

国家自然科学基金

1+阅读 · 2008年12月31日

高性能壳聚糖纳米微囊介导双基因共转染ADSCs的成骨研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员