从自然平衡的普塞多-标签中学习 (Debiased Learning from Naturally Imbalanced Pseudo-Labels) - 专知论文

会员服务 ·

0

伪标记 · 学成 · 未标记 · Extensibility · 有偏 ·

2022 年 4 月 21 日

Debiased Learning from Naturally Imbalanced Pseudo-Labels

翻译：从自然平衡的普塞多-标签中学习

Xudong Wang,Zhirong Wu,Long Lian,Stella X. Yu

from arxiv, Accepted by CVPR 2022

Pseudo-labels are confident predictions made on unlabeled target data by a classifier trained on labeled source data. They are widely used for adapting a model to unlabeled data, e.g., in a semi-supervised learning setting. Our key insight is that pseudo-labels are naturally imbalanced due to intrinsic data similarity, even when a model is trained on balanced source data and evaluated on balanced target data. If we address this previously unknown imbalanced classification problem arising from pseudo-labels instead of ground-truth training labels, we could remove model biases towards false majorities created by pseudo-labels. We propose a novel and effective debiased learning method with pseudo-labels, based on counterfactual reasoning and adaptive margins: The former removes the classifier response bias, whereas the latter adjusts the margin of each class according to the imbalance of pseudo-labels. Validated by extensive experimentation, our simple debiased learning delivers significant accuracy gains over the state-of-the-art on ImageNet-1K: 26% for semi-supervised learning with 0.2% annotations and 9% for zero-shot learning. Our code is available at: https://github.com/frank-xwang/debiased-pseudo-labeling.

翻译：在标签源数据方面受过训练的分类人员对未贴标签目标数据所作的预测是用标签源数据培训的分类人员对未贴标签的目标数据所作的自信预测。这些预测被广泛用于将模型与未贴标签数据相适应,例如半监督的学习环境。我们的关键洞察力是,伪标签由于内在数据相似性而自然地不平衡,即使一个模型是用平衡源数据培训的,并且根据平衡目标数据进行评估。如果我们解决以前未知的假标签而非地面真相培训标签产生的不平衡分类问题,我们就可以消除对假标签产生的虚假多数的模型偏见。我们基于反事实推理和适应性边际,提出一种创新和有效的不偏向学习方法。前者消除了分类者反应偏差,而后者根据假标签的不平衡性调整了每个班级的差。经过广泛的实验验证,我们简单的不偏差学习在图像Net-1K的状态上带来显著的准确性收益:以0.2%的描述值和9%的零光学准则用于半监督的半监督学习。

0

相关内容

伪标记

【干货书】机器学习设计模式，408页pdf，Machine Learning Design Patterns

【干货书】机器学习设计模式，408页pdf，Machine Learning Design Patterns

专知会员服务

138+阅读 · 2022年2月6日

【MIT】反偏差对比学习，Debiased Contrastive Learning

【MIT】反偏差对比学习，Debiased Contrastive Learning

专知会员服务

91+阅读 · 2020年7月4日

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

专知会员服务

115+阅读 · 2020年4月5日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

【ICIG2021】Latest News & Announcements of the Plenary Talk2

【ICIG2021】Latest News & Announcements of the Plenary Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年11月2日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

内质网Ca2+感受器STIM1调控糖尿病冠状动脉平滑肌细胞表型转化的机制

国家自然科学基金

0+阅读 · 2014年12月31日

长链非编码RNA AC074286.1在食管鳞癌中的生物学功能及其表观遗传机制

国家自然科学基金

0+阅读 · 2014年12月31日

知识密集型服务外包中的知识共享激励与知识资产争端协调机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

剪接蛋白SRSF10在肠癌细胞增殖和凋亡中的作用

国家自然科学基金

0+阅读 · 2013年12月31日

大肠杆菌YfiF蛋白对DNA复制起始的调控机制

国家自然科学基金

0+阅读 · 2012年12月31日

金黄色葡萄球菌AraC家族转录调节蛋白Rsp对细菌毒力的调节机制

国家自然科学基金

0+阅读 · 2012年12月31日

E3泛素连接酶Synoviolin对胰岛β细胞功能的影响

国家自然科学基金

0+阅读 · 2012年12月31日

NS5ATP9基因相关miRNAs在食管癌中的鉴定和临床研究

国家自然科学基金

0+阅读 · 2012年12月31日

福氏志贺氏菌HtrA蛋白功能研究

国家自然科学基金

0+阅读 · 2011年12月31日

民猪抗寒基因的筛选与鉴定

国家自然科学基金

0+阅读 · 2010年12月31日

Does Self-supervised Learning Really Improve Reinforcement Learning from Pixels?

Does Self-supervised Learning Really Improve Reinforcement Learning from Pixels?

Arxiv

0+阅读 · 2022年6月10日

Balanced Product of Experts for Long-Tailed Recognition

Arxiv

0+阅读 · 2022年6月10日

Balanced background and explanation data are needed in explaining deep learning models with SHAP: An empirical study on clinical decision making

Arxiv

0+阅读 · 2022年6月8日

Probabilistically Robust Learning: Balancing Average- and Worst-case Performance

Arxiv

0+阅读 · 2022年6月7日

Debiased Self-Training for Semi-Supervised Learning

Arxiv

0+阅读 · 2022年6月7日

Deep Learning Techniques for Visual Counting

Arxiv

0+阅读 · 2022年6月7日

OntoZSL: Ontology-enhanced Zero-shot Learning

Arxiv

17+阅读 · 2021年2月15日

Contrastive learning of global and local features for medical image segmentation with limited annotations

Arxiv

19+阅读 · 2020年6月18日

FocalMix: Semi-Supervised Learning for 3D Medical Image Detection

FocalMix: Semi-Supervised Learning for 3D Medical Image Detection

Arxiv

10+阅读 · 2020年3月20日

A Simple Framework for Contrastive Learning of Visual Representations

Arxiv

21+阅读 · 2020年2月13日

VIP会员

文章信息

相关主题

相关VIP内容

【干货书】机器学习设计模式，408页pdf，Machine Learning Design Patterns

【干货书】机器学习设计模式，408页pdf，Machine Learning Design Patterns

专知会员服务

138+阅读 · 2022年2月6日

【MIT】反偏差对比学习，Debiased Contrastive Learning

【MIT】反偏差对比学习，Debiased Contrastive Learning

专知会员服务

91+阅读 · 2020年7月4日

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

专知会员服务

115+阅读 · 2020年4月5日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

军事战术边缘计算的重要性

《欧洲天空盾牌倡议：应对无人机饱和攻击与高超音速导弹的多层防空演进与挑战》报告

《美军使用大语言模型技术生成领域特定文档》2025最新379页

《代理生成式人工智能与国家安全：提升竞争力的政策建议》

相关资讯

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

【ICIG2021】Latest News & Announcements of the Plenary Talk2

【ICIG2021】Latest News & Announcements of the Plenary Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年11月2日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

Does Self-supervised Learning Really Improve Reinforcement Learning from Pixels?

Does Self-supervised Learning Really Improve Reinforcement Learning from Pixels?

Arxiv

0+阅读 · 2022年6月10日

Balanced Product of Experts for Long-Tailed Recognition

Arxiv

0+阅读 · 2022年6月10日

Balanced background and explanation data are needed in explaining deep learning models with SHAP: An empirical study on clinical decision making

Arxiv

0+阅读 · 2022年6月8日

Probabilistically Robust Learning: Balancing Average- and Worst-case Performance

Arxiv

0+阅读 · 2022年6月7日

Debiased Self-Training for Semi-Supervised Learning

Arxiv

0+阅读 · 2022年6月7日

Deep Learning Techniques for Visual Counting

Arxiv

0+阅读 · 2022年6月7日

OntoZSL: Ontology-enhanced Zero-shot Learning

Arxiv

17+阅读 · 2021年2月15日

Contrastive learning of global and local features for medical image segmentation with limited annotations

Arxiv

19+阅读 · 2020年6月18日

FocalMix: Semi-Supervised Learning for 3D Medical Image Detection

FocalMix: Semi-Supervised Learning for 3D Medical Image Detection

Arxiv

10+阅读 · 2020年3月20日

A Simple Framework for Contrastive Learning of Visual Representations

Arxiv

21+阅读 · 2020年2月13日

相关基金

内质网Ca2+感受器STIM1调控糖尿病冠状动脉平滑肌细胞表型转化的机制

国家自然科学基金

0+阅读 · 2014年12月31日

长链非编码RNA AC074286.1在食管鳞癌中的生物学功能及其表观遗传机制

国家自然科学基金

0+阅读 · 2014年12月31日

知识密集型服务外包中的知识共享激励与知识资产争端协调机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

剪接蛋白SRSF10在肠癌细胞增殖和凋亡中的作用

国家自然科学基金

0+阅读 · 2013年12月31日

大肠杆菌YfiF蛋白对DNA复制起始的调控机制

国家自然科学基金

0+阅读 · 2012年12月31日

金黄色葡萄球菌AraC家族转录调节蛋白Rsp对细菌毒力的调节机制

国家自然科学基金

0+阅读 · 2012年12月31日

E3泛素连接酶Synoviolin对胰岛β细胞功能的影响

国家自然科学基金

0+阅读 · 2012年12月31日

NS5ATP9基因相关miRNAs在食管癌中的鉴定和临床研究

国家自然科学基金

0+阅读 · 2012年12月31日

福氏志贺氏菌HtrA蛋白功能研究

国家自然科学基金

0+阅读 · 2011年12月31日

民猪抗寒基因的筛选与鉴定

国家自然科学基金

0+阅读 · 2010年12月31日

微信扫码咨询专知VIP会员