基于记忆的吉特:在记忆中具有多样性的长成数据中改进视觉识别 (Memory-based Jitter: Improving Visual Recognition on Long-tailed Data with Diversity In Memory)

This paper considers deep visual recognition on long-tailed data. To be general, we consider two applied scenarios, \ie, deep classification and deep metric learning. Under the long-tailed data distribution, the majority classes (\ie, tail classes) only occupy relatively few samples and are prone to lack of within-class diversity. A radical solution is to augment the tail classes with higher diversity. To this end, we introduce a simple and reliable method named Memory-based Jitter (MBJ). We observe that during training, the deep model constantly changes its parameters after every iteration, yielding the phenomenon of \emph{weight jitters}. Consequentially, given a same image as the input, two historical editions of the model generate two different features in the deeply-embedded space, resulting in \emph{feature jitters}. Using a memory bank, we collect these (model or feature) jitters across multiple training iterations and get the so-called Memory-based Jitter. The accumulated jitters enhance the within-class diversity for the tail classes and consequentially improves long-tailed visual recognition. With slight modifications, MBJ is applicable for two fundamental visual recognition tasks, \emph{i.e.}, deep image classification and deep metric learning (on long-tailed data). Extensive experiments on five long-tailed classification benchmarks and two deep metric learning benchmarks demonstrate significant improvement. Moreover, the achieved performance are on par with the state of the art on both tasks.

翻译：本文考虑对长尾数据的深度视觉识别。一般来说, 我们考虑的是两种应用情景, \ \ \, 深分类和深度量度学习。在长尾数据分布中, 多数类( \ i, 尾类) 只占相对较少的样本, 并且容易缺少类内多样性。一个根本的解决方案是增加尾类, 多样性更大。为此, 我们引入了一个简单而可靠的方法, 名为“ 内存吉他 ” (MBJ) 。我们观察到, 在培训期间, 深型模型在每次迭代后不断改变其参数, 产生 \ emph{ 重量性急症现象。因此, 以与输入相同的图像, 多数类( \ i, 尾类, 尾类( 尾类, 尾类, 尾类, 尾类, 尾类, 尾类, 尾类) 产生类似现象。因此, 两种历史版本产生两种不同的特征特征特征特征, 。通过记忆分类, 进行微小的修改, 和深层的图像分类。

相关内容

度量学习

关注 3372

度量学习的目的为了衡量样本之间的相近程度，而这也正是模式识别的核心问题之一。大量的机器学习方法，比如K近邻、支持向量机、径向基函数网络等分类方法以及K-means聚类方法，还有一些基于图的方法，其性能好坏都主要有样本之间的相似度量方法的选择决定。度量学习通常的目标是使同类样本之间的距离尽可能缩小，不同类样本之间的距离尽可能放大。

【ICML2020-伯克利-马毅老师组】深度等距学习的视觉识别，Deep Isometric Learning for Visual Recognition

专知会员服务

25+阅读 · 2020年7月1日

【上海交大】可解释CNN的对象分类，Interpretable CNNs for Object Classification

专知会员服务

54+阅读 · 2020年3月14日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日