人群计票多关注网络 (Multi-Scale Attention Network for Crowd Counting)

In crowd counting datasets, people appear at different scales, depending on their distance from the camera. To address this issue, we propose a novel multi-branch scale-aware attention network that exploits the hierarchical structure of convolutional neural networks and generates, in a single forward pass, multi-scale density predictions from different layers of the architecture. To aggregate these maps into our final prediction, we present a new soft attention mechanism that learns a set of gating masks. Furthermore, we introduce a scale-aware loss function to regularize the training of different branches and guide them to specialize on a particular scale. As this new training requires annotations for the size of each head, we also propose a simple, yet effective technique to estimate them automatically. Finally, we present an ablation study on each of these components and compare our approach against the literature on 4 crowd counting datasets: UCF-QNRF, ShanghaiTech A & B and UCF_CC_50. Our approach achieves state-of-the-art on all them with a remarkable improvement on UCF-QNRF (+25% reduction in error).

翻译：在人群计数数据集中,人们会根据与相机的距离而出现在不同的尺度上。为了解决这一问题,我们提议建立一个新型的多部门规模关注网络,利用进化神经网络的等级结构,并在一个前传中从结构的不同层面生成多尺度的密度预测。为了将这些地图汇总到我们的最后预测中,我们提出了一个新的软关注机制,以学习一套加盖面罩。此外,我们引入了一个规模感知损失功能,以使不同分支的培训正规化,并指导他们进行特定规模的专门化。由于这一新培训需要每个头部的大小说明,我们还提出了一个简单而有效的自动估算技术。最后,我们提出了对所有这些组成部分的简单而有效的计算方法。我们提出了对每个组成部分的模拟研究,并将我们的方法与关于4个人群计数据集的文献进行比较:UCF-QNRF、上海科技A & B和UCF_CC_50。我们的方法在所有这些单元上都取得了显著的进步,在UCFC-QNRF(减少25%错误)。

相关内容

注意力机制

关注 120

Attention机制最早是在视觉图像领域提出来的，但是真正火起来应该算是google mind团队的这篇论文《Recurrent Models of Visual Attention》[14]，他们在RNN模型上使用了attention机制来进行图像分类。随后，Bahdanau等人在论文《Neural Machine Translation by Jointly Learning to Align and Translate》 [1]中，使用类似attention的机制在机器翻译任务上将翻译和对齐同时进行，他们的工作算是是第一个提出attention机制应用到NLP领域中。接着类似的基于attention机制的RNN模型扩展开始应用到各种NLP任务中。最近，如何在CNN中使用attention机制也成为了大家的研究热点。下图表示了attention研究进展的大概趋势。

深度学习搜索，Exploring Deep Learning for Search

专知会员服务

61+阅读 · 2020年5月9日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning

专知会员服务

159+阅读 · 2020年2月29日