为多标签分类提供视觉关注的无秩序 RNN, (Order-Free RNN with Visual Attention for Multi-Label Classification)

In this paper, we propose the joint learning attention and recurrent neural network (RNN) models for multi-label classification. While approaches based on the use of either model exist (e.g., for the task of image captioning), training such existing network architectures typically require pre-defined label sequences. For multi-label classification, it would be desirable to have a robust inference process, so that the prediction error would not propagate and thus affect the performance. Our proposed model uniquely integrates attention and Long Short Term Memory (LSTM) models, which not only addresses the above problem but also allows one to identify visual objects of interests with varying sizes without the prior knowledge of particular label ordering. More importantly, label co-occurrence information can be jointly exploited by our LSTM model. Finally, by advancing the technique of beam search, prediction of multiple labels can be efficiently achieved by our proposed network model.

翻译：在本文中,我们建议采用联合学习关注和经常性神经网络(RNN)模式进行多标签分类。虽然存在基于使用两种模式的方法(例如用于图像说明任务),但培训这类现有网络结构通常需要预先定义的标签序列。对于多标签分类,最好有一个强有力的推理过程,以便预测错误不会传播,从而影响性能。我们提议的模型将注意力和长短期内存(LSTM)模式(LSTM)模式(LSTM)模式(LSTM)模式(LSTM)(LSTM)模式(LSTM)(LSTM)模式(LSTM)(LSTM)(LSTM)模式(LSTM)(LSTM)模式(LSTM)(LSTM)(LSTM)模式(LSTM(LSTM)模式)(LSTM(LSTM)模式)(LSTM(LSTM(LT)模式)(LSTM(LSTM(LSTM)模式)(LSTM(LSTM(LT)模式)(LT)模式)(LSTM(LSTM(LT)模式)(LT)(LT)(LT)(LT)(LT)(LT)不仅解决上述问题,它不仅允许人们发现上述问题,它不仅可以识别问题,它不仅可以识别问题,还可以错错错错错错错错错错错,还),而且还。更重要的是别。更重要的是识别,还),而且可以有效。更重要的是,而且可以共同利用这些模式),也让错算算算算算算算算算算算算算。更重要的是,我们网络模式(LT(LT(LT)模式(LT)模式(LT)模式(LT)模式(LT)模式)模式(LT),也)。

相关内容

MoDELS

关注 43

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/