Sprassar MDOD: 培训端对端多目标检测器,无双边匹配 (Sparse MDOD: Training End-to-End Multi-Object Detector without Bipartite Matching) - 专知论文

会员服务 ·

0

端到端 · 稀疏 · 估计/估计量 · 对数似然 · 端到端学习 ·

2022 年 9 月 12 日

Sparse MDOD: Training End-to-End Multi-Object Detector without Bipartite Matching

翻译：Sprassar MDOD: 培训端对端多目标检测器,无双边匹配

Jaeyoung Yoo,Hojun Lee,Seunghyeon Seo,Inseop Chung,Nojun Kwak

from arxiv, 8 figures

Recent end-to-end multi-object detectors simplify the inference pipeline by removing the hand-crafted process such as the duplicate bounding box removal using non-maximum suppression (NMS). However, in the training, they require bipartite matching to calculate the loss from the output of the detector. Contrary to the directivity, which is at the heart of end-to-end learning, the bipartite matching makes the training of the end-to-end detector complex, heuristic, and reliant. In this paper, we propose a method to train an end-to-end multi-object detector without bipartite matching. To this end, we approach end-to-end multi-object detection as a density estimation problem using a mixture model. Our proposed detector, called Sparse Mixture Density Object Detector (Sparse MDOD), estimates the distribution of bounding boxes using a mixture model. Sparse MDOD is trained by minimizing the negative log-likelihood and our proposed regularization term, maximum component maximization (MCM) loss that prevents duplicated predictions. During training, no additional procedure such as bipartite matching is needed, and the loss is directly computed from the network outputs. Moreover, our Sparse MDOD outperforms the existing detectors on MS-COCO, a renowned multi-object detection benchmark.

翻译：最近的端到端多球探测器通过去除手动工艺,例如使用非最大抑制(NMS)来重复捆绑盒清除器等,简化导火线。但是,在培训中,它们需要双向匹配来计算探测器输出的损失。与直接性相反,这是端到端学习的核心, 双向匹配使得培训端到端检测器复杂、超度和依赖性。在本文中, 我们提议了一种方法, 用于培训一个端到端多球探测器, 没有双向匹配。为此, 我们使用混合模型, 将端到端多球探测器作为密度估计问题。我们提议的探测器, 叫做“ 分解分解分解分立器” (Sparse MDDD), 用混合模型来估计捆绑箱的分布情况。微调MDDDDD经过培训, 最大限度地减少负日志和我们提议的正规化期, 最大组成部分(MCMM) 损失, 防止重复的重复预测。在当前的IMDMD IM 测试中,, 没有额外的程序, 直接匹配现有的双级测试。

0

相关内容

端到端

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

【2022新书】高效深度学习，Efficient Deep Learning Book

【2022新书】高效深度学习，Efficient Deep Learning Book

专知会员服务

125+阅读 · 2022年4月21日

深度学习优化算法，73页ppt，Optimization Algorithms on Deep Learning

深度学习优化算法，73页ppt，Optimization Algorithms on Deep Learning

专知会员服务

135+阅读 · 2021年6月16日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

吴恩达推荐！22页「AI职业生涯发展正规之道」秘籍，AI Career Pathways: Put Yourself on the Right Track，让你不被AI失业与共建一个Work的AI团队

吴恩达推荐！22页「AI职业生涯发展正规之道」秘籍，AI Career Pathways: Put Yourself on the Right Track，让你不被AI失业与共建一个Work的AI团队

专知会员服务

53+阅读 · 2020年1月9日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【泡泡汇总】CVPR2019 SLAM Paperlist

【泡泡汇总】CVPR2019 SLAM Paperlist

泡泡机器人SLAM

14+阅读 · 2019年6月12日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【论文推荐】最新十篇目标跟踪相关论文—多帧光流跟踪、动态图学习、MV-YOLO、姿态估计、深度核相关滤波、Benchmark

【论文推荐】最新十篇目标跟踪相关论文—多帧光流跟踪、动态图学习、MV-YOLO、姿态估计、深度核相关滤波、Benchmark

专知

13+阅读 · 2018年5月26日

Capsule Networks解析

Capsule Networks解析

机器学习研究会

11+阅读 · 2017年11月12日

【推荐】YOLO实时目标检测(6fps)

【推荐】YOLO实时目标检测(6fps)

机器学习研究会

20+阅读 · 2017年11月5日

【推荐】深度学习目标检测概览

【推荐】深度学习目标检测概览

机器学习研究会

10+阅读 · 2017年9月1日

【推荐】图像分类必读开创性论文汇总

【推荐】图像分类必读开创性论文汇总

机器学习研究会

14+阅读 · 2017年8月15日

具有大线性复杂度的最优部分汉明相关跳频序列集的构造研究

国家自然科学基金

0+阅读 · 2015年12月31日

非凸稀疏正则化模型与算法的研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于分布式天线的宽带相位测距机制与方法研究

国家自然科学基金

0+阅读 · 2014年12月31日

基于智能在线虚拟参考反馈整定的控制方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

Schrodinger-Poisson方程的若干问题研究

国家自然科学基金

1+阅读 · 2012年12月31日

microRNA对NFATc1/RANKL骨免疫信号通路的调控机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于MAP的低复杂度LDPC译码算法理论和方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于超高分辨率视频的HEVC低复杂度模型和方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

非线性不适定问题的非光滑解的若干数值方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

图像匹配的新方法研究：李群约束的优化算法

国家自然科学基金

0+阅读 · 2009年12月31日

Matching Map Recovery with an Unknown Number of Outliers

Arxiv

0+阅读 · 2022年10月24日

Iterative Patch Selection for High-Resolution Image Recognition

Arxiv

0+阅读 · 2022年10月24日

Accelerating SGD for Highly Ill-Conditioned Huge-Scale Online Matrix Completion

Arxiv

0+阅读 · 2022年10月23日

H4VDM: H.264 Video Device Matching

Arxiv

0+阅读 · 2022年10月20日

Improving Data Quality with Training Dynamics of Gradient Boosting Decision Trees

Improving Data Quality with Training Dynamics of Gradient Boosting Decision Trees

Arxiv

0+阅读 · 2022年10月20日

Pruning by Active Attention Manipulation

Arxiv

0+阅读 · 2022年10月20日

Superiorized Adaptive Projected Subgradient Method with Application to MIMO Detection

Arxiv

0+阅读 · 2022年10月20日

Reverse Attention for Salient Object Detection

Arxiv

11+阅读 · 2019年4月15日

Prime Sample Attention in Object Detection

Arxiv

13+阅读 · 2019年4月9日

Convolutional Neural Networks for Aerial Multi-Label Pedestrian Detection

Convolutional Neural Networks for Aerial Multi-Label Pedestrian Detection

Arxiv

11+阅读 · 2018年7月16日

VIP会员

文章信息

相关主题

估计/估计量

端到端学习

相关VIP内容

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

【2022新书】高效深度学习，Efficient Deep Learning Book

【2022新书】高效深度学习，Efficient Deep Learning Book

专知会员服务

125+阅读 · 2022年4月21日

深度学习优化算法，73页ppt，Optimization Algorithms on Deep Learning

深度学习优化算法，73页ppt，Optimization Algorithms on Deep Learning

专知会员服务

135+阅读 · 2021年6月16日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

吴恩达推荐！22页「AI职业生涯发展正规之道」秘籍，AI Career Pathways: Put Yourself on the Right Track，让你不被AI失业与共建一个Work的AI团队

吴恩达推荐！22页「AI职业生涯发展正规之道」秘籍，AI Career Pathways: Put Yourself on the Right Track，让你不被AI失业与共建一个Work的AI团队

专知会员服务

53+阅读 · 2020年1月9日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《关于俄乌战争的系列文章》2025最新70页

《军事行动中的人机AI编队本体模型》

更智能的人工智能实现更快速的电磁辐射控制（EMCON）

《俄罗斯常规军队能力现状及重建》2025最新124页

相关资讯

【泡泡汇总】CVPR2019 SLAM Paperlist

【泡泡汇总】CVPR2019 SLAM Paperlist

泡泡机器人SLAM

14+阅读 · 2019年6月12日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【论文推荐】最新十篇目标跟踪相关论文—多帧光流跟踪、动态图学习、MV-YOLO、姿态估计、深度核相关滤波、Benchmark

【论文推荐】最新十篇目标跟踪相关论文—多帧光流跟踪、动态图学习、MV-YOLO、姿态估计、深度核相关滤波、Benchmark

专知

13+阅读 · 2018年5月26日

Capsule Networks解析

Capsule Networks解析

机器学习研究会

11+阅读 · 2017年11月12日

【推荐】YOLO实时目标检测(6fps)

【推荐】YOLO实时目标检测(6fps)

机器学习研究会

20+阅读 · 2017年11月5日

【推荐】深度学习目标检测概览

【推荐】深度学习目标检测概览

机器学习研究会

10+阅读 · 2017年9月1日

【推荐】图像分类必读开创性论文汇总

【推荐】图像分类必读开创性论文汇总

机器学习研究会

14+阅读 · 2017年8月15日

相关论文

Matching Map Recovery with an Unknown Number of Outliers

Arxiv

0+阅读 · 2022年10月24日

Iterative Patch Selection for High-Resolution Image Recognition

Arxiv

0+阅读 · 2022年10月24日

Accelerating SGD for Highly Ill-Conditioned Huge-Scale Online Matrix Completion

Arxiv

0+阅读 · 2022年10月23日

H4VDM: H.264 Video Device Matching

Arxiv

0+阅读 · 2022年10月20日

Improving Data Quality with Training Dynamics of Gradient Boosting Decision Trees

Improving Data Quality with Training Dynamics of Gradient Boosting Decision Trees

Arxiv

0+阅读 · 2022年10月20日

Pruning by Active Attention Manipulation

Arxiv

0+阅读 · 2022年10月20日

Superiorized Adaptive Projected Subgradient Method with Application to MIMO Detection

Arxiv

0+阅读 · 2022年10月20日

Reverse Attention for Salient Object Detection

Arxiv

11+阅读 · 2019年4月15日

Prime Sample Attention in Object Detection

Arxiv

13+阅读 · 2019年4月9日

Convolutional Neural Networks for Aerial Multi-Label Pedestrian Detection

Convolutional Neural Networks for Aerial Multi-Label Pedestrian Detection

Arxiv

11+阅读 · 2018年7月16日

相关基金

具有大线性复杂度的最优部分汉明相关跳频序列集的构造研究

国家自然科学基金

0+阅读 · 2015年12月31日

非凸稀疏正则化模型与算法的研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于分布式天线的宽带相位测距机制与方法研究

国家自然科学基金

0+阅读 · 2014年12月31日

基于智能在线虚拟参考反馈整定的控制方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

Schrodinger-Poisson方程的若干问题研究

国家自然科学基金

1+阅读 · 2012年12月31日

microRNA对NFATc1/RANKL骨免疫信号通路的调控机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于MAP的低复杂度LDPC译码算法理论和方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于超高分辨率视频的HEVC低复杂度模型和方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

非线性不适定问题的非光滑解的若干数值方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

图像匹配的新方法研究：李群约束的优化算法

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员