多视图探测与纸板人造模型 (Multiview Detection with Cardboard Human Modeling)

Multiview detection uses multiple calibrated cameras with overlapping fields of views to locate occluded pedestrians. In this field, existing methods typically adopt a ``human modeling - aggregation'' strategy. To find robust pedestrian representations, some intuitively incorporate 2D perception results from each frame, while others use entire frame features projected to the ground plane. However, the former does not consider the human appearance and leads to many ambiguities, and the latter suffers from projection errors due to the lack of accurate height of the human torso and head. In this paper, we propose a new pedestrian representation scheme based on human point clouds modeling. Specifically, using ray tracing for holistic human depth estimation, we model pedestrians as upright, thin cardboard point clouds on the ground. Then, we aggregate the point clouds of the pedestrian cardboard across multiple views for a final decision. Compared with existing representations, the proposed method explicitly leverages human appearance and reduces projection errors significantly by relatively accurate height estimation. On four standard evaluation benchmarks, the proposed method achieves very competitive results. Our code and data will be released at https://github.com/ZichengDuan/MvCHM.

翻译：多视图探测使用多校准相机,其视野范围重叠,以定位隐蔽行人。在这一领域,现有方法通常采用“人造模型-聚合”战略。为了找到强健的行人代表,有些直觉地将每个框架的2D感知结果纳入每个框架,而另一些则使用向地面平面预测的整个框架特征。然而,前者不考虑人的外观,导致许多模糊不清,而后者则由于人类身体和头部的高度不准确,而存在预测错误。在本文中,我们提议以人点云模型为基础的新的行人代表制方案。具体地说,我们用光线追踪来进行整体人类深度估计。我们用光线追踪来模拟行人,在地面将行人作为直立、薄的纸板点云进行模拟。然后,我们将行人纸板的点云汇集到多个角度,以便作出最后决定。与现有的表达相比,拟议方法明确利用人表和预测错误,通过相对准确的高度估计大大降低。在四个标准评价基准上,拟议方法将产生非常具有竞争性的结果。我们的代码和数据将在https://github.com/ZchhengDuan/MHMH.

相关内容

Cardboard

关注 0

Google 在 2014 年 I/O 开发者大会上公布了这款用手机当做显示屏的简易 VR 设备，价格低廉且体验良好。你甚至可以根据 Google 给出的示意图自己用纸板 DIY 制作。在 2015 年的 Google I/O 上，Cardboard 宣布兼容 iOS 设备，并且达到一百万台的出货量。
官网： Google Cardboard

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

“CVPR 2021 接受论文列表 1663篇论文都在这了

专知会员服务

32+阅读 · 2021年6月12日

【视频描述综述论文】Video Description: A Survey of Methods, Datasets, and Evaluation Metrics

专知会员服务

65+阅读 · 2020年5月12日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日