URCDC-Deph: 用于单层深度估算的含剪动平流的不确定性校正交叉蒸馏 (URCDC-Depth: Uncertainty Rectified Cross-Distillation with CutFlip for Monocular Depth Estimation) - 专知论文

会员服务 ·

0

估计/估计量 · Branch · 变换 · CNN · CLUES ·

2023 年 2 月 16 日

URCDC-Depth: Uncertainty Rectified Cross-Distillation with CutFlip for Monocular Depth Estimation

翻译：URCDC-Deph: 用于单层深度估算的含剪动平流的不确定性校正交叉蒸馏

Shuwei Shao,Zhongcai Pei,Weihai Chen,Ran Li,Zhong Liu,Zhengguo Li

from arxiv, 9 pages

This work aims to estimate a high-quality depth map from a single RGB image. Due to the lack of depth clues, making full use of the long-range correlation and the local information is critical for accurate depth estimation. Towards this end, we introduce an uncertainty rectified cross-distillation between Transformer and convolutional neural network (CNN) to learn a unified depth estimator. Specifically, we use the depth estimates derived from the Transformer branch and the CNN branch as pseudo labels to teach each other. Meanwhile, we model the pixel-wise depth uncertainty to rectify the loss weights of noisy depth labels. To avoid the large performance gap induced by the strong Transformer branch deteriorating the cross-distillation, we transfer the feature maps from Transformer to CNN and design coupling units to assist the weak CNN branch to utilize the transferred features. Furthermore, we propose a surprisingly simple yet highly effective data augmentation technique CutFlip, which enforces the model to exploit more valuable clues apart from the clue of vertical image position for depth estimation. Extensive experiments indicate that our model, termed~\textbf{URCDC-Depth}, exceeds previous state-of-the-art methods on the KITTI and NYU-Depth-v2 datasets, even with no additional computational burden at inference time. The source code is publicly available at \url{https://github.com/ShuweiShao/URCDC-Depth}.

翻译：这项工作旨在从一个 RGB 图像中估计高质量的深度地图。由于缺乏深度线索, 充分利用长距离关联和本地信息对于准确的深度估算至关重要。为此, 我们引入了一种不确定性, 纠正变异器和进化神经网络( CNN) 之间的交叉蒸馏。具体地说, 我们使用变异器分支和CNN分支的深度估算值作为假标签来相互教学。同时, 我们模拟了像素明智的深度不确定性, 以纠正噪音深度标签的损失重量。为了避免强大的变异器分支导致的大型性能差距, 我们从变异器到CNN, 并设计连接器来帮助弱的CNN分支使用所传输的特征。此外, 我们提议了一个令人惊讶而非常有效的数据增强技术CutFlip, 以模型为工具来利用比垂直图像位置更有价值的线索来进行深度估算。广泛的实验显示, 我们的模型, 被命名为“ Textbrufur CD- NYC- Deptread ” 和“ KDI- discodeal ” 方法超越了以前的状态。。我们在时间/ depal- discodeal- dal- disal_ dal_ dal_ drogard_ drodu_ drobral_ drogal_ disl_ disl_ drodd_ dislddddalddddd_ dism_ drodal_ drodal_ dal_ drodal_ disgal_ disal_dal_ drodddddddd_ drodddddddddddddddddddddddddddddd_ drodal_ droddal_ drodddddddddddddddddddddddaldalddaldddddddddddd.)。

0

相关内容

估计/估计量

估计/估计量

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知

133+阅读 · 2020年3月18日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【推荐】NiftyNet：面向医学图像分析和图像引导治疗的开源CNN平台（附代码）

【推荐】NiftyNet：面向医学图像分析和图像引导治疗的开源CNN平台（附代码）

机器学习研究会

12+阅读 · 2018年1月27日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

Lnc-TRMT2A竞争性结合miR-520a调控炎性通路在精神分裂症发病中的作用研究

国家自然科学基金

0+阅读 · 2015年12月31日

内质网应激IRE1－XBP1S通路在高糖引起肾脏及系膜细胞发生氧化应激及损伤中的机制研究

国家自然科学基金

1+阅读 · 2014年12月31日

p75ICD调控p35-p25/CDK5信号通路在脑出血诱导的神经元凋亡中的作用

国家自然科学基金

0+阅读 · 2014年12月31日

基于权重函数修正的大气CO2垂直柱浓度遥测算法研究

国家自然科学基金

0+阅读 · 2013年12月31日

P38 MAPK信号通路在S. boulardii预防DON诱导猪单核巨噬细胞凋亡的作用研究

国家自然科学基金

0+阅读 · 2013年12月31日

Numbl-TRAF6-TAB2对NF-kappa B活性的调节在小胶质细胞炎性活化中的作用

国家自然科学基金

0+阅读 · 2012年12月31日

基于Tetrolet变换的偏振遥感图像融合算法研究

国家自然科学基金

0+阅读 · 2012年12月31日

PARP-1和自噬在放射所致鼻咽癌细胞凋亡过程中的作用及其调控机制

国家自然科学基金

0+阅读 · 2011年12月31日

哮喘中T细胞活化衔接子对调节性T细胞调控研究

国家自然科学基金

0+阅读 · 2011年12月31日

拟南芥R基因介导的植物防卫反应高温敏感的分子机理

国家自然科学基金

0+阅读 · 2011年12月31日

EGA-Depth: Efficient Guided Attention for Self-Supervised Multi-Camera Depth Estimation

Arxiv

0+阅读 · 2023年4月6日

VPFusion: Towards Robust Vertical Representation Learning for 3D Object Detection

Arxiv

0+阅读 · 2023年4月6日

Joint 2D-3D Multi-Task Learning on Cityscapes-3D: 3D Detection, Segmentation, and Depth Estimation

Arxiv

0+阅读 · 2023年4月5日

Trap-Based Pest Counting: Multiscale and Deformable Attention CenterNet Integrating Internal LR and HR Joint Feature Learning

Arxiv

0+阅读 · 2023年4月5日

XKD: Cross-modal Knowledge Distillation with Domain Alignment for Video Representation Learning

Arxiv

0+阅读 · 2023年4月5日

MPCViT: Searching for Accurate and Efficient MPC-Friendly Vision Transformer with Heterogeneous Attention

Arxiv

0+阅读 · 2023年4月4日

Exploration of Lightweight Single Image Denoising with Transformers and Truly Fair Training

Arxiv

0+阅读 · 2023年4月4日

Multimodal Neural Processes for Uncertainty Estimation

Arxiv

0+阅读 · 2023年4月4日

MoLo: Motion-augmented Long-short Contrastive Learning for Few-shot Action Recognition

Arxiv

0+阅读 · 2023年4月3日

Hyperparameter Ensembles for Robustness and Uncertainty Quantification

Arxiv

12+阅读 · 2020年6月24日

VIP会员

文章信息

相关主题

估计/估计量

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【ICCV2025教程】基础模型遇见具身智能体

军事机器学习设计：关于开发自动化任务摘要系统的梯次化设计科学研究 | 2025最新93页

扩散模型中的缓存方法综述：迈向高效的多模态生成

【ICCV2025教程】《迈向视觉语言模型的全面推理》

相关资讯

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知

133+阅读 · 2020年3月18日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【推荐】NiftyNet：面向医学图像分析和图像引导治疗的开源CNN平台（附代码）

【推荐】NiftyNet：面向医学图像分析和图像引导治疗的开源CNN平台（附代码）

机器学习研究会

12+阅读 · 2018年1月27日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

相关论文

EGA-Depth: Efficient Guided Attention for Self-Supervised Multi-Camera Depth Estimation

Arxiv

0+阅读 · 2023年4月6日

VPFusion: Towards Robust Vertical Representation Learning for 3D Object Detection

Arxiv

0+阅读 · 2023年4月6日

Joint 2D-3D Multi-Task Learning on Cityscapes-3D: 3D Detection, Segmentation, and Depth Estimation

Arxiv

0+阅读 · 2023年4月5日

Trap-Based Pest Counting: Multiscale and Deformable Attention CenterNet Integrating Internal LR and HR Joint Feature Learning

Arxiv

0+阅读 · 2023年4月5日

XKD: Cross-modal Knowledge Distillation with Domain Alignment for Video Representation Learning

Arxiv

0+阅读 · 2023年4月5日

MPCViT: Searching for Accurate and Efficient MPC-Friendly Vision Transformer with Heterogeneous Attention

Arxiv

0+阅读 · 2023年4月4日

Exploration of Lightweight Single Image Denoising with Transformers and Truly Fair Training

Arxiv

0+阅读 · 2023年4月4日

Multimodal Neural Processes for Uncertainty Estimation

Arxiv

0+阅读 · 2023年4月4日

MoLo: Motion-augmented Long-short Contrastive Learning for Few-shot Action Recognition

Arxiv

0+阅读 · 2023年4月3日

Hyperparameter Ensembles for Robustness and Uncertainty Quantification

Arxiv

12+阅读 · 2020年6月24日

相关基金

Lnc-TRMT2A竞争性结合miR-520a调控炎性通路在精神分裂症发病中的作用研究

国家自然科学基金

0+阅读 · 2015年12月31日

内质网应激IRE1－XBP1S通路在高糖引起肾脏及系膜细胞发生氧化应激及损伤中的机制研究

国家自然科学基金

1+阅读 · 2014年12月31日

p75ICD调控p35-p25/CDK5信号通路在脑出血诱导的神经元凋亡中的作用

国家自然科学基金

0+阅读 · 2014年12月31日

基于权重函数修正的大气CO2垂直柱浓度遥测算法研究

国家自然科学基金

0+阅读 · 2013年12月31日

P38 MAPK信号通路在S. boulardii预防DON诱导猪单核巨噬细胞凋亡的作用研究

国家自然科学基金

0+阅读 · 2013年12月31日

Numbl-TRAF6-TAB2对NF-kappa B活性的调节在小胶质细胞炎性活化中的作用

国家自然科学基金

0+阅读 · 2012年12月31日

基于Tetrolet变换的偏振遥感图像融合算法研究

国家自然科学基金

0+阅读 · 2012年12月31日

PARP-1和自噬在放射所致鼻咽癌细胞凋亡过程中的作用及其调控机制

国家自然科学基金

0+阅读 · 2011年12月31日

哮喘中T细胞活化衔接子对调节性T细胞调控研究

国家自然科学基金

0+阅读 · 2011年12月31日

拟南芥R基因介导的植物防卫反应高温敏感的分子机理

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员