SORNet: 顺序操纵的空间物体中心表示式 (SORNet: Spatial Object-Centric Representations for Sequential Manipulation) - 专知论文

会员服务 ·

0

Learning · 正则的 · entity · 表示 · 表示学习 ·

2022 年 9 月 14 日

SORNet: Spatial Object-Centric Representations for Sequential Manipulation

翻译：SORNet: 顺序操纵的空间物体中心表示式

Wentao Yuan,Chris Paxton,Karthik Desingh,Dieter Fox

from arxiv, CoRL 2021 Best Systems Paper Finalist; Code and data available at https://github.com/wentaoyuan/sornet

Sequential manipulation tasks require a robot to perceive the state of an environment and plan a sequence of actions leading to a desired goal state. In such tasks, the ability to reason about spatial relations among object entities from raw sensor inputs is crucial in order to determine when a task has been completed and which actions can be executed. In this work, we propose SORNet (Spatial Object-Centric Representation Network), a framework for learning object-centric representations from RGB images conditioned on a set of object queries, represented as image patches called canonical object views. With only a single canonical view per object and no annotation, SORNet generalizes zero-shot to object entities whose shape and texture are both unseen during training. We evaluate SORNet on various spatial reasoning tasks such as spatial relation classification and relative direction regression in complex tabletop manipulation scenarios and show that SORNet significantly outperforms baselines including state-of-the-art representation learning techniques. We also demonstrate the application of the representation learned by SORNet on visual-servoing and task planning for sequential manipulation on a real robot.

翻译：序列操作任务要求机器人感知环境状态,并计划一系列导致预期目标状态的行动。在这种任务中,对原始传感器输入的物体实体之间的空间关系进行思考的能力至关重要,以便确定任务何时完成和可以执行哪些行动。在这项工作中,我们提议SORNet(空间物体中心代表网络),这是一个学习以一组物体查询为条件的 RGB 图像的物体中心显示框架,以一套物体查询为条件,作为称为光学天体视图的图像补丁。由于每个物体只有单一的直观,没有注解,SORNet对在训练期间形状和纹理都看不见的物体实体一般地将零光化为零。我们评估SORNet关于各种空间推理任务,例如空间关系分类和复杂桌面操作情景中相对方向回归,并显示SORNet大大超越了基线,包括状态-艺术代表学习技术。我们还演示了SORNet在视觉观察和任务规划中对真实机器人进行连续操纵方面学到的演示。

0

相关内容

Learning

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

史上最全！358篇机器学习&自然语言处理综述论文！都这儿了

专知会员服务

129+阅读 · 2020年7月18日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【跟踪Tracking】15篇论文+代码 | 中秋快乐~

【跟踪Tracking】15篇论文+代码 | 中秋快乐~

专知

18+阅读 · 2018年9月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【推荐】YOLO实时目标检测(6fps)

【推荐】YOLO实时目标检测(6fps)

机器学习研究会

20+阅读 · 2017年11月5日

Mipu1促血管新生的机制研究：对VEGF-VASH1/SVBP负反馈通路的转录调节

国家自然科学基金

0+阅读 · 2014年12月31日

酸敏感离子通道(ASICs)在过敏性紫癜患儿血管内皮细胞损伤中的调控作用

国家自然科学基金

0+阅读 · 2014年12月31日

骨髓间充质干细胞调节炎性微环境干预恒河猴糖尿病肾病免疫损伤的效应机制

国家自然科学基金

0+阅读 · 2012年12月31日

GLP-1/beta-catenin/TCF信号通路对糖尿病鼠心肌细胞凋亡的保护作用及机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

Nrf2/ARE调控的乙二醛酶1在糖尿病脑病防治中的作用及芒果苷的效应和机制

国家自然科学基金

0+阅读 · 2012年12月31日

新癌基因E3连接酶HECTD3表达调节机制的研究

国家自然科学基金

1+阅读 · 2012年12月31日

PM2.5暴露诱发胰岛素抵抗的分子作用机制

国家自然科学基金

0+阅读 · 2012年12月31日

离子通道TRPM2在血管壁内膜增生中的作用

国家自然科学基金

0+阅读 · 2011年12月31日

病理性近视易感基因研究

国家自然科学基金

0+阅读 · 2009年12月31日

糖原合酶激酶3在阿尔茨海默病突触病变中的作用及机制

国家自然科学基金

0+阅读 · 2008年12月31日

Learning Explicit Object-Centric Representations with Vision Transformers

Arxiv

0+阅读 · 2022年10月25日

Salient Object Detection via Dynamic Scale Routing

Arxiv

0+阅读 · 2022年10月25日

DeXtreme: Transfer of Agile In-hand Manipulation from Simulation to Reality

Arxiv

0+阅读 · 2022年10月25日

Benchmarking Deformable Object Manipulation with Differentiable Physics

Arxiv

0+阅读 · 2022年10月24日

Learning Neural Radiance Fields from Multi-View Geometry

Arxiv

0+阅读 · 2022年10月24日

Active Exploration for Robotic Manipulation

Arxiv

0+阅读 · 2022年10月23日

Learning Feasibility of Factored Nonlinear Programs in Robotic Manipulation Planning

Arxiv

0+阅读 · 2022年10月22日

Neural Fields for Robotic Object Manipulation from a Single Image

Arxiv

0+阅读 · 2022年10月21日

Object Goal Navigation Based on Semantics and RGB Ego View

Arxiv

0+阅读 · 2022年10月20日

Evolving Losses for Unsupervised Video Representation Learning

Arxiv

23+阅读 · 2020年2月26日

VIP会员

文章信息

相关主题

相关VIP内容

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

史上最全！358篇机器学习&自然语言处理综述论文！都这儿了

专知会员服务

129+阅读 · 2020年7月18日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

最新《扩散模型原理》新书，470页pdf

无人机作战：演进、创新与未来战场

AI 智能体简史

多模态空间推理在大模型时代：综述与基准测试

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【跟踪Tracking】15篇论文+代码 | 中秋快乐~

【跟踪Tracking】15篇论文+代码 | 中秋快乐~

专知

18+阅读 · 2018年9月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【推荐】YOLO实时目标检测(6fps)

【推荐】YOLO实时目标检测(6fps)

机器学习研究会

20+阅读 · 2017年11月5日

相关论文

Learning Explicit Object-Centric Representations with Vision Transformers

Arxiv

0+阅读 · 2022年10月25日

Salient Object Detection via Dynamic Scale Routing

Arxiv

0+阅读 · 2022年10月25日

DeXtreme: Transfer of Agile In-hand Manipulation from Simulation to Reality

Arxiv

0+阅读 · 2022年10月25日

Benchmarking Deformable Object Manipulation with Differentiable Physics

Arxiv

0+阅读 · 2022年10月24日

Learning Neural Radiance Fields from Multi-View Geometry

Arxiv

0+阅读 · 2022年10月24日

Active Exploration for Robotic Manipulation

Arxiv

0+阅读 · 2022年10月23日

Learning Feasibility of Factored Nonlinear Programs in Robotic Manipulation Planning

Arxiv

0+阅读 · 2022年10月22日

Neural Fields for Robotic Object Manipulation from a Single Image

Arxiv

0+阅读 · 2022年10月21日

Object Goal Navigation Based on Semantics and RGB Ego View

Arxiv

0+阅读 · 2022年10月20日

Evolving Losses for Unsupervised Video Representation Learning

Arxiv

23+阅读 · 2020年2月26日

相关基金

Mipu1促血管新生的机制研究：对VEGF-VASH1/SVBP负反馈通路的转录调节

国家自然科学基金

0+阅读 · 2014年12月31日

酸敏感离子通道(ASICs)在过敏性紫癜患儿血管内皮细胞损伤中的调控作用

国家自然科学基金

0+阅读 · 2014年12月31日

骨髓间充质干细胞调节炎性微环境干预恒河猴糖尿病肾病免疫损伤的效应机制

国家自然科学基金

0+阅读 · 2012年12月31日

GLP-1/beta-catenin/TCF信号通路对糖尿病鼠心肌细胞凋亡的保护作用及机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

Nrf2/ARE调控的乙二醛酶1在糖尿病脑病防治中的作用及芒果苷的效应和机制

国家自然科学基金

0+阅读 · 2012年12月31日

新癌基因E3连接酶HECTD3表达调节机制的研究

国家自然科学基金

1+阅读 · 2012年12月31日

PM2.5暴露诱发胰岛素抵抗的分子作用机制

国家自然科学基金

0+阅读 · 2012年12月31日

离子通道TRPM2在血管壁内膜增生中的作用

国家自然科学基金

0+阅读 · 2011年12月31日

病理性近视易感基因研究

国家自然科学基金

0+阅读 · 2009年12月31日

糖原合酶激酶3在阿尔茨海默病突触病变中的作用及机制

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员