通过将创制模型与现实世界数据相结合,为机器人学习提供更有力的普遍保障 (Stronger Generalization Guarantees for Robot Learning by Combining Generative Models and Real-World Data) - 专知论文

会员服务 ·

0

泛化理论 · Learning · 回合 · 生成模型 · MoDELS ·

2022 年 7 月 22 日

Stronger Generalization Guarantees for Robot Learning by Combining Generative Models and Real-World Data

翻译：通过将创制模型与现实世界数据相结合,为机器人学习提供更有力的普遍保障

Abhinav Agarwal,Sushant Veer,Allen Z. Ren,Anirudha Majumdar

We are motivated by the problem of learning policies for robotic systems with rich sensory inputs (e.g., vision) in a manner that allows us to guarantee generalization to environments unseen during training. We provide a framework for providing such generalization guarantees by leveraging a finite dataset of real-world environments in combination with a (potentially inaccurate) generative model of environments. The key idea behind our approach is to utilize the generative model in order to implicitly specify a prior over policies. This prior is updated using the real-world dataset of environments by minimizing an upper bound on the expected cost across novel environments derived via Probably Approximately Correct (PAC)-Bayes generalization theory. We demonstrate our approach on two simulated systems with nonlinear/hybrid dynamics and rich sensing modalities: (i) quadrotor navigation with an onboard vision sensor, and (ii) grasping objects using a depth sensor. Comparisons with prior work demonstrate the ability of our approach to obtain stronger generalization guarantees by utilizing generative models. We also present hardware experiments for validating our bounds for the grasping task.

翻译：我们的动力在于对具有丰富感官投入(例如视觉)的机器人系统采取学习政策的问题,这种学习政策能够保证在培训期间对不为人知的环境加以普遍化。我们提供了一个框架,通过利用现实世界环境的有限数据集,结合一种(可能不准确的)环境基因化模型,提供这种普遍化保障。我们的方法背后的关键思想是利用基因化模型,以隐含地具体说明先前的政策。前一种方法利用真实世界的环境数据集加以更新,通过“大概正确(PAC)-Bayes一般化理论”将新环境的预期成本的上限降到最低。我们展示了我们对两个模拟系统采用非线性/湿性动态和丰富感测模式的方法:(一) 利用机上视觉传感器的引力引引引引引引引引,以及(二) 利用深度传感器捕捉物体。与先前的工作进行比较表明,我们的方法有能力通过利用基因化模型获得更强的概括化保证。我们还介绍了用于验证掌握任务界限的硬件实验。

0

相关内容

泛化理论

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

专知会员服务

104+阅读 · 2022年2月10日

【ETH】最新《几何数据分析》2020课程，附PPT下载

专知会员服务

44+阅读 · 2020年12月18日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

80+阅读 · 2020年7月26日

【新书：机器学习简介】《A Concise Introduction to Machine Learning》by A.C. Faul (CRC 2019)

【新书：机器学习简介】《A Concise Introduction to Machine Learning》by A.C. Faul (CRC 2019)

专知会员服务

77+阅读 · 2020年2月8日

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

专知会员服务

244+阅读 · 2019年10月21日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

【ICIG2021】Latest News & Announcements of the Plenary Talk2

【ICIG2021】Latest News & Announcements of the Plenary Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年11月2日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

Progerin/PrelaminA诱发早老症的蛋白质组学研究

国家自然科学基金

1+阅读 · 2015年12月31日

NatD调节Slug基因表达促进肺癌细胞上皮间质转化

国家自然科学基金

0+阅读 · 2014年12月31日

木薯14-3-3蛋白磷酸化在块根淀粉积累过程中的功能研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于AOD-FLIM/CARS多模光学平台监测单个活细胞内RNA合成过程

国家自然科学基金

0+阅读 · 2013年12月31日

超短超强激光驱动的高亮度Betatron辐射光源

国家自然科学基金

1+阅读 · 2013年12月31日

Intraflagellar Transport运输纤毛蛋白的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

绿色热致相熔融纺丝工艺制备PVDF中空纤维膜成膜机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于超声波的管道流量测量及流速分布层析成像方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

Plug-In混合动力汽车能量管理及动力系统优化问题研究

国家自然科学基金

1+阅读 · 2008年12月31日

微通道内气液两相流及传质特性

国家自然科学基金

0+阅读 · 2008年12月31日

N-LIMB: Neural Limb Optimization for Efficient Morphological Design

N-LIMB: Neural Limb Optimization for Efficient Morphological Design

Arxiv

0+阅读 · 2022年9月19日

A Stack-of-Tasks Approach Combined with Behavior Trees: a New Framework for Robot Control

Arxiv

0+阅读 · 2022年9月18日

Understanding Robust Learning through the Lens of Representation Similarities

Arxiv

0+阅读 · 2022年9月15日

On the programming effort required to generate Behavior Trees and Finite State Machines for robotic applications

On the programming effort required to generate Behavior Trees and Finite State Machines for robotic applications

Arxiv

0+阅读 · 2022年9月15日

Robust explicit estimation of the log-logistic distribution with applications

Arxiv

0+阅读 · 2022年9月15日

A Discrete-Time Switching System Analysis of Q-learning

Arxiv

0+阅读 · 2022年9月15日

PROB-SLAM: Real-time Visual SLAM Based on Probabilistic Graph Optimization

Arxiv

0+阅读 · 2022年9月15日

Generative Models as a Data Source for Multiview Representation Learning

Arxiv

16+阅读 · 2021年6月9日

Attribute-Guided Adversarial Training for Robustness to Natural Perturbations

Arxiv

15+阅读 · 2020年12月3日

A Modern Introduction to Online Learning

A Modern Introduction to Online Learning

Arxiv

21+阅读 · 2019年12月31日

VIP会员

文章信息

相关主题

相关VIP内容

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

专知会员服务

104+阅读 · 2022年2月10日

【ETH】最新《几何数据分析》2020课程，附PPT下载

专知会员服务

44+阅读 · 2020年12月18日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

80+阅读 · 2020年7月26日

【新书：机器学习简介】《A Concise Introduction to Machine Learning》by A.C. Faul (CRC 2019)

【新书：机器学习简介】《A Concise Introduction to Machine Learning》by A.C. Faul (CRC 2019)

专知会员服务

77+阅读 · 2020年2月8日

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

专知会员服务

244+阅读 · 2019年10月21日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《巡飞弹药（爆炸性无人机）威胁态势分析》最新24页报告

《军用后勤无人机：破解战场运输挑战的创新方案》

人工智能战争：以色列、伊朗与新型AI战争形态

《俄乌战争：现代战争未来的启示与经验》

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

【ICIG2021】Latest News & Announcements of the Plenary Talk2

【ICIG2021】Latest News & Announcements of the Plenary Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年11月2日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

相关论文

N-LIMB: Neural Limb Optimization for Efficient Morphological Design

N-LIMB: Neural Limb Optimization for Efficient Morphological Design

Arxiv

0+阅读 · 2022年9月19日

A Stack-of-Tasks Approach Combined with Behavior Trees: a New Framework for Robot Control

Arxiv

0+阅读 · 2022年9月18日

Understanding Robust Learning through the Lens of Representation Similarities

Arxiv

0+阅读 · 2022年9月15日

On the programming effort required to generate Behavior Trees and Finite State Machines for robotic applications

On the programming effort required to generate Behavior Trees and Finite State Machines for robotic applications

Arxiv

0+阅读 · 2022年9月15日

Robust explicit estimation of the log-logistic distribution with applications

Arxiv

0+阅读 · 2022年9月15日

A Discrete-Time Switching System Analysis of Q-learning

Arxiv

0+阅读 · 2022年9月15日

PROB-SLAM: Real-time Visual SLAM Based on Probabilistic Graph Optimization

Arxiv

0+阅读 · 2022年9月15日

Generative Models as a Data Source for Multiview Representation Learning

Arxiv

16+阅读 · 2021年6月9日

Attribute-Guided Adversarial Training for Robustness to Natural Perturbations

Arxiv

15+阅读 · 2020年12月3日

A Modern Introduction to Online Learning

A Modern Introduction to Online Learning

Arxiv

21+阅读 · 2019年12月31日

相关基金

Progerin/PrelaminA诱发早老症的蛋白质组学研究

国家自然科学基金

1+阅读 · 2015年12月31日

NatD调节Slug基因表达促进肺癌细胞上皮间质转化

国家自然科学基金

0+阅读 · 2014年12月31日

木薯14-3-3蛋白磷酸化在块根淀粉积累过程中的功能研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于AOD-FLIM/CARS多模光学平台监测单个活细胞内RNA合成过程

国家自然科学基金

0+阅读 · 2013年12月31日

超短超强激光驱动的高亮度Betatron辐射光源

国家自然科学基金

1+阅读 · 2013年12月31日

Intraflagellar Transport运输纤毛蛋白的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

绿色热致相熔融纺丝工艺制备PVDF中空纤维膜成膜机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于超声波的管道流量测量及流速分布层析成像方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

Plug-In混合动力汽车能量管理及动力系统优化问题研究

国家自然科学基金

1+阅读 · 2008年12月31日

微通道内气液两相流及传质特性

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员