SE(3)行动空间的政策学习 (Policy learning in SE(3) action spaces) - 专知论文

会员服务 ·

0

DQfD · 可约的 · 状态空间 · 学成 · 机器人 ·

2020 年 10 月 6 日

Policy learning in SE(3) action spaces

翻译：SE(3)行动空间的政策学习

Dian Wang,Colin Kohler,Robert Platt

from arxiv, 16 pages, submitted to CoRL 2020

In the spatial action representation, the action space spans the space of target poses for robot motion commands, i.e. SE(2) or SE(3). This approach has been used to solve challenging robotic manipulation problems and shows promise. However, the method is often limited to a three dimensional action space and short horizon tasks. This paper proposes ASRSE3, a new method for handling higher dimensional spatial action spaces that transforms an original MDP with high dimensional action space into a new MDP with reduced action space and augmented state space. We also propose SDQfD, a variation of DQfD designed for large action spaces. ASRSE3 and SDQfD are evaluated in the context of a set of challenging block construction tasks. We show that both methods outperform standard baselines and can be used in practice on real robotics systems.

翻译：在空间行动代表中,行动空间跨越了机器人运动指令(即SE(2)或SE(3))的目标设定空间。这一方法已被用于解决具有挑战性的机器人操纵问题并显示出希望。然而,该方法往往限于三维行动空间和短期任务。本文提议ASRSE3, 这是一种处理具有高维行动空间的更高维空间行动空间的新方法,它将原来的具有高维行动空间的MDP转化为一个新的MDP,其行动空间减少,国家空间扩大。我们还提议SDQfD, 用于大型行动空间的DQfD变异。ASRSE3和SDQfD是在一套具有挑战性的块建筑任务的背景下进行评估的。我们表明,这两种方法都优于标准基线,可用于实际机器人系统。

0

相关内容

DQfD

面向知识图谱的信息抽取

专知会员服务

202+阅读 · 2020年10月14日

深度学习搜索，Exploring Deep Learning for Search

深度学习搜索，Exploring Deep Learning for Search

专知会员服务

61+阅读 · 2020年5月9日

【牛津大学】深度残差强化学习，Deep Residual Reinforcement Learning

【牛津大学】深度残差强化学习，Deep Residual Reinforcement Learning

专知会员服务

85+阅读 · 2020年2月18日

【AAAI2020教程】强化学习中的Exploration-Exploitation in Reinforcement Learning

专知会员服务

101+阅读 · 2020年2月8日

【新书】深度学习搜索，Deep Learning for Search，附327页pdf

【新书】深度学习搜索，Deep Learning for Search，附327页pdf

专知会员服务

214+阅读 · 2020年1月13日

【电子书】机器学习实战（Machine Learning in Action），附PDF

【电子书】机器学习实战（Machine Learning in Action），附PDF

专知会员服务

131+阅读 · 2019年11月25日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

【Pieter Abbeel 报告@CMU】元学习与深度强化学习机器人应用，Deep Learning to Learn，84页ppt

【Pieter Abbeel 报告@CMU】元学习与深度强化学习机器人应用，Deep Learning to Learn，84页ppt

专知会员服务

32+阅读 · 2019年10月12日

【ALT 2019 Tutorials】强化学习的探索性开发（Exploration-Exploitation in Reinforcement Learning）

【ALT 2019 Tutorials】强化学习的探索性开发（Exploration-Exploitation in Reinforcement Learning）

专知会员服务

34+阅读 · 2019年3月21日

【新书】深度学习搜索，Deep Learning for Search，327页pdf

【新书】深度学习搜索，Deep Learning for Search，327页pdf

专知

85+阅读 · 2020年1月19日

强化学习扫盲贴：从Q-learning到DQN

强化学习扫盲贴：从Q-learning到DQN

夕小瑶的卖萌屋

52+阅读 · 2019年10月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

RL 真经

CreateAMind

5+阅读 · 2018年12月28日

Hierarchical Imitation - Reinforcement Learning

Hierarchical Imitation - Reinforcement Learning

CreateAMind

19+阅读 · 2018年5月25日

推荐｜深度强化学习聊天机器人（附论文）！

推荐｜深度强化学习聊天机器人（附论文）！

全球人工智能

4+阅读 · 2018年1月30日

强化学习族谱

强化学习族谱

CreateAMind

26+阅读 · 2017年8月2日

Distilling a Hierarchical Policy for Planning and Control via Representation and Reinforcement Learning

Arxiv

0+阅读 · 2020年11月16日

PLAS: Latent Action Space for Offline Reinforcement Learning

Arxiv

0+阅读 · 2020年11月14日

Joint Space Control via Deep Reinforcement Learning

Arxiv

0+阅读 · 2020年11月12日

Robust Batch Policy Learning in Markov Decision Processes

Arxiv

0+阅读 · 2020年11月10日

Shared Experience Actor-Critic for Multi-Agent Reinforcement Learning

Arxiv

0+阅读 · 2020年11月6日

On the Search for Feedback in Reinforcement Learning

Arxiv

0+阅读 · 2020年11月4日

An On-Line POMDP Solver for Continuous Observation Spaces

Arxiv

0+阅读 · 2020年11月4日

Generalization to New Actions in Reinforcement Learning

Arxiv

0+阅读 · 2020年11月3日

NEARL: Non-Explicit Action Reinforcement Learning for Robotic Control

Arxiv

0+阅读 · 2020年11月2日

DeepPath: A Reinforcement Learning Method for Knowledge Graph Reasoning

Arxiv

20+阅读 · 2018年1月8日

VIP会员

文章信息

相关主题

相关VIP内容

面向知识图谱的信息抽取

专知会员服务

202+阅读 · 2020年10月14日

深度学习搜索，Exploring Deep Learning for Search

深度学习搜索，Exploring Deep Learning for Search

专知会员服务

61+阅读 · 2020年5月9日

【牛津大学】深度残差强化学习，Deep Residual Reinforcement Learning

【牛津大学】深度残差强化学习，Deep Residual Reinforcement Learning

专知会员服务

85+阅读 · 2020年2月18日

【AAAI2020教程】强化学习中的Exploration-Exploitation in Reinforcement Learning

专知会员服务

101+阅读 · 2020年2月8日

【新书】深度学习搜索，Deep Learning for Search，附327页pdf

【新书】深度学习搜索，Deep Learning for Search，附327页pdf

专知会员服务

214+阅读 · 2020年1月13日

【电子书】机器学习实战（Machine Learning in Action），附PDF

【电子书】机器学习实战（Machine Learning in Action），附PDF

专知会员服务

131+阅读 · 2019年11月25日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

【Pieter Abbeel 报告@CMU】元学习与深度强化学习机器人应用，Deep Learning to Learn，84页ppt

【Pieter Abbeel 报告@CMU】元学习与深度强化学习机器人应用，Deep Learning to Learn，84页ppt

专知会员服务

32+阅读 · 2019年10月12日

【ALT 2019 Tutorials】强化学习的探索性开发（Exploration-Exploitation in Reinforcement Learning）

【ALT 2019 Tutorials】强化学习的探索性开发（Exploration-Exploitation in Reinforcement Learning）

专知会员服务

34+阅读 · 2019年3月21日

热门VIP内容

开通专知VIP会员享更多权益服务

《俄乌战争背景下俄罗斯的战略性海军分析（2022-2025年）》最新100页报告

【斯坦福博士论文】数据、决策与依赖：构建可信人工智能的挑战

人工智能时代背景下的未来海战

接触战中的无人机优势：美军旅级部队面临的小型无人机系统挑战与调整

相关资讯

【新书】深度学习搜索，Deep Learning for Search，327页pdf

【新书】深度学习搜索，Deep Learning for Search，327页pdf

专知

85+阅读 · 2020年1月19日

强化学习扫盲贴：从Q-learning到DQN

强化学习扫盲贴：从Q-learning到DQN

夕小瑶的卖萌屋

52+阅读 · 2019年10月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

RL 真经

CreateAMind

5+阅读 · 2018年12月28日

Hierarchical Imitation - Reinforcement Learning

Hierarchical Imitation - Reinforcement Learning

CreateAMind

19+阅读 · 2018年5月25日

推荐｜深度强化学习聊天机器人（附论文）！

推荐｜深度强化学习聊天机器人（附论文）！

全球人工智能

4+阅读 · 2018年1月30日

强化学习族谱

强化学习族谱

CreateAMind

26+阅读 · 2017年8月2日

相关论文

Distilling a Hierarchical Policy for Planning and Control via Representation and Reinforcement Learning

Arxiv

0+阅读 · 2020年11月16日

PLAS: Latent Action Space for Offline Reinforcement Learning

Arxiv

0+阅读 · 2020年11月14日

Joint Space Control via Deep Reinforcement Learning

Arxiv

0+阅读 · 2020年11月12日

Robust Batch Policy Learning in Markov Decision Processes

Arxiv

0+阅读 · 2020年11月10日

Shared Experience Actor-Critic for Multi-Agent Reinforcement Learning

Arxiv

0+阅读 · 2020年11月6日

On the Search for Feedback in Reinforcement Learning

Arxiv

0+阅读 · 2020年11月4日

An On-Line POMDP Solver for Continuous Observation Spaces

Arxiv

0+阅读 · 2020年11月4日

Generalization to New Actions in Reinforcement Learning

Arxiv

0+阅读 · 2020年11月3日

NEARL: Non-Explicit Action Reinforcement Learning for Robotic Control

Arxiv

0+阅读 · 2020年11月2日

DeepPath: A Reinforcement Learning Method for Knowledge Graph Reasoning

Arxiv

20+阅读 · 2018年1月8日

微信扫码咨询专知VIP会员