未知环境中的等级预测控制数据驱动战略 (Data-Driven Strategies for Hierarchical Predictive Control in Unknown Environments) - 专知论文

会员服务 ·

0

控制器 · 回合 · MoDELS · 学成 · Performer ·

2021 年 7 月 14 日

Data-Driven Strategies for Hierarchical Predictive Control in Unknown Environments

翻译：未知环境中的等级预测控制数据驱动战略

Charlott Vallon,Francesco Borrelli

from arxiv, arXiv admin note: substantial text overlap with arXiv:2005.05948

This article proposes a hierarchical learning architecture for safe data-driven control in unknown environments. We consider a constrained nonlinear dynamical system and assume the availability of state-input trajectories solving control tasks in different environments. In addition to task-invariant system state and input constraints, a parameterized environment model generates task-specific state constraints, which are satisfied by the stored trajectories. Our goal is to use these trajectories to find a safe and high-performing policy for a new task in a new, unknown environment. We propose using the stored data to learn generalizable control strategies. At each time step, based on a local forecast of the new task environment, the learned strategy consists of a target region in the state space and input constraints to guide the system evolution to the target region. These target regions are used as terminal sets by a low-level model predictive controller. We show how to i) design the target sets from past data and then ii) incorporate them into a model predictive control scheme with shifting horizon that ensures safety of the closed-loop system when performing the new task. We prove the feasibility of the resulting control policy, and apply the proposed method to robotic path planning, racing, and computer game applications.

翻译：本条提出在未知环境中安全数据驱动的控制的等级学习架构。我们考虑一个限制的非线性动态系统, 并假设有国家输入轨迹可以在不同环境中解决控制任务。除了任务差异系统状态和输入限制之外, 一个参数环境模型还产生任务特定状态限制, 被存储的轨迹所满足。我们的目标是利用这些轨迹为一个新的、未知环境中的新任务寻找安全和高绩效的政策。我们提议使用存储的数据学习通用的控制战略。在每个时间步骤中, 根据对新任务环境的本地预测, 学习的战略包括州空间目标区域和指导系统向目标区域演变的输入限制。这些目标区域被一个低级别模型预测控制器用作终端组。我们展示如何从过去的数据中设计目标组, 然后将它们纳入一个具有变化视野的模型预测控制计划, 以确保闭路控制系统在执行新任务时的安全。我们证明, 由此而形成的游戏控制策略的可行性, 并应用所拟议的方法。

0

相关内容

控制器

【ICML2021】异质风险最小化，Heterogeneous Risk Minimization

专知会员服务

16+阅读 · 2021年5月21日

【新书】《图数据实践者指南》，附电子书779页pdf与代码

【新书】《图数据实践者指南》，附电子书779页pdf与代码

专知会员服务

90+阅读 · 2021年4月5日

机器学习组合优化

机器学习组合优化

专知会员服务

110+阅读 · 2021年2月16日

2020数据工程师成长路线图

专知会员服务

41+阅读 · 2020年9月6日

【DeepMind】PolyGen: 一种三维网格的自回归生成模型，PolyGen: An Autoregressive Generative Model of 3D Meshes

【DeepMind】PolyGen: 一种三维网格的自回归生成模型，PolyGen: An Autoregressive Generative Model of 3D Meshes

专知会员服务

37+阅读 · 2020年2月27日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

【KDD 2019|Tutorial】应用在交通中的强化学习 Deep Reinforcement Learning with Applications in Transportation，滴滴 AI Labs

【KDD 2019|Tutorial】应用在交通中的强化学习 Deep Reinforcement Learning with Applications in Transportation，滴滴 AI Labs

专知会员服务

65+阅读 · 2019年8月8日

量化金融强化学习论文集合

量化金融强化学习论文集合

专知

14+阅读 · 2019年12月18日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

已删除

清华大学研究生教育

3+阅读 · 2018年6月30日

Hierarchical Disentangled Representations

Hierarchical Disentangled Representations

CreateAMind

4+阅读 · 2018年4月15日

【学习】Hierarchical Softmax

【学习】Hierarchical Softmax

机器学习研究会

4+阅读 · 2017年8月6日

强化学习 cartpole_a3c

强化学习 cartpole_a3c

CreateAMind

9+阅读 · 2017年7月21日

Model Predictive Control with Environment Adaptation for Legged Locomotion

Model Predictive Control with Environment Adaptation for Legged Locomotion

Arxiv

0+阅读 · 2021年9月16日

Infusing model predictive control into meta-reinforcement learning for mobile robots in dynamic environments

Arxiv

0+阅读 · 2021年9月15日

Delay-aware Robust Control for Safe Autonomous Driving

Arxiv

0+阅读 · 2021年9月15日

DPMPC-Planner: A real-time UAV trajectory planning framework for complex static environments with dynamic obstacles

Arxiv

0+阅读 · 2021年9月14日

STORM: An Integrated Framework for Fast Joint-Space Model-Predictive Control for Reactive Manipulation

Arxiv

0+阅读 · 2021年9月14日

B-GAP: Behavior-Guided Action Prediction and Navigation for Autonomous Driving

B-GAP: Behavior-Guided Action Prediction and Navigation for Autonomous Driving

Arxiv

0+阅读 · 2021年9月14日

Design and Model Predictive Control of Mars Coaxial Quadrotor

Design and Model Predictive Control of Mars Coaxial Quadrotor

Arxiv

0+阅读 · 2021年9月14日

A Hierarchical Control Framework for Drift Maneuvering of Autonomous Vehicles

A Hierarchical Control Framework for Drift Maneuvering of Autonomous Vehicles

Arxiv

0+阅读 · 2021年9月14日

Reward learning from human preferences and demonstrations in Atari

Arxiv

8+阅读 · 2018年11月15日

Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments

Arxiv

6+阅读 · 2018年1月16日

VIP会员

文章信息

相关主题

相关VIP内容

【ICML2021】异质风险最小化，Heterogeneous Risk Minimization

专知会员服务

16+阅读 · 2021年5月21日

【新书】《图数据实践者指南》，附电子书779页pdf与代码

【新书】《图数据实践者指南》，附电子书779页pdf与代码

专知会员服务

90+阅读 · 2021年4月5日

机器学习组合优化

机器学习组合优化

专知会员服务

110+阅读 · 2021年2月16日

2020数据工程师成长路线图

专知会员服务

41+阅读 · 2020年9月6日

【DeepMind】PolyGen: 一种三维网格的自回归生成模型，PolyGen: An Autoregressive Generative Model of 3D Meshes

【DeepMind】PolyGen: 一种三维网格的自回归生成模型，PolyGen: An Autoregressive Generative Model of 3D Meshes

专知会员服务

37+阅读 · 2020年2月27日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

【KDD 2019|Tutorial】应用在交通中的强化学习 Deep Reinforcement Learning with Applications in Transportation，滴滴 AI Labs

【KDD 2019|Tutorial】应用在交通中的强化学习 Deep Reinforcement Learning with Applications in Transportation，滴滴 AI Labs

专知会员服务

65+阅读 · 2019年8月8日

热门VIP内容

开通专知VIP会员享更多权益服务

NeurIPS 2025 | 自动化所新作速览（一）

大型语言模型（LLM）赋能的知识图谱构建：综述

NeurIPS 2025 | 自动化所新作速览（二）

领域特定文本分类中的预训练语言模型新进展：系统综述

相关资讯

量化金融强化学习论文集合

量化金融强化学习论文集合

专知

14+阅读 · 2019年12月18日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

已删除

清华大学研究生教育

3+阅读 · 2018年6月30日

Hierarchical Disentangled Representations

Hierarchical Disentangled Representations

CreateAMind

4+阅读 · 2018年4月15日

【学习】Hierarchical Softmax

【学习】Hierarchical Softmax

机器学习研究会

4+阅读 · 2017年8月6日

强化学习 cartpole_a3c

强化学习 cartpole_a3c

CreateAMind

9+阅读 · 2017年7月21日

相关论文

Model Predictive Control with Environment Adaptation for Legged Locomotion

Model Predictive Control with Environment Adaptation for Legged Locomotion

Arxiv

0+阅读 · 2021年9月16日

Infusing model predictive control into meta-reinforcement learning for mobile robots in dynamic environments

Arxiv

0+阅读 · 2021年9月15日

Delay-aware Robust Control for Safe Autonomous Driving

Arxiv

0+阅读 · 2021年9月15日

DPMPC-Planner: A real-time UAV trajectory planning framework for complex static environments with dynamic obstacles

Arxiv

0+阅读 · 2021年9月14日

STORM: An Integrated Framework for Fast Joint-Space Model-Predictive Control for Reactive Manipulation

Arxiv

0+阅读 · 2021年9月14日

B-GAP: Behavior-Guided Action Prediction and Navigation for Autonomous Driving

B-GAP: Behavior-Guided Action Prediction and Navigation for Autonomous Driving

Arxiv

0+阅读 · 2021年9月14日

Design and Model Predictive Control of Mars Coaxial Quadrotor

Design and Model Predictive Control of Mars Coaxial Quadrotor

Arxiv

0+阅读 · 2021年9月14日

A Hierarchical Control Framework for Drift Maneuvering of Autonomous Vehicles

A Hierarchical Control Framework for Drift Maneuvering of Autonomous Vehicles

Arxiv

0+阅读 · 2021年9月14日

Reward learning from human preferences and demonstrations in Atari

Arxiv

8+阅读 · 2018年11月15日

Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments

Arxiv

6+阅读 · 2018年1月16日

微信扫码咨询专知VIP会员