多机器人碰撞,通过学习交流来避免多机器人碰撞 (Multi Robot Collision Avoidance by Learning Whom to Communicate) - 专知论文

会员服务 ·

0

Learning · INFORMS · Agent · Performer · Networking ·

2022 年 9 月 14 日

Multi Robot Collision Avoidance by Learning Whom to Communicate

翻译：多机器人碰撞,通过学习交流来避免多机器人碰撞

Senthil Hariharan Arul,Amrit Singh Bedi,Dinesh Manocha

Agents in decentralized multi-agent navigation lack the world knowledge to make safe and (near-)optimal plans reliably. They base their decisions on their neighbors' observable states, which hide the neighbors' navigation intent. We propose augmenting decentralized navigation with inter-agent communication to improve their performance and aid agent in making sound navigation decisions. In this regard, we present a novel reinforcement learning method for multi-agent collision avoidance using selective inter-agent communication. Our network learns to decide 'when' and with 'whom' to communicate to request additional information in an end-to-end fashion. We pose communication selection as a link prediction problem, where the network predicts if communication is necessary given the observable information. The communicated information augments the observed neighbor information to select a suitable navigation plan. As the number of neighbors for a robot varies, we use a multi-head self-attention mechanism to encode neighbor information and create a fixed-length observation vector. We validate that our proposed approach achieves safe and efficient navigation among multiple robots in challenging simulation benchmarks. Aided by learned communication, our network performs significantly better than existing decentralized methods across various metrics such as time-to-goal and collision frequency. Besides, we showcase that the network effectively learns to communicate when necessary in a situation of high complexity.

翻译：分散式多试剂导航中的代理人缺乏可靠地制定安全和(近距离)最佳计划的世界知识。他们的决定以邻居的观察状态为基础, 隐藏邻居的导航意图。我们提议增加分散式导航, 增加代理人之间的通讯, 以提高其性能, 协助代理人做出健全的导航决定。在这方面, 我们提出一种新的强化学习方法, 使用选择性的代理人间通讯来避免多试剂碰撞。我们的网络学会了“ 何时” 和“ 与谁” 进行沟通, 以最终到最后的方式要求额外信息。我们把通信选择作为连接预测问题, 网络预测通信是否有必要提供可观测的信息。传送的信息增加了观察到的邻居信息, 以选择合适的导航计划。由于机器人的邻居数量不同, 我们使用多头自省自省机制来编码邻居的信息, 并创建一个固定长度的观测矢量。我们确认我们提出的方法在挑战模拟基准中实现了安全高效的导航。我们通过学习到的通讯, 我们的网络比现有的分散式方法要好得多, 超越了各种计量标准, 以有效的时间到高频路。

0

相关内容

Learning

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

vae 相关论文表示学习 1

vae 相关论文表示学习 1

CreateAMind

12+阅读 · 2018年9月6日

原发性胆汁性肝硬化中长链非编码RNA-GACAT1对肝内胆管细胞的调控机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

神经元和星形胶质细胞特异性miRNA对神经网络发育和功能的调控机制

国家自然科学基金

1+阅读 · 2013年12月31日

基于SURE/PURE准则的图像盲反卷积算法研究

国家自然科学基金

3+阅读 · 2013年12月31日

基于"Build-and-Click"法的铂类RNA聚合酶I选择性抑制剂的构建、评价及亚细胞定位研究

国家自然科学基金

1+阅读 · 2013年12月31日

可变剪切基因REST调控SRRM3的表达在前列腺癌神经内分泌分化中的作用机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

BAFF干扰的树突状细胞参与自身免疫性关节炎免疫耐受的作用和机制

国家自然科学基金

0+阅读 · 2012年12月31日

不同基因型（p53codon72）鼻咽癌细胞放射敏感性差异的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

表观遗传学修饰NANOG基因对胶质瘤细胞生物学活性的作用及机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于RSerPool应用的SCTP传输协议传输路径的优化研究

国家自然科学基金

0+阅读 · 2011年12月31日

特异溶瘤腺病毒对前列腺癌免疫治疗的实验研究与机制

国家自然科学基金

0+阅读 · 2009年12月31日

A Task Allocation Framework for Human Multi-Robot Collaborative Settings

Arxiv

0+阅读 · 2022年10月25日

Planning Coordinated Human-Robot Motions with Neural Network Full-Body Prediction Models

Arxiv

0+阅读 · 2022年10月24日

If You Are Careful, So Am I! How Robot Communicative Motions Can Influence Human Approach in a Joint Task

Arxiv

0+阅读 · 2022年10月24日

The Design and Realization of Multi-agent Obstacle Avoidance based on Reinforcement Learning

Arxiv

0+阅读 · 2022年10月24日

Robust and Secure Resource Allocation for ISAC Systems: A Novel Optimization Framework for Variable-Length Snapshots

Arxiv

0+阅读 · 2022年10月23日

AR Point&Click: An Interface for Setting Robot Navigation Goals

Arxiv

0+阅读 · 2022年10月22日

Learning Action Duration and Synergy in Task Planning for Human-Robot Collaboration

Arxiv

0+阅读 · 2022年10月21日

Decentralized and Communication-Free Multi-Robot Navigation through Distributed Games

Arxiv

40+阅读 · 2021年9月15日

Q-value Path Decomposition for Deep Multiagent Reinforcement Learning

Q-value Path Decomposition for Deep Multiagent Reinforcement Learning

Arxiv

26+阅读 · 2020年2月10日

Order-Free RNN with Visual Attention for Multi-Label Classification

Arxiv

16+阅读 · 2017年12月20日

VIP会员

文章信息

相关主题

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《全谱战争——从拓宽工具到思考不可思考之事》

《FPV武装无人机的战斗飞行艺术与科学》最新报告

无人机作战：演进、创新与未来战场

《反无人机：用于无人机探测与定位的多输入多输出雷达》最新69页

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

vae 相关论文表示学习 1

vae 相关论文表示学习 1

CreateAMind

12+阅读 · 2018年9月6日

相关论文

A Task Allocation Framework for Human Multi-Robot Collaborative Settings

Arxiv

0+阅读 · 2022年10月25日

Planning Coordinated Human-Robot Motions with Neural Network Full-Body Prediction Models

Arxiv

0+阅读 · 2022年10月24日

If You Are Careful, So Am I! How Robot Communicative Motions Can Influence Human Approach in a Joint Task

Arxiv

0+阅读 · 2022年10月24日

The Design and Realization of Multi-agent Obstacle Avoidance based on Reinforcement Learning

Arxiv

0+阅读 · 2022年10月24日

Robust and Secure Resource Allocation for ISAC Systems: A Novel Optimization Framework for Variable-Length Snapshots

Arxiv

0+阅读 · 2022年10月23日

AR Point&Click: An Interface for Setting Robot Navigation Goals

Arxiv

0+阅读 · 2022年10月22日

Learning Action Duration and Synergy in Task Planning for Human-Robot Collaboration

Arxiv

0+阅读 · 2022年10月21日

Decentralized and Communication-Free Multi-Robot Navigation through Distributed Games

Arxiv

40+阅读 · 2021年9月15日

Q-value Path Decomposition for Deep Multiagent Reinforcement Learning

Q-value Path Decomposition for Deep Multiagent Reinforcement Learning

Arxiv

26+阅读 · 2020年2月10日

Order-Free RNN with Visual Attention for Multi-Label Classification

Arxiv

16+阅读 · 2017年12月20日

相关基金

原发性胆汁性肝硬化中长链非编码RNA-GACAT1对肝内胆管细胞的调控机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

神经元和星形胶质细胞特异性miRNA对神经网络发育和功能的调控机制

国家自然科学基金

1+阅读 · 2013年12月31日

基于SURE/PURE准则的图像盲反卷积算法研究

国家自然科学基金

3+阅读 · 2013年12月31日

基于"Build-and-Click"法的铂类RNA聚合酶I选择性抑制剂的构建、评价及亚细胞定位研究

国家自然科学基金

1+阅读 · 2013年12月31日

可变剪切基因REST调控SRRM3的表达在前列腺癌神经内分泌分化中的作用机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

BAFF干扰的树突状细胞参与自身免疫性关节炎免疫耐受的作用和机制

国家自然科学基金

0+阅读 · 2012年12月31日

不同基因型（p53codon72）鼻咽癌细胞放射敏感性差异的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

表观遗传学修饰NANOG基因对胶质瘤细胞生物学活性的作用及机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于RSerPool应用的SCTP传输协议传输路径的优化研究

国家自然科学基金

0+阅读 · 2011年12月31日

特异溶瘤腺病毒对前列腺癌免疫治疗的实验研究与机制

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员