粉碎追赶运动会 - - - - - - - - - - - - - - - - - - - - - - - 血管运动会的等级分解 (Hierarchical Decompositions of Stochastic Pursuit-Evasion Games) - 专知论文

会员服务 ·

0

可约的 · 网格世界 · Performer · 离散化 · 纳什均衡 ·

2022 年 9 月 15 日

Hierarchical Decompositions of Stochastic Pursuit-Evasion Games

翻译：粉碎追赶运动会 - - - - - - - - - - - - - - - - - - - - - - - 血管运动会的等级分解

Yue Guan,Mohammad Afshari,Qifan Zhang,Panagiotis Tsiotras

In this work we present a hierarchical framework for solving discrete stochastic pursuit-evasion games (PEGs) in large grid worlds. With a partition of the grid world into superstates (e.g., "rooms"), the proposed approach creates a two-resolution decision-making process, which consists of a set of local PEGs at the original state level and an aggregated PEG at the superstate level. Having much smaller cardinality, both the local games and the aggregated game can be easily solved to a Nash equilibrium. To connect the decision-making at the two resolutions, we use the Nash values of the local PEGs as the rewards for the aggregated game. Through numerical simulations, we show that the proposed hierarchical framework significantly reduces the computation overhead, while still maintaining a satisfactory level of performance when competing against the flat Nash policies.

翻译：在这项工作中,我们提出了一个在大网格世界中解决离散的随机逃生游戏(PEGs)的分级框架。在将网格世界分割成超级国家(例如“室”)的情况下,拟议办法产生了一个双分制的决策过程,其中包括最初州一级的一套当地PEGs和在超级国家一级的综合PEG。由于离散的基点小得多,本地的游戏和合并的游戏都可以很容易地解决成纳什平衡。为了将两个决议的决策联系起来,我们使用当地PEGs的纳什值作为综合游戏的奖励。我们通过数字模拟显示,拟议的等级框架大大降低了计算间接费用,同时在与平板的纳什政策竞争时仍然保持令人满意的业绩水平。

0

相关内容

可约的

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

柽柳Dof转录因子的耐盐调控机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

Arisandilactone A 的不对称全合成

国家自然科学基金

0+阅读 · 2012年12月31日

风轮菜黄酮类成分调控Nrf2/ARE信号通路诱导Ⅱ相解毒酶抗心肌缺血再灌注损伤的分子机制及构效关系研究

国家自然科学基金

0+阅读 · 2012年12月31日

c-Src激酶在2型糖尿病脑动脉BKCa通道功能障碍中的作用

国家自然科学基金

0+阅读 · 2011年12月31日

食管癌细胞中PI3K/AKT-HIF1α36890;路对糖酵解的影响

国家自然科学基金

0+阅读 · 2008年12月31日

Lazy Incremental Search for Efficient Replanning with Bounded Suboptimality Guarantees

Arxiv

0+阅读 · 2022年10月23日

The Stochastic Proximal Distance Algorithm

Arxiv

0+阅读 · 2022年10月21日

Independent Learning in Mean-Field Games: Satisficing Paths and Convergence to Subjective Equilibria

Arxiv

0+阅读 · 2022年10月21日

An Improved Algorithm for Clustered Federated Learning

Arxiv

0+阅读 · 2022年10月20日

Multi-Agent Cooperative Bidding Games for Multi-Objective Optimization in e-Commercial Sponsored Search

Arxiv

12+阅读 · 2021年6月8日

VIP会员

文章信息

相关主题

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

大语言模型中的事件抽取：方法、模态与未来展望的全面综述

美海军作战管理系统：变革战场空间的二十年

【MIT博士论文】以语言为中心的医学影像理解

俄罗斯“沙希德”/“天竺葵”攻击无人机

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

Lazy Incremental Search for Efficient Replanning with Bounded Suboptimality Guarantees

Arxiv

0+阅读 · 2022年10月23日

The Stochastic Proximal Distance Algorithm

Arxiv

0+阅读 · 2022年10月21日

Independent Learning in Mean-Field Games: Satisficing Paths and Convergence to Subjective Equilibria

Arxiv

0+阅读 · 2022年10月21日

An Improved Algorithm for Clustered Federated Learning

Arxiv

0+阅读 · 2022年10月20日

Multi-Agent Cooperative Bidding Games for Multi-Objective Optimization in e-Commercial Sponsored Search

Arxiv

12+阅读 · 2021年6月8日

相关基金

柽柳Dof转录因子的耐盐调控机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

Arisandilactone A 的不对称全合成

国家自然科学基金

0+阅读 · 2012年12月31日

风轮菜黄酮类成分调控Nrf2/ARE信号通路诱导Ⅱ相解毒酶抗心肌缺血再灌注损伤的分子机制及构效关系研究

国家自然科学基金

0+阅读 · 2012年12月31日

c-Src激酶在2型糖尿病脑动脉BKCa通道功能障碍中的作用

国家自然科学基金

0+阅读 · 2011年12月31日

食管癌细胞中PI3K/AKT-HIF1α36890;路对糖酵解的影响

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员